Self-Supervised Visual Representation Learning

Latest papers 151

All topics
CardsList
  1. BEAST3D: Animal behavioral analysis and neural encoding from multi-view video via Gaussian splatting

    Jun 1, 2026Yanchen Wang, Lenny Aharon, Wangshu Zhu +7Sparse-View 3D Reconstruction3D Representation Learning

  2. Multi-modal Video Representation Alignment for Robust Self-supervised Driver Distraction Detection

    Jun 1, 2026David J. Lerch, Livien Majer, Zeyun Zhong +3Multimodal RobustnessCross-Modal Alignment

  3. Paving the Way for Point Cloud Video Representation Learning Using A PDE Model

    Jun 1, 2026Zhuoxu Huang, Zhenkun Fan, Jungong Han +1Point Cloud LearningVideo Representation Learning

  4. NTR: Neural Token Reconstruction for Scene Token Bottleneck in End-to-End Driving

    May 29, 2026Jiahui Li, Jiawei Sun, Zixiang Ren +7End-to-End Autonomous DrivingSelf-Supervised Visual Representation Learning

  5. HQ-JEPA: Hybrid Quantum Joint-Embedding Predictive Architecture for Cross-Modal Remote Sensing Representation Learning

    May 29, 2026Md Aminur Hossain, Ayush V. Patel, Sanjay K. Singh +1Representation LearningCross-Modal Representation Learning

  6. Unsupervised Semantic Segmentation Facilitates Model Understanding

    May 28, 2026Xiaoyan Yu, Lisa Mais, Jannik Franzen +4Transformer InterpretabilityVision Transformer

  7. Bayesian Gated Non-Negative Contrastive Learning

    May 27, 2026Peng Cui, Jiahao Zhang, Lijie HuContrastive LearningDisentangled Representation Learning

  8. A self-supervised learning approach to deep filter banks for texture recognition

    May 27, 2026Joao B. Florindo, Lucas O. Lyra, Antonio E. FabrisAutoencodersConvolutional Autoencoder

  9. Anatomy-Anchored Self-Supervision: Distilling Vision Foundation Models for Invariant Ultrasound Representation

    May 25, 2026Chunzheng Zhu, Yijun Wang, Jianxin Lin +5Invariant Representation LearningMedical Image Representation Learning

  10. Learning from Semantic Dictionaries: Discriminative Codebook Contrastive Learning for Unified Visual Representation and Generation

    May 24, 2026Imanol G. Estepa, Jesús M Rodríguez-de-Vera, Bhalaji Nagarajan +1Contrastive LearningImage Generation

  11. The TIME Machine: On The Power of Motion for Efficient Perception

    May 21, 2026Mantas Skackauskas, Xinyue Hao, Laura Sevilla-LaraRepresentation LearningVideo Representation Learning

  12. Harnessing Self-Supervised Features for Art Classification

    May 18, 2026Federico Melis, Davide Bilardello, Emanuele Prato +2Vision Foundation ModelsSelf-Supervised Visual Representation Learning

  13. Collision-Resistant Single-Pass Method for Unsupervised Fine-Grained Image Hashing

    May 18, 2026Anh-Kiet Duong, Petra Gomez-Krämer, Jean-Michel CarozzaFine-Grained Image RetrievalImage Retrieval

  14. Self-supervised Hierarchical Visual Reasoning with World Model

    May 17, 2026Yuanfei Xu, Lin Liu, Wengang Zhou +2World Model LearningWorld Models

  15. HyperVision: A Channel-Adaptive Ground-Based Hyperspectral Vision Pre-trained Backbone

    May 17, 2026Guanyiman Fu, Jingtao Li, Zihang Cheng +8Vision Foundation Model AdaptationHyperspectral Imaging

  16. LACE: Latent Visual Representation for Cross-Embodiment Learning

    May 16, 2026Yoo Sung Jang, Kanchana Ranasinghe, Cristina Mata +3Cross-Embodiment Robot LearningRobot Imitation Learning

  17. Latent Video Prediction for World Modeling: An Evaluation Uncovering Intriguing Favorable Evidence

    May 15, 2026Ali J Alrasheed, Aryan Yazdan Parast, Basim Azam +2Video Representation LearningVideo Prediction

  18. Pretraining Objective Matters in Extreme Low-Data FGVC: A Backbone-Controlled Study

    May 15, 2026Alexander Hackett, Srikanth Thudumu, Ginny Fisher +1Contrastive LearningMasked Autoencoders

  19. Masked Next-Scale Prediction for Self-supervised Scene Text Recognition

    May 14, 2026Zhuohao Chen, Zeng Li, Yifei Zhang +2Scene Text RecognitionSelf-Supervised Visual Representation Learning

  20. Learning to Perceive "Where": Spatial Pretext Tasks for Robust Self-Supervised Learning

    May 11, 2026Yang Shen, Yusen Cai, Weronika Hryniewska-Guzik +2Visual Spatial ReasoningSelf-Supervised Visual Representation Learning

  21. CalibFree: Self-Supervised View Feature Separation for Calibration-Free Multi-Camera Multi-Object Tracking

    May 10, 2026Ruiqi Xian, Deep Patel, Iain Melvin +3Multi-View LearningInvariant Representation Learning

  22. Boosting Self-Supervised Tracking with Contextual Prompts and Noise Learning

    May 7, 2026Yaozong Zheng, Qihua Liang, Bineng Zhong +4Representation LearningSelf-Supervised Visual Representation Learning

  23. MTL-MAD: Multi-Task Learners are Effective Medical Anomaly Detectors

    May 7, 2026Bogdan Alexandru Bercean, Florinel Alin Croitoru, Vlad Hondru +3Medical Image Anomaly DetectionMulti-Task Learning

  24. Chaotic Contrastive Learning for Robust Texture Classification

    May 6, 2026Joao B FlorindoContrastive LearningSynthetic Data Augmentation