Self-Supervised Visual Representation Learning

Latest papers 151

All topics
CardsList
  1. Shared Gaussianization: What Gaussian Regularizers Certify About Contrastive Learning, and What They Miss

    Oct 7, 2026Ruoyu Zhao, Yuting Chen, Jinheng Zhang +2Contrastive LearningSelf-Supervised Visual Representation Learning

  2. Scalable Patch-Level Self-Supervised Learning

    Oct 7, 2026Maximilian Seitzer, Gabriele Trivigno, Antonín Vobecký +5Self-Supervised LearningSelf-Supervised Visual Representation Learning

  3. DisParQ: Self-Supervised Part Concepts for Interpretable Vision Foundation Models

    Oct 7, 2026Adam Pardyl, Siddhartha Gairola, Sukrut Rao +4Explainable Image ClassificationVision Foundation Models

  4. Unlocking Fine-Grained Perception in CLIP via Structurally-Aware Latent Masked Modeling

    Oct 6, 2026Juntong Li, Lingwei Dang, Haomin Wu +3VLM AdaptationDense Prediction

  5. From Pixels, Without Pre-training: Joint Generative and Self-Supervised Representation Learning in One Model

    Oct 5, 2026Vicente Balmaseda, Ching-Long Lin, Tianbao YangRepresentation LearningImage Generation

  6. CI-JEPA: A Counterfactual Analysis of Latent Representations in Joint-Embedding Predictive Architectures for Self-Supervised Learning

    Oct 4, 2026Mintu Dutta, Ritesh Vyas, Mohendra Roy *Causal Representation LearningJoint-Embedding Predictive Architecture

  7. The hidden advantage of mask resampling: a theory of masked autoencoders

    Oct 1, 2026Jorge Medina Moreira, Lorenzo Bardone, Lenka ZdeborováMasked AutoencodersMasked Language Modeling

  8. Increasing Width Allows Greedy Layer-wise Training to Rival End-to-End Backpropagation in Self-Supervised Learning

    Sep 30, 2026Syon Mansur, Joel ZylberbergDeep Learning OptimizationConvolutional Neural Networks

  9. I Have a Stream: Making Self-Supervised Learning Work on Continuous Video

    Sep 30, 2026Ivan Martinović, Lukas Knobel, Yuki M. AsanoMasked AutoencodersSelf-Supervised Learning

  10. Emergent Multi-View Geometry Through Self-Distillation

    Sep 30, 2026David Nordström, Thibaut Loiseau, Vincent Lepetit +3Multi-View GeometryMulti-View 3D Reconstruction

  11. Masked Swingers: Harnessing Data Augmentation to Advance Autoencoders for Self-Supervised Learning

    Sep 29, 2026Anthony Fuller, Scott C. Lowe, Daniel G. Kyrollos +3Masked AutoencodersInvariant Representation Learning

  12. End-to-End Self-Supervised RGB-T Tracking without Modality Misleading

    Sep 29, 2026Shenglan Li, Rui Yao, Kunyang Sun +4Multimodal RobustnessVisual Object Tracking

  13. Representation by Design in Generation: Cross-View Class-Token Alignment in Diffusion Transformers

    Sep 28, 2026Xiaoyu Wu, Yifei Wang, Chen WeiFlow MatchingDiffusion Transformer

  14. Sparse-View Interpretable 3D Animal Behavior Representations for Neural Encoding and Decoding

    Sep 28, 2026Xinming Dai, Qihang Jin, Tianshu Tan +8Representation LearningSparse-View 3D Reconstruction

  15. W2Rep: Learning Visual Representations by Watching the World Change

    Sep 28, 2026Wen Huang, Hang Guo, Jiarui Yang +3Video Representation LearningSelf-Supervised Visual Representation Learning

  16. Generative Uncertainty as a Self-supervised Signal for Semantic Similarity Learning

    Sep 28, 2026Enrico Pallotta, Sina Raoufi, Lars Doorenbos +2Video Representation LearningFeature Selection

  17. λλ-JEPA Spectral Anti-Collapse Regularization for Self-Supervised Learning

    Sep 28, 2026Berker Demirel, Clémentine Dominé, Valentino Maiorca +3Spectral RegularizationJoint-Embedding Predictive Architecture

  18. A Light Bilevel Refinement Aligns Self-Supervised Representations for Stronger Task-Specific Learning

    Sep 27, 2026Gustav Wagner Zakarias, Zheng-Hua TanFine-TuningBilevel Optimization

  19. When Does Geometric View Synthesis Help Wine Label Retrieval? A Public One-Shot Benchmark Across Self-Supervised and Vision-Language Backbones

    Sep 27, 2026Yueh-Cheng HuangVLM AdaptationNovel View Synthesis

  20. Learning a Flow to Self-Supervised Representations

    Sep 24, 2026Yuling Jiao, Wensen Ma, Houduo Qi +1Flow MatchingRepresentation Learning

  21. Two Global Crops Suffice: Locating Semantic Emergence in DINO-Style Self-Supervised Learning

    Sep 23, 2026Basavaraj Sunagad, Artur Jesslen, Adam KortylewskiMulti-View LearningSelf-Distillation

  22. Positive Pair Geometry Matters: Optimal Transport for Contrastive Learning of Visual Representations

    Sep 21, 2026Akshit Nanda, Shahzad Ahmad, Ram Prasad PadhyContrastive LearningData Augmentation

  23. Vision Transformers versus convolutional neural networks for fine-grained orchid genus identification in a species-rich, data-poor flora: a controlled benchmark on the Orchidaceae of New Guinea

    Sep 21, 2026Reza Saputra, Diah Harnoni Apriyanti, André Schuiteman +4Vision TransformerFine-Grained Image Retrieval

  24. ParticleSplat: Self-supervised Object-centric Latent Particle Splatting

    Sep 16, 2026Lyuxing He, Daniel Guo, Elizabeth Terveen +33D Gaussian SplattingObject-Centric Representation Learning

  25. CoViT: Instance-Correspondence Contrastive Learning for Vision Transformer

    Sep 1, 2026Yisen Wang, Zhirong Wu, Limin WangContrastive LearningInstance Segmentation

  26. Benchmarking Spatial, Spectral, and Self-Supervised Cues for Face Forgery Detection under Realistic Degradation

    Sep 1, 2026Lucas Cunha, Lucas Sotomaior, Lucas Gasperin +3Image Corruption RobustnessImage Forgery Detection

  27. Pix2Rep-v2: Data-Efficient Representation Learning for Dense Medical Imaging Applications

    Sep 1, 2026S. Sifaoui, E. Angelini, S. Toupin +2Representation LearningVisual Representation Learning

  28. CMRVision: A Foundation Model for Cardiac MR Image Analysis

    Sep 1, 2026Athira J. Jacob, Puneet Sharma, Daniel RueckertMedical Imaging Foundation ModelsMedical Image Classification

  29. ViTAMINS: An Empirical Study of Training Self-Supervised Vision Transformers with Synthetic Hard Negatives

    Sep 1, 2026Nikos Giakoumoglou, Andreas Floros, Kleanthis-Marios Papadopoulos +1Contrastive LearningVision Transformer