Self-Supervised Visual Representation Learning

Latest papers 151

All topics
CardsList
  1. Shared Gaussianization: What Gaussian Regularizers Certify About Contrastive Learning, and What They Miss

    Oct 7, 2026Ruoyu Zhao, Yuting Chen, Jinheng Zhang +2Contrastive LearningSelf-Supervised Visual Representation Learning

  2. Scalable Patch-Level Self-Supervised Learning

    Oct 7, 2026Maximilian Seitzer, Gabriele Trivigno, Antonín Vobecký +5Self-Supervised LearningSelf-Supervised Visual Representation Learning

  3. DisParQ: Self-Supervised Part Concepts for Interpretable Vision Foundation Models

    Oct 7, 2026Adam Pardyl, Siddhartha Gairola, Sukrut Rao +4Explainable Image ClassificationVision Foundation Models

  4. Unlocking Fine-Grained Perception in CLIP via Structurally-Aware Latent Masked Modeling

    Oct 6, 2026Juntong Li, Lingwei Dang, Haomin Wu +3VLM AdaptationDense Prediction

  5. From Pixels, Without Pre-training: Joint Generative and Self-Supervised Representation Learning in One Model

    Oct 5, 2026Vicente Balmaseda, Ching-Long Lin, Tianbao YangRepresentation LearningImage Generation

  6. CI-JEPA: A Counterfactual Analysis of Latent Representations in Joint-Embedding Predictive Architectures for Self-Supervised Learning

    Oct 4, 2026Mintu Dutta, Ritesh Vyas, Mohendra Roy *Causal Representation LearningJoint-Embedding Predictive Architecture

  7. The hidden advantage of mask resampling: a theory of masked autoencoders

    Oct 1, 2026Jorge Medina Moreira, Lorenzo Bardone, Lenka ZdeborováMasked AutoencodersMasked Language Modeling

  8. Increasing Width Allows Greedy Layer-wise Training to Rival End-to-End Backpropagation in Self-Supervised Learning

    Sep 30, 2026Syon Mansur, Joel ZylberbergDeep Learning OptimizationConvolutional Neural Networks

  9. I Have a Stream: Making Self-Supervised Learning Work on Continuous Video

    Sep 30, 2026Ivan Martinović, Lukas Knobel, Yuki M. AsanoMasked AutoencodersSelf-Supervised Learning

  10. Emergent Multi-View Geometry Through Self-Distillation

    Sep 30, 2026David Nordström, Thibaut Loiseau, Vincent Lepetit +3Multi-View GeometryMulti-View 3D Reconstruction

  11. Masked Swingers: Harnessing Data Augmentation to Advance Autoencoders for Self-Supervised Learning

    Sep 29, 2026Anthony Fuller, Scott C. Lowe, Daniel G. Kyrollos +3Masked AutoencodersInvariant Representation Learning

  12. End-to-End Self-Supervised RGB-T Tracking without Modality Misleading

    Sep 29, 2026Shenglan Li, Rui Yao, Kunyang Sun +4Multimodal RobustnessVisual Object Tracking

  13. Representation by Design in Generation: Cross-View Class-Token Alignment in Diffusion Transformers

    Sep 28, 2026Xiaoyu Wu, Yifei Wang, Chen WeiFlow MatchingDiffusion Transformer

  14. Sparse-View Interpretable 3D Animal Behavior Representations for Neural Encoding and Decoding

    Sep 28, 2026Xinming Dai, Qihang Jin, Tianshu Tan +8Representation LearningSparse-View 3D Reconstruction

  15. W2Rep: Learning Visual Representations by Watching the World Change

    Sep 28, 2026Wen Huang, Hang Guo, Jiarui Yang +3Video Representation LearningSelf-Supervised Visual Representation Learning

  16. Generative Uncertainty as a Self-supervised Signal for Semantic Similarity Learning

    Sep 28, 2026Enrico Pallotta, Sina Raoufi, Lars Doorenbos +2Video Representation LearningFeature Selection

  17. λλ-JEPA Spectral Anti-Collapse Regularization for Self-Supervised Learning

    Sep 28, 2026Berker Demirel, Clémentine Dominé, Valentino Maiorca +3Spectral RegularizationJoint-Embedding Predictive Architecture

  18. A Light Bilevel Refinement Aligns Self-Supervised Representations for Stronger Task-Specific Learning

    Sep 27, 2026Gustav Wagner Zakarias, Zheng-Hua TanFine-TuningBilevel Optimization

  19. When Does Geometric View Synthesis Help Wine Label Retrieval? A Public One-Shot Benchmark Across Self-Supervised and Vision-Language Backbones

    Sep 27, 2026Yueh-Cheng HuangVLM AdaptationNovel View Synthesis

  20. Learning a Flow to Self-Supervised Representations

    Sep 24, 2026Yuling Jiao, Wensen Ma, Houduo Qi +1Flow MatchingRepresentation Learning

  21. Two Global Crops Suffice: Locating Semantic Emergence in DINO-Style Self-Supervised Learning

    Sep 23, 2026Basavaraj Sunagad, Artur Jesslen, Adam KortylewskiMulti-View LearningSelf-Distillation

  22. Positive Pair Geometry Matters: Optimal Transport for Contrastive Learning of Visual Representations

    Sep 21, 2026Akshit Nanda, Shahzad Ahmad, Ram Prasad PadhyContrastive LearningData Augmentation

  23. Vision Transformers versus convolutional neural networks for fine-grained orchid genus identification in a species-rich, data-poor flora: a controlled benchmark on the Orchidaceae of New Guinea

    Sep 21, 2026Reza Saputra, Diah Harnoni Apriyanti, André Schuiteman +4Vision TransformerFine-Grained Image Retrieval

  24. ParticleSplat: Self-supervised Object-centric Latent Particle Splatting

    Sep 16, 2026Lyuxing He, Daniel Guo, Elizabeth Terveen +33D Gaussian SplattingObject-Centric Representation Learning

  25. CoViT: Instance-Correspondence Contrastive Learning for Vision Transformer

    Sep 1, 2026Yisen Wang, Zhirong Wu, Limin WangContrastive LearningInstance Segmentation

  26. Benchmarking Spatial, Spectral, and Self-Supervised Cues for Face Forgery Detection under Realistic Degradation

    Sep 1, 2026Lucas Cunha, Lucas Sotomaior, Lucas Gasperin +3Image Corruption RobustnessImage Forgery Detection

  27. Pix2Rep-v2: Data-Efficient Representation Learning for Dense Medical Imaging Applications

    Sep 1, 2026S. Sifaoui, E. Angelini, S. Toupin +2Representation LearningVisual Representation Learning

  28. CMRVision: A Foundation Model for Cardiac MR Image Analysis

    Sep 1, 2026Athira J. Jacob, Puneet Sharma, Daniel RueckertMedical Imaging Foundation ModelsMedical Image Classification

  29. ViTAMINS: An Empirical Study of Training Self-Supervised Vision Transformers with Synthetic Hard Negatives

    Sep 1, 2026Nikos Giakoumoglou, Andreas Floros, Kleanthis-Marios Papadopoulos +1Contrastive LearningVision Transformer

  30. Motion-Saliency Complementary Masked Modeling for Point Cloud Video Understanding

    Aug 31, 2026Wei Wang, Yiding Sun, Yuyan Wang +4Point Cloud LearningGenerative Modeling

  31. DINOcular: Self-Supervised Visuospatial Representations

    Aug 27, 2026Farkhat Almukhamedov, Sami Azirar, Hermann BlumVisual Representation LearningGeometric Representation Learning

  32. RIPE++: Reinforced Keypoint Learning from Positive Pairs Only

    Aug 20, 2026Johannes Künzel, Peter Eisert, Anna HilsmannRepresentation LearningReinforcement Learning

  33. SAR2Agri: Learning SAR Intensity Representations for Agricultural Monitoring

    Aug 11, 2026Moti Rattan Gupta, Anupam SobtiRepresentation LearningAgricultural Remote Sensing

  34. Fourier Self-Supervision for Fine-Grained Generalized Category Discovery

    Aug 9, 2026Sarah Rastegar, Mina Ghadimi Atigh, Pascal Mettes +2Frequency-Domain Feature LearningFine-Grained Image Classification

  35. Three Necessary Principles for Self-Supervised Visual Representation Learning

    Aug 8, 2026Nikos Giakoumoglou, Paschalis Giakoumoglou, Tania StathakiContrastive LearningRepresentation Collapse

  36. UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling

    Aug 7, 2026An Lanji, Dawei Liu, Jin Li +3World Model LearningJoint-Embedding Predictive Architecture

  37. Attention-Only White-Box Transformer via LeJEPA-Based Self-Supervised Pretraining

    Aug 4, 2026Yang Bai, Linyuan Wang, Haoyang Jiang +3Vision TransformerEfficient ViTs

  38. Self-supervised DXA representations encode multi-system disease risk, biological aging and heritability

    Aug 3, 2026Gil Sasson, Zachary Levine, Smadar Shilo +9Medical ImagingMedical Image Representation Learning

  39. Physics-Aligned Self-Supervised Learning for Scientific Imaging

    Jul 30, 2026Bashir Kazimi, Stefan SandfeldData AugmentationScientific ML

  40. OrganLens: Organ-Specific Representation Learning for CT Foundation Models

    Jul 28, 2026Zhixuan Ge, Anqi Li, Sadeer Al-Kindi +2Medical Imaging Foundation ModelsRepresentation Learning

  41. A Scale-adaptive Vision Model Links C. elegans Neuronal Morphology to Behavior for Neurotoxicity Assessment

    Jul 25, 2026Haochao Ying, Shenchong Lv, Yutao Sun +8Vision Foundation ModelsToxicity Detection

  42. Self-Supervised Learning of Structured Dynamics from Videos

    Jul 23, 2026Lukas Knobel, Andrew Zisserman, Yuki M. AsanoVideo Representation LearningSelf-Supervised Visual Representation Learning

  43. scMIR: a vision-language foundation model for single-cell light microscopy image representation

    Jul 21, 2026Yifan Shang, Jiahui Tan, Xiangxiang Zeng +1Cross-Modal Representation LearningVisual Representation Learning

  44. Emergent Region-Level Facial Correspondence in Frozen Vision Foundation Models

    Jul 15, 2026Izaldein Al-Zyoud, Abdulmotaleb El SaddikSemantic CorrespondenceVision Foundation Models

  45. Self-Supervised Visual Representation Learning: Pretrain-Finetuning or Joint Training?

    Jul 14, 2026Nusrat Munia, Tyler Ward, Nishat Nayla +2Supervised Fine-TuningSemi-Supervised Learning

  46. A Masked Autoencoder Approach to Unsupervised Steel Surface Defect Recognition

    Jul 14, 2026Shrey PatelMasked AutoencodersIndustrial Anomaly Detection

  47. Joint-Embedding Predictive Architecture for Solar PV Panel Fault Classification

    Jul 10, 2026Seyyedhamid Azimidokht, Mehdi Monemi, Abdelhak Kharbouch +4Thermal ImagingJoint-Embedding Predictive Architecture

  48. Lifelong Representations: A Survey on Continual Self-Supervised Learning for Vision Models

    Jul 8, 2026Sergi Masip, Alicja Dobrzeniecka, Jonathan Swinnen +4Self-Supervised LearningSelf-Supervised Visual Representation Learning

  49. `Attention-Guided Cross-Temporal Clustering for Self-Supervised Video Object Segmentation

    Jul 8, 2026Waqas Arshid, Mohammad Awrangjeb, Alan Wee-Chung Liew +1Video Object SegmentationSelf-Supervised Learning

  50. Converge to Surprise: Evolutionary Self-supervised Image Clustering

    Jul 8, 2026Canlin Zhang, Xiuwen LiuUnsupervised LearningSelf-Supervised Learning

  51. Breaking Spurious Correlations via Generative Randomization and Cross-Variant Self-Supervised Learning

    Jul 7, 2026Suraj Yadav, Anjaneya Sharma, Siddharth YadavContrastive LearningOOD Generalization

  52. Probing Geospatial SSL Representations with Environmental Signals

    Jul 6, 2026Rohita Mocharla, Vishal M. PatelRepresentation GeometryGeospatial Representation Learning

  53. HASSL: Hierarchy-Aware Self-Supervised Learning Framework for Single Cell Microscopy

    Jul 5, 2026Julius Riel, Vishwa Mohan Singh, Sai Anirudh Aryasomayajula +10Contrastive LearningHierarchical Representation Learning

  54. Mask-based Predictive Representations for Reinforcement Learning

    Jul 5, 2026Kai ZhaoRepresentation LearningPredictive Representation Learning

  55. Mask-supervised Object-centric Representation Learning with LeJEPA

    Jul 2, 2026Jakob Geusen, Ender KonukogluContrastive LearningRepresentation Learning