Visual Representation Learning

Latest papers 147

All topics
CardsList
  1. DynaFLIP: Rethinking Robotics Perception via Tri-Modal-Dynamics Guided Representation

    May 28, 2026Jusuk Lee, Seungjae Lee, Jonghun Shin +6Cross-Modal LearningCross-Modal Representation Learning

  2. Deep Psychovisual Image Representations

    May 28, 2026Wendi Ma, Aryaman Sharma, Wei Dai +1Visual Representation LearningFrequency-Domain Feature Learning

  3. Misalignment Between Backpropagation and the Hierarchy of Brain Responses to Images

    May 27, 2026Joséphine Raugel, Maximilian Seitzer, Marc Szafraniec +6BackpropagationVisual Representation Learning

  4. Structure over Pixels: Learning Variable-Length Visual Programs

    May 26, 2026Piotr Wyrwiński, Kacper Dobek, Krzysztof KrawiecVisual Representation LearningImage Tokenization

  5. Uncertainty-DTW for Sequences and Visual Tokens

    May 24, 2026Lei Wang, Syuan-Hao Li, Yongsheng Gao +1Dynamic Time WarpingVisual Representation Learning

  6. Not Too Generative, Not Too Discriminative: The Human Alignment Sweet Spot

    May 22, 2026Jorge Chang Ortega, Bastien Le Lan, Thomas Serre +1Visual Representation LearningEnergy-Based Models

  7. TextTeacher: What Can Language Teach About Images?

    May 21, 2026Tobias Christian Nauen, Stanislav Frolov, Brian Bernhard Moser +3Cross-Modal Knowledge DistillationVisual Representation Learning

  8. RiT: Vanilla Diffusion Transformers Suffice in Representation Space

    May 21, 2026Le Zhang, Ning Mang, Aishwarya AgrawalFlow MatchingVisual Representation Learning

  9. Beyond Routing: Characterising Expert Tuning and Representation in Vision Mixture-of-Experts

    May 20, 2026Gene Tangtartharakul, Katherine R. StorrsVisual Representation LearningMixture of Experts

  10. Capability ≠\neq Interpretability: Human Interpretability of Vision Foundation Models

    May 19, 2026Julien Colin, Lore Goetschalckx, Nuria Oliver +1Visual Representation LearningVision Foundation Models

  11. PEIRA: Learning Predictive Encoders through Inter-View Regressor Alignment

    May 17, 2026Michael Arbel, Basile Terver, Jean PonceUnsupervised LearningVisual Representation Learning

  12. Characterizing the visual representation of objects from the child's view

    May 14, 2026Jane Yang, Tarun Sepuri, Alvin Wei Ming Tan +3Representation LearningVisual Representation Learning

  13. Rethinking the Good Enough Embedding for Easy Few-Shot Learning

    May 13, 2026Michael Karnes, Alper YilmazVisual Representation LearningFew-Shot Learning

  14. Characterizing Universal Object Representations Across Vision Models

    May 13, 2026Florian P. Mahner, Johannes Roth, Ka Chun Lam +3Visual Representation LearningRepresentation Alignment

  15. Do Vision Transformers Need All-to-All Attention? Global Communication Through Elastic Learned Cores

    May 12, 2026Alan Z. Song, Yinjie Chen, Mu Nan +3Visual Representation LearningVision Transformer

  16. WorldComp2D: Spatio-semantic Representations of Object Identity and Location from Local Views

    May 12, 2026SeongMin Jin, Doo Seok JeongRepresentation LearningVisual Representation Learning

  17. Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery

    May 12, 2026Chi-Nguyen Tran, Dao Sy Duy Minh, Huynh Trung Kiet +3Cross-View Geo-LocalizationVisual Representation Learning

  18. FeatMap: Understanding image manipulation in the feature space and its implications for feature space geometry

    May 11, 2026Elias B. Krey, Nils Neukirch, Nils StrodthoffNeural Representation GeometryVisual Representation Learning

  19. Learning to Align Generative Appearance Priors for Fine-grained Image Retrieval

    May 11, 2026Shijie Wang, Yadan Luo, Zijian Wang +2Visual Representation LearningFine-Grained Image Retrieval

  20. Lost or Hidden? A Concept-Level Forgetting in Supervised Continual Learning

    May 10, 2026Katarzyna Filus, Kamil Faber, Roberto Corizzo +1Visual Representation LearningContinual Learning

  21. SEMASIA: A Large-Scale Dataset of Semantically Structured Latent Representations

    May 10, 2026Mario Edoardo Pandolfo, Enrico Grimaldi, Lorenzo Marinucci +4Neural Representation GeometryVisual Representation Learning

  22. 3D MRI Image Pretraining via Controllable 2D Slice Navigation Task

    May 7, 2026Yu Wang, Qingchao ChenVisual Representation LearningSelf-Supervised Pre-Training

  23. Look Beyond Saliency: Low-Attention Guided Dual Encoding for Video Semantic Search

    May 7, 2026Faisal Aljehrai, Mohammed A. Alkhrashi, Alreem Almuhrij +6Visual Representation LearningInformation Retrieval

  24. CRISP: Compositional Relations as Invariant Structural Priors for Domain Generalization

    May 7, 2026Dat Nguyen, Duc-Duy NguyenRepresentation LearningVisual Representation Learning

  25. MUSE: Resolving Manifold Misalignment in Visual Tokenization via Topological Orthogonality

    May 7, 2026Panqi Yang, Haodong Jing, Jiahao Chao +5Visual Representation LearningImage Tokenization

  26. Exploring Clustering Capability of Inpainting Model Embeddings for Pattern-based Individual Identification

    May 6, 2026Jens van Bijsterveld, Daniele Avitabile, Fons J. Verbeek +1Biometric IdentificationVisual Representation Learning