Cross-Modal Representation Learning

Momentum

14 papers in the last four weeks, up 100% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 136

All topics
CardsList
  1. SAREO-FM: Decoupled Semantic Supervision for SAR-EO Foundation Models

    Oct 7, 2026Jeonghyeok Do, Munchurl KimRemote Sensing Image UnderstandingMultimodal Pretraining

  2. UniAfford: Token-Routed Multitask Learning for Generalizable 2D-3D Affordance Perception

    Sep 29, 2026Yuhao Liu, Yiming Zhong, Hanqing Wang +9Cross-Modal Representation LearningUnified Multimodal Models

  3. Do Emotion Concepts Generalize Across Sources, Modalities, and Architectures in Vision-Language Models?

    Sep 28, 2026Bohao Xing, Xin Liu, Kaishen Yuan +5Cross-Modal LearningVision-Language Models

  4. TRACE: Expert-Aligned ECG Representation Learning with Rigorous Benchmarking and Real-World Validation in Acute Cardiac Care

    Sep 28, 2026Lovely Yeswanth Panchumarthi, Andrew Lu, Saurabh Kataria +11Representation LearningCross-Modal Representation Learning

  5. PhyMo: A Physical-Field Modality for Multimodal AI4Physics

    Sep 23, 2026Henan Sun, Haitao Hu, Jin Liu +4AI for ScienceCross-Modal Representation Learning

  6. RPA: Residual Patch-Token Adapter for Image Retrieval from EEG and MEG

    Sep 19, 2026Yuhui Jin, Yonghao Song, Bingchuan LiuCross-Subject EEG DecodingCross-Modal Representation Learning

  7. Hub-Spectral Activation of Latent Multimodal Knowledge

    Sep 15, 2026Ying Guo, Haidong Chen, Linrui Xu +7Cross-Modal Representation LearningCross-Modal Alignment

  8. Hyper-RED: Scalable Event Pre-training via Semantic Hypergraph Distillation

    Sep 15, 2026Meisen Wang, Zhiqiang Tian, Wei Bao +3Event-Based VisionCross-Modal Knowledge Distillation

  9. PACE: Progressive Angular-to-Norm Contrastive Embedding

    Sep 14, 2026Yanping Li, Wei Zhou, Yawen Liu +7Cross-Modal Representation LearningMultimodal Embedding

  10. FRIST: FMRI Representation Informed Shared-space Training Improves EEG-only Individual-Finger BCI Decoding

    Sep 14, 2026Jintao Zhang, Yidan Ding, Joshua Kosnoff +3MI ClassificationCross-Modal Representation Learning

  11. Multimodal Floorplan Encoding: Learning Dense Modality-Invariant Representations

    Sep 14, 2026Xavier Anadón, Rémi Pautrat, Rui WangRepresentation LearningCross-Modal Representation Learning

  12. MMGait: Benchmarking and Unifying Gait Recognition across Heterogeneous Modalities

    Sep 11, 2026Saihui Hou, Chenye Wang, Qingyuan Cai +2Cross-Modal Representation LearningGait Recognition

  13. Leveraging Cardiac Imaging to Improve ECG-Based Detection of Chagas Disease in Resource-Constrained Settings

    Sep 8, 2026Laura Alvarez-Florez, Daniel Uyterlinde, Samuel Ruipérez-Campillo +3Cross-Modal Representation LearningMedical Diagnosis

  14. Proprioception-Anchored Cross-Modal Pretraining for Zero-Shot Sim-to-Real Contact-Rich Assembly

    Sep 7, 2026Yuhan Wang, Yurou Chen, Hongye Jiang +2RL for RoboticsCross-Modal Representation Learning

  15. CrossFeat: Bridging Imaging Modalities in Feature Descriptor Space

    Aug 31, 2026Paul Schneider, Nazim HaouchineCross-Modal Representation LearningMultimodal Disentangled Representation Learning

  16. CardioState-JEPA: Delay-Aware Cross-Modal Learning of a Shared Cardiac Representation

    Aug 13, 2026Hamza Shafiq, Hung Manh Pham, Bin Zhu +3Cross-Modal Representation LearningJoint-Embedding Predictive Architecture

  17. Unlocking the Power of Medical Tabular Data via Semantic-Aware Multimodal Pre-training

    Aug 11, 2026Yingsheng Liu, Haiming Li, Jingmin Zhu +6Cross-Modal Representation LearningMultimodal Pretraining

  18. TeaMatch: Teachable Cross-Modal Representation Learning for 2D-3D Matching

    Aug 10, 2026Chongjian Wang, Junjie GaoCross-Modal Representation LearningImage Matching

  19. Hyperbolic Multimodal Continual Learning

    Aug 10, 2026Jiahong Liu, Ming Shen, Xiaohao Liu +4Cross-Modal Representation LearningInvariant Representation Learning

  20. GeoUniPR: A Geometry-Consistent Unified Framework for Cross-Modal Place Recognition

    Aug 10, 2026Wonbong Kim, Jiatong Xiao, Rui Li +5Cross-Modal Representation LearningLiDAR Place Recognition

  21. Learning Deep Modality-Shared Self-Expressiveness for Image Clustering with Textual Information

    Aug 9, 2026Xianghan Meng, Wei He, Zhiyuan Huang +1Cross-Modal Representation LearningClustering

  22. Consistency-Driven Co-Evolution for Self-Supervised Cross-Representation Learning

    Aug 5, 2026Xuehang Guo, Pengyuan Li, Tom Hope +3Multi-View ConsistencyRepresentation Learning

  23. Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language

    Aug 5, 2026Xinran Feng, Yi Xie, Chao Zhang +4Cross-Modal Representation LearningMultimodal Pretraining

  24. Learning Molecular Representations from Cellular Phenotypes with Structure Preservation

    Aug 3, 2026Xuan Lin, Jingyu Sheng, Tengfei Ma +2Disentangled Representation LearningCross-Modal Representation Learning

  25. Generic Vision and Cross-Attention for Reaction Yield Prediction

    Aug 1, 2026Qiwei Han, Chi ZhouCross-Modal Representation LearningCross-Modal Attention