Multimodal Representations

Momentum

10 papers in the last four weeks, against 2 the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 44

All topics
CardsList
  1. Multimodal Flow: Unified Flow Modeling of Language and Vision in Embedding Spaces

    Sep 30, 2026Hongyuan Tao, Xinggang Wang, Lianghui Zhu +7Multimodal PretrainingMultimodal Representations

  2. Spherical Interpolation for Backward-Compatible Multimodal Representations

    Sep 30, 2026Simone Ricci, Niccolò Biondi, Federico PerniciMultimodal RepresentationsCross-Modal

  3. Mutual Equilibrium: Multimodal Representation Learning through Reciprocal Feedback

    Sep 30, 2026Ho-min Park, Byungkon KangMultimodal RepresentationsMultimodal Classification

  4. Multi-Marginal Inverse Optimal Transport for Contrastive Learning Via Explicit Anchor-Positive-Negative Coupling

    Sep 27, 2026Ngoc-Hai Nguyen, Thuan Nguyen, Prakash Ishwar +1Contrastive LearningOptimal Transport Approach

  5. ICE: Task-Aligned Clifford Latent Fields for Multimodal Graph Foundation Models

    Sep 24, 2026Xunkai Li, Xu Wang, Yinlin Zhu +4Graph Foundation ModelMultimodal Representations

  6. LAYERSCOPE: A Layerwise Characterization of Video and Multimodal Learned Representations

    Sep 23, 2026Sandra Arcos-Holzinger, Debashish Chakraborty, Rohita Mocharla +7Multimodal RepresentationsFine-Grained Video Understanding

  7. PhyMo: A Physical-Field Modality for Multimodal AI4Physics

    Sep 23, 2026Henan Sun, Haitao Hu, Jin Liu +4Multimodal RepresentationsModalities

  8. Hub-Spectral Activation of Latent Multimodal Knowledge

    Sep 15, 2026Ying Guo, Haidong Chen, Linrui Xu +7Multimodal RepresentationsCross-Modal

  9. Cross-Regional Grapevine Cold Hardiness Prediction via Learned Multimodal Latent Representations

    Aug 31, 2026William Solow, Paola Pesantez-Cabrera, Markus Keller +3Agricultural PredictionFuture Latent Representations

  10. Multimodal Shared Latent Representation of Narration, Microscope and iOCT Images for Phase Recognition in Vitreoretinal Surgery

    Aug 31, 2026Onur Izmitlioglu, Shervin Dehghani, Tarek Ghannoum +2OphthalmologyOptical Coherence Tomography

  11. Learning Deep Modality-Shared Self-Expressiveness for Image Clustering with Textual Information

    Aug 9, 2026Xianghan Meng, Wei He, Zhiyuan Huang +1Multimodal RepresentationsModalities

  12. GALA: Generative Aligned Learning for Adaptive Multimodal Representation in the Taobao Shangou Recommender System

    Jul 31, 2026Jiping Liu, Zhongmin Zhang, Zisen Sang +7Multimodal Recommendation ModelMultimodal Representations

  13. OmniStyle-INR: Universal and Multimodal Style Transfer for INRs

    Jul 17, 2026Rafał Kajca, Michał Miziołek, Kornel Howil +2Photorealistic Style TransferMultimodal Representations

  14. AlphaWiSE: Adaptive Weight Interpolation for Continual Multimodal Representation Learning

    Jul 16, 2026Sarthak Jain, Qiran Hu, Zhen Zhu +1Multimodal RepresentationsCross-Modal

  15. The Hyperspherical Geometry of CLIP Latent Space: A Semantic Mixture Model

    Jul 15, 2026Zijie Yu, Gaowen Liu, Ramana Rao Kompella +2Contrastive Language-Image Pre-Training ModelSpherical Latent Space

  16. Multi-Agent Collaborative Reasoning with Tool-Augmented Evidence for Urban Region Profiling

    Jul 15, 2026Xixuan Hao, Yutian Jiang, Jiabo Liu +4Multi-Agent ReasoningUrban Environments

  17. Time Imprint: Learning Time-Aware Representations in Multi-Modal Knowledge Graphs

    Jul 8, 2026Pengyu Zhang, Klim Zaporojets, Congfeng Cao +2Knowledge Graph EmbeddingsMultimodal Representations

  18. What Images Cannot Say: Language-Guided Olfactory Representation Learning

    Jul 7, 2026Eleftherios Tsonis, Xi Wang, Vicky KalogeitonSmellsMultimodal Representations

  19. Forewarned is Forearmed: When Non-Sequential Embedding Turns Into an Anomaly Detector

    Jun 29, 2026Elys Allesiardo, Antoine Caubrière, Valentin VielzeufMultimodal EmbeddingsMultimodal Representations

  20. Meta-learning as a principle for human-like visual representations

    Jun 24, 2026Can Demircan, Marcel Binz, Alireza Modirshanechi +1Meta-LearningVisual Representations

  21. Morphology-Aware Multimodal Representation Learning for Insect Phylogenetic Reconstruction

    Jun 20, 2026Zixuan Liu, Kaijie Yu, Chun He +5Phylogenetic InferenceMorphology

  22. Generative-Model Predictive Planning for Navigation in Partially Observable Environments

    Jun 17, 2026Thomas Quilter, Yifan Zhu, Guorui Quan +2Diffusion PlanningPartially Observable Markov Decision Process

  23. When to Align, When to Predict: A Phase Diagram for Multimodal Learning

    Jun 9, 2026Ilay Kamai, Hugues Van Assel, Aviv Regev +2Multimodal LearningMultimodal Representations

  24. Before Fusion, Ask What to Keep: Contextual Calibration of Multimodal Signals

    Jun 1, 2026Jiyuan Liu, Liangwei Nathan Zheng, Wei Emma Zhang +2Multimodal FusionMultimodal Representations

  25. MIC: Maximizing Informational Capacity in Adaptive Representations via Isotropic Subspace Alignment

    May 28, 2026Dang Nguyen Hong, Nhi Ngoc-Yen Nguyen, Huy-Hieu PhamSubspaceMultiscale

  26. UniNote: A Unified Embedding Model for Multimodal Representation and Ranking

    May 28, 2026Jinghan Zhao, Wenwei Jin, Anqi Li +5Multimodal RetrievalMultimodal Representations

  27. The Wittgensteinian Representation Hypothesis: Is Language the Attractor of Multimodal Convergence?

    May 10, 2026Zhaoyang Zhang, Run Shao, Dongyue Wu +4Multimodal RepresentationsAttractors

  28. Learning Generalizable Multimodal Representations for Software Vulnerability Detection

    Apr 28, 2026Zeming Dong, Yuejun Guo, Qiang Hu +5Multimodal RepresentationsSoftware

  29. Modular Representation Compression: Adapting LLMs for Efficient and Effective Recommendations

    Apr 20, 2026Yunjia Xi, Menghui Zhu, Jianghao Lin +4Multimodal RepresentationsLayer-Wise

  30. Hyperbolic Enhanced Representation Learning for Incomplete Multi-view Clustering

    Apr 18, 2026Tianyi Chen, Haobo Wang, Kai Tang +5Hyperbolic LearningMulti-View

  31. Semantic Purification for Conditional Representation Learning

    Feb 5, 2026Jiaquan Wang, Yan Lyu, Chen Li +1Representation LearningMultimodal Representations

  32. Multimodal Feature Prototype Learning for Interpretable and Discriminative Cancer Survival Prediction

    Oct 7, 2025Shuo Jiang, Zhuwen Chen, Liaoman Xu +6Survival PredictionsLearnable Prototypes

  33. The Platonic Universe: Do Foundation Models See the Same Sky?

    Sep 23, 2025UniverseTBD, :, Trinidad Borrell +11Astronomical ImagesFoundation Model

  34. A quantitative analysis of semantic information in deep representations of text and images

    May 21, 2025Santiago Acevedo, Andrea Mascaretti, Riccardo Rende +3Semantic RepresentationsMultimodal Representations

  35. CLIP Embeddings for AI-Generated Image Detection: A Few-Shot Study with Lightweight Classifier

    May 15, 2025Ziyang OuAi-Generated Image DetectionContrastive Language-Image Pre-Training Model