Cross-Modal Feature Matching

Latest papers 41

All topics
CardsList
  1. ARCH-B: Architectural Representation, Comprehension and Hierarchy Benchmark

    Sep 28, 2026Kieran Sagar Parikh, Jose Luis Garcia del Castillo y LopezSemantic CorrespondenceCross-Modal Alignment

  2. MatcherCompass: A Deployment-Aware Benchmark to Guide Image Matcher Selection in the Wild

    Sep 22, 2026Hyunwoo Kim, Giseop KimImage MatchingCross-Modal Feature Matching

  3. Mask 2D-3D: Adaptive Dual-Masked Autoencoder Network for Image-to-Point Cloud Registration

    Sep 16, 2026Zhixin Cheng, Jiacheng Deng, Xiaotian Yin +3Masked AutoencodersPoint Cloud Registration

  4. Multimodal Floorplan Encoding: Learning Dense Modality-Invariant Representations

    Sep 14, 2026Xavier Anadón, Rémi Pautrat, Rui WangRepresentation LearningCross-Modal Representation Learning

  5. Beyond Ambiguous Visual Cues: Studying Physiological Disruptions and Cross-Modal Inconsistencies in Deepfake Videos

    Sep 14, 2026Chenxi Yang, Yassine Ouzar, Larbi BoubchirRemote PhotoplethysmographyDeepfake Detection

  6. RoMa-ΩΩ: What Feed-Forward 3D Models Know About Image Matching

    Sep 8, 2026David Nordström, Xinyue Zhang, Thibaut Loiseau +2Image Matching3D Representation Learning

  7. DXPR: Depth-Based Vision-LiDAR Cross-Modal Place Recognition Using Vision Foundation Models

    Sep 8, 2026Yungsoo Han, Youngseok Jang, Seungwon Roh +2LiDAR Place RecognitionVisual Place Recognition

  8. CrossFeat: Bridging Imaging Modalities in Feature Descriptor Space

    Aug 31, 2026Paul Schneider, Nazim HaouchineCross-Modal Representation LearningMultimodal Disentangled Representation Learning

  9. TeaMatch: Teachable Cross-Modal Representation Learning for 2D-3D Matching

    Aug 10, 2026Chongjian Wang, Junjie GaoCross-Modal Representation LearningImage Matching

  10. SLAP: Selective Local Vision-Language Alignment for Fish Re-Identification via Partial Optimal Transport

    Aug 9, 2026Cigdem Beyan, Tonje Knutsen Sordalen, Kim Tallaksen HalvorsenInverse Optimal TransportAnimal Re-Identification

  11. Face and Voice Cross-modal Association with Learning Convex Feature Embedding

    Jul 30, 2026Taewan Kim, Jiwoo KangCross-Modal Feature Matching

  12. XMatchAD: A Cross-Modal Matching Perspective on Reconstruction-based Anomaly Detection

    Jul 26, 2026Mingxiu Cai, Zhe Zhang, Gaochang Wu +1Industrial Anomaly DetectionAnomaly Localization

  13. Cross-Coordinate Correspondence Pruning for Image-to-Point Cloud Registration

    Jul 19, 2026Xin Liu, Rong Qin, Huipeng Lin +5Point Cloud RegistrationCross-Modal Feature Matching

  14. SUFLECA: Scaling Up Feature Learning for CAD-to-image Alignment

    Jul 16, 2026Saad Ejaz, Miguel Fernandez-Cortizas, Javier Civera +2Object Pose EstimationGeometric Representation Learning

  15. Label-Free Target-Domain Adaptation for Unconstrained Event-Image Feature Matching via Dual-Stage Distillation

    Jul 11, 2026Zhonghua Yi, Hao Shi, Qi Jiang +3Event-Based VisionCross-Modal Knowledge Distillation

  16. PLGSA-Transformer: Periocular Landmark-Guided Attention with Occlusion-Adaptive Cosine Thresholding for Cross-Modal Masked and Unmasked Face Recognition

    Jul 3, 2026Dana A AbdullahCross-Modal Feature MatchingFace Recognition

  17. DetailAnywhere: Fashion Detail Generation via Cross-Modal Feature Alignment Distillation

    Jul 2, 2026Zijun Li, Yimin Zhou, Jia Sun +12Diffusion Model DistillationConditional Image Generation

  18. Cross4D-JEPA: Dense Cross-modal Correspondence Distillation for 4D Point Cloud Representation Learning

    Jul 1, 2026Trung Thanh Nguyen, Hai Nguyen-Truong, Tu Vo +2Cross-Modal Knowledge DistillationPoint Cloud Learning

  19. AnyMatch: Supercharging Universal Multi-Modal Image Matching with Large-Scale Single-View Images

    Jun 30, 2026Meng Yang, Zizhuo Li, Linfeng Tang +2Synthetic Data AugmentationCross-Modal Feature Matching

  20. Clinical Risk-Aware Multi-Level Grading for Coronary Artery Stenosis through Curved Feature Reconstruction

    Jun 29, 2026Shishuang Zhao, Hongtai Li, Junjie Hou +13D Medical ImagingCross-Modal Feature Matching

  21. Cross-Spectral Stereo Inertial Odometry

    Jun 29, 2026Seungsang Yun, Hyunsoo Jang, Tai Hyoung Rhee +3Visual-Inertial OdometryCross-Modal Feature Matching

  22. CMDS-AD: Cross-Modal Dual-Stream Decoupling for Few-Shot Anomaly Detection

    Jun 18, 2026Junhao Cai, Junyu Chen, Deyu Zeng +4Few-Shot Anomaly DetectionMultimodal Anomaly Detection

  23. G2IA: Geometry-Guided Instance-Aware Retrieval and Refinement for Cross-Modal Place Recognition

    Jun 13, 2026Xianyun Jiao, Jingyi Xu, Zhongmiao Yan +2LiDAR Place RecognitionVisual Place Recognition

  24. FIGMA: Towards FIne-Grained Music retrievAl

    Jun 4, 2026Nishit Anand, Ashish Seth, Sreyan Ghosh +2Audio UnderstandingAudio-Text Retrieval

  25. Geometry-Preserving Unsupervised Alignment for Heterogeneous Foundation Models

    Jun 3, 2026Shuwen Yu, Zhanxuan Hu, Yi Zhao +2Cross-Modal Representation LearningVision Foundation Model Adaptation

  26. Cross-Modality Feature Fusion Based on Structured State Space Duality for Multimodal Image Registration Network

    Jun 2, 2026Zhikang Li, Yan Wu, Xin Hu +2Multi-Scale Feature FusionCross-Modal Feature Matching

  27. Best Segmentation Buddies for Image-Shape Correspondence

    May 18, 2026Itai Lang, Dongwei Lyu, Dale Decatur +13D Semantic SegmentationCross-Modal Feature Matching

  28. Mind the Gap: Learning Modality-Agnostic Representations with a Cross-Modality UNet

    May 16, 2026Xin Niu, Enyi Li, Jinchao Liu +3Cross-Modal LearningCross-Modal Representation Learning

  29. VoxCor: Training-Free Volumetric Features for Multimodal Voxel Correspondence

    May 13, 2026Guney Tombak, Ertunc Erdil, Ender KonukogluMedical Imaging Foundation ModelsCross-Modal Representation Learning

  30. TAR: Text Semantic Assisted Cross-modal Image Registration Framework for Optical and SAR Images

    May 12, 2026Zhuoyu Cai, Dou Quan, Ning Huyan +3Remote SensingSynthetic Aperture Radar

  31. Zero-Shot Chinese Character Recognition via Global-Local Dual-Branch Alignment and Hierarchical Inference

    May 9, 2026Wei Cao, Hao Xu, Xiaolei DiaoZero-Shot LearningEfficient VLM Inference

  32. FS-I2P:A Hierarchical Focus-Sweep Registration Network with Dynamically Allocated Depth

    May 8, 2026Zhixin Cheng, Yujia Chen, Xujing Tao +4Cross-Modal AttentionPoint Cloud Registration

  33. Angle-I2P: Angle-Consistent-Aware Hierarchical Attention for Cross-Modality Outlier Rejection

    May 6, 2026Muyao Peng, Shun Zou, Pei An +2Point Cloud RegistrationCross-Modal Feature Matching

  34. UniCorrn: Unified Correspondence Transformer Across 2D and 3D

    May 5, 2026Prajnan Goswami, Tianye Ding, Feng Liu +1Image MatchingPoint Cloud Registration

  35. REALM: An RGB- and Event-Aligned Latent Manifold for Cross-Modal Perception

    Apr 30, 2026Vincenzo Polizzi, David B. Lindell, Jonathan KellyCross-Modal LearningEvent-Based Vision

  36. Beyond Visual Cues: Semantic-Driven Token Filtering and Expert Routing for Anytime Person ReID

    Apr 16, 2026Jiaxuan Li, Xin Wen, Zhihang LiCross-Modal Feature MatchingPerson Re-Identification

  37. SEPS: Semantic-enhanced Patch Slimming Framework for fine-grained cross-modal alignment

    Nov 3, 2025Xinyu Mao, Junsi Li, Haoji Zhang +2Cross-Modal AlignmentCross-Modal Feature Matching