cs.CVSep 29, 2026

Structured Visual Target Learning For Cross-Subject eeg-to-image retrieval

Authors: Salini Yadav, Taveena Lotey, Mickaël Coustaty, Pravendra Singh, Partha Pratim Roy

Organizations: Indian Institute of Technology Roorkee, India · University of La Rochelle, France · Indian Institute of Technology (ISM) Dhanbad, India

Abstract

Cross-subject EEG-to-image retrieval requires a neural represen- tation trained on source subjects to remain aligned with a visual embedding space for an unseen subject. Whereas existing methods primarily focus on the EEG side, we address this problem from the perspective of the visual target. Our approach preserves the spatial information of the Perception Encoder, converts its patch grid into a compact set of learned visual views, and aggregates them for each image with a block-structured, content-dependent router. The target is learned jointly with the EEG encoder through contrastive learning with MMD regularization across source subjects. For deployment, we propose a training-free representation refinement that aligns frozen embeddings without updating either encoder. Under leave- one-subject-out evaluation on THINGS-EEG2, the structured target achieves 35.3%/65.6% Top-1/Top-5 accuracy, the best among com- pared methods. Refinement raises this to 48.1%/77.1%, an 18.5% Top-1 gain over the strongest compared method, improving all ten held-out subjects.

Figures & tables

Explore similar work

CardsList
  1. Subject-Aware Multi-Granularity Alignment for Zero-Shot EEG-to-Image Retrieval

    Apr 20, 2026Lin Jiang, Qingshan She, Jiale Xu +3Zero-Shot Composed Image RetrievalVisual Embeddings

  2. Adaptive Cortically Constrained EEG-Vision Alignment for Zero-Shot Brain-to-Image Retrieval

    Sep 21, 2026Ye Wang, Haokun Ren, Wei Wu +4Electroencephalography DecodingHuman Visual Cortex

  3. What Does the Brain See? Multiview Neural Representations to Demystify the Brain-Visual Alignment

    Jun 24, 2026Salini Yadav, Taveena Lotey, Pravendra Singh +1Electroencephalography DecodingVisual Perception