cs.CVSep 28, 2026

Natural Image Autoencoder-Based fMRI Representations for Trait and State Prediction

Authors: Juhyeon Park, Yeonwoo Kim, Peter Yongho Kim, Yansen Wang, Mingqing Xiao, Dongqi Han, Dongsheng Li, Taesup Moon

Organizations: IPAI, Seoul National University · Microsoft Research · ECE, Seoul National University · ASRI / INMC / AIIS, Seoul National University

Abstract

Foundation models pre-trained on large-scale fMRI datasets have shown strong downstream performance, but at substantial data and computation cost. To investigate how much fMRI-specific pre-training is actually needed for such performance, we introduce FReD, which derives fMRI representations from a frozen Deep Compression AutoEncoder (DCAE) pre-trained exclusively on natural images and pairs them with a task specific readout. For trait prediction, FReD summarizes frame-wise representations by their temporal mean and log-standard deviation and applies linear probing, with late fusion across two normalization schemes. For state prediction, it represents each frame as a single token and models temporal dependencies with a shallow Transformer. Across four resting-state datasets spanning six trait-prediction targets, linear probes on frozen DCAE features generally outperform those on fMRI foundation model representations and remain competitive with fully fine-tuned fMRI foundation models. On three task-fMRI state-prediction tasks, a temporal readout on DCAE features performs comparably to the strongest foundation models evaluated. A Gaussian injection analysis further shows that localized signal changes are recovered more accurately from the frozen DCAE features than from the evaluated foundation-model representations. Together, these results show that strong performance on current fMRI benchmarks is possible without fMRI-specific representation pre-training, making frozen natural-image features as a useful baseline for assessing its added value.

Figures & tables

Appendix figures & tables8 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Scaling Vision Transformers for Functional MRI with Flat Maps

    Oct 15, 2025Connor Lane, Mihir Tripathy, Leema Krishna Murali +15Functional Magnetic Resonance ImagingVision Transformer

  2. Brain-DiT: A Universal Multi-state fMRI Foundation Model with Metadata-Conditioned Pretraining

    Apr 14, 2026Junfeng Xia, Wenhao Ye, Xuanye Pan +3PretrainingFoundation Model

  3. Frozen Brain-MRI Foundation Models Are Site Fingerprints

    Aug 10, 2026Saman RahbarFunctional Magnetic Resonance ImagingFingerprint