FReD is introduced, which derives fMRI representations from a frozen Deep Compression AutoEncoder pre-trained exclusively on natural images and pairs them with a task specific readout, making frozen natural-image features as a useful baseline for assessing its added value on current fMRI benchmarks.
Abstract
Foundation models pre-trained on large-scale fMRI datasets have shown strong downstream performance, but at substantial data and computation cost. To investigate how much fMRI-specific pre-training is actually needed for such performance, we introduce FReD, which derives fMRI representations from a frozen Deep Compression AutoEncoder (DCAE) pre-trained exclusively on natural images and pairs them with a task specific readout. For trait prediction, FReD summarizes frame-wise representations by their temporal mean and log-standard deviation and applies linear probing, with late fusion across two normalization schemes. For state prediction, it represents each frame as a single token and models temporal dependencies with a shallow Transformer. Across four resting-state datasets spanning six trait-prediction targets, linear probes on frozen DCAE features generally outperform those on fMRI foundation model representations and remain competitive with fully fine-tuned fMRI foundation models. On three task-fMRI state-prediction tasks, a temporal readout on DCAE features performs comparably to the strongest foundation models evaluated. A Gaussian injection analysis further shows that localized signal changes are recovered more accurately from the frozen DCAE features than from the evaluated foundation-model representations. Together, these results show that strong performance on current fMRI benchmarks is possible without fMRI-specific representation pre-training, making frozen natural-image features as a useful baseline for assessing its added value.
We present a dynamical-systems based model for resting-state functional magnetic resonance imaging (rs-fMRI), trained on a dataset of roughly 40K rs-fMRI sequences covering a wide variety of public and available-by-permission datasets. While most existing proposals use transformer backbones, we utilize multi-resolution...
Sourav Pal, V. Lương, Hoseok Lee et al.· 1 citation
Resting-state connectivity can predict task-evoked fMRI activation, but correspondence with an individual task map may partly reflect a shared population pattern. We evaluated the Variational Resting-state-to-Task Prediction TransformeR (VRPTR), a three-dimensional encoder-decoder combining a compressed Transformer bot...
D. Di Giovanni, D. L. Collins· bioRxiv· 0 citations
Self-supervised pretraining reshaped prediction in language and vision, and brain foundation models (BFMs) inherited its promise. Representations learned from large unlabelled corpora should capture individual functional dynamics and generalise across cohorts. However, kernel ridge regression (KRR) fitted on functional...
G. Marraffini, Victoria Shevchenko, Carlo Alberto Barbano et al.· 0 citations
EEG foundation models increasingly use masked prediction to learn from unlabeled recordings, but optimizing this objective does not ensure transferable neural representations. A central challenge is that stable positional cues and local correlations can make masked regions predictable without integrating distributed ne...
Kieren Yu, Zi-Yang Liu, Chang Huang et al.· 0 citations
Findings support organizing fMRI pretraining and adaptation by measured learning relations rather than treating domains and tasks as independent flat sets.
Jun-Feng Xia, Wen-Hao Ye, Jun-Xiang Zhang et al.· 1 citation
A general fMRI sequence prediction model, the Frequency-Filtered Attention Transformer (FFAformer), which models low-frequency variations in the frequency domain to capture long-range dependencies and incorporates FC consistency constraints to preserve brain network structure.
Chengcheng Du, Fei-Fei Zhao, Yin-Qian Sun et al.· iScience· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.