Skip to content

The Capacity of Thought: Benchmarking Llama 3.2 in Semantic fMRI Neural Language Decoding and Improving the Huth Encoding-Model Baseline

Jul 2026 · arXiv.org · Vol abs/2607.12079 · 0 citations · 6 references
Computer Science

TL;DR

The results demonstrate that high-capacity language models do not inherently improve fMRI decoding and can actively obscure failures without rigorous blind-control evaluation.

Abstract

Decoding continuous language from fMRI signals remains a core challenge in non-invasive brain-computer interface research. We present two complementary investigations. First, we improve the Huth et al. ridge regression encoding pipeline through expanded voxel selection (10K->15K), substitution of GPT-2 medium for GPT-1 as the beam-search proposal model, and GPU-accelerated bootstrap training, achieving mean METEOR = 0.149 and BLEU-1 = 0.200 across three held-out narratives for subject UTS03 -- an 11% relative METEOR gain over our replication baseline. Second, we introduce fMRIFlamingo, which maps BOLD activity to a frozen Llama-3.2-1B with trainable gated cross-attention layers via a learned brain tokenizer and a Perceiver Resampler. Despite achieving 42.86% Top-1 accuracy on a 1-in-100 ranking task, well above chance, a blind control ablation with zeroed fMRI inputs yields near-identical scores, revealing that apparent decoding success is driven primarily by the frozen language prior rather than by neural input. These results demonstrate that high-capacity language models do not inherently improve fMRI decoding and can actively obscure failures without rigorous blind-control evaluation.

View source

Similar papers

Preprint Sep 2026

Predictor Construction Can Reverse Multimodal Neural Contrasts

These results show that nested neural contrasts do not identify represented content by themselves: predictor construction is part of the experimental design, and matched controls are required for representational claims.

Lucas Nadolskis, Galen Pogoncheff, Michael Beyeler · 0 citations
Open access Aug 2026

Brain–Language Alignment During Naturalistic Reading and Its Disruption by Mind-Wandering

Encoding models offer a principled framework for linking computational representations of language to neural activity, but most electroencephalography (EEG) evidence for brain–language alignment comes from tightly controlled, word-by-word reading paradigms. Whether such alignment is detectable during naturalistic readi...

Haorui Sun, D. Jangraw · 0 citations

Pretraining for Sample-Efficient Neural Interfaces

This work proposes MAPA, an otherwise vanilla masked autoencoder with two spatial encodings, an anatomical region embedding and a relative positional encoding that together enable it to learn neural representations that transfer to unseen subjects and across various tasks.

Ben-Ting Tang, Z. Spalding, G. Cogan · 0 citations
Preprint Aug 2026

When Is a Steerable Concept Representation Real? Measurement Confounds in a Cross-Family Audit of Neuroscience Parallels in LLMs

Large language models (LLMs) are increasingly reported to exhibit human-like neural and cognitive signatures, including concept cells, mental number lines, and cognitive maps. These claims often rely on linear probing and activation steering applied to a single model, yet both methods are highly sensitive to measuremen...

Yuqi Wu, Shengming Zhao, Jie Chen · 3 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.