Skip to content

Contextual Semantic Relevance and Word Surprisal Predict N400 and P600 Dynamics During Naturalistic Reading

Jul 2026 · arXiv.org · Vol abs/2607.04107 · 1 citation
Computer Science

Abstract

Word surprisal is a well-established computational predictor of human neural responses during language comprehension, but it remains less clear whether local semantic fit explains neural response variation beyond lexical expectation during naturalistic reading. Using the Dublin EEG-based Reading Experiment Corpus (DERCo), this study examined whether contextual semantic relevance predicts word-locked EEG activity in the N400 and P600 windows. Contextual semantic relevance was computed as an attention-aware measure of how strongly a target word is semantically connected to its recent discourse context, and it was compared with GPT-based word surprisal. Across 22 participants and 32 EEG channels, we tested both predictors using regression-based ERP analyses and generalized additive mixed models while controlling for lexical variables and repeated observations. Both predictors were reliably associated with EEG responses, but they showed partly different temporal and scalp-level patterns. Surprisal captured expectancy-related variation, whereas contextual semantic relevance showed robust effects across N400- and P600-window mean voltages, with particularly strong explanatory support in the P600 window. Model comparisons indicated that contextual semantic relevance contributed explanatory value beyond lexical controls and surprisal. These findings suggest that naturalistic reading depends on both lexical expectation and local semantic integration, and that contextual semantic relevance offers an interpretable computational link between discourse semantic fit and ERP dynamics.

View source

Similar papers

Open access Aug 2026

Cognitive Processes of Probabilistic Prediction in Reading: Language Model Surprisal Across Model Sizes, Token Granularity, and Reading Paradigms

Surprisal, the negative log probability of a word given its context, is the dominant computational metric for quantifying reading difficulty and a common item difficulty estimator in reading research. Yet how the language model family, surprisal granularity, and corpus type jointly shape the surprisal–reading time link...

Shu-Ting Liu, Yong Mei · 0 citations
Open access Sep 2026

Neural tracking of surprisal and semantic distance in naturalistic movie viewing

Understanding speech requires listeners to integrate incoming input with prior linguistic and thematic knowledge to access meaning, a task greatly aided by prediction. Surprisal and related phenomena (e.g., next word prediction) tend to be associated with broad activation of language regions during listening. A major c...

Ryan M. O'Leary, Hailey C. Smith, Emily B. Myers et al. · 0 citations
Open access Aug 2026

Attention or prediction? Characterizing the top-down influence of predictive context on speech encoding

Theories of predictive coding propose that perception is the process of inferring the causes of our sensory input by comparing that input with predictions derived from our internal models of the world. Such predictive processes are thought to play a central role in language comprehension, however, robust neurophysiolog...

Alyssa Horng, Wei-Ching Lin, Ilaria Benciolini et al. · 0 citations
Review Open access Sep 2026

Evaluating the compatibility of predictive coding as a model of linguistic predictability effects in reading

Emerging work suggests that the computational architecture of predictive coding (PC), widely studied in visual perception, may be an appropriate model for understanding predictability effects in real-time reading (i. e., the influence of a preceding sentence context on current word processing via putative anticipatory...

Allyson Copeland, B. Payne · 0 citations
Open access Aug 2026

Brain–Language Alignment During Naturalistic Reading and Its Disruption by Mind-Wandering

Encoding models offer a principled framework for linking computational representations of language to neural activity, but most electroencephalography (EEG) evidence for brain–language alignment comes from tightly controlled, word-by-word reading paradigms. Whether such alignment is detectable during naturalistic readi...

Haorui Sun, D. Jangraw · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.