bioRxiv · 10.64898/2026.07.10.737662
Surprisal contributes little beyond contextual embeddings in high-gamma ECoG encoding
Abstract
Surprisal and contextual embeddings are both derived from large language models and are widely used to predict neural responses during language comprehension, but it is unclear whether surprisal adds information beyond embeddings. We test this directly: does word surprisal improve out-of-sample prediction of high-gamma ECoG responses after GPT-2 XL contextual embeddings are included? Using public ECoG recordings from natural speech, we fit word-aligned ridge encoding models with baseline stimulus features, GPT-2 XL embeddings, and GPT-2 XL surprisal. Adding one surprisal predictor left held-out correlation essentially unchanged at the center lag, and the effect remained within a prespecified equivalence margin across alternative lags and sensitivity analyses. This near-zero increment suggests that surprisal does not act as an independent predictor. It is better understood as a compressed readout of the same broader predictive state that the embeddings already capture.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Sakuma, T.. 2026-07-16. Surprisal contributes little beyond contextual embeddings in high-gamma ECoG encoding. https://doi.org/10.64898/2026.07.10.737662
Cite the original work for its findings. Save a collection to share your selection of sources.