Brain Score Tracks Shared Properties of Languages: Evidence from Many Natural Languages and Structured Sequences

Jingnong Qu; Ashvin Ranjan; Shane Steinert-Threlkeld

arXiv:2604.15503·cs.CL·April 20, 2026

Brain Score Tracks Shared Properties of Languages: Evidence from Many Natural Languages and Structured Sequences

Jingnong Qu, Ashvin Ranjan, Shane Steinert-Threlkeld

PDF

TL;DR

This study evaluates how language models trained on diverse natural and structured data correlate with human brain activity, revealing shared structural properties but questioning the specificity of Brain Score as a measure of human-like processing.

Contribution

The paper demonstrates that Brain Score performance is similar across models trained on various natural and structured datasets, highlighting shared structural features.

Findings

01

Models trained on different natural languages have similar Brain Score performance.

02

Structured data like the genome and code also achieve comparable Brain Score results.

03

High Brain Score does not necessarily indicate human-like language processing.

Abstract

Recent breakthroughs in language models (LMs) using neural networks have raised the question: how similar are these models' processing to human language processing? Results using a framework called Brain Score (BS) -- predicting fMRI activations during reading from LM activations -- have been used to argue for a high degree of similarity. To understand this similarity, we conduct experiments by training LMs on various types of input data and evaluate them on BS. We find that models trained on various natural languages from many different language families have very similar BS performance. LMs trained on other structured data -- the human genome, Python, and pure hierarchical structure (nested parentheses) -- also perform reasonably well and close to natural languages in some cases. These findings suggest that BS can highlight language models' ability to extract common structure across…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.