From Language to Cognition: How LLMs Outgrow the Human Language Network

Badr Alkhamissi; Greta Tuckute; Yingtian Tang; Taha Osama A Binhuraib; Antoine Bosselut; Martin Schrimpf

From Language to Cognition: How LLMs Outgrow the Human Language Network

Badr AlKhamissi, Greta Tuckute, Yingtian Tang, Taha Osama A Binhuraib, Antoine Bosselut, Martin Schrimpf

Abstract

Large language models (LLMs) exhibit remarkable similarity to neural activity in the human language network. However, the key properties of language underlying this alignment—and how brain-like representations emerge and change across training—remain unclear. We here benchmark 34 training checkpoints spanning 300B tokens across 8 different model sizes to analyze how brain alignment relates to linguistic competence. Specifically, we find that brain alignment tracks the development of formal linguistic competence—i.e., knowledge of linguistic rules—more closely than functional linguistic competence. While functional competence, which involves world knowledge and reasoning, continues to develop throughout training, its relationship with brain alignment is weaker, suggesting that the human language network primarily encodes formal linguistic structure rather than broader cognitive functions. Notably, we find that the correlation between next-word prediction, behavioral alignment, and brain alignment fades once models surpass human language proficiency. We further show that model size is not a reliable predictor of brain alignment when controlling for the number of features. Finally, using the largest set of rigorous neural language benchmarks to date, we show that language brain alignment benchmarks remain unsaturated, highlighting opportunities for improving future models. Taken together, our findings suggest that the human language network is best modeled by formal, rather than functional, aspects of language.

Anthology ID:: 2025.emnlp-main.1237
Volume:: Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing
Month:: November
Year:: 2025
Address:: Suzhou, China
Editors:: Christos Christodoulopoulos, Tanmoy Chakraborty, Carolyn Rose, Violet Peng
Venue:: EMNLP
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 24332–24350
Language:
URL:: https://preview.aclanthology.org/ingest-emnlp/2025.emnlp-main.1237/
DOI:
Bibkey:
Cite (ACL):: Badr AlKhamissi, Greta Tuckute, Yingtian Tang, Taha Osama A Binhuraib, Antoine Bosselut, and Martin Schrimpf. 2025. From Language to Cognition: How LLMs Outgrow the Human Language Network. In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing, pages 24332–24350, Suzhou, China. Association for Computational Linguistics.
Cite (Informal):: From Language to Cognition: How LLMs Outgrow the Human Language Network (AlKhamissi et al., EMNLP 2025)
Copy Citation:
PDF:: https://preview.aclanthology.org/ingest-emnlp/2025.emnlp-main.1237.pdf
Checklist:: 2025.emnlp-main.1237.checklist.pdf

PDF Cite Search Checklist Fix data