Incremental Sentence Processing Mechanisms in Autoregressive Transformer Language Models

Michael Hanna, Aaron Mueller


Abstract
Autoregressive transformer language models (LMs) possess strong syntactic abilities, often successfully handling phenomena from agreement to NPI licensing. However, the features they use to incrementally process their linguistic input are not well understood. In this paper, we fill this gap by studying the mechanisms underlying garden path sentence processing in LMs. Specifically, we ask: (1) Do LMs use syntactic features or shallow heuristics to perform incremental sentence processing? (2) Do LMs represent only one potential interpretation, or multiple? and (3) Do LMs reanalyze or repair their initial incorrect representations? To address these questions, we use sparse autoencoders to identify interpretable features that determine which continuation—and thus which reading—of a garden path sentence the LM prefers. We find that while many important features relate to syntactic structure, some reflect syntactically irrelevant heuristics. Moreover, though most active features correspond to one reading of the sentence, some features correspond to the other, suggesting that LMs assign weight to both possibilities. Finally, LMs fail to re-use features to answer follow-up questions.
Anthology ID:
2025.naacl-long.164
Volume:
Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers)
Month:
April
Year:
2025
Address:
Albuquerque, New Mexico
Editors:
Luis Chiruzzo, Alan Ritter, Lu Wang
Venue:
NAACL
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
3181–3203
Language:
URL:
https://preview.aclanthology.org/moar-dois/2025.naacl-long.164/
DOI:
10.18653/v1/2025.naacl-long.164
Bibkey:
Cite (ACL):
Michael Hanna and Aaron Mueller. 2025. Incremental Sentence Processing Mechanisms in Autoregressive Transformer Language Models. In Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers), pages 3181–3203, Albuquerque, New Mexico. Association for Computational Linguistics.
Cite (Informal):
Incremental Sentence Processing Mechanisms in Autoregressive Transformer Language Models (Hanna & Mueller, NAACL 2025)
Copy Citation:
PDF:
https://preview.aclanthology.org/moar-dois/2025.naacl-long.164.pdf