Evaluating the Adaptability of Large Language Models to Linguistic Variation

Ziyan Xu, Marina Seghier, Alice Millour, Carlos-Emiliano Gonzalez-Gallardo, Jean-Yves Antoine


Abstract
Large language models (LLMs) are often assumed to generalize easily across linguistic contexts, yet their ability to adapt to genre variation remains underexplored. This study examines that question through a French Named Entity Recognition (NER) task conducted on NEM.fr, a multi-genre corpus annotated with gold named entities (NEs) spanning 11 text types, from juridical and encyclopedic prose to poetry, political speech, and online discourse. We evaluate the reasoning-oriented model DeepSeek R1 across six prompting configurations (zero-, one-, and few-shot, with and without chain-of-thought reasoning), while keeping the annotation scheme, prompting format, and evaluation pipeline constant to isolate the role of genre. Performance is measured using both strict and fuzzy F1-based metrics. The results show that prompting choices have little effect once the model has learned the task format, but that genre differences strongly influence outcomes: fuzzy F1 scores range from about 0.85 in formal genres to below 0.20 in informal ones. Even under tightly controlled conditions, LLM behaviour proves highly sensitive to textual regularity and stylistic variation, highlighting genre as a key factor in assessing model robustness.
Anthology ID:
2026.lrec-1.183
Volume:
Proceedings of the Fifteenth Language Resources and Evaluation Conference
Month:
May
Year:
2026
Address:
Palma de Mallorca, Spain
Editors:
Stelios Piperidis, Núria Bel, Henk van den Heuvel, Nancy Ide, Simon Krek, Antonio Toral
Venue:
LREC
SIG:
Publisher:
ELRA Language Resource Association
Note:
Pages:
2334–2343
Language:
External URL:
https://lrec.elra.info/lrec2026-main-183
DOI:
10.63317/57bpwacmcpr2
Bibkey:
Cite (ACL):
Ziyan Xu, Marina Seghier, Alice Millour, Carlos-Emiliano Gonzalez-Gallardo, and Jean-Yves Antoine. 2026. Evaluating the Adaptability of Large Language Models to Linguistic Variation. In Proceedings of the Fifteenth Language Resources and Evaluation Conference, pages 2334–2343, Palma de Mallorca, Spain. ELRA Language Resource Association.
Cite (Informal):
Evaluating the Adaptability of Large Language Models to Linguistic Variation (Xu et al., LREC 2026)
Copy Citation: