OLA: Output Language Alignment in Code-Switched LLM Interactions

Juhyun Oh; Haneul Yoo; Faiz Ghifari Haznitrama; Alice Oh

OLA: Output Language Alignment in Code-Switched LLM Interactions

Juhyun Oh, Haneul Yoo, Faiz Ghifari Haznitrama, Alice Oh

Abstract

Code-switching, alternating between languages within a conversation, is natural for multilingual users, yet poses fundamental challenges for large language models (LLMs). When a user code-switches in their prompt to an LLM, they typically do not specify the expected language of the LLM response, and thus LLMs must infer the output language from contextual and pragmatic cues. We find that current LLMs systematically fail to align with this expectation, responding in undesired languages even when cues are clear to humans. We introduce OLA, a benchmark to evaluate LLMs’ Output Language Alignment in code-switched interactions. OLA focuses on Korean–English code-switching and spans simple intra-sentential mixing to instruction–content mismatches. Even frontier models frequently misinterpret implicit language expectation, exhibiting a systematic bias toward non-English responses. We further show this bias generalizes beyond Korean to Chinese and Indonesian pairs. Models also show instability through mid-response switching and language intrusions. Chain-of-Thought prompting fails to resolve these errors, indicating weak pragmatic reasoning about output language. However, Code-Switching Aware DPO with minimal data (~1K examples) substantially reduces misalignment, suggesting these failures stem from insufficient alignment rather than fundamental limitations. Our results highlight the need to align multilingual LLMs with users’ implicit expectations in real-world code-switched interactions.

Anthology ID:: 2026.acl-long.2162
Volume:: Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Month:: July
Year:: 2026
Address:: San Diego, California, United States
Editors:: Maria Liakata, Viviane P. Moreira, Jiajun Zhang, David Jurgens
Venue:: ACL
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 46603–46624
Language:
URL:: https://preview.aclanthology.org/ingest-acl/2026.acl-long.2162/
DOI:
Bibkey:
Cite (ACL):: Juhyun Oh, Haneul Yoo, Faiz Ghifari Haznitrama, and Alice Oh. 2026. OLA: Output Language Alignment in Code-Switched LLM Interactions. In Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 46603–46624, San Diego, California, United States. Association for Computational Linguistics.
Cite (Informal):: OLA: Output Language Alignment in Code-Switched LLM Interactions (Oh et al., ACL 2026)
Copy Citation:
PDF:: https://preview.aclanthology.org/ingest-acl/2026.acl-long.2162.pdf
Checklist:: 2026.acl-long.2162.checklist.pdf

PDF Cite Search Checklist Fix data