A Systematic Assessment of Language Models with Linguistic Minimal Pairs in Chinese
Yikang Liu, Yeting Shen, Hongao Zhu, Lilong Xu, Zhiheng Qian, Siyuan Song, Kejia Zhang, Jialong Tang, Pei Zhang, Baosong Yang, Rui Wang, Hai Hu
Abstract
We present ZhoBLiMP, the largest linguistic minimal pair benchmark for Chinese, with over 100 paradigms, ranging from topicalization to the Ba construction. We then train from scratch a suite of Chinese language models (LMs) with different tokenizers, parameter sizes, and token volumes, to study the learning curves of LMs on Chinese. To mitigate the biases introduced by unequal lengths of the sentences in a minimal pair, we propose a new metric named sub-linear length normalized log-probabilities (SLLN-LP). Using SLLN-LP as the metric, our results show that Anaphor, Quantifiers, and Ellipsis in Chinese are difficult for LMs even up to 32B parameters, and that SLLN-LP successfully mitigates biases in ZhoBLiMP, JBLiMP and BLiMP. We conclude that future evaluations should be more carefully designed to consider the intricate relations between linking functions, LMs, and targeted minimal pairs.- Anthology ID:
- 2026.tacl-1.34
- Volume:
- Transactions of the Association for Computational Linguistics, Volume 14
- Month:
- Year:
- 2026
- Address:
- Cambridge, MA
- Venue:
- TACL
- SIG:
- Publisher:
- MIT Press
- Note:
- Pages:
- 755–771
- Language:
- URL:
- https://preview.aclanthology.org/ingest-latest-mitpress-cl-tacl/2026.tacl-1.34/
- DOI:
- 10.1162/tacl.a.648
- Cite (ACL):
- Yikang Liu, Yeting Shen, Hongao Zhu, Lilong Xu, Zhiheng Qian, Siyuan Song, Kejia Zhang, Jialong Tang, Pei Zhang, Baosong Yang, Rui Wang, and Hai Hu. 2026. A Systematic Assessment of Language Models with Linguistic Minimal Pairs in Chinese. Transactions of the Association for Computational Linguistics, 14:755–771.
- Cite (Informal):
- A Systematic Assessment of Language Models with Linguistic Minimal Pairs in Chinese (Liu et al., TACL 2026)
- PDF:
- https://preview.aclanthology.org/ingest-latest-mitpress-cl-tacl/2026.tacl-1.34.pdf