Abstract
Recent advancements in Chinese Spelling Correction (CSC) predominantly leverage pre-trained language models (PLMs). However, a notable challenge with fine-tuned PLM-based CSC models is their tendency to over-correct, leading to poor generalization for error patterns outside the standard distribution. To address this, we developed a teacher network guided by prior knowledge for distillation learning of CSC models. Unlike traditional teacher networks, which depend on task-related pre-training, our method infuses task-related prior information into the teacher network, offering guidance beyond mere labels to the student network. This strategy significantly enhances the CSC model’s language modeling capabilities, crucial for minimizing over-correction. Importantly, our approach is model-independent and the teacher network does not require task-related pre-training, making it broadly applicable for enhancing various PLM-based CSC models with minimal additional computational resources. Extensive experiments on widely used benchmarks demonstrate that our method achieves new state-of-the-art results. Additionally, we explored the potential of generalizing our method to other non-autoregressive text-generation tasks.- Anthology ID:
- 2024.findings-acl.806
- Volume:
- Findings of the Association for Computational Linguistics ACL 2024
- Month:
- August
- Year:
- 2024
- Address:
- Bangkok, Thailand and virtual meeting
- Editors:
- Lun-Wei Ku, Andre Martins, Vivek Srikumar
- Venue:
- Findings
- SIG:
- Publisher:
- Association for Computational Linguistics
- Note:
- Pages:
- 13578–13589
- Language:
- URL:
- https://aclanthology.org/2024.findings-acl.806
- DOI:
- Cite (ACL):
- Chi Wei, Shaobin Huang, Rongsheng Li, Naiyu Yan, and Rui Wang. 2024. Training a Better Chinese Spelling Correction Model via Prior-knowledge Guided Teacher. In Findings of the Association for Computational Linguistics ACL 2024, pages 13578–13589, Bangkok, Thailand and virtual meeting. Association for Computational Linguistics.
- Cite (Informal):
- Training a Better Chinese Spelling Correction Model via Prior-knowledge Guided Teacher (Wei et al., Findings 2024)
- PDF:
- https://preview.aclanthology.org/nschneid-patch-4/2024.findings-acl.806.pdf