Plug-and-Play Knowledge Injection for Pre-trained Language Models

Zhengyan Zhang; Zhiyuan Zeng; Yankai Lin; Huadong Wang; Deming Ye; Chaojun Xiao; Xu Han; Zhiyuan Liu; Peng Li; Maosong Sun; Jie Zhou

doi:10.18653/v1/2023.acl-long.594

Plug-and-Play Knowledge Injection for Pre-trained Language Models

Zhengyan Zhang, Zhiyuan Zeng, Yankai Lin, Huadong Wang, Deming Ye, Chaojun Xiao, Xu Han, Zhiyuan Liu, Peng Li, Maosong Sun, Jie Zhou

Abstract

Injecting external knowledge can improve the performance of pre-trained language models (PLMs) on various downstream NLP tasks. However, massive retraining is required to deploy new knowledge injection methods or knowledge bases for downstream tasks. In this work, we are the first to study how to improve the flexibility and efficiency of knowledge injection by reusing existing downstream models. To this end, we explore a new paradigm plug-and-play knowledge injection, where knowledge bases are injected into frozen existing downstream models by a knowledge plugin. Correspondingly, we propose a plug-and-play injection method map-tuning, which trains a mapping of knowledge embeddings to enrich model inputs with mapped embeddings while keeping model parameters frozen. Experimental results on three knowledge-driven NLP tasks show that existing injection methods are not suitable for the new paradigm, while map-tuning effectively improves the performance of downstream models. Moreover, we show that a frozen downstream model can be well adapted to different domains with different mapping networks of domain knowledge. Our code and models are available at https://github.com/THUNLP/Knowledge-Plugin.

Anthology ID:: 2023.acl-long.594
Original:: 2023.acl-long.594v1
Version 2:: 2023.acl-long.594v2
Volume:: Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Month:: July
Year:: 2023
Address:: Toronto, Canada
Editors:: Anna Rogers, Jordan Boyd-Graber, Naoaki Okazaki
Venue:: ACL
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 10641–10658
Language:
URL:: https://aclanthology.org/2023.acl-long.594
DOI:: 10.18653/v1/2023.acl-long.594
Bibkey:
Cite (ACL):: Zhengyan Zhang, Zhiyuan Zeng, Yankai Lin, Huadong Wang, Deming Ye, Chaojun Xiao, Xu Han, Zhiyuan Liu, Peng Li, Maosong Sun, and Jie Zhou. 2023. Plug-and-Play Knowledge Injection for Pre-trained Language Models. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 10641–10658, Toronto, Canada. Association for Computational Linguistics.
Cite (Informal):: Plug-and-Play Knowledge Injection for Pre-trained Language Models (Zhang et al., ACL 2023)
Copy Citation:
PDF:: https://preview.aclanthology.org/naacl24-info/2023.acl-long.594.pdf
Video:: https://preview.aclanthology.org/naacl24-info/2023.acl-long.594.mp4

PDF (v2) PDF (v1) Search Video