Parsing Chinese Sentences with Grammatical Relations

Weiwei Sun; Yufei Chen; Xiaojun Wan; Meichun Liu

doi:10.1162/coli_a_00343

Parsing Chinese Sentences with Grammatical Relations

Weiwei Sun, Yufei Chen, Xiaojun Wan, Meichun Liu

Abstract

We report our work on building linguistic resources and data-driven parsers in the grammatical relation (GR) analysis for Mandarin Chinese. Chinese, as an analytic language, encodes grammatical information in a highly configurational rather than morphological way. Accordingly, it is possible and reasonable to represent almost all grammatical relations as bilexical dependencies. In this work, we propose to represent grammatical information using general directed dependency graphs. Both only-local and rich long-distance dependencies are explicitly represented. To create high-quality annotations, we take advantage of an existing TreeBank, namely, Chinese TreeBank (CTB), which is grounded on the Government and Binding theory. We define a set of linguistic rules to explore CTB’s implicit phrase structural information and build deep dependency graphs. The reliability of this linguistically motivated GR extraction procedure is highlighted by manual evaluation. Based on the converted corpus, data-driven, including graph- and transition-based, models are explored for Chinese GR parsing. For graph-based parsing, a new perspective, graph merging, is proposed for building flexible dependency graphs: constructing complex graphs via constructing simple subgraphs. Two key problems are discussed in this perspective: (1) how to decompose a complex graph into simple subgraphs, and (2) how to combine subgraphs into a coherent complex graph. For transition-based parsing, we introduce a neural parser based on a list-based transition system. We also discuss several other key problems, including dynamic oracle and beam search for neural transition-based parsing. Evaluation gauges how successful GR parsing for Chinese can be by applying data-driven models. The empirical analysis suggests several directions for future study.

Anthology ID:: J19-1003
Volume:: Computational Linguistics, Volume 45, Issue 1 - March 2019
Month:: March
Year:: 2019
Address:: Cambridge, MA
Venue:: CL
SIG:
Publisher:: MIT Press
Note:
Pages:: 95–136
Language:
URL:: https://aclanthology.org/J19-1003
DOI:: 10.1162/coli_a_00343
Bibkey:
Cite (ACL):: Weiwei Sun, Yufei Chen, Xiaojun Wan, and Meichun Liu. 2019. Parsing Chinese Sentences with Grammatical Relations. Computational Linguistics, 45(1):95–136.
Cite (Informal):: Parsing Chinese Sentences with Grammatical Relations (Sun et al., CL 2019)
Copy Citation:
PDF:: https://preview.aclanthology.org/nschneid-patch-4/J19-1003.pdf

PDF Search