Dun Deng


2017

pdf bib
Translation Divergences in Chinese–English Machine Translation: An Empirical Investigation
Dun Deng | Nianwen Xue
Computational Linguistics, Volume 43, Issue 3 - September 2017

In this article, we conduct an empirical investigation of translation divergences between Chinese and English relying on a parallel treebank. To do this, we first devise a hierarchical alignment scheme where Chinese and English parse trees are aligned in a way that eliminates conflicts and redundancies between word alignments and syntactic parses to prevent the generation of spurious translation divergences. Using this Hierarchically Aligned Chinese–English Parallel Treebank (HACEPT), we are able to semi-automatically identify and categorize the translation divergences between the two languages and quantify each type of translation divergence. Our results show that the translation divergences are much broader than described in previous studies that are largely based on anecdotal evidence and linguistic knowledge. The distribution of the translation divergences also shows that some high-profile translation divergences that motivate previous research are actually very rare in our data, whereas other translation divergences that have previously received little attention actually exist in large quantities. We also show that HACEPT allows the extraction of syntax-based translation rules, most of which are expressive enough to capture the translation divergences, and point out that the syntactic annotation in existing treebanks is not optimal for extracting such translation rules. We also discuss the implications of our study for attempts to bridge translation divergences by devising shared semantic representations across languages. Our quantitative results lend further support to the observation that although it is possible to bridge some translation divergences with semantic representations, other translation divergences are open-ended, thus building a semantic representation that captures all possible translation divergences may be impractical.

2015

pdf bib
Harmonizing word alignments and syntactic structures for extracting phrasal translation equivalents
Dun Deng | Nianwen Xue | Shiman Guo
Proceedings of the Ninth Workshop on Syntax, Semantics and Structure in Statistical Translation

2014

pdf bib
Aligning Chinese-English Parallel Parse Trees: Is it Feasible?
Dun Deng | Nianwen Xue
Proceedings of LAW VIII - The 8th Linguistic Annotation Workshop

pdf bib
Building a Hierarchically Aligned Chinese-English Parallel Treebank
Dun Deng | Nianwen Xue
Proceedings of COLING 2014, the 25th International Conference on Computational Linguistics: Technical Papers