Elaine Uí Dhonnchadha
Also published as: Elaine Uí Dhonnchadha
Papers on this page may belong to the following people: Elaine Uí Dhonnchadha, Elaine Uí Dhonnchadha
2026
Creating a Hybrid Rule and Neural Network Based Semantic Tagger Using Silver Standard Data: The PyMUSAS Framework for Multilingual Semantic Annotation
Andrew Moore | Paul Rayson | Dawn Archer | Tim Czerniak | Dawn Knight | Daisy Monika Lal | Gearóid Ó Donnchadha | Mícheál J. Ó Meachair | Scott Piao | Elaine Uí Dhonnchadha | Johanna Vuorinen | Yan Yabo | Xiaobin Yang
Proceedings of the Fifteenth Language Resources and Evaluation Conference
Andrew Moore | Paul Rayson | Dawn Archer | Tim Czerniak | Dawn Knight | Daisy Monika Lal | Gearóid Ó Donnchadha | Mícheál J. Ó Meachair | Scott Piao | Elaine Uí Dhonnchadha | Johanna Vuorinen | Yan Yabo | Xiaobin Yang
Proceedings of the Fifteenth Language Resources and Evaluation Conference
Word Sense Disambiguation (WSD) has been widely evaluated using the semantic frameworks of WordNet, BabelNet, and the Oxford Dictionary of English. However, for the UCREL Semantic Analysis System (USAS) framework, no open extensive evaluation has been performed beyond lexical coverage or single language evaluation. In this work, we perform the largest semantic tagging evaluation of the rule based system that uses the lexical resources in the USAS framework covering five different languages using four existing datasets and one novel Chinese dataset. We create a new silver labelled English dataset, to overcome the lack of manually tagged training data, that we train and evaluate various mono and multilingual neural models in both mono and cross-lingual evaluation setups with comparisons to their rule based counterparts, and show how a rule based system can be enhanced with a neural network model. The resulting neural network models, including the data they were trained on, the Chinese evaluation dataset, and all of the code will be released as open resources.
Grammar Engineering Meets LLMs: Development of Cantonese and Irish ParGram Treebanks
Chit-Fung Lam | Elaine Uí Dhonnchadha
Proceedings of the Third Workshop on the Bridges and Gaps between Formal and Computational Linguistics (BriGap-3)
Chit-Fung Lam | Elaine Uí Dhonnchadha
Proceedings of the Third Workshop on the Bridges and Gaps between Formal and Computational Linguistics (BriGap-3)
Grammar engineering requires expertise in linguistic formalism and computational implementation, particularly in parallel grammar projects that balance cross-linguistic consistency with language-specific properties. This paper presents the development of Cantonese and Irish treebanks within the Parallel Grammar (ParGram) Project, where linguistic parallelism is maintained at an abstract functional level. We also investigated the methodological potential and limitations of using multilingual LLMs to support grammar engineering, focusing on Cantonese–Irish translation and the generation of formal syntactic structures using OpenAI’s gpt-oss-120b. The results showed that translation performance was generally unsatisfactory and unaffected by prompt language. For syntactic structure generation, the model produced some structurally meaningful outputs, but performed poorly on tasks requiring cross-linguistic abstraction. Nonetheless, LLM-generated outputs may still offer some reference value by suggesting alternative analyses and (partially) capturing predicate–argument relations. Overall, our findings highlight both the potential and limitations of using LLMs in collaborative grammar engineering, while underscoring the continued importance of expert-driven analysis and verification.
2025
Proceedings of the 5th Celtic Language Technology Workshop
Brian Davis | Theodorus Fransen | Elaine Uí Dhonnchadha | Abigail Walsh
Proceedings of the 5th Celtic Language Technology Workshop
Brian Davis | Theodorus Fransen | Elaine Uí Dhonnchadha | Abigail Walsh
Proceedings of the 5th Celtic Language Technology Workshop
2022
How NLP Can Strengthen Digital Game Based Language Learning Resources for Less Resourced Languages
Monica Ward | Liang Xu | Elaine Uí Dhonnchadha
Proceedings of the 9th Workshop on Games and Natural Language Processing within the 13th Language Resources and Evaluation Conference
Monica Ward | Liang Xu | Elaine Uí Dhonnchadha
Proceedings of the 9th Workshop on Games and Natural Language Processing within the 13th Language Resources and Evaluation Conference
This paper provides an overview of the Cipher engine which enables the development of a Digital Educational Game (DEG) based on noticing ciphers or patterns in texts. The Cipher engine was used to develop the Cipher: Faoi Gheasa, a digital educational game for Irish, which incorporates NLP resources and is informed by Digital Game-Based Language Learning (DGBLL) and Computer-Assisted Language Learning (CALL) research. The paper outlines six phases where NLP has strengthened the Cipher: Faoi Gheasa game. It shows how the Cipher engine can be used to build a Cipher game for other languages, particularly low-resourced and endangered languages in which NLP resources are under-developed or few in number.