Cecilia Domingo
2025
Mention detection with LLMs in pair-programming dialogue
Cecilia Domingo
|
Paul Piwek
|
Svetlana Stoyanchev
|
Michel Wermelinger
Proceedings of the Eighth Workshop on Computational Models of Reference, Anaphora and Coreference
We tackle the task of mention detection for pair-programming dialogue, a setting which adds several challenges to the task due to the characteristics of natural dialogue, the dynamic environment of the dialogue task, and the domain-specific vocabulary and structures. We compare recent variants of the Llama and GPT families and explore different prompt and context engineering approaches. While aspects like hesitations and references to read-out code and variable names made the task challenging, GPT 4.1 approximated human performance when we provided few-shot examples similar to the inference text and corrected formatting errors.
2022
Discourse annotation — Towards a dialogue system for pair programming
Cecilia Domingo
|
Paul Piwek
|
Svetlana Stoyanchev
|
Michel Wermelinger
Traitement Automatique des Langues, Volume 63, Numéro 3 : Etats de l'art en TAL [Review articles in NLP]
2021
What is on Social Media that is not in WordNet? A Preliminary Analysis on the TwitterAAE Corpus
Cecilia Domingo
|
Tatiana Gonzalez-Ferrero
|
Itziar Gonzalez-Dios
Proceedings of the 11th Global Wordnet Conference
Natural Language Processing tools and resources have been so far mainly created and trained for standard varieties of language. Nowadays, with the use of large amounts of data gathered from social media, other varieties and registers need to be processed, which may present other challenges and difficulties. In this work, we focus on English and we present a preliminary analysis by comparing the TwitterAAE corpus, which is annotated for ethnicity, and WordNet by quantifying and explaining the online language that WordNet misses.