@inproceedings{mervaala-kousa-2024-order,
    title = "Order Up! Micromanaging Inconsistencies in {C}hat{GPT}-4o Text Analyses",
    author = "Mervaala, Erkki  and
      Kousa, Ilona",
    editor = {H{\"a}m{\"a}l{\"a}inen, Mika  and
      {\"O}hman, Emily  and
      Miyagawa, So  and
      Alnajjar, Khalid  and
      Bizzoni, Yuri},
    booktitle = "Proceedings of the 4th International Conference on Natural Language Processing for Digital Humanities",
    month = nov,
    year = "2024",
    address = "Miami, USA",
    publisher = "Association for Computational Linguistics",
    url = "https://preview.aclanthology.org/sigedu-bea-out-of-sync-correction/2024.nlp4dh-1.51/",
    doi = "10.18653/v1/2024.nlp4dh-1.51",
    pages = "521--535",
    abstract = "Large language model (LLM) applications have taken the world by storm in the past two years, and the academic sphere has not been an exception. One common, cumbersome task for researchers to attempt to automatise has been text annotation and, to an extent, analysis. Popular LLMs such as ChatGPT have been examined as a research assistant and as an analysis tool, and several discrepancies regarding both transparency and the generative content have been uncovered. Our research approaches the usability and trustworthiness of ChatGPT for text analysis from the point of view of an ``out-of-the-box'' zero-shot or few-shot setting, focusing on how the context window and mixed text types affect the analyses generated. Results from our testing indicate that both the types of the texts and the ordering of different kinds of texts do affect the ChatGPT analysis, but also that the context-building is less likely to cause analysis deterioration when analysing similar texts. Though some of these issues are at the core of how LLMs function, many of these caveats can be addressed by transparent research planning."
}Markdown (Informal)
[Order Up! Micromanaging Inconsistencies in ChatGPT-4o Text Analyses](https://preview.aclanthology.org/sigedu-bea-out-of-sync-correction/2024.nlp4dh-1.51/) (Mervaala & Kousa, NLP4DH 2024)
ACL