A Multilingual Dataset of Racial Stereotypes in Social Media Conversational Threads

Tom Bourgeade; Alessandra Teresa Cignarella; Simona Frenda; Mario Laurent; Wolfgang Schmeisser-Nieto; Farah Benamara; Cristina Bosco; Véronique Moriceau; Viviana Patti; Mariona Taulé

doi:10.18653/v1/2023.findings-eacl.51

A Multilingual Dataset of Racial Stereotypes in Social Media Conversational Threads

Tom Bourgeade, Alessandra Teresa Cignarella, Simona Frenda, Mario Laurent, Wolfgang Schmeisser-Nieto, Farah Benamara, Cristina Bosco, Véronique Moriceau, Viviana Patti, Mariona Taulé

Abstract

In this paper, we focus on the topics of misinformation and racial hoaxes from a perspective derived from both social psychology and computational linguistics. In particular, we consider the specific case of anti-immigrant feeling as a first case study for addressing racial stereotypes. We describe the first corpus-based study for multilingual racial stereotype identification in social media conversational threads. Our contributions are: (i) a multilingual corpus of racial hoaxes, (ii) a set of common guidelines for the annotation of racial stereotypes in social media texts, and a multi-layered, fine-grained scheme, psychologically grounded on the work by Fiske, including not only stereotype presence, but also contextuality, implicitness, and forms of discredit, (iii) a multilingual dataset in Italian, Spanish, and French annotated following the aforementioned guidelines, and cross-lingual comparative analyses taking into account racial hoaxes and stereotypes in online discussions. The analysis and results show the usefulness of our methodology and resources, shedding light on how racial hoaxes are spread, and enable the identification of negative stereotypes that reinforce them.

Anthology ID:: 2023.findings-eacl.51
Volume:: Findings of the Association for Computational Linguistics: EACL 2023
Month:: May
Year:: 2023
Address:: Dubrovnik, Croatia
Editors:: Andreas Vlachos, Isabelle Augenstein
Venue:: Findings
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 686–696
Language:
URL:: https://aclanthology.org/2023.findings-eacl.51
DOI:: 10.18653/v1/2023.findings-eacl.51
Bibkey:
Cite (ACL):: Tom Bourgeade, Alessandra Teresa Cignarella, Simona Frenda, Mario Laurent, Wolfgang Schmeisser-Nieto, Farah Benamara, Cristina Bosco, Véronique Moriceau, Viviana Patti, and Mariona Taulé. 2023. A Multilingual Dataset of Racial Stereotypes in Social Media Conversational Threads. In Findings of the Association for Computational Linguistics: EACL 2023, pages 686–696, Dubrovnik, Croatia. Association for Computational Linguistics.
Cite (Informal):: A Multilingual Dataset of Racial Stereotypes in Social Media Conversational Threads (Bourgeade et al., Findings 2023)
Copy Citation:
PDF:: https://preview.aclanthology.org/naacl24-info/2023.findings-eacl.51.pdf
Video:: https://preview.aclanthology.org/naacl24-info/2023.findings-eacl.51.mp4

PDF Search Video