A Multilingual Dataset of Racial Stereotypes in Social Media Conversational Threads
Contributo in Atti di convegno
Data di Pubblicazione:
2023
Abstract:
In this paper, we focus on the topics of misinformation and racial hoaxes from a perspective derived from both social psychology and computational linguistics. In particular, we consider the specific case of anti-immigrant feeling as a first case study for addressing racial stereotypes. We describe the first corpus-based study for multilingual racial stereotype identification in social media conversational threads. Our contributions are: (i) a multilingual corpus of racial hoaxes, (ii) a set of common guidelines for the annotation of racial stereotypes in social media texts, and a multi-layered, fine-grained scheme, psychologically grounded on the work by Fiske, including not only stereotype presence, but also contextuality, implicitness, and forms of discredit, (iii) a multilingual dataset in Italian, Spanish, and French annotated following the aforementioned guidelines, and cross-lingual comparative analyses taking into account racial hoaxes and stereotypes in online discussions. The analysis and results show the usefulness of our methodology and resources, shedding light on how racial hoaxes are spread, and enable the identification of negative stereotypes that reinforce them.
Tipologia CRIS:
04A-Conference paper in volume
Keywords:
racial stereotypes, multilingual corpus, racial hoaxes, social media
Elenco autori:
Bourgeade T.; Cignarella A.T.; Frenda S.; Laurent M.; Schmeisser-Nieto W.S.; Benamara F.; Bosco C.; Moriceau V.; Patti V.; Taule M.
Link alla scheda completa:
Link al Full Text:
Titolo del libro:
Findings of the Association for Computational Linguistics: EACL 2023