UPR at SemEval-2026 Task 9: Multi-Label Classification of Polarization Across Social Dimensions and Manifestation Identification in Urdu

Mtayyaba Shahzad, Inzmam Khadam, Zaufishan Mahmood, Junaid Rashid, Shamaila Hayat, Fakhar Ayub


Abstract
The analysis of polarized content on social networks is crucial for understanding public discourse; however, research on low-resource languages such as Urdu remains limited. In this work, we address two complementary subtasks of polarization analysis in Urdu social media text. First, we formulate polarization classification across multiple social dimensions as a multi-label task, including political, religious, racial/ethnic, gender/sexual, and other. We fine-tune XLM-RoBERTa for multi-label classification with language-specific preprocessing, duplicate filtering, and data augmentation to handle class imbalance. The proposed model achieves a Macro F1-score of 0.758 for social-dimension polarization classification.Second, we perform polarization manifestation identification, focusing on how polarization is expressed in text through six manifestations: stereotype, vilification, dehumanization, extreme language, lack of empathy, and invalidation. Using the same transformer-based framework with imbalance-aware training, our system achieves a Macro F1-score of 0.72 on the official test set. These results demonstrate the effectiveness of multilingual transformer models for multi-dimensional polarization analysis in low-resource Urdu text.
Anthology ID:
2026.semeval-1.333
Volume:
Proceedings of the 20th International Workshop on Semantic Evaluation (2026)
Month:
July
Year:
2026
Address:
San Diego, California, USA
Editors:
Ekaterina Kochmar, Debanjan Ghosh, Kai North, Mamoru Komachi
Venues:
SemEval | WS
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
2642–2647
Language:
URL:
https://preview.aclanthology.org/ingest-acl-workshops/2026.semeval-1.333/
DOI:
Bibkey:
Cite (ACL):
Mtayyaba Shahzad, Inzmam Khadam, Zaufishan Mahmood, Junaid Rashid, Shamaila Hayat, and Fakhar Ayub. 2026. UPR at SemEval-2026 Task 9: Multi-Label Classification of Polarization Across Social Dimensions and Manifestation Identification in Urdu. In Proceedings of the 20th International Workshop on Semantic Evaluation (2026), pages 2642–2647, San Diego, California, USA. Association for Computational Linguistics.
Cite (Informal):
UPR at SemEval-2026 Task 9: Multi-Label Classification of Polarization Across Social Dimensions and Manifestation Identification in Urdu (Shahzad et al., SemEval 2026)
Copy Citation:
PDF:
https://preview.aclanthology.org/ingest-acl-workshops/2026.semeval-1.333.pdf