|国家预印本平台
首页|POLAR: A Benchmark for Multilingual, Multicultural, and Multi-Event Online Polarization

POLAR: A Benchmark for Multilingual, Multicultural, and Multi-Event Online Polarization

POLAR: A Benchmark for Multilingual, Multicultural, and Multi-Event Online Polarization

来源:Arxiv_logoArxiv
英文摘要

Online polarization poses a growing challenge for democratic discourse, yet most computational social science research remains monolingual, culturally narrow, or event-specific. We introduce POLAR, a multilingual, multicultural, and multievent dataset with over 23k instances in seven languages from diverse online platforms and real-world events. Polarization is annotated along three axes: presence, type, and manifestation, using a variety of annotation platforms adapted to each cultural context. We conduct two main experiments: (1) we fine-tune six multilingual pretrained language models in both monolingual and cross-lingual setups; and (2) we evaluate a range of open and closed large language models (LLMs) in few-shot and zero-shot scenarios. Results show that while most models perform well on binary polarization detection, they achieve substantially lower scores when predicting polarization types and manifestations. These findings highlight the complex, highly contextual nature of polarization and the need for robust, adaptable approaches in NLP and computational social science. All resources will be released to support further research and effective mitigation of digital polarization globally.

Usman Naseem、Juan Ren、Saba Anwar、Sarah Kohail、Rudy Alexandro Garrido Veliz、Robert Geislinger、Aisha Jabr、Idris Abdulmumin、Laiba Qureshi、Aarushi Ajay Borkar、Maryam Ibrahim Mukhtar、Abinew Ali Ayele、Ibrahim Said Ahmad、Adem Ali、Martin Semmann、Shamsuddeen Hassan Muhammad、Seid Muhie Yimam

语言学计算技术、计算机技术

Usman Naseem,Juan Ren,Saba Anwar,Sarah Kohail,Rudy Alexandro Garrido Veliz,Robert Geislinger,Aisha Jabr,Idris Abdulmumin,Laiba Qureshi,Aarushi Ajay Borkar,Maryam Ibrahim Mukhtar,Abinew Ali Ayele,Ibrahim Said Ahmad,Adem Ali,Martin Semmann,Shamsuddeen Hassan Muhammad,Seid Muhie Yimam.POLAR: A Benchmark for Multilingual, Multicultural, and Multi-Event Online Polarization[EB/OL].(2025-05-26)[2025-07-16].https://arxiv.org/abs/2505.20624.点此复制

评论