Search papers, labs, and topics across Lattice.
This paper introduces the Mawqif-v2 Extension, a dataset comprising 996 manually annotated Arabic tweets focused on three public targets: Women Driving, E-Cars, and Trimester System, aimed at enhancing cross-target stance detection. By providing a held-out evaluation set, it allows researchers to assess model generalization to both related and novel targets, addressing the scarcity of Arabic datasets in this domain. Baseline results from various Arabic and multilingual transformer models, including zero-shot LLMs, are established to support reproducible evaluations of stance detection performance.
Mawqif-v2 reveals that existing models struggle with cross-target generalization in Arabic stance detection, highlighting a critical gap in current datasets.
Publicly available Arabic datasets for target-specific stance detection remain limited, particularly for evaluating cross-target generalization. This paper presents the Mawqif-v2 Extension, consisting of 996 manually annotated Arabic tweets collected from three public targets: Women Driving, E-Cars, and Trimester System. Each tweet is annotated with stance, sentiment, and sarcasm labels following the original Mawqif annotation scheme. The released extension is intended as a held-out evaluation set for assessing model generalization to both semantically related and previously unseen targets, while the original Mawqif dataset is used for training and development. In addition, we establish baseline results using several Arabic and multilingual transformer models, as well as zero-shot large language models (LLMs), to facilitate reproducible evaluation. Together with the original Mawqif dataset, the Mawqif-v2 Extension provides a benchmark for evaluating cross-target generalization in Arabic stance detection.