Mawqif-v2: An Arabic Benchmark Dataset for Cross-Target Stance Detection

📅 2026-08-10
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
This study addresses the scarcity of publicly available datasets for evaluating cross-target generalization in Arabic stance detection. To bridge this gap, the authors introduce Mawqif-v2, an expanded dataset comprising 996 manually annotated Arabic tweets spanning three distinct topics: women driving, electric vehicles, and the trimester academic system. This work presents the first benchmark specifically designed for cross-target stance detection in Arabic, with each tweet labeled for stance, sentiment, and sarcasm. The dataset enables systematic evaluation of both zero-shot large language models and Arabic/multilingual Transformer-based approaches. By establishing reproducible baseline performance metrics, the study provides a standardized evaluation framework to advance research on cross-target generalization in Arabic stance detection.
📝 Abstract
Publicly available Arabic datasets for target-specific stance detection remain limited, particularly for evaluating cross-target generalization. This paper presents the Mawqif-v2 Extension, consisting of 996 manually annotated Arabic tweets collected from three public targets: Women Driving, E-Cars, and Trimester System. Each tweet is annotated with stance, sentiment, and sarcasm labels following the original Mawqif annotation scheme. The released extension is intended as a held-out evaluation set for assessing model generalization to both semantically related and previously unseen targets, while the original Mawqif dataset is used for training and development. In addition, we establish baseline results using several Arabic and multilingual transformer models, as well as zero-shot large language models (LLMs), to facilitate reproducible evaluation. Together with the original Mawqif dataset, the Mawqif-v2 Extension provides a benchmark for evaluating cross-target generalization in Arabic stance detection.
Problem

Research questions and friction points this paper is trying to address.

Arabic stance detection
cross-target generalization
benchmark dataset
target-specific stance
model generalization
Innovation

Methods, ideas, or system contributions that make the work stand out.

cross-target stance detection
Arabic benchmark dataset
zero-shot LLMs
stance generalization
manual annotation
🔎 Similar Papers
No similar papers found.
R
Rasha Albalawi
Information and Computer Science Department, KFUPM, Saudi Arabia; SDAIA-KFUPM JRC for AI, KFUPM, Saudi Arabia; University of Tabuk, Saudi Arabia
N
Nuha Albadi
Information and Computer Science Department, KFUPM, Saudi Arabia; SDAIA-KFUPM JRC for AI, KFUPM, Saudi Arabia
Hamzah Luqman
Hamzah Luqman
Associate Professor, King Fahd university for Petroleum and Minerals (KFUPM)
Computer VisionArabic Natural Language Processing
M
Maram Kurdi
Information and Computer Science Department, KFUPM, Saudi Arabia
S
Saad Ezzini
Information and Computer Science Department, KFUPM, Saudi Arabia
A
Asma Yamani
Information and Computer Science Department, KFUPM, Saudi Arabia
Ahmed Ashraf
Ahmed Ashraf
Master of Science in Computer Science, KFUPM
Natural Language Processing (NLP)Arabic NLPAI safetyAI Alignment