Cohesion-6K: An Arabic Dataset for Analyzing Social Cohesion and Conflict in Online Discourse
Fuente:
arXiv
Saved in:
| Main Authors: | Al-Athba, Aisha Ali, Zaghouani, Wajdi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AraHopeCorpus: Annotation Guidelines and Dataset for Hope Speech in Arabic Social Media Crisis Discourse
by: Sharqawi, Esra'a, et al.
Published: (2026)
by: Sharqawi, Esra'a, et al.
Published: (2026)
Cultural Adaptation in Large Language Models for Political Discourse
by: Zaghouani, Wajdi
Published: (2026)
by: Zaghouani, Wajdi
Published: (2026)
Building Arabic NLP from the Ground Up: Twenty Years of Lessons, Failures, and Open Problems
by: Zaghouani, Wajdi
Published: (2026)
by: Zaghouani, Wajdi
Published: (2026)
EmoHopeSpeech: An Annotated Dataset of Emotions and Hope Speech in English and Arabic
by: Zaghouani, Wajdi, et al.
Published: (2025)
by: Zaghouani, Wajdi, et al.
Published: (2025)
Toward Responsible and Epistemically Grounded Multilingual LLMs for Computational Social Science and Humanities
by: Zaghouani, Wajdi
Published: (2026)
by: Zaghouani, Wajdi
Published: (2026)
An Annotated Corpus of Arabic Tweets for Hate Speech Analysis
by: Zaghouani, Wajdi, et al.
Published: (2025)
by: Zaghouani, Wajdi, et al.
Published: (2025)
ArPoMeme: An Annotated Arabic Multimodal Dataset for Political Ideology and Polarization
by: Zaghouani, Wajdi, et al.
Published: (2026)
by: Zaghouani, Wajdi, et al.
Published: (2026)
JobArabi: An Arabic Corpus and Analysis of Job Announcements from Social Media
by: Zaghouani, Wajdi, et al.
Published: (2026)
by: Zaghouani, Wajdi, et al.
Published: (2026)
Audience Engagement with Arabic Women's Social Empowerment and Wellbeing: A Decadal Corpus
by: Zaghouani, Wajdi, et al.
Published: (2026)
by: Zaghouani, Wajdi, et al.
Published: (2026)
MemeMind at ArAIEval Shared Task: Spotting Persuasive Spans in Arabic Text with Persuasion Techniques Identification
by: Biswas, Md Rafiul, et al.
Published: (2024)
by: Biswas, Md Rafiul, et al.
Published: (2024)
ArabDiscrim: A Decade-Long Arabic Facebook Corpus on Racism and Discrimination
by: Zaghouani, Wajdi, et al.
Published: (2026)
by: Zaghouani, Wajdi, et al.
Published: (2026)
AlbanianLLMSafety: A Safety Evaluation Dataset for Large Language Models in Albanian
by: Zaghouani, Wajdi, et al.
Published: (2026)
by: Zaghouani, Wajdi, et al.
Published: (2026)
MARSAD: A Multi-Functional Tool for Real-Time Social Media Analysis
by: Biswas, Md. Rafiul, et al.
Published: (2025)
by: Biswas, Md. Rafiul, et al.
Published: (2025)
Evaluating Discourse Cohesion in Pre-trained Language Models
by: He, Jie, et al.
Published: (2025)
by: He, Jie, et al.
Published: (2025)
Transformers and Ensemble methods: A solution for Hate Speech Detection in Arabic languages
by: de Paula, Angel Felipe Magnossão, et al.
Published: (2023)
by: de Paula, Angel Felipe Magnossão, et al.
Published: (2023)
Chinese Offensive Language Detection:Current Status and Future Directions
by: Xiao, Yunze, et al.
Published: (2024)
by: Xiao, Yunze, et al.
Published: (2024)
ThatiAR: Subjectivity Detection in Arabic News Sentences
by: Suwaileh, Reem, et al.
Published: (2024)
by: Suwaileh, Reem, et al.
Published: (2024)
Beyond English and Evasion: A Human-Annotated Multi-Domain Benchmark for High-Stakes LLM Safety Evaluation in Chinese
by: Zaghouani, Wajdi, et al.
Published: (2026)
by: Zaghouani, Wajdi, et al.
Published: (2026)
KZ-SafetyPrompts: A Kazakh Safety Evaluation Prompt Dataset for Large Language Models
by: Zaghouani, Wajdi, et al.
Published: (2026)
by: Zaghouani, Wajdi, et al.
Published: (2026)
Nullpointer at CheckThat! 2024: Identifying Subjectivity from Multilingual Text Sequence
by: Biswas, Md. Rafiul, et al.
Published: (2024)
by: Biswas, Md. Rafiul, et al.
Published: (2024)
Propaganda to Hate: A Multimodal Analysis of Arabic Memes with Multi-Agent LLMs
by: Alam, Firoj, et al.
Published: (2024)
by: Alam, Firoj, et al.
Published: (2024)
Parametricity via Cohesion
by: Aberlé, C. B.
Published: (2024)
by: Aberlé, C. B.
Published: (2024)
ArAIEval Shared Task: Propagandistic Techniques Detection in Unimodal and Multimodal Arabic Content
by: Hasanain, Maram, et al.
Published: (2024)
by: Hasanain, Maram, et al.
Published: (2024)
Semantically Cohesive Word Grouping in Indian Languages
by: Karthika, N J, et al.
Published: (2025)
by: Karthika, N J, et al.
Published: (2025)
Rater Cohesion and Quality from a Vicarious Perspective
by: Pandita, Deepak, et al.
Published: (2024)
by: Pandita, Deepak, et al.
Published: (2024)
Cohesive Conversations: Enhancing Authenticity in Multi-Agent Simulated Dialogues
by: Chu, KuanChao, et al.
Published: (2024)
by: Chu, KuanChao, et al.
Published: (2024)
SEE: Strategic Exploration and Exploitation for Cohesive In-Context Prompt Optimization
by: Cui, Wendi, et al.
Published: (2024)
by: Cui, Wendi, et al.
Published: (2024)
Multi-task Learning with Active Learning for Arabic Offensive Speech Detection
by: Alansari, Aisha, et al.
Published: (2025)
by: Alansari, Aisha, et al.
Published: (2025)
AraHalluEval: A Fine-grained Hallucination Evaluation Framework for Arabic LLMs
by: Alansari, Aisha, et al.
Published: (2025)
by: Alansari, Aisha, et al.
Published: (2025)
Zero-Shot Detection of LLM-Generated Text using Token Cohesiveness
by: Ma, Shixuan, et al.
Published: (2024)
by: Ma, Shixuan, et al.
Published: (2024)
STORYTELLER: An Enhanced Plot-Planning Framework for Coherent and Cohesive Story Generation
by: Li, Jiaming, et al.
Published: (2025)
by: Li, Jiaming, et al.
Published: (2025)
GLARE: Google Apps Arabic Reviews Dataset
by: AlGhamdi, Fatima, et al.
Published: (2024)
by: AlGhamdi, Fatima, et al.
Published: (2024)
`Keep it Together': Enforcing Cohesion in Extractive Summaries by Simulating Human Memory
by: Cardenas, Ronald, et al.
Published: (2024)
by: Cardenas, Ronald, et al.
Published: (2024)
Noor-Ghateh: A Benchmark Dataset for Evaluating Arabic Word Segmenters in Hadith Domain
by: AlShuhayeb, Huda, et al.
Published: (2023)
by: AlShuhayeb, Huda, et al.
Published: (2023)
CLAPNQ: Cohesive Long-form Answers from Passages in Natural Questions for RAG systems
by: Rosenthal, Sara, et al.
Published: (2024)
by: Rosenthal, Sara, et al.
Published: (2024)
Arabic Little STT: Arabic Children Speech Recognition Dataset
by: Alkadri, Mouhand, et al.
Published: (2025)
by: Alkadri, Mouhand, et al.
Published: (2025)
Understanding and Analyzing Inappropriately Targeting Language in Online Discourse: A Comparative Annotation Study
by: Barbarestani, Baran, et al.
Published: (2025)
by: Barbarestani, Baran, et al.
Published: (2025)
Enhancing Essay Cohesion Assessment: A Novel Item Response Theory Approach
by: Rosa, Bruno Alexandre, et al.
Published: (2025)
by: Rosa, Bruno Alexandre, et al.
Published: (2025)
Pearl: A Multimodal Culturally-Aware Arabic Instruction Dataset
by: Alwajih, Fakhraddin, et al.
Published: (2025)
by: Alwajih, Fakhraddin, et al.
Published: (2025)
MultiProSE: A Multi-label Arabic Dataset for Propaganda, Sentiment, and Emotion Detection
by: Al-Henaki, Lubna, et al.
Published: (2025)
by: Al-Henaki, Lubna, et al.
Published: (2025)
Similar Items
-
AraHopeCorpus: Annotation Guidelines and Dataset for Hope Speech in Arabic Social Media Crisis Discourse
by: Sharqawi, Esra'a, et al.
Published: (2026) -
Cultural Adaptation in Large Language Models for Political Discourse
by: Zaghouani, Wajdi
Published: (2026) -
Building Arabic NLP from the Ground Up: Twenty Years of Lessons, Failures, and Open Problems
by: Zaghouani, Wajdi
Published: (2026) -
EmoHopeSpeech: An Annotated Dataset of Emotions and Hope Speech in English and Arabic
by: Zaghouani, Wajdi, et al.
Published: (2025) -
Toward Responsible and Epistemically Grounded Multilingual LLMs for Computational Social Science and Humanities
by: Zaghouani, Wajdi
Published: (2026)