CARMA: Comprehensive Automatically-annotated Reddit Mental Health Dataset for Arabic
Fuente:
arXiv
Saved in:
| Main Authors: | Mankarious, Saad, Zirikly, Ayah |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Style Transfer as Bias Mitigation: Diffusion Models for Synthetic Mental Health Text for Arabic
by: Mankarious, Saad, et al.
Published: (2026)
by: Mankarious, Saad, et al.
Published: (2026)
MindSET: Advancing Mental Health Benchmarking through Large-Scale Social Media Data
by: Mankarious, Saad, et al.
Published: (2025)
by: Mankarious, Saad, et al.
Published: (2025)
Mirroring Minds: Asymmetric Linguistic Accommodation and Diagnostic Identity in ADHD and Autism Reddit Communities
by: Mankarious, Saad, et al.
Published: (2026)
by: Mankarious, Saad, et al.
Published: (2026)
A Comprehensive Review of Datasets for Clinical Mental Health AI Systems
by: Mandal, Aishik, et al.
Published: (2025)
by: Mandal, Aishik, et al.
Published: (2025)
RedditESS: A Mental Health Social Support Interaction Dataset -- Understanding Effective Social Support to Refine AI-Driven Support Tools
by: Alghamdi, Zeyad, et al.
Published: (2025)
by: Alghamdi, Zeyad, et al.
Published: (2025)
A Comprehensive Evaluation of Large Language Models on Mental Illnesses in Arabic Context
by: Zahran, Noureldin, et al.
Published: (2025)
by: Zahran, Noureldin, et al.
Published: (2025)
Predicting stock prices with ChatGPT-annotated Reddit sentiment
by: Kmak, Mateusz, et al.
Published: (2025)
by: Kmak, Mateusz, et al.
Published: (2025)
Quantum Frog: Emergent Cooperation and Difficulty Scaling in a Quantized-Time Cooperative Game
by: Mankarious, Saad
Published: (2026)
by: Mankarious, Saad
Published: (2026)
Ar-Spider: Text-to-SQL in Arabic
by: Almohaimeed, Saleh, et al.
Published: (2024)
by: Almohaimeed, Saleh, et al.
Published: (2024)
Arabic Automatic Story Generation with Large Language Models
by: El-Shangiti, Ahmed Oumar, et al.
Published: (2024)
by: El-Shangiti, Ahmed Oumar, et al.
Published: (2024)
Dialect2SQL: A Novel Text-to-SQL Dataset for Arabic Dialects with a Focus on Moroccan Darija
by: Chafik, Salmane, et al.
Published: (2025)
by: Chafik, Salmane, et al.
Published: (2025)
AI Text Detectors and the Misclassification of Slightly Polished Arabic Text
by: Almohaimeed, Saleh, et al.
Published: (2025)
by: Almohaimeed, Saleh, et al.
Published: (2025)
ArabIcros: AI-Powered Arabic Crossword Puzzle Generation for Educational Applications
by: Zeinalipour, Kamyar, et al.
Published: (2023)
by: Zeinalipour, Kamyar, et al.
Published: (2023)
AHaSIS: Shared Task on Sentiment Analysis for Arabic Dialects
by: Alharbi, Maram, et al.
Published: (2025)
by: Alharbi, Maram, et al.
Published: (2025)
A New Benchmark for Evaluating Automatic Speech Recognition in the Arabic Call Domain
by: Obaidah, Qusai Abo, et al.
Published: (2024)
by: Obaidah, Qusai Abo, et al.
Published: (2024)
Promoting the Responsible Development of Speech Datasets for Mental Health and Neurological Disorders Research
by: Mancini, Eleonora, et al.
Published: (2024)
by: Mancini, Eleonora, et al.
Published: (2024)
Jawaher: A Multidialectal Dataset of Arabic Proverbs for LLM Benchmarking
by: Magdy, Samar M., et al.
Published: (2025)
by: Magdy, Samar M., et al.
Published: (2025)
An AI-Based Behavioral Health Safety Filter and Dataset for Identifying Mental Health Crises in Text-Based Conversations
by: Nelson, Benjamin W., et al.
Published: (2025)
by: Nelson, Benjamin W., et al.
Published: (2025)
Are Clinical T5 Models Better for Clinical Text?
by: Li, Yahan, et al.
Published: (2024)
by: Li, Yahan, et al.
Published: (2024)
Palm: A Culturally Inclusive and Linguistically Diverse Dataset for Arabic LLMs
by: Alwajih, Fakhraddin, et al.
Published: (2025)
by: Alwajih, Fakhraddin, et al.
Published: (2025)
TrustMH-Bench: A Comprehensive Benchmark for Evaluating the Trustworthiness of Large Language Models in Mental Health
by: Xiong, Zixin, et al.
Published: (2026)
by: Xiong, Zixin, et al.
Published: (2026)
Estimating the Level of Dialectness Predicts Interannotator Agreement in Multi-dialect Arabic Datasets
by: Keleg, Amr, et al.
Published: (2024)
by: Keleg, Amr, et al.
Published: (2024)
ATHAR: A High-Quality and Diverse Dataset for Classical Arabic to English Translation
by: Khalil, Mohammed, et al.
Published: (2024)
by: Khalil, Mohammed, et al.
Published: (2024)
PoPreRo: A New Dataset for Popularity Prediction of Romanian Reddit Posts
by: Rogoz, Ana-Cristina, et al.
Published: (2024)
by: Rogoz, Ana-Cristina, et al.
Published: (2024)
LLM_annotate: A Python package for annotating and analyzing fiction characters
by: Rosenbusch, Hannes
Published: (2025)
by: Rosenbusch, Hannes
Published: (2025)
Arabic Little STT: Arabic Children Speech Recognition Dataset
by: Alkadri, Mouhand, et al.
Published: (2025)
by: Alkadri, Mouhand, et al.
Published: (2025)
The Arabic Generality Score: Another Dimension of Modeling Arabic Dialectness
by: Shaban, Sanad, et al.
Published: (2025)
by: Shaban, Sanad, et al.
Published: (2025)
Opioid Named Entity Recognition (ONER-2025) from Reddit
by: Ahmad, Muhammad, et al.
Published: (2025)
by: Ahmad, Muhammad, et al.
Published: (2025)
Tibyan Corpus: Balanced and Comprehensive Error Coverage Corpus Using ChatGPT for Arabic Grammatical Error Correction
by: Alrehili, Ahlam, et al.
Published: (2024)
by: Alrehili, Ahlam, et al.
Published: (2024)
DialectalArabicMMLU: Benchmarking Dialectal Capabilities in Arabic and Multilingual Language Models
by: Altakrori, Malik H., et al.
Published: (2025)
by: Altakrori, Malik H., et al.
Published: (2025)
ArabicNumBench: Evaluating Arabic Number Reading in Large Language Models
by: Alhumud, Anas, et al.
Published: (2026)
by: Alhumud, Anas, et al.
Published: (2026)
Zero-shot Explainable Mental Health Analysis on Social Media by Incorporating Mental Scales
by: Li, Wenyu, et al.
Published: (2024)
by: Li, Wenyu, et al.
Published: (2024)
Automatic Generation of Inference Making Questions for Reading Comprehension Assessments
by: Ma, Wanjing Anya, et al.
Published: (2025)
by: Ma, Wanjing Anya, et al.
Published: (2025)
multiMentalRoBERTa: A Fine-tuned Multiclass Classifier for Mental Health Disorder
by: Islam, K M Sajjadul, et al.
Published: (2025)
by: Islam, K M Sajjadul, et al.
Published: (2025)
Harf-Speech: A Clinically Aligned Framework for Arabic Phoneme-Level Speech Assessment
by: Azad, Asif, et al.
Published: (2026)
by: Azad, Asif, et al.
Published: (2026)
Mind the Gap: A Review of Arabic Post-Training Datasets and Their Limitations
by: Alkhowaiter, Mohammed, et al.
Published: (2025)
by: Alkhowaiter, Mohammed, et al.
Published: (2025)
MentalChat16K: A Benchmark Dataset for Conversational Mental Health Assistance
by: Xu, Jia, et al.
Published: (2025)
by: Xu, Jia, et al.
Published: (2025)
ALBA: Adaptive Language-based Assessments for Mental Health
by: Varadarajan, Vasudha, et al.
Published: (2023)
by: Varadarajan, Vasudha, et al.
Published: (2023)
Automatic Dataset Generation for Knowledge Intensive Question Answering Tasks
by: Yuen, Sizhe, et al.
Published: (2025)
by: Yuen, Sizhe, et al.
Published: (2025)
Killkan: The Automatic Speech Recognition Dataset for Kichwa with Morphosyntactic Information
by: Taguchi, Chihiro, et al.
Published: (2024)
by: Taguchi, Chihiro, et al.
Published: (2024)
Similar Items
-
Style Transfer as Bias Mitigation: Diffusion Models for Synthetic Mental Health Text for Arabic
by: Mankarious, Saad, et al.
Published: (2026) -
MindSET: Advancing Mental Health Benchmarking through Large-Scale Social Media Data
by: Mankarious, Saad, et al.
Published: (2025) -
Mirroring Minds: Asymmetric Linguistic Accommodation and Diagnostic Identity in ADHD and Autism Reddit Communities
by: Mankarious, Saad, et al.
Published: (2026) -
A Comprehensive Review of Datasets for Clinical Mental Health AI Systems
by: Mandal, Aishik, et al.
Published: (2025) -
RedditESS: A Mental Health Social Support Interaction Dataset -- Understanding Effective Social Support to Refine AI-Driven Support Tools
by: Alghamdi, Zeyad, et al.
Published: (2025)