AraTrust: An Evaluation of Trustworthiness for LLMs in Arabic
Fuente:
arXiv
Saved in:
| Main Authors: | Alghamdi, Emad A., Masoud, Reem I., Alnuhait, Deema, Alomairi, Afnan Y., Ashraf, Ahmed, Zaytoon, Mohamed |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mind the Gap: A Review of Arabic Post-Training Datasets and Their Limitations
by: Alkhowaiter, Mohammed, et al.
Published: (2025)
by: Alkhowaiter, Mohammed, et al.
Published: (2025)
AraReasoner: Evaluating Reasoning-Based LLMs for Arabic NLP
by: Hasanaath, Ahmed, et al.
Published: (2025)
by: Hasanaath, Ahmed, et al.
Published: (2025)
AraSpider: Democratizing Arabic-to-SQL
by: Heakl, Ahmed, et al.
Published: (2024)
by: Heakl, Ahmed, et al.
Published: (2024)
CIDAR: Culturally Relevant Instruction Dataset For Arabic
by: Alyafeai, Zaid, et al.
Published: (2024)
by: Alyafeai, Zaid, et al.
Published: (2024)
FactCheckmate: Preemptively Detecting and Mitigating Hallucinations in LMs
by: Alnuhait, Deema, et al.
Published: (2024)
by: Alnuhait, Deema, et al.
Published: (2024)
AraS2P: Arabic Speech-to-Phonemes System
by: Matar, Bassam, et al.
Published: (2025)
by: Matar, Bassam, et al.
Published: (2025)
AraHalluEval: A Fine-grained Hallucination Evaluation Framework for Arabic LLMs
by: Alansari, Aisha, et al.
Published: (2025)
by: Alansari, Aisha, et al.
Published: (2025)
AraFinNews: Arabic Financial Summarisation with Domain-Adapted LLMs
by: El-Haj, Mo, et al.
Published: (2025)
by: El-Haj, Mo, et al.
Published: (2025)
!MSA at AraHealthQA 2025 Shared Task: Enhancing LLM Performance for Arabic Clinical Question Answering through Prompt Engineering and Ensemble Learning
by: Tarek, Mohamed, et al.
Published: (2025)
by: Tarek, Mohamed, et al.
Published: (2025)
AraTable: Benchmarking LLMs' Reasoning and Understanding of Arabic Tabular Data
by: Alshaikh, Rana, et al.
Published: (2025)
by: Alshaikh, Rana, et al.
Published: (2025)
Ara-HOPE: Human-Centric Post-Editing Evaluation for Dialectal Arabic to Modern Standard Arabic Translation
by: Alabdullah, Abdullah, et al.
Published: (2025)
by: Alabdullah, Abdullah, et al.
Published: (2025)
Ara-Best-RQ: Multi Dialectal Arabic SSL
by: Elleuch, Haroun, et al.
Published: (2026)
by: Elleuch, Haroun, et al.
Published: (2026)
AraSpot: Arabic Spoken Command Spotting
by: Salhab, Mahmoud, et al.
Published: (2023)
by: Salhab, Mahmoud, et al.
Published: (2023)
AraSTEM: A Native Arabic Multiple Choice Question Benchmark for Evaluating LLMs Knowledge In STEM Subjects
by: Mustapha, Ahmad, et al.
Published: (2024)
by: Mustapha, Ahmad, et al.
Published: (2024)
CSLRConformer: A Data-Centric Conformer Approach for Continuous Arabic Sign Language Recognition on the Isharah Datase
by: Elden, Fatimah Mohamed Emad
Published: (2025)
by: Elden, Fatimah Mohamed Emad
Published: (2025)
AraFinNLP 2024: The First Arabic Financial NLP Shared Task
by: Malaysha, Sanad, et al.
Published: (2024)
by: Malaysha, Sanad, et al.
Published: (2024)
Arabic Dataset for LLM Safeguard Evaluation
by: Ashraf, Yasser, et al.
Published: (2024)
by: Ashraf, Yasser, et al.
Published: (2024)
AraSpell: A Deep Learning Approach for Arabic Spelling Correction
by: Salhab, Mahmoud, et al.
Published: (2024)
by: Salhab, Mahmoud, et al.
Published: (2024)
MedAraBench: Large-Scale Arabic Medical Question Answering Dataset and Benchmark
by: Abu-Daoud, Mouath, et al.
Published: (2026)
by: Abu-Daoud, Mouath, et al.
Published: (2026)
AraHealthQA 2025: The First Shared Task on Arabic Health Question Answering
by: Alhuzali, Hassan, et al.
Published: (2025)
by: Alhuzali, Hassan, et al.
Published: (2025)
Second Language (Arabic) Acquisition of LLMs via Progressive Vocabulary Expansion
by: Zhu, Jianqing, et al.
Published: (2024)
by: Zhu, Jianqing, et al.
Published: (2024)
AraPoemBERT: A Pretrained Language Model for Arabic Poetry Analysis
by: Qarah, Faisal
Published: (2024)
by: Qarah, Faisal
Published: (2024)
AraModernBERT: Transtokenized Initialization and Long-Context Encoder Modeling for Arabic
by: Elshehy, Omar, et al.
Published: (2026)
by: Elshehy, Omar, et al.
Published: (2026)
What makes video‐based academic lectures difficult for language learners to comprehend? The role of multimodal complexity
by: Emad A. Alghamdi
Published: (2024)
by: Emad A. Alghamdi
Published: (2024)
Creating Arabic LLM Prompts at Scale
by: El-Sheikh, Abdelrahman, et al.
Published: (2024)
by: El-Sheikh, Abdelrahman, et al.
Published: (2024)
TrustScore: Reference-Free Evaluation of LLM Response Trustworthiness
by: Zheng, Danna, et al.
Published: (2024)
by: Zheng, Danna, et al.
Published: (2024)
Efficient Intent-Based Filtering for Multi-Party Conversations Using Knowledge Distillation from LLMs
by: Gody, Reem, et al.
Published: (2025)
by: Gody, Reem, et al.
Published: (2025)
How Well Do LLMs Understand Tunisian Arabic?
by: Mahdi, Mohamed
Published: (2025)
by: Mahdi, Mohamed
Published: (2025)
AraHopeCorpus: Annotation Guidelines and Dataset for Hope Speech in Arabic Social Media Crisis Discourse
by: Sharqawi, Esra'a, et al.
Published: (2026)
by: Sharqawi, Esra'a, et al.
Published: (2026)
EmoAra: Emotion-Preserving English Speech Transcription and Cross-Lingual Translation with Arabic Text-to-Speech
by: Hassan, Besher, et al.
Published: (2026)
by: Hassan, Besher, et al.
Published: (2026)
Advancing Complex Medical Communication in Arabic with Sporo AraSum: Surpassing Existing Large Language Models
by: Lee, Chanseo, et al.
Published: (2024)
by: Lee, Chanseo, et al.
Published: (2024)
AraLingBench A Human-Annotated Benchmark for Evaluating Arabic Linguistic Capabilities of Large Language Models
by: Zbeeb, Mohammad, et al.
Published: (2025)
by: Zbeeb, Mohammad, et al.
Published: (2025)
Decoding Compressed Trust: Scrutinizing the Trustworthiness of Efficient LLMs Under Compression
by: Hong, Junyuan, et al.
Published: (2024)
by: Hong, Junyuan, et al.
Published: (2024)
ArabicNLU 2024: The First Arabic Natural Language Understanding Shared Task
by: Khalilia, Mohammed, et al.
Published: (2024)
by: Khalilia, Mohammed, et al.
Published: (2024)
!MSA at BAREC Shared Task 2025: Ensembling Arabic Transformers for Readability Assessment
by: Basem, Mohamed, et al.
Published: (2025)
by: Basem, Mohamed, et al.
Published: (2025)
The SAMER Arabic Text Simplification Corpus
by: Alhafni, Bashar, et al.
Published: (2024)
by: Alhafni, Bashar, et al.
Published: (2024)
Cultural Alignment in Large Language Models Using Soft Prompt Tuning
by: Masoud, Reem I., et al.
Published: (2025)
by: Masoud, Reem I., et al.
Published: (2025)
dzFinNlp at AraFinNLP: Improving Intent Detection in Financial Conversational Agents
by: Lichouri, Mohamed, et al.
Published: (2024)
by: Lichouri, Mohamed, et al.
Published: (2024)
AraToken: Optimizing Arabic Tokenization with Normalization Pipeline and Language Extension for Qwen3
by: Kashirskiy, Mark, et al.
Published: (2025)
by: Kashirskiy, Mark, et al.
Published: (2025)
TrustLLM: Trustworthiness in Large Language Models
by: Huang, Yue, et al.
Published: (2024)
by: Huang, Yue, et al.
Published: (2024)
Similar Items
-
Mind the Gap: A Review of Arabic Post-Training Datasets and Their Limitations
by: Alkhowaiter, Mohammed, et al.
Published: (2025) -
AraReasoner: Evaluating Reasoning-Based LLMs for Arabic NLP
by: Hasanaath, Ahmed, et al.
Published: (2025) -
AraSpider: Democratizing Arabic-to-SQL
by: Heakl, Ahmed, et al.
Published: (2024) -
CIDAR: Culturally Relevant Instruction Dataset For Arabic
by: Alyafeai, Zaid, et al.
Published: (2024) -
FactCheckmate: Preemptively Detecting and Mitigating Hallucinations in LMs
by: Alnuhait, Deema, et al.
Published: (2024)