Noor-Ghateh: A Benchmark Dataset for Evaluating Arabic Word Segmenters in Hadith Domain
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | AlShuhayeb, Huda, Minaei-Bidgoli, Behrouz, Shenassa, Mohammad E., Hossayni, Sayyed-Ali |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rezwan: Leveraging Large Language Models for Comprehensive Hadith Text Processing: A 1.2M Corpus Development
von: Asgari-Bidhendi, Majid, et al.
Veröffentlicht: (2025)
von: Asgari-Bidhendi, Majid, et al.
Veröffentlicht: (2025)
PESTS: Persian_English Cross Lingual Corpus for Semantic Textual Similarity
von: Abdous, Mohammad, et al.
Veröffentlicht: (2023)
von: Abdous, Mohammad, et al.
Veröffentlicht: (2023)
A New Method for Cross-Lingual-based Semantic Role Labeling
von: Ebrahimi, Mohammad, et al.
Veröffentlicht: (2024)
von: Ebrahimi, Mohammad, et al.
Veröffentlicht: (2024)
Bactrainus: Optimizing Large Language Models for Multi-hop Complex Question Answering Tasks
von: Barati, Iman, et al.
Veröffentlicht: (2025)
von: Barati, Iman, et al.
Veröffentlicht: (2025)
DREaM: Drug-Drug Relation Extraction via Transfer Learning Method
von: Fata, Ali, et al.
Veröffentlicht: (2025)
von: Fata, Ali, et al.
Veröffentlicht: (2025)
An $O(n^3)$ time algorithm for the maximum-weight limited-capacity many-to-many matching
von: Rajabi-Alni, Fatemeh, et al.
Veröffentlicht: (2014)
von: Rajabi-Alni, Fatemeh, et al.
Veröffentlicht: (2014)
Efficient Many-To-Many Matching of Points with Demands in One Dimension
von: Rajabi-Alni, Fatemeh, et al.
Veröffentlicht: (2019)
von: Rajabi-Alni, Fatemeh, et al.
Veröffentlicht: (2019)
FARSIQA: Faithful and Advanced RAG System for Islamic Question Answering
von: Asl, Mohammad Aghajani, et al.
Veröffentlicht: (2025)
von: Asl, Mohammad Aghajani, et al.
Veröffentlicht: (2025)
Approximation Algorithms for the Freeze Tag Problem inside Polygons
von: Rajabi-Alni, Fatemeh, et al.
Veröffentlicht: (2024)
von: Rajabi-Alni, Fatemeh, et al.
Veröffentlicht: (2024)
FAIR-RAG: Faithful Adaptive Iterative Refinement for Retrieval-Augmented Generation
von: Asl, Mohammad Aghajani, et al.
Veröffentlicht: (2025)
von: Asl, Mohammad Aghajani, et al.
Veröffentlicht: (2025)
Grounding Arabic LLMs in the Doha Historical Dictionary: Retrieval-Augmented Understanding of Quran and Hadith
von: Eltanbouly, Somaya, et al.
Veröffentlicht: (2026)
von: Eltanbouly, Somaya, et al.
Veröffentlicht: (2026)
101 Billion Arabic Words Dataset
von: Aloui, Manel, et al.
Veröffentlicht: (2024)
von: Aloui, Manel, et al.
Veröffentlicht: (2024)
ArEEG_Words: Dataset for Envisioned Speech Recognition using EEG for Arabic Words
von: Darwish, Hazem, et al.
Veröffentlicht: (2024)
von: Darwish, Hazem, et al.
Veröffentlicht: (2024)
A New Benchmark for Evaluating Automatic Speech Recognition in the Arabic Call Domain
von: Obaidah, Qusai Abo, et al.
Veröffentlicht: (2024)
von: Obaidah, Qusai Abo, et al.
Veröffentlicht: (2024)
Cohesion-6K: An Arabic Dataset for Analyzing Social Cohesion and Conflict in Online Discourse
von: Al-Athba, Aisha Ali, et al.
Veröffentlicht: (2026)
von: Al-Athba, Aisha Ali, et al.
Veröffentlicht: (2026)
ASCAT: An Arabic Scientific Corpus and Benchmark for Advanced Translation Evaluation
von: Sibaee, Serry, et al.
Veröffentlicht: (2026)
von: Sibaee, Serry, et al.
Veröffentlicht: (2026)
Arabic Dataset for LLM Safeguard Evaluation
von: Ashraf, Yasser, et al.
Veröffentlicht: (2024)
von: Ashraf, Yasser, et al.
Veröffentlicht: (2024)
MedAraBench: Large-Scale Arabic Medical Question Answering Dataset and Benchmark
von: Abu-Daoud, Mouath, et al.
Veröffentlicht: (2026)
von: Abu-Daoud, Mouath, et al.
Veröffentlicht: (2026)
ADAB: Arabic Dataset for Automated Politeness Benchmarking -- A Large-Scale Resource for Computational Sociopragmatics
von: Al-Khalifa, Hend, et al.
Veröffentlicht: (2026)
von: Al-Khalifa, Hend, et al.
Veröffentlicht: (2026)
GLARE: Google Apps Arabic Reviews Dataset
von: AlGhamdi, Fatima, et al.
Veröffentlicht: (2024)
von: AlGhamdi, Fatima, et al.
Veröffentlicht: (2024)
Are Arabic Benchmarks Reliable? QIMMA's Quality-First Approach to LLM Evaluation
von: AlQadi, Leen, et al.
Veröffentlicht: (2026)
von: AlQadi, Leen, et al.
Veröffentlicht: (2026)
A Strategy for Implementing description Temporal Dynamic Algorithms in Dynamic Knowledge Graphs by SPIN
von: Shahbazi, Alireza, et al.
Veröffentlicht: (2024)
von: Shahbazi, Alireza, et al.
Veröffentlicht: (2024)
Proper Noun Diacritization for Arabic Wikipedia: A Benchmark Dataset
von: Bondok, Rawan, et al.
Veröffentlicht: (2025)
von: Bondok, Rawan, et al.
Veröffentlicht: (2025)
ALHD: A Large-Scale and Multigenre Benchmark Dataset for Arabic LLM-Generated Text Detection
von: Khairallah, Ali, et al.
Veröffentlicht: (2025)
von: Khairallah, Ali, et al.
Veröffentlicht: (2025)
Arabic Little STT: Arabic Children Speech Recognition Dataset
von: Alkadri, Mouhand, et al.
Veröffentlicht: (2025)
von: Alkadri, Mouhand, et al.
Veröffentlicht: (2025)
The Qiyas Benchmark: Measuring ChatGPT Mathematical and Language Understanding in Arabic
von: Al-Khalifa, Shahad, et al.
Veröffentlicht: (2024)
von: Al-Khalifa, Shahad, et al.
Veröffentlicht: (2024)
Evaluating Arabic Large Language Models: A Survey of Benchmarks, Methods, and Gaps
von: Alzubaidi, Ahmed, et al.
Veröffentlicht: (2025)
von: Alzubaidi, Ahmed, et al.
Veröffentlicht: (2025)
ALARB: An Arabic Legal Argument Reasoning Benchmark
von: Shairah, Harethah Abu, et al.
Veröffentlicht: (2025)
von: Shairah, Harethah Abu, et al.
Veröffentlicht: (2025)
CIDAR: Culturally Relevant Instruction Dataset For Arabic
von: Alyafeai, Zaid, et al.
Veröffentlicht: (2024)
von: Alyafeai, Zaid, et al.
Veröffentlicht: (2024)
MURAD: A Large-Scale Multi-Domain Unified Reverse Arabic Dictionary Dataset
von: Sibaee, Serry, et al.
Veröffentlicht: (2026)
von: Sibaee, Serry, et al.
Veröffentlicht: (2026)
Jawaher: A Multidialectal Dataset of Arabic Proverbs for LLM Benchmarking
von: Magdy, Samar M., et al.
Veröffentlicht: (2025)
von: Magdy, Samar M., et al.
Veröffentlicht: (2025)
MultiProSE: A Multi-label Arabic Dataset for Propaganda, Sentiment, and Emotion Detection
von: Al-Henaki, Lubna, et al.
Veröffentlicht: (2025)
von: Al-Henaki, Lubna, et al.
Veröffentlicht: (2025)
Alexandria: A Multi-Domain Dialectal Arabic Machine Translation Dataset for Culturally Inclusive and Linguistically Diverse LLMs
von: Mekki, Abdellah El, et al.
Veröffentlicht: (2026)
von: Mekki, Abdellah El, et al.
Veröffentlicht: (2026)
A novel feature selection method in the categorization of imbalanced textual data
von: Pouramini, Jafar, et al.
Veröffentlicht: (2018)
von: Pouramini, Jafar, et al.
Veröffentlicht: (2018)
AraSTEM: A Native Arabic Multiple Choice Question Benchmark for Evaluating LLMs Knowledge In STEM Subjects
von: Mustapha, Ahmad, et al.
Veröffentlicht: (2024)
von: Mustapha, Ahmad, et al.
Veröffentlicht: (2024)
Advancing the Arabic WordNet: Elevating Content Quality
von: Freihat, Abed Alhakim, et al.
Veröffentlicht: (2024)
von: Freihat, Abed Alhakim, et al.
Veröffentlicht: (2024)
AraLingBench A Human-Annotated Benchmark for Evaluating Arabic Linguistic Capabilities of Large Language Models
von: Zbeeb, Mohammad, et al.
Veröffentlicht: (2025)
von: Zbeeb, Mohammad, et al.
Veröffentlicht: (2025)
Abjad-Kids: An Arabic Speech Classification Dataset for Primary Education
von: Snoubara, Abdul Aziz, et al.
Veröffentlicht: (2026)
von: Snoubara, Abdul Aziz, et al.
Veröffentlicht: (2026)
Cultural Benchmarking of LLMs in Standard and Dialectal Arabic Dialogues
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2026)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2026)
Evaluation of Semantic Search and its Role in Retrieved-Augmented-Generation (RAG) for Arabic Language
von: Mahboub, Ali, et al.
Veröffentlicht: (2024)
von: Mahboub, Ali, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Rezwan: Leveraging Large Language Models for Comprehensive Hadith Text Processing: A 1.2M Corpus Development
von: Asgari-Bidhendi, Majid, et al.
Veröffentlicht: (2025) -
PESTS: Persian_English Cross Lingual Corpus for Semantic Textual Similarity
von: Abdous, Mohammad, et al.
Veröffentlicht: (2023) -
A New Method for Cross-Lingual-based Semantic Role Labeling
von: Ebrahimi, Mohammad, et al.
Veröffentlicht: (2024) -
Bactrainus: Optimizing Large Language Models for Multi-hop Complex Question Answering Tasks
von: Barati, Iman, et al.
Veröffentlicht: (2025) -
DREaM: Drug-Drug Relation Extraction via Transfer Learning Method
von: Fata, Ali, et al.
Veröffentlicht: (2025)