MedArabiQ: Benchmarking Large Language Models on Arabic Medical Tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Daoud, Mouath Abu, Abouzahir, Chaimae, Kharouf, Leen, Al-Eisawi, Walid, Habash, Nizar, Shamout, Farah E. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AraHealthQA 2025: The First Shared Task on Arabic Health Question Answering
von: Alhuzali, Hassan, et al.
Veröffentlicht: (2025)
von: Alhuzali, Hassan, et al.
Veröffentlicht: (2025)
Cross-Lingual Empirical Evaluation of Large Language Models for Arabic Medical Tasks
von: Abouzahir, Chaimae, et al.
Veröffentlicht: (2026)
von: Abouzahir, Chaimae, et al.
Veröffentlicht: (2026)
MedAraBench: Large-Scale Arabic Medical Question Answering Dataset and Benchmark
von: Abu-Daoud, Mouath, et al.
Veröffentlicht: (2026)
von: Abu-Daoud, Mouath, et al.
Veröffentlicht: (2026)
Desk2Desk: Optimization-based Mixed Reality Workspace Integration for Remote Side-by-side Collaboration
von: Sidenmark, Ludwig, et al.
Veröffentlicht: (2024)
von: Sidenmark, Ludwig, et al.
Veröffentlicht: (2024)
Measuring Large Language Models Dependency: Validating the Arabic Version of the LLM-D12 Scale
von: AlShakhsi, Sameha, et al.
Veröffentlicht: (2025)
von: AlShakhsi, Sameha, et al.
Veröffentlicht: (2025)
K-QA: A Real-World Medical Q&A Benchmark
von: Manes, Itay, et al.
Veröffentlicht: (2024)
von: Manes, Itay, et al.
Veröffentlicht: (2024)
Making It Work Is the Work: Engineering Maturity as Epistemic Work
von: Leen, Danny, et al.
Veröffentlicht: (2026)
von: Leen, Danny, et al.
Veröffentlicht: (2026)
Design of a visual environment for programming by direct data manipulation
von: Adam, Michel, et al.
Veröffentlicht: (2025)
von: Adam, Michel, et al.
Veröffentlicht: (2025)
Arabic Little STT: Arabic Children Speech Recognition Dataset
von: Alkadri, Mouhand, et al.
Veröffentlicht: (2025)
von: Alkadri, Mouhand, et al.
Veröffentlicht: (2025)
Developing and Validating the Arabic Version of the Attitudes Toward Large Language Models Scale
von: Barajeeh, Basad, et al.
Veröffentlicht: (2025)
von: Barajeeh, Basad, et al.
Veröffentlicht: (2025)
MedBike: A Cardiac Patient Monitoring System Enhanced through Gamification
von: Hossain, Tahmim, et al.
Veröffentlicht: (2024)
von: Hossain, Tahmim, et al.
Veröffentlicht: (2024)
Combining Automation and Expertise: A Semi-automated Approach to Correcting Eye Tracking Data in Reading Tasks
von: Madi, Naser Al, et al.
Veröffentlicht: (2025)
von: Madi, Naser Al, et al.
Veröffentlicht: (2025)
SyriSign: A Parallel Corpus for Arabic Text to Syrian Arabic Sign Language Translation
von: Khalil, Mohammad Amer, et al.
Veröffentlicht: (2026)
von: Khalil, Mohammad Amer, et al.
Veröffentlicht: (2026)
MedFoundationHub: A Lightweight and Secure Toolkit for Deploying Medical Vision Language Foundation Models
von: Li, Xiao, et al.
Veröffentlicht: (2025)
von: Li, Xiao, et al.
Veröffentlicht: (2025)
ArEEG_Chars: Dataset for Envisioned Speech Recognition using EEG for Arabic Characters
von: Darwish, Hazem, et al.
Veröffentlicht: (2024)
von: Darwish, Hazem, et al.
Veröffentlicht: (2024)
ArEEG_Words: Dataset for Envisioned Speech Recognition using EEG for Arabic Words
von: Darwish, Hazem, et al.
Veröffentlicht: (2024)
von: Darwish, Hazem, et al.
Veröffentlicht: (2024)
Adjusting B‐Tree for Better Usability: Nodes as Files Instead of Disk Blocks
von: Majed AbuSafiya
Veröffentlicht: (2026)
von: Majed AbuSafiya
Veröffentlicht: (2026)
MedDialogRubrics: A Comprehensive Benchmark and Evaluation Framework for Multi-turn Medical Consultations in Large Language Models
von: Gong, Lecheng, et al.
Veröffentlicht: (2026)
von: Gong, Lecheng, et al.
Veröffentlicht: (2026)
Experimental Interface for Multimodal and Large Language Model Based Explanations of Educational Recommender Systems
von: Abu-Rasheed, Hasan, et al.
Veröffentlicht: (2024)
von: Abu-Rasheed, Hasan, et al.
Veröffentlicht: (2024)
Bias Beneath the Tone: Empirical Characterisation of Tone Bias in LLM-Driven UX Systems
von: Bodara, Heet, et al.
Veröffentlicht: (2025)
von: Bodara, Heet, et al.
Veröffentlicht: (2025)
Abjad-Kids: An Arabic Speech Classification Dataset for Primary Education
von: Snoubara, Abdul Aziz, et al.
Veröffentlicht: (2026)
von: Snoubara, Abdul Aziz, et al.
Veröffentlicht: (2026)
Investigating the Task Load of Investigating the Task Load in Visualization Studies
von: Pahr, Daniel, et al.
Veröffentlicht: (2025)
von: Pahr, Daniel, et al.
Veröffentlicht: (2025)
Lemmatization as a Classification Task: Results from Arabic across Multiple Genres
von: Saeed, Mostafa, et al.
Veröffentlicht: (2025)
von: Saeed, Mostafa, et al.
Veröffentlicht: (2025)
Prompt2Task: Automating UI Tasks on Smartphones from Textual Prompts
von: Huang, Tian, et al.
Veröffentlicht: (2024)
von: Huang, Tian, et al.
Veröffentlicht: (2024)
Task Mode: Dynamic Filtering for Task-Specific Web Navigation using LLMs
von: Mohanbabu, Ananya Gubbi, et al.
Veröffentlicht: (2025)
von: Mohanbabu, Ananya Gubbi, et al.
Veröffentlicht: (2025)
Does GenAI Make Usability Testing Obsolete?
von: Pourasad, Ali Ebrahimi, et al.
Veröffentlicht: (2024)
von: Pourasad, Ali Ebrahimi, et al.
Veröffentlicht: (2024)
RFM-HRI : A Multimodal Dataset of Medical Robot Failure, User Reaction and Recovery Preferences for Item Retrieval Tasks
von: Batra, Yashika, et al.
Veröffentlicht: (2026)
von: Batra, Yashika, et al.
Veröffentlicht: (2026)
TaskLens: Generating Task-Conditioned Scaffolded Interfaces for Learning Professional Creative Software
von: Liu, Yimeng, et al.
Veröffentlicht: (2025)
von: Liu, Yimeng, et al.
Veröffentlicht: (2025)
CPS-TaskForge: Generating Collaborative Problem Solving Environments for Diverse Communication Tasks
von: Haduong, Nikita, et al.
Veröffentlicht: (2024)
von: Haduong, Nikita, et al.
Veröffentlicht: (2024)
Exploring diversity perceptions in a community through a Q&A chatbot
von: Kun, Peter, et al.
Veröffentlicht: (2024)
von: Kun, Peter, et al.
Veröffentlicht: (2024)
VisionTasker: Mobile Task Automation Using Vision Based UI Understanding and LLM Task Planning
von: Song, Yunpeng, et al.
Veröffentlicht: (2023)
von: Song, Yunpeng, et al.
Veröffentlicht: (2023)
From Following to Understanding: Investigating the Role of Reflective Prompts in AR-Guided Tasks to Promote Task Understanding
von: Zhang, Nandi, et al.
Veröffentlicht: (2025)
von: Zhang, Nandi, et al.
Veröffentlicht: (2025)
Expert-Generated Privacy Q&A Dataset for Conversational AI and User Study Insights
von: Leschanowsky, Anna, et al.
Veröffentlicht: (2025)
von: Leschanowsky, Anna, et al.
Veröffentlicht: (2025)
Management and Visualization Tools for Emergency Medical Services
von: Guigues, Vincent, et al.
Veröffentlicht: (2024)
von: Guigues, Vincent, et al.
Veröffentlicht: (2024)
PyZoBot: A Platform for Conversational Information Extraction and Synthesis from Curated Zotero Reference Libraries through Advanced Retrieval-Augmented Generation
von: Alshammari, Suad, et al.
Veröffentlicht: (2024)
von: Alshammari, Suad, et al.
Veröffentlicht: (2024)
ErgoGlide: A Wearable Trackball Device for Ergonomic Text Entry in Virtual Reality
von: Bakar, Muhammad Abu, et al.
Veröffentlicht: (2026)
von: Bakar, Muhammad Abu, et al.
Veröffentlicht: (2026)
ContextQ: Generated Questions to Support Meaningful Parent-Child Dialogue While Co-Reading
von: Smith, Griffin Dietz, et al.
Veröffentlicht: (2024)
von: Smith, Griffin Dietz, et al.
Veröffentlicht: (2024)
Task-Aware Delegation Cues for LLM Agents
von: Gu, Xingrui
Veröffentlicht: (2026)
von: Gu, Xingrui
Veröffentlicht: (2026)
A Typology of Decision-Making Tasks for Visualization
von: Brumar, Camelia D., et al.
Veröffentlicht: (2024)
von: Brumar, Camelia D., et al.
Veröffentlicht: (2024)
Haptic VR Simulation for Surgery Procedures in Medical Training
von: Jie, Lim Zheng, et al.
Veröffentlicht: (2024)
von: Jie, Lim Zheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AraHealthQA 2025: The First Shared Task on Arabic Health Question Answering
von: Alhuzali, Hassan, et al.
Veröffentlicht: (2025) -
Cross-Lingual Empirical Evaluation of Large Language Models for Arabic Medical Tasks
von: Abouzahir, Chaimae, et al.
Veröffentlicht: (2026) -
MedAraBench: Large-Scale Arabic Medical Question Answering Dataset and Benchmark
von: Abu-Daoud, Mouath, et al.
Veröffentlicht: (2026) -
Desk2Desk: Optimization-based Mixed Reality Workspace Integration for Remote Side-by-side Collaboration
von: Sidenmark, Ludwig, et al.
Veröffentlicht: (2024) -
Measuring Large Language Models Dependency: Validating the Arabic Version of the LLM-D12 Scale
von: AlShakhsi, Sameha, et al.
Veröffentlicht: (2025)