HARNESS: Lightweight Distilled Arabic Speech Foundation Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | sukhadia, Vrunda N., Chowdhury, Shammur Absar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HARNESS: Lightweight Distilled Arabic Speech Foundation Models
von: Sukhadia, Vrunda N., et al.
Veröffentlicht: (2026)
von: Sukhadia, Vrunda N., et al.
Veröffentlicht: (2026)
Children's Speech Recognition through Discrete Token Enhancement
von: Sukhadia, Vrunda N., et al.
Veröffentlicht: (2024)
von: Sukhadia, Vrunda N., et al.
Veröffentlicht: (2024)
Multi-Task Instruction Tuning via Data Scheduling for Low-Resource Arabic AudioLLMs
von: Bhatti, Hunzalah Hassan, et al.
Veröffentlicht: (2026)
von: Bhatti, Hunzalah Hassan, et al.
Veröffentlicht: (2026)
SpokenNativQA: Multilingual Everyday Spoken Queries for LLMs
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
WASIL: In-the-Wild Arabic Spoken Interactions with LLMs
von: Ali, Zien Sheikh, et al.
Veröffentlicht: (2026)
von: Ali, Zien Sheikh, et al.
Veröffentlicht: (2026)
Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages
von: Alam, Firoj, et al.
Veröffentlicht: (2026)
von: Alam, Firoj, et al.
Veröffentlicht: (2026)
Beyond LLM-as-a-Judge: Deterministic Metrics for Multilingual Generative Text Evaluation
von: Alam, Firoj, et al.
Veröffentlicht: (2026)
von: Alam, Firoj, et al.
Veröffentlicht: (2026)
NativQA: Multilingual Culturally-Aligned Natural Query for LLMs
von: Hasan, Md. Arid, et al.
Veröffentlicht: (2024)
von: Hasan, Md. Arid, et al.
Veröffentlicht: (2024)
From Words to Waves: Analyzing Concept Formation in Speech and Text-Based Foundation Models
von: Ersoy, Asım, et al.
Veröffentlicht: (2025)
von: Ersoy, Asım, et al.
Veröffentlicht: (2025)
Speech Representation Analysis based on Inter- and Intra-Model Similarities
von: Kheir, Yassine El, et al.
Veröffentlicht: (2024)
von: Kheir, Yassine El, et al.
Veröffentlicht: (2024)
Beyond Orthography: Automatic Recovery of Short Vowels and Dialectal Sounds in Arabic
von: Kheir, Yassine El, et al.
Veröffentlicht: (2024)
von: Kheir, Yassine El, et al.
Veröffentlicht: (2024)
CAFE A Novel Code switching Dataset for Algerian Dialect French and English
von: Lachemat, Houssam Eddine-Othman, et al.
Veröffentlicht: (2024)
von: Lachemat, Houssam Eddine-Othman, et al.
Veröffentlicht: (2024)
Once Correct, Still Wrong: Counterfactual Hallucination in Multilingual Vision-Language Models
von: Mousi, Basel, et al.
Veröffentlicht: (2026)
von: Mousi, Basel, et al.
Veröffentlicht: (2026)
LAraBench: Benchmarking Arabic AI with Large Language Models
von: Abdelali, Ahmed, et al.
Veröffentlicht: (2023)
von: Abdelali, Ahmed, et al.
Veröffentlicht: (2023)
NativQA Framework: Enabling LLMs and VLMs with Native, Local, and Everyday Knowledge
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
MENASpeechBank: A Reference Voice Bank with Persona-Conditioned Multi-Turn Conversations for AudioLLMs
von: Ali, Zien Sheikh, et al.
Veröffentlicht: (2026)
von: Ali, Zien Sheikh, et al.
Veröffentlicht: (2026)
Casablanca: Data and Models for Multidialectal Arabic Speech Recognition
von: Talafha, Bashar, et al.
Veröffentlicht: (2024)
von: Talafha, Bashar, et al.
Veröffentlicht: (2024)
STaR: Distilling Speech Temporal Relation for Lightweight Speech Self-Supervised Learning Models
von: Jang, Kangwook, et al.
Veröffentlicht: (2023)
von: Jang, Kangwook, et al.
Veröffentlicht: (2023)
GenAI Content Detection Task 2: AI vs. Human -- Academic Essay Authenticity Challenge
von: Chowdhury, Shammur Absar, et al.
Veröffentlicht: (2024)
von: Chowdhury, Shammur Absar, et al.
Veröffentlicht: (2024)
Habibi: Laying the Open-Source Foundation of Unified-Dialectal Arabic Speech Synthesis
von: Chen, Yushen, et al.
Veröffentlicht: (2026)
von: Chen, Yushen, et al.
Veröffentlicht: (2026)
Community Needs and Assets: A Computational Analysis of Community Conversations
von: Chowdhury, Md Towhidul Absar, et al.
Veröffentlicht: (2024)
von: Chowdhury, Md Towhidul Absar, et al.
Veröffentlicht: (2024)
Comprehensive and Efficient Distillation for Lightweight Sentiment Analysis Models
von: Xie, Guangyu, et al.
Veröffentlicht: (2025)
von: Xie, Guangyu, et al.
Veröffentlicht: (2025)
EmoHopeSpeech: An Annotated Dataset of Emotions and Hope Speech in English and Arabic
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2025)
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2025)
An Annotated Corpus of Arabic Tweets for Hate Speech Analysis
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2025)
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2025)
AraS2P: Arabic Speech-to-Phonemes System
von: Matar, Bassam, et al.
Veröffentlicht: (2025)
von: Matar, Bassam, et al.
Veröffentlicht: (2025)
AraDiCE: Benchmarks for Dialectal and Cultural Capabilities in LLMs
von: Mousi, Basel, et al.
Veröffentlicht: (2024)
von: Mousi, Basel, et al.
Veröffentlicht: (2024)
BnTTS: Few-Shot Speaker Adaptation in Low-Resource Setting
von: Basher, Mohammad Jahid Ibna, et al.
Veröffentlicht: (2025)
von: Basher, Mohammad Jahid Ibna, et al.
Veröffentlicht: (2025)
Arabic Little STT: Arabic Children Speech Recognition Dataset
von: Alkadri, Mouhand, et al.
Veröffentlicht: (2025)
von: Alkadri, Mouhand, et al.
Veröffentlicht: (2025)
DistilQwen2.5: Industrial Practices of Training Distilled Open Lightweight Language Models
von: Wang, Chengyu, et al.
Veröffentlicht: (2025)
von: Wang, Chengyu, et al.
Veröffentlicht: (2025)
Arab Voices: Mapping Standard and Dialectal Arabic Speech Technology
von: Sullivan, Peter, et al.
Veröffentlicht: (2026)
von: Sullivan, Peter, et al.
Veröffentlicht: (2026)
Arabic Tweet Act: A Weighted Ensemble Pre-Trained Transformer Model for Classifying Arabic Speech Acts on Twitter
von: Alshehri, Khadejaa, et al.
Veröffentlicht: (2024)
von: Alshehri, Khadejaa, et al.
Veröffentlicht: (2024)
ArabEmoNet: A Lightweight Hybrid 2D CNN-BiLSTM Model with Attention for Robust Arabic Speech Emotion Recognition
von: Abouzeid, Ali, et al.
Veröffentlicht: (2025)
von: Abouzeid, Ali, et al.
Veröffentlicht: (2025)
State-of-the-Art Arabic Language Modeling with Sparse MoE Fine-Tuning and Chain-of-Thought Distillation
von: Singh, Navan Preet, et al.
Veröffentlicht: (2026)
von: Singh, Navan Preet, et al.
Veröffentlicht: (2026)
Multi-task Learning with Active Learning for Arabic Offensive Speech Detection
von: Alansari, Aisha, et al.
Veröffentlicht: (2025)
von: Alansari, Aisha, et al.
Veröffentlicht: (2025)
ZAEBUC-Spoken: A Multilingual Multidialectal Arabic-English Speech Corpus
von: Hamed, Injy, et al.
Veröffentlicht: (2024)
von: Hamed, Injy, et al.
Veröffentlicht: (2024)
SAGE: Spliced-Audio Generated Data for Enhancing Foundational Models in Low-Resource Arabic-English Code-Switched Speech Recognition
von: Farooq, Muhammad Umar, et al.
Veröffentlicht: (2025)
von: Farooq, Muhammad Umar, et al.
Veröffentlicht: (2025)
EmoAra: Emotion-Preserving English Speech Transcription and Cross-Lingual Translation with Arabic Text-to-Speech
von: Hassan, Besher, et al.
Veröffentlicht: (2026)
von: Hassan, Besher, et al.
Veröffentlicht: (2026)
Speech Translation with Speech Foundation Models and Large Language Models: What is There and What is Missing?
von: Gaido, Marco, et al.
Veröffentlicht: (2024)
von: Gaido, Marco, et al.
Veröffentlicht: (2024)
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
von: Luu, Nam, et al.
Veröffentlicht: (2025)
von: Luu, Nam, et al.
Veröffentlicht: (2025)
Dialectal Coverage And Generalization in Arabic Speech Recognition
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2024)
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
HARNESS: Lightweight Distilled Arabic Speech Foundation Models
von: Sukhadia, Vrunda N., et al.
Veröffentlicht: (2026) -
Children's Speech Recognition through Discrete Token Enhancement
von: Sukhadia, Vrunda N., et al.
Veröffentlicht: (2024) -
Multi-Task Instruction Tuning via Data Scheduling for Low-Resource Arabic AudioLLMs
von: Bhatti, Hunzalah Hassan, et al.
Veröffentlicht: (2026) -
SpokenNativQA: Multilingual Everyday Spoken Queries for LLMs
von: Alam, Firoj, et al.
Veröffentlicht: (2025) -
WASIL: In-the-Wild Arabic Spoken Interactions with LLMs
von: Ali, Zien Sheikh, et al.
Veröffentlicht: (2026)