ELYADATA & LIA at NADI 2025: ASR and ADI Subtasks
Fuente:
arXiv
Saved in:
| Main Authors: | Elleuch, Haroun, Saidi, Youssef, Mdhaffar, Salima, Estève, Yannick, Bougares, Fethi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ADI-20: Arabic Dialect Identification dataset and models
by: Elleuch, Haroun, et al.
Published: (2025)
by: Elleuch, Haroun, et al.
Published: (2025)
Performance Analysis of Speech Encoders for Low-Resource SLU and ASR in Tunisian Dialect
by: Mdhaffar, Salima, et al.
Published: (2024)
by: Mdhaffar, Salima, et al.
Published: (2024)
SLURP-TN : Resource for Tunisian Dialect Spoken Language Understanding
by: Elleuch, Haroun, et al.
Published: (2026)
by: Elleuch, Haroun, et al.
Published: (2026)
TEDxTN: A Three-way Speech Translation Corpus for Code-Switched Tunisian Arabic - English
by: Bougares, Fethi, et al.
Published: (2025)
by: Bougares, Fethi, et al.
Published: (2025)
Ara-Best-RQ: Multi Dialectal Arabic SSL
by: Elleuch, Haroun, et al.
Published: (2026)
by: Elleuch, Haroun, et al.
Published: (2026)
CV-18 NER: Augmented Common Voice for Named Entity Recognition from Arabic Speech
by: Saidi, Youssef, et al.
Published: (2026)
by: Saidi, Youssef, et al.
Published: (2026)
SENSE models: an open source solution for multilingual and multimodal semantic-based tasks
by: Mdhaffar, Salima, et al.
Published: (2025)
by: Mdhaffar, Salima, et al.
Published: (2025)
Learning Multiple Utterance-Level Attribute Representations with a Unified Speech Encoder
by: Bouziane, Maryem, et al.
Published: (2026)
by: Bouziane, Maryem, et al.
Published: (2026)
Using Multimodal and Language-Agnostic Sentence Embeddings for Abstractive Summarization
by: Chellaf, Chaimae, et al.
Published: (2026)
by: Chellaf, Chaimae, et al.
Published: (2026)
In-domain SSL pre-training and streaming ASR
by: Duret, Jarod, et al.
Published: (2025)
by: Duret, Jarod, et al.
Published: (2025)
Sonos Voice Control Bias Assessment Dataset: A Methodology for Demographic Bias Assessment in Voice Assistants
by: Sekkat, Chloé, et al.
Published: (2024)
by: Sekkat, Chloé, et al.
Published: (2024)
NADI 2025: The First Multidialectal Arabic Speech Processing Shared Task
by: Talafha, Bashar, et al.
Published: (2025)
by: Talafha, Bashar, et al.
Published: (2025)
Munsit at NADI 2025 Shared Task 2: Pushing the Boundaries of Multidialectal Arabic ASR with Weakly Supervised Pretraining and Continual Supervised Fine-tuning
by: Salhab, Mahmoud, et al.
Published: (2025)
by: Salhab, Mahmoud, et al.
Published: (2025)
Abjad AI at NADI 2025: CATT-Whisper: Multimodal Diacritic Restoration Using Text and Speech Representations
by: Ghannam, Ahmad, et al.
Published: (2025)
by: Ghannam, Ahmad, et al.
Published: (2025)
An Ultra-Low Latency, End-to-End Streaming Speech Synthesis Architecture via Block-Wise Generation and Depth-Wise Codec Decoding
by: Su, Tianhui, et al.
Published: (2026)
by: Su, Tianhui, et al.
Published: (2026)
NADI 2024: The Fifth Nuanced Arabic Dialect Identification Shared Task
by: Abdul-Mageed, Muhammad, et al.
Published: (2024)
by: Abdul-Mageed, Muhammad, et al.
Published: (2024)
dzNLP at NADI 2024 Shared Task: Multi-Classifier Ensemble with Weighted Voting and TF-IDF Features
by: Lichouri, Mohamed, et al.
Published: (2024)
by: Lichouri, Mohamed, et al.
Published: (2024)
Investigating Low-Cost LLM Annotation for~Spoken Dialogue Understanding Datasets
by: Druart, Lucas, et al.
Published: (2024)
by: Druart, Lucas, et al.
Published: (2024)
Assessing Open-Source Large Language Models on Argumentation Mining Subtasks
by: Abkenar, Mohammad Yeghaneh, et al.
Published: (2024)
by: Abkenar, Mohammad Yeghaneh, et al.
Published: (2024)
Open Implementation and Study of BEST-RQ for Speech Processing
by: Whetten, Ryan, et al.
Published: (2024)
by: Whetten, Ryan, et al.
Published: (2024)
Analyzing Speech Unit Selection for Textless Speech-to-Speech Translation
by: Duret, Jarod, et al.
Published: (2024)
by: Duret, Jarod, et al.
Published: (2024)
Is one brick enough to break the wall of spoken dialogue state tracking?
by: Druart, Lucas, et al.
Published: (2023)
by: Druart, Lucas, et al.
Published: (2023)
Neural Bandit Based Optimal LLM Selection for a Pipeline of Subtasks
by: Atalar, Baran, et al.
Published: (2025)
by: Atalar, Baran, et al.
Published: (2025)
A dual task learning approach to fine-tune a multilingual semantic speech encoder for Spoken Language Understanding
by: Laperrière, Gaëlle, et al.
Published: (2024)
by: Laperrière, Gaëlle, et al.
Published: (2024)
Internal Chain-of-Thought: Empirical Evidence for Layer-wise Subtask Scheduling in LLMs
by: Yang, Zhipeng, et al.
Published: (2025)
by: Yang, Zhipeng, et al.
Published: (2025)
Enhancing Complex Causality Extraction via Improved Subtask Interaction and Knowledge Fusion
by: Gao, Jinglong, et al.
Published: (2024)
by: Gao, Jinglong, et al.
Published: (2024)
A Training-free LLM Framework with Interaction between Contextually Related Subtasks in Solving Complex Tasks
by: Liu, Hongjia, et al.
Published: (2025)
by: Liu, Hongjia, et al.
Published: (2025)
SoRFT: Issue Resolving with Subtask-oriented Reinforced Fine-Tuning
by: Ma, Zexiong, et al.
Published: (2025)
by: Ma, Zexiong, et al.
Published: (2025)
Simultaneous Speech-to-Speech Translation Without Aligned Data
by: Labiausse, Tom, et al.
Published: (2026)
by: Labiausse, Tom, et al.
Published: (2026)
Semantic enrichment towards efficient speech representations
by: Laperrière, Gaëlle, et al.
Published: (2023)
by: Laperrière, Gaëlle, et al.
Published: (2023)
RAGVUE: A Diagnostic View for Explainable and Automated Evaluation of Retrieval-Augmented Generation
by: Murugaraj, Keerthana, et al.
Published: (2025)
by: Murugaraj, Keerthana, et al.
Published: (2025)
Pantagruel: Unified Self-Supervised Encoders for French Text and Speech
by: Le, Phuong-Hang, et al.
Published: (2026)
by: Le, Phuong-Hang, et al.
Published: (2026)
Towards Early Prediction of Self-Supervised Speech Model Performance
by: Whetten, Ryan, et al.
Published: (2025)
by: Whetten, Ryan, et al.
Published: (2025)
An Analysis of Linear Complexity Attention Substitutes with BEST-RQ
by: Whetten, Ryan, et al.
Published: (2024)
by: Whetten, Ryan, et al.
Published: (2024)
LeBenchmark 2.0: a Standardized, Replicable and Enhanced Framework for Self-supervised Representations of French Speech
by: Parcollet, Titouan, et al.
Published: (2023)
by: Parcollet, Titouan, et al.
Published: (2023)
On the Robust Approximation of ASR Metrics
by: Waheed, Abdul, et al.
Published: (2025)
by: Waheed, Abdul, et al.
Published: (2025)
CantoASR: Prosody-Aware ASR-LALM Collaboration for Low-Resource Cantonese
by: Chen, Dazhong, et al.
Published: (2025)
by: Chen, Dazhong, et al.
Published: (2025)
Polyglot-Lion: Efficient Multilingual ASR for Singapore via Balanced Fine-Tuning of Qwen3-ASR
by: Dang, Quy-Anh, et al.
Published: (2026)
by: Dang, Quy-Anh, et al.
Published: (2026)
Open Universal Arabic ASR Leaderboard
by: Wang, Yingzhi, et al.
Published: (2024)
by: Wang, Yingzhi, et al.
Published: (2024)
PromptASR for contextualized ASR with controllable style
by: Yang, Xiaoyu, et al.
Published: (2023)
by: Yang, Xiaoyu, et al.
Published: (2023)
Similar Items
-
ADI-20: Arabic Dialect Identification dataset and models
by: Elleuch, Haroun, et al.
Published: (2025) -
Performance Analysis of Speech Encoders for Low-Resource SLU and ASR in Tunisian Dialect
by: Mdhaffar, Salima, et al.
Published: (2024) -
SLURP-TN : Resource for Tunisian Dialect Spoken Language Understanding
by: Elleuch, Haroun, et al.
Published: (2026) -
TEDxTN: A Three-way Speech Translation Corpus for Code-Switched Tunisian Arabic - English
by: Bougares, Fethi, et al.
Published: (2025) -
Ara-Best-RQ: Multi Dialectal Arabic SSL
by: Elleuch, Haroun, et al.
Published: (2026)