A Benchmark of French ASR Systems Based on Error Severity
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tholly, Antoine, Wottawa, Jane, Rouvier, Mickael, Dufour, Richard |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Universal-2-TF: Robust All-Neural Text Formatting for ASR
von: Khare, Yash, et al.
Veröffentlicht: (2025)
von: Khare, Yash, et al.
Veröffentlicht: (2025)
Low-resource neural machine translation with morphological modeling
von: Nzeyimana, Antoine
Veröffentlicht: (2024)
von: Nzeyimana, Antoine
Veröffentlicht: (2024)
Dialect Matters: Cross-Lingual ASR Transfer for Low-Resource Indic Language Varieties
von: Dhasmana, Akriti, et al.
Veröffentlicht: (2026)
von: Dhasmana, Akriti, et al.
Veröffentlicht: (2026)
Hard to Be Heard: Phoneme-Level ASR Analysis of Phonologically Complex, Low-Resource Endangered Languages
von: Akavarapu, V. S. D. S. Mahesh, et al.
Veröffentlicht: (2026)
von: Akavarapu, V. S. D. S. Mahesh, et al.
Veröffentlicht: (2026)
KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation
von: Nzeyimana, Antoine, et al.
Veröffentlicht: (2025)
von: Nzeyimana, Antoine, et al.
Veröffentlicht: (2025)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
Luth: Efficient French Specialization for Small Language Models and Cross-Lingual Transfer
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2025)
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
Distinguishing Ignorance from Error in LLM Hallucinations
von: Simhi, Adi, et al.
Veröffentlicht: (2024)
von: Simhi, Adi, et al.
Veröffentlicht: (2024)
ASR Error Correction in Low-Resource Burmese with Alignment-Enhanced Transformers using Phonetic Features
von: Lin, Ye Bhone, et al.
Veröffentlicht: (2025)
von: Lin, Ye Bhone, et al.
Veröffentlicht: (2025)
OpenFactCheck: Building, Benchmarking Customized Fact-Checking Systems and Evaluating the Factuality of Claims and LLMs
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
A Multi-Pass Large Language Model Framework for Precise and Efficient Radiology Report Error Detection
von: Kim, Songsoo, et al.
Veröffentlicht: (2025)
von: Kim, Songsoo, et al.
Veröffentlicht: (2025)
MAWARITH: A Dataset and Benchmark for Legal Inheritance Reasoning with LLMs
von: Bouchekif, Abdessalam, et al.
Veröffentlicht: (2026)
von: Bouchekif, Abdessalam, et al.
Veröffentlicht: (2026)
GroUSE: A Benchmark to Evaluate Evaluators in Grounded Question Answering
von: Muller, Sacha, et al.
Veröffentlicht: (2024)
von: Muller, Sacha, et al.
Veröffentlicht: (2024)
A Multi-Task Benchmark for Abusive Language Detection in Low-Resource Settings
von: Gaim, Fitsum, et al.
Veröffentlicht: (2025)
von: Gaim, Fitsum, et al.
Veröffentlicht: (2025)
LLM-GLOBE: A Benchmark Evaluating the Cultural Values Embedded in LLM Output
von: Karinshak, Elise, et al.
Veröffentlicht: (2024)
von: Karinshak, Elise, et al.
Veröffentlicht: (2024)
RAID: A Shared Benchmark for Robust Evaluation of Machine-Generated Text Detectors
von: Dugan, Liam, et al.
Veröffentlicht: (2024)
von: Dugan, Liam, et al.
Veröffentlicht: (2024)
Constructing Benchmarks and Interventions for Combating Hallucinations in LLMs
von: Simhi, Adi, et al.
Veröffentlicht: (2024)
von: Simhi, Adi, et al.
Veröffentlicht: (2024)
EnDive: A Cross-Dialect Benchmark for Fairness and Performance in Large Language Models
von: Gupta, Abhay, et al.
Veröffentlicht: (2025)
von: Gupta, Abhay, et al.
Veröffentlicht: (2025)
EmoS: A High-Fidelity Multimodal Benchmark for Fine-grained Streaming Emotional Understanding
von: Guo, Pengze, et al.
Veröffentlicht: (2026)
von: Guo, Pengze, et al.
Veröffentlicht: (2026)
PL-Guard: Benchmarking Language Model Safety for Polish
von: Krasnodębska, Aleksandra, et al.
Veröffentlicht: (2025)
von: Krasnodębska, Aleksandra, et al.
Veröffentlicht: (2025)
EfficientQA : a RoBERTa Based Phrase-Indexed Question-Answering System
von: Chaybouti, Sofian, et al.
Veröffentlicht: (2021)
von: Chaybouti, Sofian, et al.
Veröffentlicht: (2021)
Assessing Latency in ASR Systems: A Methodological Perspective for Real-Time Use
von: Arriaga, Carlos, et al.
Veröffentlicht: (2024)
von: Arriaga, Carlos, et al.
Veröffentlicht: (2024)
LCFO: Long Context and Long Form Output Dataset and Benchmarking
von: Costa-jussà, Marta R., et al.
Veröffentlicht: (2024)
von: Costa-jussà, Marta R., et al.
Veröffentlicht: (2024)
Syntactic Blind Spots: How Misalignment Leads to LLMs Mathematical Errors
von: Williamson, Dane, et al.
Veröffentlicht: (2025)
von: Williamson, Dane, et al.
Veröffentlicht: (2025)
EduGuardBench: A Holistic Benchmark for Evaluating the Pedagogical Fidelity and Adversarial Safety of LLMs as Simulated Teachers
von: Jiang, Yilin, et al.
Veröffentlicht: (2025)
von: Jiang, Yilin, et al.
Veröffentlicht: (2025)
BOUQuET: dataset, Benchmark and Open initiative for Universal Quality Evaluation in Translation
von: The Omnilingual MT Team, et al.
Veröffentlicht: (2025)
von: The Omnilingual MT Team, et al.
Veröffentlicht: (2025)
HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns
von: Wang, Xintao, et al.
Veröffentlicht: (2026)
von: Wang, Xintao, et al.
Veröffentlicht: (2026)
The GDN-CC Dataset: Automatic Corpus Clarification for AI-enhanced Democratic Citizen Consultations
von: Lequeu, Pierre-Antoine, et al.
Veröffentlicht: (2026)
von: Lequeu, Pierre-Antoine, et al.
Veröffentlicht: (2026)
Graphemic Normalization of the Perso-Arabic Script
von: Doctor, Raiomond, et al.
Veröffentlicht: (2022)
von: Doctor, Raiomond, et al.
Veröffentlicht: (2022)
Beyond Arabic: Software for Perso-Arabic Script Manipulation
von: Gutkin, Alexander, et al.
Veröffentlicht: (2023)
von: Gutkin, Alexander, et al.
Veröffentlicht: (2023)
Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews
von: Li, Bowen, et al.
Veröffentlicht: (2026)
von: Li, Bowen, et al.
Veröffentlicht: (2026)
UrduFactCheck: An Agentic Fact-Checking Framework for Urdu with Evidence Boosting and Benchmarking
von: Ahmad, Sarfraz, et al.
Veröffentlicht: (2025)
von: Ahmad, Sarfraz, et al.
Veröffentlicht: (2025)
EVM-QuestBench: An Execution-Grounded Benchmark for Natural-Language Transaction Code Generation
von: Yang, Pei, et al.
Veröffentlicht: (2026)
von: Yang, Pei, et al.
Veröffentlicht: (2026)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
von: Saji, Alan, et al.
Veröffentlicht: (2025)
von: Saji, Alan, et al.
Veröffentlicht: (2025)
SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic Grading
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024)
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024)
HalluScan: A Systematic Benchmark for Detecting and Mitigating Hallucinations in Instruction-Following LLMs
von: Cherif, Ahmed
Veröffentlicht: (2026)
von: Cherif, Ahmed
Veröffentlicht: (2026)
Grade Guard: A Smart System for Short Answer Automated Grading
von: Dadu, Niharika, et al.
Veröffentlicht: (2025)
von: Dadu, Niharika, et al.
Veröffentlicht: (2025)
A Comprehensive Analysis of Tokenization and Self-Supervised Learning in End-to-End Automatic Speech Recognition applied on French Language
von: Bañeras-Roux, Thibault, et al.
Veröffentlicht: (2026)
von: Bañeras-Roux, Thibault, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Universal-2-TF: Robust All-Neural Text Formatting for ASR
von: Khare, Yash, et al.
Veröffentlicht: (2025) -
Low-resource neural machine translation with morphological modeling
von: Nzeyimana, Antoine
Veröffentlicht: (2024) -
Dialect Matters: Cross-Lingual ASR Transfer for Low-Resource Indic Language Varieties
von: Dhasmana, Akriti, et al.
Veröffentlicht: (2026) -
Hard to Be Heard: Phoneme-Level ASR Analysis of Phonologically Complex, Low-Resource Endangered Languages
von: Akavarapu, V. S. D. S. Mahesh, et al.
Veröffentlicht: (2026) -
KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation
von: Nzeyimana, Antoine, et al.
Veröffentlicht: (2025)