Beyond Single Ground Truth: Reference Monism as Epistemic Injustice in ASR Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Choi, Anna Seo Gyeong, Teleki, Maria, Caverlee, James, del Rio, Miguel, Miller, Corey, Choi, Hoon |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fairness of Automatic Speech Recognition: Looking Through a Philosophical Lens
by: Choi, Anna Seo Gyeong, et al.
Published: (2025)
by: Choi, Anna Seo Gyeong, et al.
Published: (2025)
A Survey on LLM Inference-Time Self-Improvement
by: Dong, Xiangjue, et al.
Published: (2024)
by: Dong, Xiangjue, et al.
Published: (2024)
Masculine Defaults via Gendered Discourse in Podcasts and Large Language Models
by: Teleki, Maria, et al.
Published: (2025)
by: Teleki, Maria, et al.
Published: (2025)
CHOIR: Collaborative Harmonization fOr Inference Robustness
by: Dong, Xiangjue, et al.
Published: (2025)
by: Dong, Xiangjue, et al.
Published: (2025)
Quantification of stylistic differences in human- and ASR-produced transcripts of African American English
by: Heuser, Annika, et al.
Published: (2024)
by: Heuser, Annika, et al.
Published: (2024)
Style-agnostic evaluation of ASR using multiple reference transcripts
by: McNamara, Quinten, et al.
Published: (2024)
by: McNamara, Quinten, et al.
Published: (2024)
Data-Driven Mispronunciation Pattern Discovery for Robust Speech Recognition
by: Choi, Anna Seo Gyeong, et al.
Published: (2025)
by: Choi, Anna Seo Gyeong, et al.
Published: (2025)
Single Ground Truth Is Not Enough: Adding Flexibility to Aspect-Based Sentiment Analysis Evaluation
by: Yang, Soyoung, et al.
Published: (2024)
by: Yang, Soyoung, et al.
Published: (2024)
PromptHelper: A Prompt Recommender System for Encouraging Creativity in AI Chatbot Interactions
by: Kim, Jason, et al.
Published: (2026)
by: Kim, Jason, et al.
Published: (2026)
Analyzing Dialectical Biases in LLMs for Knowledge and Reasoning Benchmarks
by: Pan, Eileen, et al.
Published: (2025)
by: Pan, Eileen, et al.
Published: (2025)
Careless Whisper: Speech-to-Text Hallucination Harms
by: Koenecke, Allison, et al.
Published: (2024)
by: Koenecke, Allison, et al.
Published: (2024)
Reverb: Open-Source ASR and Diarization from Rev
by: Bhandari, Nishchal, et al.
Published: (2024)
by: Bhandari, Nishchal, et al.
Published: (2024)
Truth Knows No Language: Evaluating Truthfulness Beyond English
by: Figueras, Blanca Calvo, et al.
Published: (2025)
by: Figueras, Blanca Calvo, et al.
Published: (2025)
Addressing Pitfalls in Auditing Practices of Automatic Speech Recognition Technologies: A Case Study of People with Aphasia
by: Mei, Katelyn Xiaoying, et al.
Published: (2025)
by: Mei, Katelyn Xiaoying, et al.
Published: (2025)
The Voice Behind the Words: Quantifying Intersectional Bias in SpeechLLMs
by: Satish, Shree Harsha Bokkahalli, et al.
Published: (2026)
by: Satish, Shree Harsha Bokkahalli, et al.
Published: (2026)
Analyzing Error Propagation in Korean Spoken QA with ASR-LLM Cascades
by: Jung, Donghyuk, et al.
Published: (2026)
by: Jung, Donghyuk, et al.
Published: (2026)
Epistemic Injustice and Violence
Published: (2026)
Published: (2026)
Literature and Epistemic Injustice
by: Colvin, Sarah
Published: (2025)
by: Colvin, Sarah
Published: (2025)
Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth
by: Gur-Arieh, Yoav, et al.
Published: (2026)
by: Gur-Arieh, Yoav, et al.
Published: (2026)
MaTableGPT: GPT-based Table Data Extractor from Materials Science Literature
by: Yi, Gyeong Hoon, et al.
Published: (2024)
by: Yi, Gyeong Hoon, et al.
Published: (2024)
$DA^3$: A Distribution-Aware Adversarial Attack against Language Models
by: Wang, Yibo, et al.
Published: (2023)
by: Wang, Yibo, et al.
Published: (2023)
Disclosure and Mitigation of Gender Bias in LLMs
by: Dong, Xiangjue, et al.
Published: (2024)
by: Dong, Xiangjue, et al.
Published: (2024)
Beyond Single Embeddings: Capturing Diverse Targets with Multi-Query Retrieval
by: Chen, Hung-Ting, et al.
Published: (2025)
by: Chen, Hung-Ting, et al.
Published: (2025)
The Epistemic Injustice of Genocide Denialism
by: Altanian, Melanie
Published: (2025)
by: Altanian, Melanie
Published: (2025)
Epistemic Injustice in Generative AI
by: Kay, Jackie, et al.
Published: (2024)
by: Kay, Jackie, et al.
Published: (2024)
Epistemic Injustice and Indigenous Epistemology
by: V. Hari Narayanan
Published: (2026)
by: V. Hari Narayanan
Published: (2026)
Toward Reliable Scientific Hypothesis Generation: Evaluating Truthfulness and Hallucination in Large Language Models
by: Xiong, Guangzhi, et al.
Published: (2025)
by: Xiong, Guangzhi, et al.
Published: (2025)
Meaning Beyond Truth Conditions: Evaluating Discourse Level Understanding via Anaphora Accessibility
by: Zhu, Xiaomeng, et al.
Published: (2025)
by: Zhu, Xiaomeng, et al.
Published: (2025)
Beyond Words: Evaluating and Bridging Epistemic Divergence in User-Agent Interaction via Theory of Mind
by: Ruan, Minyuan, et al.
Published: (2026)
by: Ruan, Minyuan, et al.
Published: (2026)
Bayesian Evolutionary Swarm Architecture: A Formal Epistemic System Grounded in Truth-Based Competition
by: Wright, Craig Steven
Published: (2025)
by: Wright, Craig Steven
Published: (2025)
Language Models as Semantic Augmenters for Sequential Recommenders
by: Valizadeh, Mahsa, et al.
Published: (2025)
by: Valizadeh, Mahsa, et al.
Published: (2025)
Chapter 10 Healthcare Practice, Epistemic Injustice, and Naturalism
by: Kidd, Ian James, et al.
Published: (2021)
by: Kidd, Ian James, et al.
Published: (2021)
Beyond Oversmoothing: Evaluating DDPM and MSE for Scalable Speech Synthesis in ASR
by: Minixhofer, Christoph, et al.
Published: (2024)
by: Minixhofer, Christoph, et al.
Published: (2024)
Elderly-Contextual Data Augmentation via Speech Synthesis for Elderly ASR
by: Lee, Minsik, et al.
Published: (2026)
by: Lee, Minsik, et al.
Published: (2026)
KatotohananQA: Evaluating Truthfulness of Large Language Models in Filipino
by: Nery, Lorenzo Alfred, et al.
Published: (2025)
by: Nery, Lorenzo Alfred, et al.
Published: (2025)
VeriEquivBench: An Equivalence Score for Ground-Truth-Free Evaluation of Formally Verifiable Code
by: Zeng, Lingfei, et al.
Published: (2025)
by: Zeng, Lingfei, et al.
Published: (2025)
Distributed Partial Information Puzzles: Examining Common Ground Construction Under Epistemic Asymmetry
by: Zhu, Yifan, et al.
Published: (2026)
by: Zhu, Yifan, et al.
Published: (2026)
SocialPulse: An Open-Source Subreddit Sensemaking Toolkit
by: Birkelbach, Stephanie, et al.
Published: (2026)
by: Birkelbach, Stephanie, et al.
Published: (2026)
Ground-Truth Subgraphs for Better Training and Evaluation of Knowledge Graph Augmented LLMs
by: Cattaneo, Alberto, et al.
Published: (2025)
by: Cattaneo, Alberto, et al.
Published: (2025)
TruthEval: A Dataset to Evaluate LLM Truthfulness and Reliability
by: Khatun, Aisha, et al.
Published: (2024)
by: Khatun, Aisha, et al.
Published: (2024)
Similar Items
-
Fairness of Automatic Speech Recognition: Looking Through a Philosophical Lens
by: Choi, Anna Seo Gyeong, et al.
Published: (2025) -
A Survey on LLM Inference-Time Self-Improvement
by: Dong, Xiangjue, et al.
Published: (2024) -
Masculine Defaults via Gendered Discourse in Podcasts and Large Language Models
by: Teleki, Maria, et al.
Published: (2025) -
CHOIR: Collaborative Harmonization fOr Inference Robustness
by: Dong, Xiangjue, et al.
Published: (2025) -
Quantification of stylistic differences in human- and ASR-produced transcripts of African American English
by: Heuser, Annika, et al.
Published: (2024)