Beyond WER: Probing Whisper's Sub-token Decoder Across Diverse Language Resource Levels
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Liang, Siyu, Ballier, Nicolas, Levow, Gina-Anne, Wright, Richard |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
The Limits of Data Scaling: Sub-token Utilization and Acoustic Saturation in Multilingual ASR
par: Liang, Siyu, et autres
Publié: (2025)
par: Liang, Siyu, et autres
Publié: (2025)
Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages
par: Liang, Siyu, et autres
Publié: (2025)
par: Liang, Siyu, et autres
Publié: (2025)
Hybrid Neural-LLM Pipeline for Morphological Glossing in Endangered Language Documentation: A Case Study of Jungar Tuvan
par: Liang, Siyu, et autres
Publié: (2026)
par: Liang, Siyu, et autres
Publié: (2026)
A Sociophonetic Analysis of Racial Bias in Commercial ASR Systems Using the Pacific Northwest English Corpus
par: Scott, Michael, et autres
Publié: (2025)
par: Scott, Michael, et autres
Publié: (2025)
TEII: Think, Explain, Interact and Iterate with Large Language Models to Solve Cross-lingual Emotion Detection
par: Cheng, Long, et autres
Publié: (2024)
par: Cheng, Long, et autres
Publié: (2024)
SN-WER: Script-Normalized WER for Multi-Script Indic ASR Evaluation
par: Pattnayak, Priyaranjan
Publié: (2026)
par: Pattnayak, Priyaranjan
Publié: (2026)
Deepfake Word Detection by Next-token Prediction using Fine-tuned Whisper
par: Tran, Hoan My, et autres
Publié: (2026)
par: Tran, Hoan My, et autres
Publié: (2026)
SubTokenTest: A Practical Benchmark for Real-World Sub-token Understanding
par: Hou, Shuyang, et autres
Publié: (2026)
par: Hou, Shuyang, et autres
Publié: (2026)
Beyond Decodability: Reconstructing Language Model Representations with an Encoding Probe
par: Shen, Gaofei, et autres
Publié: (2026)
par: Shen, Gaofei, et autres
Publié: (2026)
WER We Stand: Benchmarking Urdu ASR Models
par: Arif, Samee, et autres
Publié: (2024)
par: Arif, Samee, et autres
Publié: (2024)
Speculative Decoding Across Languages
par: Paudel, Nirajan, et autres
Publié: (2026)
par: Paudel, Nirajan, et autres
Publié: (2026)
Fine-tuning Whisper on Low-Resource Languages for Real-World Applications
par: Timmel, Vincenzo, et autres
Publié: (2024)
par: Timmel, Vincenzo, et autres
Publié: (2024)
Whispering in Amharic: Fine-tuning Whisper for Low-resource Language
par: Gete, Dawit Ketema, et autres
Publié: (2025)
par: Gete, Dawit Ketema, et autres
Publié: (2025)
Adapting Whisper for Code-Switching through Encoding Refining and Language-Aware Decoding
par: Zhao, Jiahui, et autres
Publié: (2024)
par: Zhao, Jiahui, et autres
Publié: (2024)
Plug, Play, and Fuse: Zero-Shot Joint Decoding via Word-Level Re-ranking Across Diverse Vocabularies
par: Koneru, Sai, et autres
Publié: (2024)
par: Koneru, Sai, et autres
Publié: (2024)
Whisper Finetuning on Nepali Language
par: Rijal, Sanjay, et autres
Publié: (2024)
par: Rijal, Sanjay, et autres
Publié: (2024)
Assessing the Feasibility of Lightweight Whisper Models for Low-Resource Urdu Transcription
par: Antall, Abdul Rehman, et autres
Publié: (2025)
par: Antall, Abdul Rehman, et autres
Publié: (2025)
To Words and Beyond: Probing Large Language Models for Sentence-Level Psycholinguistic Norms of Memorability and Reading Times
par: Clark, Thomas Hikaru, et autres
Publié: (2026)
par: Clark, Thomas Hikaru, et autres
Publié: (2026)
WER is Unaware: Assessing How ASR Errors Distort Clinical Understanding in Patient Facing Dialogue
par: Ellis, Zachary, et autres
Publié: (2025)
par: Ellis, Zachary, et autres
Publié: (2025)
COMPACT: Common-token Optimized Model Pruning Across Channels and Tokens
par: Kwek, Eugene, et autres
Publié: (2025)
par: Kwek, Eugene, et autres
Publié: (2025)
CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference
par: Liu, Dong, et autres
Publié: (2025)
par: Liu, Dong, et autres
Publié: (2025)
Order-Level Attention Similarity Across Language Models: A Latent Commonality
par: Liang, Jinglin, et autres
Publié: (2025)
par: Liang, Jinglin, et autres
Publié: (2025)
Beyond Quality: Unlocking Diversity in Ad Headline Generation with Large Language Models
par: Wang, Chang, et autres
Publié: (2025)
par: Wang, Chang, et autres
Publié: (2025)
From Ghazals to Sonnets: Decoding the Polysemous Expressions of Love Across Languages
par: Ali, Syed Mohammad Sualeh
Publié: (2025)
par: Ali, Syed Mohammad Sualeh
Publié: (2025)
Loss Masking Is Not Needed in Decoder-only Transformer for Discrete-token-based ASR
par: Chen, Qian, et autres
Publié: (2023)
par: Chen, Qian, et autres
Publié: (2023)
Better & Faster Large Language Models via Multi-token Prediction
par: Gloeckle, Fabian, et autres
Publié: (2024)
par: Gloeckle, Fabian, et autres
Publié: (2024)
LBPE: Long-token-first Tokenization to Improve Large Language Models
par: Lian, Haoran, et autres
Publié: (2024)
par: Lian, Haoran, et autres
Publié: (2024)
Beyond Tokens: Semantic-Aware Speculative Decoding for Efficient Inference by Probing Internal States
par: Dong, Ximing, et autres
Publié: (2026)
par: Dong, Ximing, et autres
Publié: (2026)
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation
par: Hu, Rui, et autres
Publié: (2025)
par: Hu, Rui, et autres
Publié: (2025)
Jacobian Scopes: token-level causal attributions in LLMs
par: Liu, Toni J. B., et autres
Publié: (2026)
par: Liu, Toni J. B., et autres
Publié: (2026)
LLM Probe: Evaluating LLMs for Low-Resource Languages
par: Teklehaymanot, Hailay Kidu, et autres
Publié: (2026)
par: Teklehaymanot, Hailay Kidu, et autres
Publié: (2026)
Assessing the validity of new paradigmatic complexity measures as criterial features for proficiency in L2 writings in English
par: Mallart, Cyriel, et autres
Publié: (2025)
par: Mallart, Cyriel, et autres
Publié: (2025)
Beyond Confidence: Adaptive and Coherent Decoding for Diffusion Language Models
par: Chen, Kecheng, et autres
Publié: (2025)
par: Chen, Kecheng, et autres
Publié: (2025)
Probing Context Localization of Polysemous Words in Pre-trained Language Model Sub-Layers
par: Vijayakumar, Soniya, et autres
Publié: (2024)
par: Vijayakumar, Soniya, et autres
Publié: (2024)
Calm-Whisper: Reduce Whisper Hallucination On Non-Speech By Calming Crazy Heads Down
par: Wang, Yingzhi, et autres
Publié: (2025)
par: Wang, Yingzhi, et autres
Publié: (2025)
Moral Reasoning Across Languages: The Critical Role of Low-Resource Languages in LLMs
par: Zhou, Huichi, et autres
Publié: (2025)
par: Zhou, Huichi, et autres
Publié: (2025)
Semantic-guided Diverse Decoding for Large Language Model
par: Shi, Weijie, et autres
Publié: (2025)
par: Shi, Weijie, et autres
Publié: (2025)
Beyond the Surface: Probing the Ideological Depth of Large Language Models
par: Kabir, Shariar, et autres
Publié: (2025)
par: Kabir, Shariar, et autres
Publié: (2025)
Languages in Whisper-Style Speech Encoders Align Both Phonetically and Semantically
par: Shim, Ryan Soh-Eun, et autres
Publié: (2025)
par: Shim, Ryan Soh-Eun, et autres
Publié: (2025)
WhisperPipe: A Resource-Efficient Streaming Architecture for Real-Time Automatic Speech Recognition
par: Ramezani, Erfan, et autres
Publié: (2026)
par: Ramezani, Erfan, et autres
Publié: (2026)
Documents similaires
-
The Limits of Data Scaling: Sub-token Utilization and Acoustic Saturation in Multilingual ASR
par: Liang, Siyu, et autres
Publié: (2025) -
Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages
par: Liang, Siyu, et autres
Publié: (2025) -
Hybrid Neural-LLM Pipeline for Morphological Glossing in Endangered Language Documentation: A Case Study of Jungar Tuvan
par: Liang, Siyu, et autres
Publié: (2026) -
A Sociophonetic Analysis of Racial Bias in Commercial ASR Systems Using the Pacific Northwest English Corpus
par: Scott, Michael, et autres
Publié: (2025) -
TEII: Think, Explain, Interact and Iterate with Large Language Models to Solve Cross-lingual Emotion Detection
par: Cheng, Long, et autres
Publié: (2024)