Beyond WER: Probing Whisper's Sub-token Decoder Across Diverse Language Resource Levels
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liang, Siyu, Ballier, Nicolas, Levow, Gina-Anne, Wright, Richard |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Limits of Data Scaling: Sub-token Utilization and Acoustic Saturation in Multilingual ASR
von: Liang, Siyu, et al.
Veröffentlicht: (2025)
von: Liang, Siyu, et al.
Veröffentlicht: (2025)
Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages
von: Liang, Siyu, et al.
Veröffentlicht: (2025)
von: Liang, Siyu, et al.
Veröffentlicht: (2025)
Hybrid Neural-LLM Pipeline for Morphological Glossing in Endangered Language Documentation: A Case Study of Jungar Tuvan
von: Liang, Siyu, et al.
Veröffentlicht: (2026)
von: Liang, Siyu, et al.
Veröffentlicht: (2026)
A Sociophonetic Analysis of Racial Bias in Commercial ASR Systems Using the Pacific Northwest English Corpus
von: Scott, Michael, et al.
Veröffentlicht: (2025)
von: Scott, Michael, et al.
Veröffentlicht: (2025)
TEII: Think, Explain, Interact and Iterate with Large Language Models to Solve Cross-lingual Emotion Detection
von: Cheng, Long, et al.
Veröffentlicht: (2024)
von: Cheng, Long, et al.
Veröffentlicht: (2024)
SN-WER: Script-Normalized WER for Multi-Script Indic ASR Evaluation
von: Pattnayak, Priyaranjan
Veröffentlicht: (2026)
von: Pattnayak, Priyaranjan
Veröffentlicht: (2026)
Deepfake Word Detection by Next-token Prediction using Fine-tuned Whisper
von: Tran, Hoan My, et al.
Veröffentlicht: (2026)
von: Tran, Hoan My, et al.
Veröffentlicht: (2026)
SubTokenTest: A Practical Benchmark for Real-World Sub-token Understanding
von: Hou, Shuyang, et al.
Veröffentlicht: (2026)
von: Hou, Shuyang, et al.
Veröffentlicht: (2026)
Beyond Decodability: Reconstructing Language Model Representations with an Encoding Probe
von: Shen, Gaofei, et al.
Veröffentlicht: (2026)
von: Shen, Gaofei, et al.
Veröffentlicht: (2026)
WER We Stand: Benchmarking Urdu ASR Models
von: Arif, Samee, et al.
Veröffentlicht: (2024)
von: Arif, Samee, et al.
Veröffentlicht: (2024)
Speculative Decoding Across Languages
von: Paudel, Nirajan, et al.
Veröffentlicht: (2026)
von: Paudel, Nirajan, et al.
Veröffentlicht: (2026)
Fine-tuning Whisper on Low-Resource Languages for Real-World Applications
von: Timmel, Vincenzo, et al.
Veröffentlicht: (2024)
von: Timmel, Vincenzo, et al.
Veröffentlicht: (2024)
Whispering in Amharic: Fine-tuning Whisper for Low-resource Language
von: Gete, Dawit Ketema, et al.
Veröffentlicht: (2025)
von: Gete, Dawit Ketema, et al.
Veröffentlicht: (2025)
Adapting Whisper for Code-Switching through Encoding Refining and Language-Aware Decoding
von: Zhao, Jiahui, et al.
Veröffentlicht: (2024)
von: Zhao, Jiahui, et al.
Veröffentlicht: (2024)
Plug, Play, and Fuse: Zero-Shot Joint Decoding via Word-Level Re-ranking Across Diverse Vocabularies
von: Koneru, Sai, et al.
Veröffentlicht: (2024)
von: Koneru, Sai, et al.
Veröffentlicht: (2024)
Whisper Finetuning on Nepali Language
von: Rijal, Sanjay, et al.
Veröffentlicht: (2024)
von: Rijal, Sanjay, et al.
Veröffentlicht: (2024)
Assessing the Feasibility of Lightweight Whisper Models for Low-Resource Urdu Transcription
von: Antall, Abdul Rehman, et al.
Veröffentlicht: (2025)
von: Antall, Abdul Rehman, et al.
Veröffentlicht: (2025)
To Words and Beyond: Probing Large Language Models for Sentence-Level Psycholinguistic Norms of Memorability and Reading Times
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
WER is Unaware: Assessing How ASR Errors Distort Clinical Understanding in Patient Facing Dialogue
von: Ellis, Zachary, et al.
Veröffentlicht: (2025)
von: Ellis, Zachary, et al.
Veröffentlicht: (2025)
COMPACT: Common-token Optimized Model Pruning Across Channels and Tokens
von: Kwek, Eugene, et al.
Veröffentlicht: (2025)
von: Kwek, Eugene, et al.
Veröffentlicht: (2025)
CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference
von: Liu, Dong, et al.
Veröffentlicht: (2025)
von: Liu, Dong, et al.
Veröffentlicht: (2025)
Order-Level Attention Similarity Across Language Models: A Latent Commonality
von: Liang, Jinglin, et al.
Veröffentlicht: (2025)
von: Liang, Jinglin, et al.
Veröffentlicht: (2025)
Beyond Quality: Unlocking Diversity in Ad Headline Generation with Large Language Models
von: Wang, Chang, et al.
Veröffentlicht: (2025)
von: Wang, Chang, et al.
Veröffentlicht: (2025)
From Ghazals to Sonnets: Decoding the Polysemous Expressions of Love Across Languages
von: Ali, Syed Mohammad Sualeh
Veröffentlicht: (2025)
von: Ali, Syed Mohammad Sualeh
Veröffentlicht: (2025)
Loss Masking Is Not Needed in Decoder-only Transformer for Discrete-token-based ASR
von: Chen, Qian, et al.
Veröffentlicht: (2023)
von: Chen, Qian, et al.
Veröffentlicht: (2023)
Better & Faster Large Language Models via Multi-token Prediction
von: Gloeckle, Fabian, et al.
Veröffentlicht: (2024)
von: Gloeckle, Fabian, et al.
Veröffentlicht: (2024)
LBPE: Long-token-first Tokenization to Improve Large Language Models
von: Lian, Haoran, et al.
Veröffentlicht: (2024)
von: Lian, Haoran, et al.
Veröffentlicht: (2024)
Beyond Tokens: Semantic-Aware Speculative Decoding for Efficient Inference by Probing Internal States
von: Dong, Ximing, et al.
Veröffentlicht: (2026)
von: Dong, Ximing, et al.
Veröffentlicht: (2026)
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation
von: Hu, Rui, et al.
Veröffentlicht: (2025)
von: Hu, Rui, et al.
Veröffentlicht: (2025)
Jacobian Scopes: token-level causal attributions in LLMs
von: Liu, Toni J. B., et al.
Veröffentlicht: (2026)
von: Liu, Toni J. B., et al.
Veröffentlicht: (2026)
LLM Probe: Evaluating LLMs for Low-Resource Languages
von: Teklehaymanot, Hailay Kidu, et al.
Veröffentlicht: (2026)
von: Teklehaymanot, Hailay Kidu, et al.
Veröffentlicht: (2026)
Assessing the validity of new paradigmatic complexity measures as criterial features for proficiency in L2 writings in English
von: Mallart, Cyriel, et al.
Veröffentlicht: (2025)
von: Mallart, Cyriel, et al.
Veröffentlicht: (2025)
Beyond Confidence: Adaptive and Coherent Decoding for Diffusion Language Models
von: Chen, Kecheng, et al.
Veröffentlicht: (2025)
von: Chen, Kecheng, et al.
Veröffentlicht: (2025)
Probing Context Localization of Polysemous Words in Pre-trained Language Model Sub-Layers
von: Vijayakumar, Soniya, et al.
Veröffentlicht: (2024)
von: Vijayakumar, Soniya, et al.
Veröffentlicht: (2024)
Calm-Whisper: Reduce Whisper Hallucination On Non-Speech By Calming Crazy Heads Down
von: Wang, Yingzhi, et al.
Veröffentlicht: (2025)
von: Wang, Yingzhi, et al.
Veröffentlicht: (2025)
Moral Reasoning Across Languages: The Critical Role of Low-Resource Languages in LLMs
von: Zhou, Huichi, et al.
Veröffentlicht: (2025)
von: Zhou, Huichi, et al.
Veröffentlicht: (2025)
Semantic-guided Diverse Decoding for Large Language Model
von: Shi, Weijie, et al.
Veröffentlicht: (2025)
von: Shi, Weijie, et al.
Veröffentlicht: (2025)
Beyond the Surface: Probing the Ideological Depth of Large Language Models
von: Kabir, Shariar, et al.
Veröffentlicht: (2025)
von: Kabir, Shariar, et al.
Veröffentlicht: (2025)
Languages in Whisper-Style Speech Encoders Align Both Phonetically and Semantically
von: Shim, Ryan Soh-Eun, et al.
Veröffentlicht: (2025)
von: Shim, Ryan Soh-Eun, et al.
Veröffentlicht: (2025)
WhisperPipe: A Resource-Efficient Streaming Architecture for Real-Time Automatic Speech Recognition
von: Ramezani, Erfan, et al.
Veröffentlicht: (2026)
von: Ramezani, Erfan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
The Limits of Data Scaling: Sub-token Utilization and Acoustic Saturation in Multilingual ASR
von: Liang, Siyu, et al.
Veröffentlicht: (2025) -
Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages
von: Liang, Siyu, et al.
Veröffentlicht: (2025) -
Hybrid Neural-LLM Pipeline for Morphological Glossing in Endangered Language Documentation: A Case Study of Jungar Tuvan
von: Liang, Siyu, et al.
Veröffentlicht: (2026) -
A Sociophonetic Analysis of Racial Bias in Commercial ASR Systems Using the Pacific Northwest English Corpus
von: Scott, Michael, et al.
Veröffentlicht: (2025) -
TEII: Think, Explain, Interact and Iterate with Large Language Models to Solve Cross-lingual Emotion Detection
von: Cheng, Long, et al.
Veröffentlicht: (2024)