Adapting Whisper for Streaming Speech Recognition via Two-Pass Decoding
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Haoran, Song, Xingchen, Fahy, Brendan, Song, Qiaochu, Zhang, Binbin, Peng, Zhendong, Wadhawan, Anshul, Jiang, Denglin, Verma, Apurv, Ramesh, Vinay, Prasad, Srivas, Franceschini, Michele M. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HydraFormer: One Encoder For All Subsampling Rates
by: Xu, Yaoxun, et al.
Published: (2024)
by: Xu, Yaoxun, et al.
Published: (2024)
Adapting Whisper for Code-Switching through Encoding Refining and Language-Aware Decoding
by: Zhao, Jiahui, et al.
Published: (2024)
by: Zhao, Jiahui, et al.
Published: (2024)
Adapting Diarization-Conditioned Whisper for End-to-End Multi-Talker Speech Recognition
by: Kocour, Martin, et al.
Published: (2025)
by: Kocour, Martin, et al.
Published: (2025)
SProBench: Stream Processing Benchmark for High Performance Computing Infrastructure
by: Kulkarni, Apurv Deepak, et al.
Published: (2025)
by: Kulkarni, Apurv Deepak, et al.
Published: (2025)
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation
by: Hu, Rui, et al.
Published: (2025)
by: Hu, Rui, et al.
Published: (2025)
Whisper-SV: Adapting Whisper for Low-data-resource Speaker Verification
by: Zhang, Li, et al.
Published: (2024)
by: Zhang, Li, et al.
Published: (2024)
Adapting Whisper for Lightweight and Efficient Automatic Speech Recognition of Children for On-device Edge Applications
by: Dutta, Satwik, et al.
Published: (2025)
by: Dutta, Satwik, et al.
Published: (2025)
WhispEar: A Bi-directional Framework for Scaling Whispered Speech Conversion via Pseudo-Parallel Whisper Generation
by: Fang, Zihao, et al.
Published: (2026)
by: Fang, Zihao, et al.
Published: (2026)
Approximation Of Logarithm, Factorial And Euler Mascheroni Constant Using Odd Harmonic Series
by: Wadhawan, Narinder Kumar, et al.
Published: (2025)
by: Wadhawan, Narinder Kumar, et al.
Published: (2025)
Representation Of Integers By Sum Of Three Cubes, A New Approach Based On Seed Equation
by: Wadhawan, Narinder Kumar, et al.
Published: (2025)
by: Wadhawan, Narinder Kumar, et al.
Published: (2025)
WhisperRT -- Turning Whisper into a Causal Streaming Model
by: Krichli, Tomer, et al.
Published: (2025)
by: Krichli, Tomer, et al.
Published: (2025)
Simul-Whisper: Attention-Guided Streaming Whisper with Truncation Detection
by: Wang, Haoyu, et al.
Published: (2024)
by: Wang, Haoyu, et al.
Published: (2024)
BrainWhisperer: Leveraging Large-Scale ASR Models for Neural Speech Decoding
by: Boccato, Tommaso, et al.
Published: (2026)
by: Boccato, Tommaso, et al.
Published: (2026)
U2++ MoE: Scaling 4.7x parameters with minimal impact on RTF
by: Song, Xingchen, et al.
Published: (2024)
by: Song, Xingchen, et al.
Published: (2024)
Adapting Whisper for Parameter-efficient Code-Switching Speech Recognition via Soft Prompt Tuning
by: Yang, Hongli, et al.
Published: (2025)
by: Yang, Hongli, et al.
Published: (2025)
TouchASP: Elastic Automatic Speech Perception that Everyone Can Touch
by: Song, Xingchen, et al.
Published: (2024)
by: Song, Xingchen, et al.
Published: (2024)
Watermarking Degrades Alignment in Language Models: Analysis and Mitigation
by: Verma, Apurv, et al.
Published: (2025)
by: Verma, Apurv, et al.
Published: (2025)
Evaluating Performance and Bias of Negative Sampling in Large-Scale Sequential Recommendation Models
by: Prakash, Arushi, et al.
Published: (2024)
by: Prakash, Arushi, et al.
Published: (2024)
Adaptive Modified Weak Galerkin Method for Obstacle Problem
by: Wadhawan, Tanvi
Published: (2025)
by: Wadhawan, Tanvi
Published: (2025)
Factors Influence Dementia Treatment‐Seeking Behavior among Asian Indian Americans
by: Anju Wadhawan
Published: (2025)
by: Anju Wadhawan
Published: (2025)
Mapping state interventions for international labour migration from India
by: Neha Wadhawan
Published: (2020)
by: Neha Wadhawan
Published: (2020)
L'assurance maladie en Inde: une réforme s'impose
by: S. K. Wadhawan
Published: (1987)
by: S. K. Wadhawan
Published: (1987)
Health insurance in India: the case for reform
by: S. K. Wadhawan
Published: (1987)
by: S. K. Wadhawan
Published: (1987)
Theory Of Smoothpath where it connects to various branches
by: Sarangi, Apurv
Published: (2025)
by: Sarangi, Apurv
Published: (2025)
Differentially Private High Dimensional Bandits
by: Shukla, Apurv
Published: (2024)
by: Shukla, Apurv
Published: (2024)
WhisperPipe: A Resource-Efficient Streaming Architecture for Real-Time Automatic Speech Recognition
by: Ramezani, Erfan, et al.
Published: (2026)
by: Ramezani, Erfan, et al.
Published: (2026)
WhisperD: Dementia Speech Recognition and Filler Word Detection with Whisper
by: Akinrintoyo, Emmanuel, et al.
Published: (2025)
by: Akinrintoyo, Emmanuel, et al.
Published: (2025)
Whisper-GPT: A Hybrid Representation Audio Large Language Model
by: Verma, Prateek
Published: (2024)
by: Verma, Prateek
Published: (2024)
Whisper-CD: Accurate Long-Form Speech Recognition using Multi-Negative Contrastive Decoding
by: Ahn, Hoseong, et al.
Published: (2026)
by: Ahn, Hoseong, et al.
Published: (2026)
LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
by: Pradhan, Anu, et al.
Published: (2025)
by: Pradhan, Anu, et al.
Published: (2025)
VERSE: Virtual-Gradient Aware Streaming Lifelong Learning with Anytime Inference
by: Banerjee, Soumya, et al.
Published: (2023)
by: Banerjee, Soumya, et al.
Published: (2023)
Adopting Whisper for Confidence Estimation
by: Aggarwal, Vaibhav, et al.
Published: (2025)
by: Aggarwal, Vaibhav, et al.
Published: (2025)
Evaluation of the Efficacy of Telemedicine for Pre‐Anesthetic Check‐Up in Pediatric Patients Undergoing Elective Surgery: A Pilot Randomized Controlled Trial
by: Yukti Shah, et al.
Published: (2025)
by: Yukti Shah, et al.
Published: (2025)
WhisperMask: A Noise Suppressive Mask-Type Microphone for Whisper Speech
by: Hiraki, Hirotaka, et al.
Published: (2024)
by: Hiraki, Hirotaka, et al.
Published: (2024)
Whispy: Adapting STT Whisper Models to Real-Time Environments
by: Bevilacqua, Antonio, et al.
Published: (2024)
by: Bevilacqua, Antonio, et al.
Published: (2024)
One Pass Streaming Algorithm for Super Long Token Attention Approximation in Sublinear Space
by: Addanki, Raghav, et al.
Published: (2023)
by: Addanki, Raghav, et al.
Published: (2023)
Scientific Calculator With The Aid Of Geometry And Based Upon It, A Mechanical Calculator
by: Wadhawan, Narinder Kumar
Published: (2025)
by: Wadhawan, Narinder Kumar
Published: (2025)
Numerical Approximation of Lambert W Function For Real Values By Unique Method of Quadratic Approximation
by: Wadhawan, Narinder Kumar
Published: (2025)
by: Wadhawan, Narinder Kumar
Published: (2025)
Numerical Approximation In Real Domain Of Special Function Of Product Of A Variable And Its Double Exponential
by: Wadhawan, Narinder Kumar
Published: (2025)
by: Wadhawan, Narinder Kumar
Published: (2025)
Dynamical 4-D Gauss-Bonnet action from matter-graviton interactions in a curved background
by: Keer, Apurv, et al.
Published: (2026)
by: Keer, Apurv, et al.
Published: (2026)
Similar Items
-
HydraFormer: One Encoder For All Subsampling Rates
by: Xu, Yaoxun, et al.
Published: (2024) -
Adapting Whisper for Code-Switching through Encoding Refining and Language-Aware Decoding
by: Zhao, Jiahui, et al.
Published: (2024) -
Adapting Diarization-Conditioned Whisper for End-to-End Multi-Talker Speech Recognition
by: Kocour, Martin, et al.
Published: (2025) -
SProBench: Stream Processing Benchmark for High Performance Computing Infrastructure
by: Kulkarni, Apurv Deepak, et al.
Published: (2025) -
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation
by: Hu, Rui, et al.
Published: (2025)