Fine-tuning Whisper for Pashto ASR: strategies and scale
Fuente:
arXiv
Salvato in:
| Autore principale: | Rahman, Hanif |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Benchmarking Multilingual Speech Models on Pashto: Zero-Shot ASR, Script Failure, and Cross-Domain Evaluation
di: Rahman, Hanif
Pubblicazione: (2026)
di: Rahman, Hanif
Pubblicazione: (2026)
PashtoTTS-Bench: automated screening for low-resource non-Latin-script text-to-speech
di: Rahman, Hanif
Pubblicazione: (2026)
di: Rahman, Hanif
Pubblicazione: (2026)
PashtoCorp: A 1.25-Billion-Word Corpus, Evaluation Suite, and Reproducible Pipeline for Low-Resource Language Development
di: Rahman, Hanif
Pubblicazione: (2026)
di: Rahman, Hanif
Pubblicazione: (2026)
Pashto Common Voice: Building the First Open Speech Corpus for a 60-Million-Speaker Low-Resource Language
di: Rahman, Hanif, et al.
Pubblicazione: (2026)
di: Rahman, Hanif, et al.
Pubblicazione: (2026)
Extending Whisper with prompt tuning to target-speaker ASR
di: Ma, Hao, et al.
Pubblicazione: (2023)
di: Ma, Hao, et al.
Pubblicazione: (2023)
Whispering in Amharic: Fine-tuning Whisper for Low-resource Language
di: Gete, Dawit Ketema, et al.
Pubblicazione: (2025)
di: Gete, Dawit Ketema, et al.
Pubblicazione: (2025)
On the Role of Encoder Depth: Pruning Whisper and LoRA Fine-Tuning in SLAM-ASR
di: Kolluri, Ganesh Pavan Kartikeya Bharadwaj, et al.
Pubblicazione: (2026)
di: Kolluri, Ganesh Pavan Kartikeya Bharadwaj, et al.
Pubblicazione: (2026)
Fine-tuning Whisper on Low-Resource Languages for Real-World Applications
di: Timmel, Vincenzo, et al.
Pubblicazione: (2024)
di: Timmel, Vincenzo, et al.
Pubblicazione: (2024)
Overcoming Data Scarcity in Multi-Dialectal Arabic ASR via Whisper Fine-Tuning
di: Özyilmaz, Ömer Tarik, et al.
Pubblicazione: (2025)
di: Özyilmaz, Ömer Tarik, et al.
Pubblicazione: (2025)
Deepfake Word Detection by Next-token Prediction using Fine-tuned Whisper
di: Tran, Hoan My, et al.
Pubblicazione: (2026)
di: Tran, Hoan My, et al.
Pubblicazione: (2026)
Improving the Inclusivity of Dutch Speech Recognition by Fine-tuning Whisper on the JASMIN-CGN Corpus
di: Shekoufandeh, Golshid, et al.
Pubblicazione: (2025)
di: Shekoufandeh, Golshid, et al.
Pubblicazione: (2025)
Whisper: Courtside Edition Enhancing ASR Performance Through LLM-Driven Context Generation
di: Ron, Yonathan, et al.
Pubblicazione: (2026)
di: Ron, Yonathan, et al.
Pubblicazione: (2026)
LoRA-Whisper: Parameter-Efficient and Extensible Multilingual ASR
di: Song, Zheshu, et al.
Pubblicazione: (2024)
di: Song, Zheshu, et al.
Pubblicazione: (2024)
Adaptability of ASR Models on Low-Resource Language: A Comparative Study of Whisper and Wav2Vec-BERT on Bangla
di: Ridoy, Md Sazzadul Islam, et al.
Pubblicazione: (2025)
di: Ridoy, Md Sazzadul Islam, et al.
Pubblicazione: (2025)
Fast Streaming Transducer ASR Prototyping via Knowledge Distillation with Whisper
di: Thorbecke, Iuliia, et al.
Pubblicazione: (2024)
di: Thorbecke, Iuliia, et al.
Pubblicazione: (2024)
Quantizing Whisper-small: How design choices affect ASR performance
di: Söhler, Arthur, et al.
Pubblicazione: (2025)
di: Söhler, Arthur, et al.
Pubblicazione: (2025)
WhisperKit: On-device Real-time ASR with Billion-Scale Transformers
di: Orhon, Atila, et al.
Pubblicazione: (2025)
di: Orhon, Atila, et al.
Pubblicazione: (2025)
From Scarcity to Scale: A Release-Level Analysis of the Pashto Common Voice Dataset
di: Jahani, Jandad, et al.
Pubblicazione: (2026)
di: Jahani, Jandad, et al.
Pubblicazione: (2026)
Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages
di: Liang, Siyu, et al.
Pubblicazione: (2025)
di: Liang, Siyu, et al.
Pubblicazione: (2025)
Evaluating ASR robustness to spontaneous speech errors: A study of WhisperX using a Speech Error Database
di: Alderete, John, et al.
Pubblicazione: (2025)
di: Alderete, John, et al.
Pubblicazione: (2025)
Fine-tuning large language models for domain adaptation: Exploration of training strategies, scaling, model merging and synergistic capabilities
di: Lu, Wei, et al.
Pubblicazione: (2024)
di: Lu, Wei, et al.
Pubblicazione: (2024)
Noise-Robust AV-ASR Using Visual Features Both in the Whisper Encoder and Decoder
di: Li, Zhengyang, et al.
Pubblicazione: (2026)
di: Li, Zhengyang, et al.
Pubblicazione: (2026)
Polyglot-Lion: Efficient Multilingual ASR for Singapore via Balanced Fine-Tuning of Qwen3-ASR
di: Dang, Quy-Anh, et al.
Pubblicazione: (2026)
di: Dang, Quy-Anh, et al.
Pubblicazione: (2026)
Towards Rehearsal-Free Multilingual ASR: A LoRA-based Case Study on Whisper
di: Xu, Tianyi, et al.
Pubblicazione: (2024)
di: Xu, Tianyi, et al.
Pubblicazione: (2024)
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
di: Zhou, Huichi, et al.
Pubblicazione: (2025)
di: Zhou, Huichi, et al.
Pubblicazione: (2025)
Fine-tuned Large Language Models (LLMs): Improved Prompt Injection Attacks Detection
di: Rahman, Md Abdur, et al.
Pubblicazione: (2024)
di: Rahman, Md Abdur, et al.
Pubblicazione: (2024)
Be Careful When Fine-tuning On Open-Source LLMs: Your Fine-tuning Data Could Be Secretly Stolen!
di: Zhang, Zhexin, et al.
Pubblicazione: (2025)
di: Zhang, Zhexin, et al.
Pubblicazione: (2025)
Munsit at NADI 2025 Shared Task 2: Pushing the Boundaries of Multidialectal Arabic ASR with Weakly Supervised Pretraining and Continual Supervised Fine-tuning
di: Salhab, Mahmoud, et al.
Pubblicazione: (2025)
di: Salhab, Mahmoud, et al.
Pubblicazione: (2025)
A Comparative Study of LLM-based ASR and Whisper in Low Resource and Code Switching Scenario
di: Song, Zheshu, et al.
Pubblicazione: (2024)
di: Song, Zheshu, et al.
Pubblicazione: (2024)
Evaluating Standard and Dialectal Frisian ASR: Multilingual Fine-tuning and Language Identification for Improved Low-resource Performance
di: Amooie, Reihaneh, et al.
Pubblicazione: (2025)
di: Amooie, Reihaneh, et al.
Pubblicazione: (2025)
Whisper Turns Stronger: Augmenting Wav2Vec 2.0 for Superior ASR in Low-Resource Languages
di: Anidjar, Or Haim, et al.
Pubblicazione: (2024)
di: Anidjar, Or Haim, et al.
Pubblicazione: (2024)
Fine-tuning Done Right in Model Editing
di: Yang, Wanli, et al.
Pubblicazione: (2025)
di: Yang, Wanli, et al.
Pubblicazione: (2025)
AI Generated Text Detection Using Instruction Fine-tuned Large Language and Transformer-Based Models
di: Guggilla, Chinnappa, et al.
Pubblicazione: (2025)
di: Guggilla, Chinnappa, et al.
Pubblicazione: (2025)
Calm-Whisper: Reduce Whisper Hallucination On Non-Speech By Calming Crazy Heads Down
di: Wang, Yingzhi, et al.
Pubblicazione: (2025)
di: Wang, Yingzhi, et al.
Pubblicazione: (2025)
Memorization of Named Entities in Fine-tuned BERT Models
di: Diera, Andor, et al.
Pubblicazione: (2022)
di: Diera, Andor, et al.
Pubblicazione: (2022)
Fine-tuning Large Language Models with Sequential Instructions
di: Hu, Hanxu, et al.
Pubblicazione: (2024)
di: Hu, Hanxu, et al.
Pubblicazione: (2024)
Sparse Matrix in Large Language Model Fine-tuning
di: He, Haoze, et al.
Pubblicazione: (2024)
di: He, Haoze, et al.
Pubblicazione: (2024)
Learning or Self-aligning? Rethinking Instruction Fine-tuning
di: Ren, Mengjie, et al.
Pubblicazione: (2024)
di: Ren, Mengjie, et al.
Pubblicazione: (2024)
Data-efficient LLM Fine-tuning for Code Generation
di: Lv, Weijie, et al.
Pubblicazione: (2025)
di: Lv, Weijie, et al.
Pubblicazione: (2025)
LoFiT: Localized Fine-tuning on LLM Representations
di: Yin, Fangcong, et al.
Pubblicazione: (2024)
di: Yin, Fangcong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Benchmarking Multilingual Speech Models on Pashto: Zero-Shot ASR, Script Failure, and Cross-Domain Evaluation
di: Rahman, Hanif
Pubblicazione: (2026) -
PashtoTTS-Bench: automated screening for low-resource non-Latin-script text-to-speech
di: Rahman, Hanif
Pubblicazione: (2026) -
PashtoCorp: A 1.25-Billion-Word Corpus, Evaluation Suite, and Reproducible Pipeline for Low-Resource Language Development
di: Rahman, Hanif
Pubblicazione: (2026) -
Pashto Common Voice: Building the First Open Speech Corpus for a 60-Million-Speaker Low-Resource Language
di: Rahman, Hanif, et al.
Pubblicazione: (2026) -
Extending Whisper with prompt tuning to target-speaker ASR
di: Ma, Hao, et al.
Pubblicazione: (2023)