Continual Learning with Embedding Layer Surgery and Task-wise Beam Search using Whisper
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kwok, Chin Yuen, Yip, Jia Qi, Chng, Eng Siong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improving Synthetic Data Training for Contextual Biasing Models with a Keyword-Aware Cost Function
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2025)
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2025)
Continual Learning Optimizations for Auto-regressive Decoder of Multilingual ASR systems
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2024)
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2024)
Efficient Trie-based Biasing using K-step Prediction for Rare Word Recognition
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2025)
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2025)
Speech Enhancement Using Continuous Embeddings of Neural Audio Codec
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
Speech Separation using Neural Audio Codecs with Embedding Loss
von: Yip, Jia Qi, et al.
Veröffentlicht: (2024)
von: Yip, Jia Qi, et al.
Veröffentlicht: (2024)
Bona fide Cross Testing Reveals Weak Spot in Audio Deepfake Detection Systems
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2025)
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2025)
Hierarchical Self-Supervised Representation Learning for Depression Detection from Speech
von: Li, Yuxin, et al.
Veröffentlicht: (2025)
von: Li, Yuxin, et al.
Veröffentlicht: (2025)
Large Language Models Meet Contrastive Learning: Zero-Shot Emotion Recognition Across Languages
von: Zou, Heqing, et al.
Veröffentlicht: (2025)
von: Zou, Heqing, et al.
Veröffentlicht: (2025)
DepFlow: Disentangled Speech Generation to Mitigate Semantic Bias in Depression Detection
von: Li, Yuxin, et al.
Veröffentlicht: (2026)
von: Li, Yuxin, et al.
Veröffentlicht: (2026)
Listen Again and Choose the Right Answer: A New Paradigm for Automatic Speech Recognition with Large Language Models
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
Chronological Thinking in Full-Duplex Spoken Dialogue Language Models
von: Wu, Donghang, et al.
Veröffentlicht: (2025)
von: Wu, Donghang, et al.
Veröffentlicht: (2025)
BaldWhisper: Faster Whisper with Head Shearing and Layer Merging
von: Sy, Yaya, et al.
Veröffentlicht: (2025)
von: Sy, Yaya, et al.
Veröffentlicht: (2025)
Zero-shot Context Biasing with Trie-based Decoding using Synthetic Multi-Pronunciation
von: Liu, Changsong, et al.
Veröffentlicht: (2025)
von: Liu, Changsong, et al.
Veröffentlicht: (2025)
Layer-wise Regularized Dropout for Neural Language Models
von: Ni, Shiwen, et al.
Veröffentlicht: (2024)
von: Ni, Shiwen, et al.
Veröffentlicht: (2024)
Evolutionary Feature-wise Thresholding for Binary Representation of NLP Embeddings
von: Sinha, Soumen, et al.
Veröffentlicht: (2025)
von: Sinha, Soumen, et al.
Veröffentlicht: (2025)
Layer-wise Positional Bias in Short-Context Language Modeling
von: Rahimi, Maryam, et al.
Veröffentlicht: (2026)
von: Rahimi, Maryam, et al.
Veröffentlicht: (2026)
Unsupervised Layer-wise Score Aggregation for Textual OOD Detection
von: Darrin, Maxime, et al.
Veröffentlicht: (2023)
von: Darrin, Maxime, et al.
Veröffentlicht: (2023)
MULTI-Bench: A Multi-Turn Interactive Benchmark for Assessing Emotional Intelligence ability of Spoken Dialogue Models
von: Deng, Yayue, et al.
Veröffentlicht: (2025)
von: Deng, Yayue, et al.
Veröffentlicht: (2025)
Efficient Layer-wise LLM Fine-tuning for Revision Intention Prediction
von: Liu, Zhexiong, et al.
Veröffentlicht: (2025)
von: Liu, Zhexiong, et al.
Veröffentlicht: (2025)
Whisper Finetuning on Nepali Language
von: Rijal, Sanjay, et al.
Veröffentlicht: (2024)
von: Rijal, Sanjay, et al.
Veröffentlicht: (2024)
GenTranslate: Large Language Models are Generative Multilingual Speech and Machine Translators
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
Mechanistic Steering of LLMs Reveals Layer-wise Feature Vulnerabilities in Adversarial Settings
von: Das, Nilanjana, et al.
Veröffentlicht: (2026)
von: Das, Nilanjana, et al.
Veröffentlicht: (2026)
Punctuation Restoration for Singaporean Spoken Languages: English, Malay, and Mandarin
von: Rao, Abhinav, et al.
Veröffentlicht: (2022)
von: Rao, Abhinav, et al.
Veröffentlicht: (2022)
Unlabeled Debiasing in Downstream Tasks via Class-wise Low Variance Regularization
von: Masoudian, Shahed, et al.
Veröffentlicht: (2024)
von: Masoudian, Shahed, et al.
Veröffentlicht: (2024)
Search or Accelerate: Confidence-Switched Position Beam Search for Diffusion Language Models
von: Cao, Mingyu, et al.
Veröffentlicht: (2026)
von: Cao, Mingyu, et al.
Veröffentlicht: (2026)
Self-Taught Recognizer: Toward Unsupervised Adaptation for Speech Foundation Models
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
Improve Decoding Factuality by Token-wise Cross Layer Entropy of Large Language Models
von: Wu, Jialiang, et al.
Veröffentlicht: (2025)
von: Wu, Jialiang, et al.
Veröffentlicht: (2025)
AuthorMix: Modular Authorship Style Transfer via Layer-wise Adapter Mixing
von: Thillainathan, Sarubi, et al.
Veröffentlicht: (2026)
von: Thillainathan, Sarubi, et al.
Veröffentlicht: (2026)
Bi-directional Context-Enhanced Speech Large Language Models for Multilingual Conversational ASR
von: Peng, Yizhou, et al.
Veröffentlicht: (2025)
von: Peng, Yizhou, et al.
Veröffentlicht: (2025)
Optimizing Sentence Embedding with Pseudo-Labeling and Model Ensembles: A Hierarchical Framework for Enhanced NLP Tasks
von: Liu, Ziwei, et al.
Veröffentlicht: (2025)
von: Liu, Ziwei, et al.
Veröffentlicht: (2025)
Template-assisted Contrastive Learning of Task-oriented Dialogue Sentence Embeddings
von: Oh, Minsik, et al.
Veröffentlicht: (2023)
von: Oh, Minsik, et al.
Veröffentlicht: (2023)
Where meaning lives: Layer-wise accessibility of psycholinguistic features in encoder and decoder language models
von: Tikhomirova, Taisiia, et al.
Veröffentlicht: (2026)
von: Tikhomirova, Taisiia, et al.
Veröffentlicht: (2026)
LRP4RAG: Detecting Hallucinations in Retrieval-Augmented Generation via Layer-wise Relevance Propagation
von: Hu, Haichuan, et al.
Veröffentlicht: (2024)
von: Hu, Haichuan, et al.
Veröffentlicht: (2024)
PropRAG: Guiding Retrieval with Beam Search over Proposition Paths
von: Wang, Jingjin, et al.
Veröffentlicht: (2025)
von: Wang, Jingjin, et al.
Veröffentlicht: (2025)
Automatic Dataset Generation for Knowledge Intensive Question Answering Tasks
von: Yuen, Sizhe, et al.
Veröffentlicht: (2025)
von: Yuen, Sizhe, et al.
Veröffentlicht: (2025)
Exploring RWKV for Sentence Embeddings: Layer-wise Analysis and Baseline Comparison for Semantic Similarity
von: Pan, Xinghan
Veröffentlicht: (2025)
von: Pan, Xinghan
Veröffentlicht: (2025)
Whispering Context: Distilling Syntax and Semantics for Long Speech Transcripts
von: Altinok, Duygu
Veröffentlicht: (2025)
von: Altinok, Duygu
Veröffentlicht: (2025)
Whispering Experts: Neural Interventions for Toxicity Mitigation in Language Models
von: Suau, Xavier, et al.
Veröffentlicht: (2024)
von: Suau, Xavier, et al.
Veröffentlicht: (2024)
LEAP: Layer-wise Exit-Aware Pretraining for Efficient Transformer Inference
von: Kapadia, Shashank, et al.
Veröffentlicht: (2026)
von: Kapadia, Shashank, et al.
Veröffentlicht: (2026)
PEML: Parameter-efficient Multi-Task Learning with Optimized Continuous Prompts
von: Chowdhury, Anjir Ahmed, et al.
Veröffentlicht: (2026)
von: Chowdhury, Anjir Ahmed, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Improving Synthetic Data Training for Contextual Biasing Models with a Keyword-Aware Cost Function
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2025) -
Continual Learning Optimizations for Auto-regressive Decoder of Multilingual ASR systems
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2024) -
Efficient Trie-based Biasing using K-step Prediction for Rare Word Recognition
von: Kwok, Chin Yuen, et al.
Veröffentlicht: (2025) -
Speech Enhancement Using Continuous Embeddings of Neural Audio Codec
von: Li, Haoyang, et al.
Veröffentlicht: (2025) -
Speech Separation using Neural Audio Codecs with Embedding Loss
von: Yip, Jia Qi, et al.
Veröffentlicht: (2024)