Hypernetworks for Personalizing ASR to Atypical Speech
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Müller-Eberstein, Max, Yee, Dianna, Yang, Karren, Mantena, Gautam Varma, Lea, Colin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PolyPythias: Stability and Outliers across Fifty Language Model Pre-Training Runs
von: van der Wal, Oskar, et al.
Veröffentlicht: (2025)
von: van der Wal, Oskar, et al.
Veröffentlicht: (2025)
Voice Quality Dimensions as Interpretable Primitives for Speaking Style for Atypical Speech and Affect
von: Narain, Jaya, et al.
Veröffentlicht: (2025)
von: Narain, Jaya, et al.
Veröffentlicht: (2025)
HyperTTS: Parameter Efficient Adaptation in Text to Speech using Hypernetworks
von: Li, Yingting, et al.
Veröffentlicht: (2024)
von: Li, Yingting, et al.
Veröffentlicht: (2024)
Interpretable Next-token Prediction via the Generalized Induction Head
von: Kim, Eunji, et al.
Veröffentlicht: (2024)
von: Kim, Eunji, et al.
Veröffentlicht: (2024)
Zhyper: Factorized Hypernetworks for Conditioned LLM Fine-Tuning
von: Abdalla, M. H. I., et al.
Veröffentlicht: (2025)
von: Abdalla, M. H. I., et al.
Veröffentlicht: (2025)
Moonshine v2: Ergodic Streaming Encoder ASR for Latency-Critical Speech Applications
von: Kudlur, Manjunath, et al.
Veröffentlicht: (2026)
von: Kudlur, Manjunath, et al.
Veröffentlicht: (2026)
HyperEdit: Unlocking Instruction-based Text Editing in LLMs via Hypernetworks
von: Zeng, Yiming, et al.
Veröffentlicht: (2025)
von: Zeng, Yiming, et al.
Veröffentlicht: (2025)
Codec-ASR: Training Performant Automatic Speech Recognition Systems with Discrete Speech Representations
von: Dhawan, Kunal, et al.
Veröffentlicht: (2024)
von: Dhawan, Kunal, et al.
Veröffentlicht: (2024)
HyperSteer: Activation Steering at Scale with Hypernetworks
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
PropMEND: Hypernetworks for Knowledge Propagation in LLMs
von: Liu, Zeyu Leo, et al.
Veröffentlicht: (2025)
von: Liu, Zeyu Leo, et al.
Veröffentlicht: (2025)
Learn-to-learn on Arbitrary Textual Conditioning: A Hypernetwork-Driven Meta-Gated LLM
von: Ji, Luo, et al.
Veröffentlicht: (2026)
von: Ji, Luo, et al.
Veröffentlicht: (2026)
RO-N3WS: Enhancing Generalization in Low-Resource ASR with Diverse Romanian Speech Benchmarks
von: Diaconu, Alexandra, et al.
Veröffentlicht: (2026)
von: Diaconu, Alexandra, et al.
Veröffentlicht: (2026)
HyperDAS: Towards Automating Mechanistic Interpretability with Hypernetworks
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
Affect Models Have Weak Generalizability to Atypical Speech
von: Narain, Jaya, et al.
Veröffentlicht: (2025)
von: Narain, Jaya, et al.
Veröffentlicht: (2025)
Leveraging Allophony in Self-Supervised Speech Models for Atypical Pronunciation Assessment
von: Choi, Kwanghee, et al.
Veröffentlicht: (2025)
von: Choi, Kwanghee, et al.
Veröffentlicht: (2025)
When Meanings Meet: Investigating the Emergence and Quality of Shared Concept Spaces during Multilingual Language Model Training
von: Körner, Felicia, et al.
Veröffentlicht: (2026)
von: Körner, Felicia, et al.
Veröffentlicht: (2026)
HyperAdaLoRA: Accelerating LoRA Rank Allocation During Training via Hypernetworks without Sacrificing Performance
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
OLMoASR: Open Models and Data for Training Robust Speech Recognition Models
von: Ngo, Huong, et al.
Veröffentlicht: (2025)
von: Ngo, Huong, et al.
Veröffentlicht: (2025)
New Insights into Optimal Alignment of Acoustic and Linguistic Representations for Knowledge Transfer in ASR
von: Lu, Xugang, et al.
Veröffentlicht: (2025)
von: Lu, Xugang, et al.
Veröffentlicht: (2025)
In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions
von: Fan, Xulin, et al.
Veröffentlicht: (2026)
von: Fan, Xulin, et al.
Veröffentlicht: (2026)
HYPEROFA: Expanding LLM Vocabulary to New Languages via Hypernetwork-Based Embedding Initialization
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
Flavors of Moonshine: Tiny Specialized ASR Models for Edge Devices
von: King, Evan, et al.
Veröffentlicht: (2025)
von: King, Evan, et al.
Veröffentlicht: (2025)
Training-Free Cross-Lingual Dysarthria Severity Assessment via Phonological Subspace Analysis in Self-Supervised Speech Representations
von: Muller, Bernard, et al.
Veröffentlicht: (2026)
von: Muller, Bernard, et al.
Veröffentlicht: (2026)
Exploring Pathological Speech Quality Assessment with ASR-Powered Wav2Vec2 in Data-Scarce Context
von: Nguyen, Tuan, et al.
Veröffentlicht: (2024)
von: Nguyen, Tuan, et al.
Veröffentlicht: (2024)
EuroSpeech: A Multilingual Speech Corpus
von: Pfisterer, Samuel, et al.
Veröffentlicht: (2025)
von: Pfisterer, Samuel, et al.
Veröffentlicht: (2025)
SnakModel: Lessons Learned from Training an Open Danish Large Language Model
von: Zhang, Mike, et al.
Veröffentlicht: (2024)
von: Zhang, Mike, et al.
Veröffentlicht: (2024)
Hypernetworks for Model-Heterogeneous Personalized Federated Learning
von: Zhang, Chen, et al.
Veröffentlicht: (2025)
von: Zhang, Chen, et al.
Veröffentlicht: (2025)
Skipformer: A Skip-and-Recover Strategy for Efficient Speech Recognition
von: Zhu, Wenjing, et al.
Veröffentlicht: (2024)
von: Zhu, Wenjing, et al.
Veröffentlicht: (2024)
Toward expanding the scope of radiology report summarization to multiple anatomies and modalities
von: Chen, Zhihong, et al.
Veröffentlicht: (2022)
von: Chen, Zhihong, et al.
Veröffentlicht: (2022)
Unknown Unknowns: Why Hidden Intentions in LLMs Evade Detection
von: Srivastav, Devansh, et al.
Veröffentlicht: (2026)
von: Srivastav, Devansh, et al.
Veröffentlicht: (2026)
Confronting LLMs with Traditional ML: Rethinking the Fairness of Large Language Models in Tabular Classifications
von: Liu, Yanchen, et al.
Veröffentlicht: (2023)
von: Liu, Yanchen, et al.
Veröffentlicht: (2023)
Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization
von: Ma, Qiyao, et al.
Veröffentlicht: (2026)
von: Ma, Qiyao, et al.
Veröffentlicht: (2026)
Are More Tokens Rational? Inference-Time Scaling in Language Models as Adaptive Resource Rationality
von: Hu, Zhimin, et al.
Veröffentlicht: (2026)
von: Hu, Zhimin, et al.
Veröffentlicht: (2026)
Diagnosing Hate Speech Classification: Where Do Humans and Machines Disagree, and Why?
von: Yang, Xilin
Veröffentlicht: (2024)
von: Yang, Xilin
Veröffentlicht: (2024)
Examining Test-Time Adaptation for Personalized Child Speech Recognition
von: Shi, Zhonghao, et al.
Veröffentlicht: (2024)
von: Shi, Zhonghao, et al.
Veröffentlicht: (2024)
Anatomy of Industrial Scale Multilingual ASR
von: Ramirez, Francis McCann, et al.
Veröffentlicht: (2024)
von: Ramirez, Francis McCann, et al.
Veröffentlicht: (2024)
Beyond Transcription: Mechanistic Interpretability in ASR
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs
von: Zhao, Siyan, et al.
Veröffentlicht: (2025)
von: Zhao, Siyan, et al.
Veröffentlicht: (2025)
Predicting Compact Phrasal Rewrites with Large Language Models for ASR Post Editing
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
SemCAFE: When Named Entities make the Difference Assessing Web Source Reliability through Entity-level Analytics
von: Shahi, Gautam Kishore, et al.
Veröffentlicht: (2025)
von: Shahi, Gautam Kishore, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PolyPythias: Stability and Outliers across Fifty Language Model Pre-Training Runs
von: van der Wal, Oskar, et al.
Veröffentlicht: (2025) -
Voice Quality Dimensions as Interpretable Primitives for Speaking Style for Atypical Speech and Affect
von: Narain, Jaya, et al.
Veröffentlicht: (2025) -
HyperTTS: Parameter Efficient Adaptation in Text to Speech using Hypernetworks
von: Li, Yingting, et al.
Veröffentlicht: (2024) -
Interpretable Next-token Prediction via the Generalized Induction Head
von: Kim, Eunji, et al.
Veröffentlicht: (2024) -
Zhyper: Factorized Hypernetworks for Conditioned LLM Fine-Tuning
von: Abdalla, M. H. I., et al.
Veröffentlicht: (2025)