Voice Conversion for Lombard Speaking Style with Implicit and Explicit Acoustic Feature Conditioning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Woszczyk, Dominika, Ribeiro, Manuel Sam, Merritt, Thomas, Korzekwa, Daniel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Factor-Conditioned Speaking-Style Captioning
von: Ando, Atsushi, et al.
Veröffentlicht: (2024)
von: Ando, Atsushi, et al.
Veröffentlicht: (2024)
Grapheme-Coherent Phonemic and Prosodic Annotation of Speech by Implicit and Explicit Grapheme Conditioning
von: Ohnaka, Hien, et al.
Veröffentlicht: (2025)
von: Ohnaka, Hien, et al.
Veröffentlicht: (2025)
StyleSinger: Style Transfer for Out-of-Domain Singing Voice Synthesis
von: Zhang, Yu, et al.
Veröffentlicht: (2023)
von: Zhang, Yu, et al.
Veröffentlicht: (2023)
ClaritySpeech: Dementia Obfuscation in Speech
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2025)
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2025)
Prosody-Driven Privacy-Preserving Dementia Detection
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2024)
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2024)
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style
von: Kang, Wonjune, et al.
Veröffentlicht: (2025)
von: Kang, Wonjune, et al.
Veröffentlicht: (2025)
TCSinger: Zero-Shot Singing Voice Synthesis with Style Transfer and Multi-Level Style Control
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias
von: Satish, Shree Harsha Bokkahalli, et al.
Veröffentlicht: (2025)
von: Satish, Shree Harsha Bokkahalli, et al.
Veröffentlicht: (2025)
Maestro-EVC: Controllable Emotional Voice Conversion Guided by References and Explicit Prosody
von: Yoon, Jinsung, et al.
Veröffentlicht: (2025)
von: Yoon, Jinsung, et al.
Veröffentlicht: (2025)
Revisiting Acoustic Features for Robust ASR
von: Shah, Muhammad A., et al.
Veröffentlicht: (2024)
von: Shah, Muhammad A., et al.
Veröffentlicht: (2024)
AdaptVC: High Quality Voice Conversion with Adaptive Learning
von: Kim, Jaehun, et al.
Veröffentlicht: (2025)
von: Kim, Jaehun, et al.
Veröffentlicht: (2025)
Stepback: Enhanced Disentanglement for Voice Conversion via Multi-Task Learning
von: Yang, Qian, et al.
Veröffentlicht: (2025)
von: Yang, Qian, et al.
Veröffentlicht: (2025)
StableVC: Style Controllable Zero-Shot Voice Conversion with Conditional Flow Matching
von: Yao, Jixun, et al.
Veröffentlicht: (2024)
von: Yao, Jixun, et al.
Veröffentlicht: (2024)
Voice Conversion Improves Cross-Domain Robustness for Spoken Arabic Dialect Identification
von: Abdullah, Badr M., et al.
Veröffentlicht: (2025)
von: Abdullah, Badr M., et al.
Veröffentlicht: (2025)
Conan: A Chunkwise Online Network for Zero-Shot Adaptive Voice Conversion
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
Noro: Noise-Robust One-shot Voice Conversion with Hidden Speaker Representation Learning
von: He, Haorui, et al.
Veröffentlicht: (2024)
von: He, Haorui, et al.
Veröffentlicht: (2024)
SEF-VC: Speaker Embedding Free Zero-Shot Voice Conversion with Cross Attention
von: Li, Junjie, et al.
Veröffentlicht: (2023)
von: Li, Junjie, et al.
Veröffentlicht: (2023)
Towards Inclusive ASR: Investigating Voice Conversion for Dysarthric Speech Recognition in Low-Resource Languages
von: Li, Chin-Jou, et al.
Veröffentlicht: (2025)
von: Li, Chin-Jou, et al.
Veröffentlicht: (2025)
Custom Data Augmentation for low resource ASR using Bark and Retrieval-Based Voice Conversion
von: Kamble, Anand, et al.
Veröffentlicht: (2023)
von: Kamble, Anand, et al.
Veröffentlicht: (2023)
Attention Is Not Always the Answer: Optimizing Voice Activity Detection with Simple Feature Fusion
von: Tripathi, Kumud, et al.
Veröffentlicht: (2025)
von: Tripathi, Kumud, et al.
Veröffentlicht: (2025)
LibriTTS-P: A Corpus with Speaking Style and Speaker Identity Prompts for Text-to-Speech and Style Captioning
von: Kawamura, Masaya, et al.
Veröffentlicht: (2024)
von: Kawamura, Masaya, et al.
Veröffentlicht: (2024)
EMALG: An Enhanced Mandarin Lombard Grid Corpus with Meaningful Sentences
von: Li, Baifeng, et al.
Veröffentlicht: (2023)
von: Li, Baifeng, et al.
Veröffentlicht: (2023)
Building Tailored Speech Recognizers for Japanese Speaking Assessment
von: Kubo, Yotaro, et al.
Veröffentlicht: (2025)
von: Kubo, Yotaro, et al.
Veröffentlicht: (2025)
Improving Acoustic Word Embeddings through Correspondence Training of Self-supervised Speech Representations
von: Meghanani, Amit, et al.
Veröffentlicht: (2024)
von: Meghanani, Amit, et al.
Veröffentlicht: (2024)
A Pilot Study of Applying Sequence-to-Sequence Voice Conversion to Evaluate the Intelligibility of L2 Speech Using a Native Speaker's Shadowings
von: Geng, Haopeng, et al.
Veröffentlicht: (2024)
von: Geng, Haopeng, et al.
Veröffentlicht: (2024)
VStyle: A Benchmark for Voice Style Adaptation with Spoken Instructions
von: Zhan, Jun, et al.
Veröffentlicht: (2025)
von: Zhan, Jun, et al.
Veröffentlicht: (2025)
VoiceCraft-X: Unifying Multilingual, Voice-Cloning Speech Synthesis and Speech Editing
von: Zheng, Zhisheng, et al.
Veröffentlicht: (2025)
von: Zheng, Zhisheng, et al.
Veröffentlicht: (2025)
The Third VoicePrivacy Challenge: Preserving Emotional Expressiveness and Linguistic Content in Voice Anonymization
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2026)
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2026)
Infusing Acoustic Pause Context into Text-Based Dementia Assessment
von: Braun, Franziska, et al.
Veröffentlicht: (2024)
von: Braun, Franziska, et al.
Veröffentlicht: (2024)
Voice Adaptation for Swiss German
von: Stucki, Samuel, et al.
Veröffentlicht: (2025)
von: Stucki, Samuel, et al.
Veröffentlicht: (2025)
Marco-Voice Technical Report
von: Tian, Fengping, et al.
Veröffentlicht: (2025)
von: Tian, Fengping, et al.
Veröffentlicht: (2025)
Advancing Automated Speaking Assessment Leveraging Multifaceted Relevance and Grammar Information
von: Lu, Hao-Chien, et al.
Veröffentlicht: (2025)
von: Lu, Hao-Chien, et al.
Veröffentlicht: (2025)
A Novel Data Augmentation Approach for Automatic Speaking Assessment on Opinion Expressions
von: Wang, Chung-Chun, et al.
Veröffentlicht: (2025)
von: Wang, Chung-Chun, et al.
Veröffentlicht: (2025)
Alethia: A Foundational Encoder for Voice Deepfakes
von: Zhu, Yi, et al.
Veröffentlicht: (2026)
von: Zhu, Yi, et al.
Veröffentlicht: (2026)
StarVC: A Unified Auto-Regressive Framework for Joint Text and Speech Generation in Voice Conversion
von: Li, Fengjin, et al.
Veröffentlicht: (2025)
von: Li, Fengjin, et al.
Veröffentlicht: (2025)
A Theoretical Framework for Acoustic Neighbor Embeddings
von: Jeon, Woojay
Veröffentlicht: (2024)
von: Jeon, Woojay
Veröffentlicht: (2024)
Exploring the Benefits of Tokenization of Discrete Acoustic Units
von: Dekel, Avihu, et al.
Veröffentlicht: (2024)
von: Dekel, Avihu, et al.
Veröffentlicht: (2024)
Scalable Offline ASR for Command-Style Dictation in Courtrooms
von: Nethil, Kumarmanas, et al.
Veröffentlicht: (2025)
von: Nethil, Kumarmanas, et al.
Veröffentlicht: (2025)
Converting Anyone's Voice: End-to-End Expressive Voice Conversion with a Conditional Diffusion Model
von: Du, Zongyang, et al.
Veröffentlicht: (2024)
von: Du, Zongyang, et al.
Veröffentlicht: (2024)
Salmon: A Suite for Acoustic Language Model Evaluation
von: Maimon, Gallil, et al.
Veröffentlicht: (2024)
von: Maimon, Gallil, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Factor-Conditioned Speaking-Style Captioning
von: Ando, Atsushi, et al.
Veröffentlicht: (2024) -
Grapheme-Coherent Phonemic and Prosodic Annotation of Speech by Implicit and Explicit Grapheme Conditioning
von: Ohnaka, Hien, et al.
Veröffentlicht: (2025) -
StyleSinger: Style Transfer for Out-of-Domain Singing Voice Synthesis
von: Zhang, Yu, et al.
Veröffentlicht: (2023) -
ClaritySpeech: Dementia Obfuscation in Speech
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2025) -
Prosody-Driven Privacy-Preserving Dementia Detection
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2024)