Gender-ambiguous voice generation through feminine speaking style transfer in male voices
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Koutsogiannaki, Maria, Dowall, Shafel Mc, Agiomyrgiannakis, Ioannis |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adversarial speech for voice privacy protection from Personalized Speech generation
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
Comparison of fundamental frequency estimators with subharmonic voice signals
von: Ikuma, Takeshi, et al.
Veröffentlicht: (2025)
von: Ikuma, Takeshi, et al.
Veröffentlicht: (2025)
Towards detecting the pathological subharmonic voicing with fully convolutional neural networks
von: Ikuma, Takeshi, et al.
Veröffentlicht: (2025)
von: Ikuma, Takeshi, et al.
Veröffentlicht: (2025)
Subjective quality evaluation of personalized own voice reconstruction systems
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2025)
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2025)
DNN-based ensemble singing voice synthesis with interactions between singers
von: Hyodo, Hiroaki, et al.
Veröffentlicht: (2024)
von: Hyodo, Hiroaki, et al.
Veröffentlicht: (2024)
Using voice analysis as an early indicator of risk for depression in young adults
von: Scherer, Klaus R., et al.
Veröffentlicht: (2024)
von: Scherer, Klaus R., et al.
Veröffentlicht: (2024)
Resource-constrained stereo singing voice cancellation
von: Borrelli, Clara, et al.
Veröffentlicht: (2024)
von: Borrelli, Clara, et al.
Veröffentlicht: (2024)
Building speech corpus with diverse voice characteristics for its prompt-based representation
von: Watanabe, Aya, et al.
Veröffentlicht: (2024)
von: Watanabe, Aya, et al.
Veröffentlicht: (2024)
Exploring synthetic data for cross-speaker style transfer in style representation based TTS
von: Ueda, Lucas H., et al.
Veröffentlicht: (2024)
von: Ueda, Lucas H., et al.
Veröffentlicht: (2024)
Introducing voice timbre attribute detection
von: He, Jinghao, et al.
Veröffentlicht: (2025)
von: He, Jinghao, et al.
Veröffentlicht: (2025)
SponTTS: modeling and transferring spontaneous style for TTS
von: Li, Hanzhao, et al.
Veröffentlicht: (2023)
von: Li, Hanzhao, et al.
Veröffentlicht: (2023)
Audiovisual angle and voice incongruence do not affect audiovisual verbal short-term memory in virtual reality
von: Ermert, Cosima A., et al.
Veröffentlicht: (2024)
von: Ermert, Cosima A., et al.
Veröffentlicht: (2024)
SelfTTS: cross-speaker style transfer through explicit embedding disentanglement and self-refinement using self-augmentation
von: Ueda, Lucas H., et al.
Veröffentlicht: (2026)
von: Ueda, Lucas H., et al.
Veröffentlicht: (2026)
Easy, Interpretable, Effective: openSMILE for voice deepfake detection
von: Pascu, Octavian, et al.
Veröffentlicht: (2024)
von: Pascu, Octavian, et al.
Veröffentlicht: (2024)
Accurate analysis of the pitch pulse-based magnitude/phase structure of natural vowels and assessment of three lightweight time/frequency voicing restoration methods
von: Ferreira, Aníbal J. S., et al.
Veröffentlicht: (2025)
von: Ferreira, Aníbal J. S., et al.
Veröffentlicht: (2025)
Non-autoregressive real-time Accent Conversion model with voice cloning
von: Nechaev, Vladimir, et al.
Veröffentlicht: (2024)
von: Nechaev, Vladimir, et al.
Veröffentlicht: (2024)
Voxceleb-ESP: preliminary experiments detecting Spanish celebrities from their voices
von: Labrador, Beltrán, et al.
Veröffentlicht: (2023)
von: Labrador, Beltrán, et al.
Veröffentlicht: (2023)
Can we reconstruct a dysarthric voice with the large speech model Parler TTS?
von: Sanchez, Ariadna, et al.
Veröffentlicht: (2025)
von: Sanchez, Ariadna, et al.
Veröffentlicht: (2025)
Combining audio control and style transfer using latent diffusion
von: Demerlé, Nils, et al.
Veröffentlicht: (2024)
von: Demerlé, Nils, et al.
Veröffentlicht: (2024)
A multi-speaker multi-lingual voice cloning system based on vits2 for limmits 2024 challenge
von: Wang, Xiaopeng, et al.
Veröffentlicht: (2024)
von: Wang, Xiaopeng, et al.
Veröffentlicht: (2024)
Developing vocal system impaired patient-aimed voice quality assessment approach using ASR representation-included multiple features
von: Dang, Shaoxiang, et al.
Veröffentlicht: (2024)
von: Dang, Shaoxiang, et al.
Veröffentlicht: (2024)
Voice Passing : a Non-Binary Voice Gender Prediction System for evaluating Transgender voice transition
von: Doukhan, David, et al.
Veröffentlicht: (2024)
von: Doukhan, David, et al.
Veröffentlicht: (2024)
PiCoGen2: Piano cover generation with transfer learning approach and weakly aligned data
von: Tan, Chih-Pin, et al.
Veröffentlicht: (2024)
von: Tan, Chih-Pin, et al.
Veröffentlicht: (2024)
Screening method for early dementia using sound objects as voice biomarkers
von: Pluta, Adam, et al.
Veröffentlicht: (2024)
von: Pluta, Adam, et al.
Veröffentlicht: (2024)
VECL-TTS: Voice identity and Emotional style controllable Cross-Lingual Text-to-Speech
von: Gudmalwar, Ashishkumar, et al.
Veröffentlicht: (2024)
von: Gudmalwar, Ashishkumar, et al.
Veröffentlicht: (2024)
Face-voice Association in Multilingual Environments (FAME) Challenge 2024 Evaluation Plan
von: Saeed, Muhammad Saad, et al.
Veröffentlicht: (2024)
von: Saeed, Muhammad Saad, et al.
Veröffentlicht: (2024)
Hear Your Face: Face-based voice conversion with F0 estimation
von: Lee, Jaejun, et al.
Veröffentlicht: (2024)
von: Lee, Jaejun, et al.
Veröffentlicht: (2024)
SynthCloner: Synthesizer-style Audio Transfer via Factorized Codec with ADSR Envelope Control
von: Liu, Jeng-Yue, et al.
Veröffentlicht: (2025)
von: Liu, Jeng-Yue, et al.
Veröffentlicht: (2025)
MusicGen-Stem: Multi-stem music generation and edition through autoregressive modeling
von: Rouard, Simon, et al.
Veröffentlicht: (2025)
von: Rouard, Simon, et al.
Veröffentlicht: (2025)
Accent-VITS:accent transfer for end-to-end TTS
von: Ma, Linhan, et al.
Veröffentlicht: (2023)
von: Ma, Linhan, et al.
Veröffentlicht: (2023)
Real-time implementation of vibrato transfer as an audio effect
von: Hyrkas, Jeremy
Veröffentlicht: (2025)
von: Hyrkas, Jeremy
Veröffentlicht: (2025)
Improved symbolic drum style classification with grammar-based hierarchical representations
von: Géré, Léo, et al.
Veröffentlicht: (2024)
von: Géré, Léo, et al.
Veröffentlicht: (2024)
Complexity of frequency fluctuations and the interpretive style in the bass viola da gamba
von: Lugo, Igor, et al.
Veröffentlicht: (2025)
von: Lugo, Igor, et al.
Veröffentlicht: (2025)
Monaural speech enhancement on drone via Adapter based transfer learning
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
von: Chen, Xingyu, et al.
Veröffentlicht: (2024)
PERSONA: An Application for Emotion Recognition, Gender Recognition and Age Estimation
von: Koshal, Devyani, et al.
Veröffentlicht: (2024)
von: Koshal, Devyani, et al.
Veröffentlicht: (2024)
On the Language and Gender Biases in PSTN, VoIP and Neural Audio Codecs
von: Altwlkany, Kemal, et al.
Veröffentlicht: (2025)
von: Altwlkany, Kemal, et al.
Veröffentlicht: (2025)
Evaluating Multichannel Speech Enhancement Algorithms at the Phoneme Scale Across Genders
von: Monir, Nasser-Eddine, et al.
Veröffentlicht: (2025)
von: Monir, Nasser-Eddine, et al.
Veröffentlicht: (2025)
TBDM-Net: Bidirectional Dense Networks with Gender Information for Speech Emotion Recognition
von: Striletchi, Vlad, et al.
Veröffentlicht: (2024)
von: Striletchi, Vlad, et al.
Veröffentlicht: (2024)
When Voice Matters: Evidence of Gender Disparity in Positional Bias of SpeechLLMs
von: Satish, Shree Harsha Bokkahalli, et al.
Veröffentlicht: (2025)
von: Satish, Shree Harsha Bokkahalli, et al.
Veröffentlicht: (2025)
PromptASR for contextualized ASR with controllable style
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2023)
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Adversarial speech for voice privacy protection from Personalized Speech generation
von: Chen, Shihao, et al.
Veröffentlicht: (2024) -
Comparison of fundamental frequency estimators with subharmonic voice signals
von: Ikuma, Takeshi, et al.
Veröffentlicht: (2025) -
Towards detecting the pathological subharmonic voicing with fully convolutional neural networks
von: Ikuma, Takeshi, et al.
Veröffentlicht: (2025) -
Subjective quality evaluation of personalized own voice reconstruction systems
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2025) -
DNN-based ensemble singing voice synthesis with interactions between singers
von: Hyodo, Hiroaki, et al.
Veröffentlicht: (2024)