Saved in:
| Main Authors: | Yu, Fan, Wang, Tao, Wu, You, Zhu, Lin, Deng, Wei, Han, Weisheng, Wang, Wenchao, Hu, Lin, Liang, Xiangyu, He, Xiaodong, Huang, Yankun, Gu, Yu, Liu, Yuan, Wang, Yuxuan, Xiao, Zhangyu, Wang, Ziteng, Dong, Boya, Dang, Feng, Chen, Jinming, Li, Jingdong, Wang, Jun, Jin, Yechen, Zhang, Yuan, Sheng, Zhengyan, Wang, Xin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.19090 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generating Novel and Realistic Speakers for Voice Conversion
by: Chen, Meiying Melissa, et al.
Published: (2025)
by: Chen, Meiying Melissa, et al.
Published: (2025)
JoyTTS: LLM-based Spoken Chatbot With Voice Cloning
by: Zhou, Fangru, et al.
Published: (2025)
by: Zhou, Fangru, et al.
Published: (2025)
Learning Equilibrium Fluctuation Expansions from Overdamped Langevin Dynamics
by: Wang, Lin, et al.
Published: (2026)
by: Wang, Lin, et al.
Published: (2026)
Probabilistic Approaches to The Energy Equality in Forced Surface Quasi-Geostrophic Equations
by: Wang, Lin, et al.
Published: (2024)
by: Wang, Lin, et al.
Published: (2024)
Multi-level Temporal-channel Speaker Retrieval for Zero-shot Voice Conversion
by: Wang, Zhichao, et al.
Published: (2023)
by: Wang, Zhichao, et al.
Published: (2023)
Residual Speaker Representation for One-Shot Voice Conversion
by: Xu, Le, et al.
Published: (2023)
by: Xu, Le, et al.
Published: (2023)
ML-SAN: Multi-Level Speaker-Adaptive Network for Emotion Recognition in Conversations
by: Wang, Kexue, et al.
Published: (2026)
by: Wang, Kexue, et al.
Published: (2026)
ReFlow-VC: Zero-shot Voice Conversion Based on Rectified Flow and Speaker Feature Optimization
by: Ren, Pengyu, et al.
Published: (2025)
by: Ren, Pengyu, et al.
Published: (2025)
Noro: Noise-Robust One-shot Voice Conversion with Hidden Speaker Representation Learning
by: He, Haorui, et al.
Published: (2024)
by: He, Haorui, et al.
Published: (2024)
Unifying Speech Recognition, Synthesis and Conversion with Autoregressive Transformers
by: Cai, Runyuan, et al.
Published: (2026)
by: Cai, Runyuan, et al.
Published: (2026)
Virasoro constraints for K3 surfaces and monodromy operators
by: Wang, Weisheng
Published: (2024)
by: Wang, Weisheng
Published: (2024)
StreamVoice+: Evolving into End-to-end Streaming Zero-shot Voice Conversion
by: Wang, Zhichao, et al.
Published: (2024)
by: Wang, Zhichao, et al.
Published: (2024)
DualMap: Enabling Both Cache Affinity and Load Balancing for Distributed LLM Serving
by: Yuan, Ying, et al.
Published: (2026)
by: Yuan, Ying, et al.
Published: (2026)
Talking Spell: A Wearable System Enabling Real-Time Anthropomorphic Voice Interaction with Everyday Objects
by: Wang, Xuetong, et al.
Published: (2025)
by: Wang, Xuetong, et al.
Published: (2025)
Humanlike AI for Corporate Social Responsibility Communication: How Perceived Anthropomorphism Shapes Stakeholder Acceptance of Chatbots
by: Yangzhi (Nicole) Jiang, et al.
Published: (2026)
by: Yangzhi (Nicole) Jiang, et al.
Published: (2026)
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant
by: Hou, Yixuan, et al.
Published: (2025)
by: Hou, Yixuan, et al.
Published: (2025)
Human-Inspired Soft Anthropomorphic Hand System for Neuromorphic Object and Pose Recognition Using Multimodal Signals
by: Wang, Fengyi, et al.
Published: (2025)
by: Wang, Fengyi, et al.
Published: (2025)
Object Classification Utilizing Neuromorphic Proprioceptive Signals in Active Exploration: Validated on a Soft Anthropomorphic Hand
by: Wang, Fengyi, et al.
Published: (2025)
by: Wang, Fengyi, et al.
Published: (2025)
One-variable equations over the lamplighter group
by: Ushakov, Alexander, et al.
Published: (2026)
by: Ushakov, Alexander, et al.
Published: (2026)
One variable equations over the lamplighter group
by: Ushakov, Alexander, et al.
Published: (2025)
by: Ushakov, Alexander, et al.
Published: (2025)
Online Prediction of Operating Temperature for Permanent Magnet Motor Used in EVs Based on Parameter Identification
by: Wang Yankun, et al.
Published: (2025)
by: Wang Yankun, et al.
Published: (2025)
From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench
by: Xu, Ke, et al.
Published: (2026)
by: Xu, Ke, et al.
Published: (2026)
StreamVoice: Streamable Context-Aware Language Modeling for Real-time Zero-Shot Voice Conversion
by: Wang, Zhichao, et al.
Published: (2024)
by: Wang, Zhichao, et al.
Published: (2024)
Underwater Acoustic Target Recognition based on Smoothness-inducing Regularization and Spectrogram-based Data Augmentation
by: Xu, Ji, et al.
Published: (2023)
by: Xu, Ji, et al.
Published: (2023)
SVSNet+: Enhancing Speaker Voice Similarity Assessment Models with Representations from Speech Foundation Models
by: Yin, Chun, et al.
Published: (2024)
by: Yin, Chun, et al.
Published: (2024)
JoyAgent-JDGenie: Technical Report on the GAIA
by: Liu, Jiarun, et al.
Published: (2025)
by: Liu, Jiarun, et al.
Published: (2025)
Expressive Prompting: Improving Emotion Intensity and Speaker Consistency in Zero-Shot TTS
by: Wang, Haoyu, et al.
Published: (2024)
by: Wang, Haoyu, et al.
Published: (2024)
An Investigation on Speaker Augmentation for End-to-End Speaker Extraction
by: You, Zhenghai, et al.
Published: (2025)
by: You, Zhenghai, et al.
Published: (2025)
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
by: Qi, Tianhua, et al.
Published: (2024)
by: Qi, Tianhua, et al.
Published: (2024)
DisfluencySpeech -- Single-Speaker Conversational Speech Dataset with Paralanguage
by: Wang, Kyra, et al.
Published: (2024)
by: Wang, Kyra, et al.
Published: (2024)
DreamVoice: Text-Guided Voice Conversion
by: Hai, Jiarui, et al.
Published: (2024)
by: Hai, Jiarui, et al.
Published: (2024)
VITRIX-CLIPIN: Enhancing Fine-Grained Visual Understanding in CLIP via Instruction Editing Data and Long Captions
by: Wang, Ziteng, et al.
Published: (2025)
by: Wang, Ziteng, et al.
Published: (2025)
SEF-VC: Speaker Embedding Free Zero-Shot Voice Conversion with Cross Attention
by: Li, Junjie, et al.
Published: (2023)
by: Li, Junjie, et al.
Published: (2023)
An Optimally Accurate Lanczos Algorithm in the Matrix Product State Representation
by: Wang, Yu, et al.
Published: (2025)
by: Wang, Yu, et al.
Published: (2025)
Analysis of Adherence Status and Influencing Factors Among Middle‐Aged and Elderly Hypertension Patients in Rural Areas of Northeast China
by: Xinyuan Lu, et al.
Published: (2025)
by: Xinyuan Lu, et al.
Published: (2025)
Effect of Comprehensive Health Management on Medication Adherence and Healthy Lifestyle Behavior of Patients With Hypertension
by: Xinyuan Lu, et al.
Published: (2025)
by: Xinyuan Lu, et al.
Published: (2025)
Strong limit theorems / Lin Zhengyan and Lu Chuanrong
by: Zhengyan, Lin
by: Zhengyan, Lin
Cavitation‐Induced Shear Failure Mechanism of Fractured Plugging Zone and Structure Strengthening Method for Lost Circulation Control in High‐Temperature and High‐Pressure Fractured Gas Reservoirs
by: Xiaoming Su, et al.
Published: (2024)
by: Xiaoming Su, et al.
Published: (2024)
On a partial data inverse problem for the semi-linear wave equation
by: Liu, Boya, et al.
Published: (2025)
by: Liu, Boya, et al.
Published: (2025)
USM-VC: Mitigating Timbre Leakage with Universal Semantic Mapping Residual Block for Voice Conversion
by: Li, Na, et al.
Published: (2025)
by: Li, Na, et al.
Published: (2025)
Similar Items
-
Generating Novel and Realistic Speakers for Voice Conversion
by: Chen, Meiying Melissa, et al.
Published: (2025) -
JoyTTS: LLM-based Spoken Chatbot With Voice Cloning
by: Zhou, Fangru, et al.
Published: (2025) -
Learning Equilibrium Fluctuation Expansions from Overdamped Langevin Dynamics
by: Wang, Lin, et al.
Published: (2026) -
Probabilistic Approaches to The Energy Equality in Forced Surface Quasi-Geostrophic Equations
by: Wang, Lin, et al.
Published: (2024) -
Multi-level Temporal-channel Speaker Retrieval for Zero-shot Voice Conversion
by: Wang, Zhichao, et al.
Published: (2023)