On the Emotion Understanding of Synthesized Speech
Fuente:
arXiv
Saved in:
| Main Authors: | Ge, Yuan, Zhao, Haishu, Hao, Aokai, Zhang, Junxiang, Li, Bei, Liu, Xiaoqian, Wang, Chenglong, Wang, Jianjin, Zhou, Bingsen, Liu, Bingyu, Zhu, Jingbo, Yu, Zhengtao, Xiao, Tong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
StyleBench: Evaluating Speech Language Models on Conversational Speaking Style Control
by: Zhao, Haishu, et al.
Published: (2026)
by: Zhao, Haishu, et al.
Published: (2026)
MTP-S2UT: Enhancing Speech-to-Speech Translation Quality with Multi-token Prediction
by: Wang, Jianjin, et al.
Published: (2025)
by: Wang, Jianjin, et al.
Published: (2025)
A Modular-based Strategy for Mitigating Gradient Conflicts in Simultaneous Speech Translation
by: Liu, Xiaoqian, et al.
Published: (2024)
by: Liu, Xiaoqian, et al.
Published: (2024)
When Scaling Fails: Mitigating Audio Perception Decay of LALMs via Multi-Step Perception-Aware Reasoning
by: Mao, Ruixiang, et al.
Published: (2026)
by: Mao, Ruixiang, et al.
Published: (2026)
FLEXI: Benchmarking Full-duplex Human-LLM Speech Interaction
by: Ge, Yuan, et al.
Published: (2025)
by: Ge, Yuan, et al.
Published: (2025)
SageLM: A Multi-aspect and Explainable Large Language Model for Speech Judgement
by: Ge, Yuan, et al.
Published: (2025)
by: Ge, Yuan, et al.
Published: (2025)
APR: Penalizing Structural Redundancy in Large Reasoning Models via Anchor-based Process Rewards
by: Chang, Kaiyan, et al.
Published: (2026)
by: Chang, Kaiyan, et al.
Published: (2026)
Attention2Probability: Attention-Driven Terminology Probability Estimation for Robust Speech-to-Text System
by: Du, Yanfan, et al.
Published: (2025)
by: Du, Yanfan, et al.
Published: (2025)
Efficient Prompting Methods for Large Language Models: A Survey
by: Chang, Kaiyan, et al.
Published: (2024)
by: Chang, Kaiyan, et al.
Published: (2024)
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models
by: Chang, Kaiyan, et al.
Published: (2025)
by: Chang, Kaiyan, et al.
Published: (2025)
Hybrid Alignment Training for Large Language Models
by: Wang, Chenglong, et al.
Published: (2024)
by: Wang, Chenglong, et al.
Published: (2024)
Leveraging Unit Language Guidance to Advance Speech Modeling in Textless Speech-to-Speech Translation
by: Zhang, Yuhao, et al.
Published: (2025)
by: Zhang, Yuhao, et al.
Published: (2025)
NDP: Next Distribution Prediction as a More Broad Target
by: Ruan, Junhao, et al.
Published: (2024)
by: Ruan, Junhao, et al.
Published: (2024)
Revisiting Interpolation Augmentation for Speech-to-Text Generation
by: Xu, Chen, et al.
Published: (2024)
by: Xu, Chen, et al.
Published: (2024)
Recent Advances in End-to-End Simultaneous Speech Translation
by: Liu, Xiaoqian, et al.
Published: (2024)
by: Liu, Xiaoqian, et al.
Published: (2024)
MRO: Enhancing Reasoning in Diffusion Language Models via Multi-Reward Optimization
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
Social Cognitive Domain Coordination in LeftBehind Children: A Comparative Study of LeftBehind and Non-Left-Behind Children in Rural China
by: Jianjin Liu
Published: (2016)
by: Jianjin Liu
Published: (2016)
Chinese Adolescents’ Conceptions of Teacher Authority and Their Relations to Rule Violations in School
by: Jianjin Liu
Published: (2018)
by: Jianjin Liu
Published: (2018)
EfficientGraph-RAG: Structured Retrieval-State Management for Cross-Task Retrieval-Augmented Generation
by: Niu, Miaohe, et al.
Published: (2026)
by: Niu, Miaohe, et al.
Published: (2026)
MemoSight: Unifying Context Compression and Multi Token Prediction for Reasoning Acceleration
by: Liu, Xinyu, et al.
Published: (2026)
by: Liu, Xinyu, et al.
Published: (2026)
Prior Constraints-based Reward Model Training for Aligning Large Language Models
by: Zhou, Hang, et al.
Published: (2024)
by: Zhou, Hang, et al.
Published: (2024)
RouteLMT: Learned Sample Routing for Hybrid LLM Translation Deployment
by: Luo, Yingfeng, et al.
Published: (2026)
by: Luo, Yingfeng, et al.
Published: (2026)
MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks
by: Ruan, Junhao, et al.
Published: (2026)
by: Ruan, Junhao, et al.
Published: (2026)
Probing Preference Representations: A Multi-Dimensional Evaluation and Analysis Method for Reward Models
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
SPS: Steering Probability Squeezing for Better Exploration in Reinforcement Learning for Large Language Models
by: Huo, Yifu, et al.
Published: (2026)
by: Huo, Yifu, et al.
Published: (2026)
MSRL: Scaling Generative Multimodal Reward Modeling via Multi-Stage Reinforcement Learning
by: Wang, Chenglong, et al.
Published: (2026)
by: Wang, Chenglong, et al.
Published: (2026)
EIT: Enhanced Interactive Transformer
by: Zheng, Tong, et al.
Published: (2022)
by: Zheng, Tong, et al.
Published: (2022)
Forgetting Curve: A Reliable Method for Evaluating Memorization Capability for Long-context Models
by: Liu, Xinyu, et al.
Published: (2024)
by: Liu, Xinyu, et al.
Published: (2024)
Optimizing Speech Multi-View Feature Fusion through Conditional Computation
by: Shan, Weiqiao, et al.
Published: (2025)
by: Shan, Weiqiao, et al.
Published: (2025)
Autoencoding-Free Context Compression for LLMs via Contextual Semantic Anchors
by: Liu, Xin, et al.
Published: (2025)
by: Liu, Xin, et al.
Published: (2025)
One Size Does Not Fit All: A Distribution-Aware Sparsification for More Precise Model Merging
by: Luo, Yingfeng, et al.
Published: (2025)
by: Luo, Yingfeng, et al.
Published: (2025)
M-CIF: Multi-Scale Alignment For CIF-Based Non-Autoregressive ASR
by: Mao, Ruixiang, et al.
Published: (2025)
by: Mao, Ruixiang, et al.
Published: (2025)
Foundations of Large Language Models
by: Xiao, Tong, et al.
Published: (2025)
by: Xiao, Tong, et al.
Published: (2025)
Beyond Decoder-only: Large Language Models Can be Good Encoders for Machine Translation
by: Luo, Yingfeng, et al.
Published: (2025)
by: Luo, Yingfeng, et al.
Published: (2025)
PartialFormer: Modeling Part Instead of Whole for Machine Translation
by: Zheng, Tong, et al.
Published: (2023)
by: Zheng, Tong, et al.
Published: (2023)
Revealing the Parallel Multilingual Learning within Large Language Models
by: Mu, Yongyu, et al.
Published: (2024)
by: Mu, Yongyu, et al.
Published: (2024)
GRAM: A Generative Foundation Reward Model for Reward Generalization
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
Learning Evaluation Models from Large Language Models for Sequence Generation
by: Wang, Chenglong, et al.
Published: (2023)
by: Wang, Chenglong, et al.
Published: (2023)
GRAM-R$^2$: Self-Training Generative Foundation Reward Models for Reward Reasoning
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
Observation of Ballistic Thermal Transport in a Nonintegrable Classical Many-Body System
by: Wang, Jianjin, et al.
Published: (2024)
by: Wang, Jianjin, et al.
Published: (2024)
Similar Items
-
StyleBench: Evaluating Speech Language Models on Conversational Speaking Style Control
by: Zhao, Haishu, et al.
Published: (2026) -
MTP-S2UT: Enhancing Speech-to-Speech Translation Quality with Multi-token Prediction
by: Wang, Jianjin, et al.
Published: (2025) -
A Modular-based Strategy for Mitigating Gradient Conflicts in Simultaneous Speech Translation
by: Liu, Xiaoqian, et al.
Published: (2024) -
When Scaling Fails: Mitigating Audio Perception Decay of LALMs via Multi-Step Perception-Aware Reasoning
by: Mao, Ruixiang, et al.
Published: (2026) -
FLEXI: Benchmarking Full-duplex Human-LLM Speech Interaction
by: Ge, Yuan, et al.
Published: (2025)