SageLM: A Multi-aspect and Explainable Large Language Model for Speech Judgement
Fuente:
arXiv
Saved in:
| Main Authors: | Ge, Yuan, Zhang, Junxiang, Liu, Xiaoqian, Li, Bei, Ma, Xiangnan, Wang, Chenglong, Ye, Kaiyang, Du, Yangfan, Zhang, Linfeng, Huang, Yuxin, Xiao, Tong, Yu, Zhengtao, Zhu, JingBo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Emotion Understanding of Synthesized Speech
by: Ge, Yuan, et al.
Published: (2026)
by: Ge, Yuan, et al.
Published: (2026)
BIASEDTALES-ML: A Multilingual Dataset for Analyzing Narrative Attribute Distributions in LLM-Generated Stories
by: Ouyang, Yuxuan, et al.
Published: (2026)
by: Ouyang, Yuxuan, et al.
Published: (2026)
Leveraging Unit Language Guidance to Advance Speech Modeling in Textless Speech-to-Speech Translation
by: Zhang, Yuhao, et al.
Published: (2025)
by: Zhang, Yuhao, et al.
Published: (2025)
PRJ: Perception-Retrieval-Judgement for Generated Images
by: Fu, Qiang, et al.
Published: (2025)
by: Fu, Qiang, et al.
Published: (2025)
A Modular-based Strategy for Mitigating Gradient Conflicts in Simultaneous Speech Translation
by: Liu, Xiaoqian, et al.
Published: (2024)
by: Liu, Xiaoqian, et al.
Published: (2024)
Offline Exploration-Aware Fine-Tuning for Long-Chain Mathematical Reasoning
by: Mu, Yongyu, et al.
Published: (2026)
by: Mu, Yongyu, et al.
Published: (2026)
FLEXI: Benchmarking Full-duplex Human-LLM Speech Interaction
by: Ge, Yuan, et al.
Published: (2025)
by: Ge, Yuan, et al.
Published: (2025)
MTP-S2UT: Enhancing Speech-to-Speech Translation Quality with Multi-token Prediction
by: Wang, Jianjin, et al.
Published: (2025)
by: Wang, Jianjin, et al.
Published: (2025)
Recent Advances in End-to-End Simultaneous Speech Translation
by: Liu, Xiaoqian, et al.
Published: (2024)
by: Liu, Xiaoqian, et al.
Published: (2024)
Reasoning Beyond Majority Vote: An Explainable SpeechLM Framework for Speech Emotion Recognition
by: Su, Bo-Hao, et al.
Published: (2025)
by: Su, Bo-Hao, et al.
Published: (2025)
Causal Autoregressive Diffusion Language Model
by: Ruan, Junhao, et al.
Published: (2026)
by: Ruan, Junhao, et al.
Published: (2026)
RLAIF-SPA: Structured AI Feedback for Semantic-Prosodic Alignment in Speech Synthesis
by: Yang, Qing, et al.
Published: (2025)
by: Yang, Qing, et al.
Published: (2025)
Attention2Probability: Attention-Driven Terminology Probability Estimation for Robust Speech-to-Text System
by: Du, Yanfan, et al.
Published: (2025)
by: Du, Yanfan, et al.
Published: (2025)
SageAttention2++: A More Efficient Implementation of SageAttention2
by: Zhang, Jintao, et al.
Published: (2025)
by: Zhang, Jintao, et al.
Published: (2025)
Emotion-Aligned Generation in Diffusion Text to Speech Models via Preference-Guided Optimization
by: Shi, Jiacheng, et al.
Published: (2025)
by: Shi, Jiacheng, et al.
Published: (2025)
DTRT: Enhancing Human Intent Estimation and Role Allocation for Physical Human-Robot Collaboration
by: Liu, Haotian, et al.
Published: (2025)
by: Liu, Haotian, et al.
Published: (2025)
EchoX: Towards Mitigating Acoustic-Semantic Gap via Echo Training for Speech-to-Speech LLMs
by: Zhang, Yuhao, et al.
Published: (2025)
by: Zhang, Yuhao, et al.
Published: (2025)
APR: Penalizing Structural Redundancy in Large Reasoning Models via Anchor-based Process Rewards
by: Chang, Kaiyan, et al.
Published: (2026)
by: Chang, Kaiyan, et al.
Published: (2026)
FinSage: A Multi-aspect RAG System for Financial Filings Question Answering
by: Wang, Xinyu, et al.
Published: (2025)
by: Wang, Xinyu, et al.
Published: (2025)
Towards Explainability and Fairness in Swiss Judgement Prediction: Benchmarking on a Multilingual Dataset
by: S, Santosh T. Y. S., et al.
Published: (2024)
by: S, Santosh T. Y. S., et al.
Published: (2024)
LM-SPT: LM-Aligned Semantic Distillation for Speech Tokenization
by: Jo, Daejin, et al.
Published: (2025)
by: Jo, Daejin, et al.
Published: (2025)
ShieldLM: Empowering LLMs as Aligned, Customizable and Explainable Safety Detectors
by: Zhang, Zhexin, et al.
Published: (2024)
by: Zhang, Zhexin, et al.
Published: (2024)
Revealing the molecular structures of a-Al2O3(0001)-water interface by machine learning based computational vibrational spectroscopy
by: Du, Xianglong, et al.
Published: (2024)
by: Du, Xianglong, et al.
Published: (2024)
A telomere‐to‐telomere haplotype‐resolved genome of white‐fruited strawberry reveals the complexity of fruit colour formation of cultivated strawberry
by: Junxiang Zhang, et al.
Published: (2024)
by: Junxiang Zhang, et al.
Published: (2024)
SECodec: Structural Entropy-based Compressive Speech Representation Codec for Speech Language Models
by: Wang, Linqin, et al.
Published: (2024)
by: Wang, Linqin, et al.
Published: (2024)
Evaluating Gender Bias of LLMs in Making Morality Judgements
by: Bajaj, Divij, et al.
Published: (2024)
by: Bajaj, Divij, et al.
Published: (2024)
Probabilistic Fatigue Life Prediction Framework for Natural Rubber Considering Ambient Temperatures
by: Xiangnan Liu, et al.
Published: (2026)
by: Xiangnan Liu, et al.
Published: (2026)
Leveraging Entailment Judgements in Cross-Lingual Summarisation
by: Zhang, Huajian, et al.
Published: (2024)
by: Zhang, Huajian, et al.
Published: (2024)
M-CIF: Multi-Scale Alignment For CIF-Based Non-Autoregressive ASR
by: Mao, Ruixiang, et al.
Published: (2025)
by: Mao, Ruixiang, et al.
Published: (2025)
SA-WavLM: Speaker-Aware Self-Supervised Pre-training for Mixture Speech
by: Lin, Jingru, et al.
Published: (2024)
by: Lin, Jingru, et al.
Published: (2024)
Head-Pose-Aware Visual Speech Recognition with FiLM Modulation
by: Teng, Matthew Kit Khinn, et al.
Published: (2026)
by: Teng, Matthew Kit Khinn, et al.
Published: (2026)
ESPnet-SpeechLM: An Open Speech Language Model Toolkit
by: Tian, Jinchuan, et al.
Published: (2025)
by: Tian, Jinchuan, et al.
Published: (2025)
Auditory‐Perceptual Evaluation of Situationally‐Bound Judgements of Listener Comfort for Postlaryngectomy Voice and Speech
by: Natalie Smith, et al.
Published: (2025)
by: Natalie Smith, et al.
Published: (2025)
Rethinking Political Judgement
by: Mrovlje, Maša
Published: (2022)
by: Mrovlje, Maša
Published: (2022)
Context, Judgement, Deduction
by: Coraglia, Greta, et al.
Published: (2021)
by: Coraglia, Greta, et al.
Published: (2021)
Adapting WavLM for Speech Emotion Recognition
by: Diatlova, Daria, et al.
Published: (2024)
by: Diatlova, Daria, et al.
Published: (2024)
Law’s Abstract Judgement (LAJ) and Intelligent Fidelity: On William Lucy’s Law’s Judgement
by: Imer B. Flores
Published: (2019)
by: Imer B. Flores
Published: (2019)
KEDRec-LM: A Knowledge-distilled Explainable Drug Recommendation Large Language Model
by: Zhang, Kai, et al.
Published: (2025)
by: Zhang, Kai, et al.
Published: (2025)
Interdisciplinarity in Transition: The Formation and Transformation of the Committee on Human Development, 1930s–1950s
by: Liping Wang, et al.
Published: (2026)
by: Liping Wang, et al.
Published: (2026)
CHITNet: A Complementary to Harmonious Information Transfer Network for Infrared and Visible Image Fusion
by: Du, Keying, et al.
Published: (2023)
by: Du, Keying, et al.
Published: (2023)
Similar Items
-
On the Emotion Understanding of Synthesized Speech
by: Ge, Yuan, et al.
Published: (2026) -
BIASEDTALES-ML: A Multilingual Dataset for Analyzing Narrative Attribute Distributions in LLM-Generated Stories
by: Ouyang, Yuxuan, et al.
Published: (2026) -
Leveraging Unit Language Guidance to Advance Speech Modeling in Textless Speech-to-Speech Translation
by: Zhang, Yuhao, et al.
Published: (2025) -
PRJ: Perception-Retrieval-Judgement for Generated Images
by: Fu, Qiang, et al.
Published: (2025) -
A Modular-based Strategy for Mitigating Gradient Conflicts in Simultaneous Speech Translation
by: Liu, Xiaoqian, et al.
Published: (2024)