Enregistré dans:
| Auteurs principaux: | Yang, Qing, Liu, Zhenghao, Du, Yangfan, Huang, Pengcheng, Xiao, Tong |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2510.14628 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
par: Lee, Harrison, et autres
Publié: (2023)
par: Lee, Harrison, et autres
Publié: (2023)
Empirical Analysis of Decoding Biases in Masked Diffusion Models
par: Huang, Pengcheng, et autres
Publié: (2025)
par: Huang, Pengcheng, et autres
Publié: (2025)
Speech Emotion Recognition with Distilled Prosodic and Linguistic Affect Representations
par: Shome, Debaditya, et autres
Publié: (2023)
par: Shome, Debaditya, et autres
Publié: (2023)
Reverse Constitutional AI: A Framework for Controllable Toxic Data Generation via Probability-Clamped RLAIF
par: Fang, Yuan, et autres
Publié: (2026)
par: Fang, Yuan, et autres
Publié: (2026)
Recent Advances in End-to-End Simultaneous Speech Translation
par: Liu, Xiaoqian, et autres
Publié: (2024)
par: Liu, Xiaoqian, et autres
Publié: (2024)
SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Model
par: Zhang, Yongting, et autres
Publié: (2024)
par: Zhang, Yongting, et autres
Publié: (2024)
Curriculum-RLAIF: Curriculum Alignment with Reinforcement Learning from AI Feedback
par: Lin, Jiaye, et autres
Publié: (2025)
par: Lin, Jiaye, et autres
Publié: (2025)
ParamMute: Suppressing Knowledge-Critical FFNs for Faithful Retrieval-Augmented Generation
par: Huang, Pengcheng, et autres
Publié: (2025)
par: Huang, Pengcheng, et autres
Publié: (2025)
Intent-conditioned and Non-toxic Counterspeech Generation using Multi-Task Instruction Tuning with RLAIF
par: Hengle, Amey, et autres
Publié: (2024)
par: Hengle, Amey, et autres
Publié: (2024)
Emotion-Aligned Generation in Diffusion Text to Speech Models via Preference-Guided Optimization
par: Shi, Jiacheng, et autres
Publié: (2025)
par: Shi, Jiacheng, et autres
Publié: (2025)
Prosodic Structure Beyond Lexical Content: A Study of Self-Supervised Learning
par: Wallbridge, Sarenne, et autres
Publié: (2025)
par: Wallbridge, Sarenne, et autres
Publié: (2025)
Direct Language Model Alignment from Online AI Feedback
par: Guo, Shangmin, et autres
Publié: (2024)
par: Guo, Shangmin, et autres
Publié: (2024)
Towards Fine-Grained Code-Switch Speech Translation with Semantic Space Alignment
par: Gao, Yan, et autres
Publié: (2025)
par: Gao, Yan, et autres
Publié: (2025)
Translate-and-Revise: Boosting Large Language Models for Constrained Translation
par: Huang, Pengcheng, et autres
Publié: (2024)
par: Huang, Pengcheng, et autres
Publié: (2024)
WhiSPA: Semantically and Psychologically Aligned Whisper with Self-Supervised Contrastive and Student-Teacher Learning
par: Rao, Rajath, et autres
Publié: (2025)
par: Rao, Rajath, et autres
Publié: (2025)
Extensive Self-Contrast Enables Feedback-Free Language Model Alignment
par: Liu, Xiao, et autres
Publié: (2024)
par: Liu, Xiao, et autres
Publié: (2024)
SPA: A Simple but Tough-to-Beat Baseline for Knowledge Injection
par: Tang, Kexian, et autres
Publié: (2026)
par: Tang, Kexian, et autres
Publié: (2026)
Entity Alignment with Noisy Annotations from Large Language Models
par: Chen, Shengyuan, et autres
Publié: (2024)
par: Chen, Shengyuan, et autres
Publié: (2024)
Understand Then Memory: A Cognitive Gist-Driven RAG Framework with Global Semantic Diffusion
par: Zhou, Pengcheng, et autres
Publié: (2026)
par: Zhou, Pengcheng, et autres
Publié: (2026)
Semantic XPath: Structured Agentic Memory Access for Conversational AI
par: Liu, Yifan Simon, et autres
Publié: (2026)
par: Liu, Yifan Simon, et autres
Publié: (2026)
Chunks as Arms: Multi-Armed Bandit-Guided Sampling for Long-Context LLM Preference Optimization
par: Duan, Shaohua, et autres
Publié: (2025)
par: Duan, Shaohua, et autres
Publié: (2025)
LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety
par: Yang, Junxiao, et autres
Publié: (2026)
par: Yang, Junxiao, et autres
Publié: (2026)
EchoX: Towards Mitigating Acoustic-Semantic Gap via Echo Training for Speech-to-Speech LLMs
par: Zhang, Yuhao, et autres
Publié: (2025)
par: Zhang, Yuhao, et autres
Publié: (2025)
Advancing Mathematical Reasoning in Language Models: The Impact of Problem-Solving Data, Data Synthesis Methods, and Training Stages
par: Chen, Zui, et autres
Publié: (2025)
par: Chen, Zui, et autres
Publié: (2025)
The Impact of Prosodic Segmentation on Speech Synthesis of Spontaneous Speech
par: Galdino, Julio Cesar, et autres
Publié: (2025)
par: Galdino, Julio Cesar, et autres
Publié: (2025)
GlobeSumm: A Challenging Benchmark Towards Unifying Multi-lingual, Cross-lingual and Multi-document News Summarization
par: Ye, Yangfan, et autres
Publié: (2024)
par: Ye, Yangfan, et autres
Publié: (2024)
SLAM: Towards Efficient Multilingual Reasoning via Selective Language Alignment
par: Fan, Yuchun, et autres
Publié: (2025)
par: Fan, Yuchun, et autres
Publié: (2025)
Soundwave: Less is More for Speech-Text Alignment in LLMs
par: Zhang, Yuhao, et autres
Publié: (2025)
par: Zhang, Yuhao, et autres
Publié: (2025)
Modeling Sarcastic Speech: Semantic and Prosodic Cues in a Speech Synthesis Framework
par: Li, Zhu, et autres
Publié: (2025)
par: Li, Zhu, et autres
Publié: (2025)
Attention-MoA: Enhancing Mixture-of-Agents via Inter-Agent Semantic Attention and Deep Residual Synthesis
par: Wen, Jianyu, et autres
Publié: (2026)
par: Wen, Jianyu, et autres
Publié: (2026)
Ustnlp16 at SemEval-2025 Task 9: Improving Model Performance through Imbalance Handling and Focal Loss
par: Cai, Zhuoang, et autres
Publié: (2025)
par: Cai, Zhuoang, et autres
Publié: (2025)
Toolink: Linking Toolkit Creation and Using through Chain-of-Solving on Open-Source Model
par: Qian, Cheng, et autres
Publié: (2023)
par: Qian, Cheng, et autres
Publié: (2023)
Speech-based Psychological Crisis Assessment using LLMs
par: Chiba, Terumi, et autres
Publié: (2026)
par: Chiba, Terumi, et autres
Publié: (2026)
The Molecular Structure of Thought: Mapping the Topology of Long Chain-of-Thought Reasoning
par: Chen, Qiguang, et autres
Publié: (2026)
par: Chen, Qiguang, et autres
Publié: (2026)
Leveraging Unit Language Guidance to Advance Speech Modeling in Textless Speech-to-Speech Translation
par: Zhang, Yuhao, et autres
Publié: (2025)
par: Zhang, Yuhao, et autres
Publié: (2025)
Semantic Differentiation in Speech Emotion Recognition: Insights from Descriptive and Expressive Speech Roles
par: Guo, Rongchen, et autres
Publié: (2025)
par: Guo, Rongchen, et autres
Publié: (2025)
Understanding the Modality Gap: An Empirical Study on the Speech-Text Alignment Mechanism of Large Speech Language Models
par: Xiang, Bajian, et autres
Publié: (2025)
par: Xiang, Bajian, et autres
Publié: (2025)
RLTHF: Targeted Human Feedback for LLM Alignment
par: Xu, Yifei, et autres
Publié: (2025)
par: Xu, Yifei, et autres
Publié: (2025)
MORALISE: A Structured Benchmark for Moral Alignment in Visual Language Models
par: Lin, Xiao, et autres
Publié: (2025)
par: Lin, Xiao, et autres
Publié: (2025)
Improving Speech Recognition of Named Entities in Classroom Speech with LLM Revision and Phonetic-Semantic Context
par: Trinh, Viet Anh, et autres
Publié: (2025)
par: Trinh, Viet Anh, et autres
Publié: (2025)
Documents similaires
-
RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
par: Lee, Harrison, et autres
Publié: (2023) -
Empirical Analysis of Decoding Biases in Masked Diffusion Models
par: Huang, Pengcheng, et autres
Publié: (2025) -
Speech Emotion Recognition with Distilled Prosodic and Linguistic Affect Representations
par: Shome, Debaditya, et autres
Publié: (2023) -
Reverse Constitutional AI: A Framework for Controllable Toxic Data Generation via Probability-Clamped RLAIF
par: Fang, Yuan, et autres
Publié: (2026) -
Recent Advances in End-to-End Simultaneous Speech Translation
par: Liu, Xiaoqian, et autres
Publié: (2024)