From Prediction to Justification: Aligning Sentiment Reasoning with Human Rationale via Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Shihao, Wang, Ziwei, Zhou, Jie, Wu, Yulan, Chen, Qin, Lei, Zhikai, Yu, Liyang, Dou, Liang, He, Liang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Boosting Large Language Models with Continual Learning for Aspect-based Sentiment Analysis
by: Ding, Xuanwen, et al.
Published: (2024)
by: Ding, Xuanwen, et al.
Published: (2024)
Learning Intrinsic Dimension via Information Bottleneck for Explainable Aspect-based Sentiment Analysis
by: Cheng, Zhenxiao, et al.
Published: (2024)
by: Cheng, Zhenxiao, et al.
Published: (2024)
Enhanced Multimodal Aspect-Based Sentiment Analysis by LLM-Generated Rationales
by: Cao, Jun, et al.
Published: (2025)
by: Cao, Jun, et al.
Published: (2025)
Optimizing Question Semantic Space for Dynamic Retrieval-Augmented Multi-hop Question Answering
by: Ye, Linhao, et al.
Published: (2025)
by: Ye, Linhao, et al.
Published: (2025)
Enhancing Event Causality Identification with Rationale and Structure-Aware Causal Question Answering
by: Zhang, Baiyan, et al.
Published: (2024)
by: Zhang, Baiyan, et al.
Published: (2024)
A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law
by: Pan, Qianjun, et al.
Published: (2025)
by: Pan, Qianjun, et al.
Published: (2025)
Boosting Conversational Question Answering with Fine-Grained Retrieval-Augmentation and Self-Check
by: Ye, Linhao, et al.
Published: (2024)
by: Ye, Linhao, et al.
Published: (2024)
StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
Reinforced Interactive Continual Learning via Real-time Noisy Human Feedback
by: Yang, Yutao, et al.
Published: (2025)
by: Yang, Yutao, et al.
Published: (2025)
Rewarding Creativity: A Human-Aligned Generative Reward Model for Reinforcement Learning in Storytelling
by: Li, Zhaoyan, et al.
Published: (2026)
by: Li, Zhaoyan, et al.
Published: (2026)
VisualQuality-R1: Reasoning-Induced Image Quality Assessment via Reinforcement Learning to Rank
by: Wu, Tianhe, et al.
Published: (2025)
by: Wu, Tianhe, et al.
Published: (2025)
LLM-KT: Aligning Large Language Models with Knowledge Tracing using a Plug-and-Play Instruction
by: Wang, Ziwei, et al.
Published: (2025)
by: Wang, Ziwei, et al.
Published: (2025)
Reasoning Pattern Matters: Learning to Reason without Human Rationales
by: Pang, Chaoxu, et al.
Published: (2025)
by: Pang, Chaoxu, et al.
Published: (2025)
From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought
by: Tan, Wentao, et al.
Published: (2025)
by: Tan, Wentao, et al.
Published: (2025)
Code-Driven Inductive Synthesis: Enhancing Reasoning Abilities of Large Language Models with Sequences
by: Chen, Kedi, et al.
Published: (2025)
by: Chen, Kedi, et al.
Published: (2025)
Structural Rationale Distillation via Reasoning Space Compression
by: Yang, Jialin, et al.
Published: (2026)
by: Yang, Jialin, et al.
Published: (2026)
Rationale-Grounded In-Context Learning for Time Series Reasoning with Multimodal Large Language Models
by: Liu, Qingxiang, et al.
Published: (2026)
by: Liu, Qingxiang, et al.
Published: (2026)
Learning User Interests via Reasoning and Distillation for Cross-Domain News Recommendation
by: Zhu, Mengdan, et al.
Published: (2026)
by: Zhu, Mengdan, et al.
Published: (2026)
SHARP: Synthesizing High-quality Aligned Reasoning Problems for Large Reasoning Models Reinforcement Learning
by: Wu, Xiong Jun, et al.
Published: (2025)
by: Wu, Xiong Jun, et al.
Published: (2025)
SymbolicThought: Integrating Language Models and Symbolic Reasoning for Consistent and Interpretable Human Relationship Understanding
by: Zhao, Runcong, et al.
Published: (2025)
by: Zhao, Runcong, et al.
Published: (2025)
COPR: Continual Human Preference Learning via Optimal Policy Regularization
by: Zhang, Han, et al.
Published: (2024)
by: Zhang, Han, et al.
Published: (2024)
TempoGPT: Enhancing Time Series Reasoning via Quantizing Embedding
by: Zhang, Haochuan, et al.
Published: (2025)
by: Zhang, Haochuan, et al.
Published: (2025)
Aligning Attention with Human Rationales for Self-Explaining Hate Speech Detection
by: Eilertsen, Brage, et al.
Published: (2025)
by: Eilertsen, Brage, et al.
Published: (2025)
SenseAI: A Human-in-the-Loop Dataset for RLHF-Aligned Financial Sentiment Reasoning
by: Kabalisa, Berny
Published: (2026)
by: Kabalisa, Berny
Published: (2026)
Let's Rectify Step by Step: Improving Aspect-based Sentiment Analysis with Diffusion Models
by: Liu, Shunyu, et al.
Published: (2024)
by: Liu, Shunyu, et al.
Published: (2024)
Towards Rationale-Answer Alignment of LVLMs via Self-Rationale Calibration
by: Wu, Yuanchen, et al.
Published: (2025)
by: Wu, Yuanchen, et al.
Published: (2025)
Code-driven Number Sequence Calculation: Enhancing the inductive Reasoning Abilities of Large Language Models
by: Chen, Kedi, et al.
Published: (2025)
by: Chen, Kedi, et al.
Published: (2025)
Calibrating LLMs with Preference Optimization on Thought Trees for Generating Rationale in Science Question Scoring
by: Li, Jiazheng, et al.
Published: (2024)
by: Li, Jiazheng, et al.
Published: (2024)
Set-Aligning Framework for Auto-Regressive Event Temporal Graph Generation
by: Tan, Xingwei, et al.
Published: (2024)
by: Tan, Xingwei, et al.
Published: (2024)
From Illusion to Intention: Visual Rationale Learning for Vision-Language Reasoning
by: Wang, Changpeng, et al.
Published: (2025)
by: Wang, Changpeng, et al.
Published: (2025)
Beyond the Trade-off: Self-Supervised Reinforcement Learning for Reasoning Models' Instruction Following
by: Ren, Qingyu, et al.
Published: (2025)
by: Ren, Qingyu, et al.
Published: (2025)
Towards a Unified Textual Graph Framework for Spectral Reasoning via Physical and Chemical Information Fusion
by: Liang, Jiheng, et al.
Published: (2025)
by: Liang, Jiheng, et al.
Published: (2025)
AlignSAM: Aligning Segment Anything Model to Open Context via Reinforcement Learning
by: Huang, Duojun, et al.
Published: (2024)
by: Huang, Duojun, et al.
Published: (2024)
A Regularization-based Transfer Learning Method for Information Extraction via Instructed Graph Decoder
by: Chen, Kedi, et al.
Published: (2024)
by: Chen, Kedi, et al.
Published: (2024)
Too Correct to Learn: Reinforcement Learning on Saturated Reasoning Data
by: Liang, Zhenwen, et al.
Published: (2026)
by: Liang, Zhenwen, et al.
Published: (2026)
MetaXCR: Reinforcement-Based Meta-Transfer Learning for Cross-Lingual Commonsense Reasoning
by: He, Jie, et al.
Published: (2025)
by: He, Jie, et al.
Published: (2025)
Mitigating Strategy Preference Bias in Emotional Support Conversation via Uncertainty Estimations
by: Zhou, Yougen, et al.
Published: (2025)
by: Zhou, Yougen, et al.
Published: (2025)
Tap-to-Adapt: Learning User-Aligned Response Timing for Speech Agents
by: He, Zihong, et al.
Published: (2026)
by: He, Zihong, et al.
Published: (2026)
DeepTrans: Deep Reasoning Translation via Reinforcement Learning
by: Wang, Jiaan, et al.
Published: (2025)
by: Wang, Jiaan, et al.
Published: (2025)
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning
by: Wang, Shaojie, et al.
Published: (2026)
by: Wang, Shaojie, et al.
Published: (2026)
Similar Items
-
Boosting Large Language Models with Continual Learning for Aspect-based Sentiment Analysis
by: Ding, Xuanwen, et al.
Published: (2024) -
Learning Intrinsic Dimension via Information Bottleneck for Explainable Aspect-based Sentiment Analysis
by: Cheng, Zhenxiao, et al.
Published: (2024) -
Enhanced Multimodal Aspect-Based Sentiment Analysis by LLM-Generated Rationales
by: Cao, Jun, et al.
Published: (2025) -
Optimizing Question Semantic Space for Dynamic Retrieval-Augmented Multi-hop Question Answering
by: Ye, Linhao, et al.
Published: (2025) -
Enhancing Event Causality Identification with Rationale and Structure-Aware Causal Question Answering
by: Zhang, Baiyan, et al.
Published: (2024)