DialogueReason: Rule-Based RL Sparks Dialogue Reasoning in LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shu, Yubo, Huang, Zhewei, Wu, Xin, Hu, Chen, Zhou, Shuchang, Jiang, Daxin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reasoning or Not? A Comprehensive Evaluation of Reasoning LLMs for Dialogue Summarization
von: Jin, Keyan, et al.
Veröffentlicht: (2025)
von: Jin, Keyan, et al.
Veröffentlicht: (2025)
Tailored Conversations beyond LLMs: A RL-Based Dialogue Manager
von: Galland, Lucie, et al.
Veröffentlicht: (2025)
von: Galland, Lucie, et al.
Veröffentlicht: (2025)
Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
von: Xie, Tian, et al.
Veröffentlicht: (2025)
von: Xie, Tian, et al.
Veröffentlicht: (2025)
TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL
von: Yang, Tianze, et al.
Veröffentlicht: (2026)
von: Yang, Tianze, et al.
Veröffentlicht: (2026)
Towards Dialogues for Joint Human-AI Reasoning and Value Alignment
von: Bezou-Vrakatseli, Elfia, et al.
Veröffentlicht: (2024)
von: Bezou-Vrakatseli, Elfia, et al.
Veröffentlicht: (2024)
Rethinking and Benchmarking Large Language Models for Graph Reasoning
von: Hu, Yuwei, et al.
Veröffentlicht: (2025)
von: Hu, Yuwei, et al.
Veröffentlicht: (2025)
Mix-Ecom: Towards Mixed-Type E-Commerce Dialogues with Complex Domain Rules
von: Zhou, Chenyu, et al.
Veröffentlicht: (2025)
von: Zhou, Chenyu, et al.
Veröffentlicht: (2025)
TReMu: Towards Neuro-Symbolic Temporal Reasoning for LLM-Agents with Memory in Multi-Session Dialogues
von: Ge, Yubin, et al.
Veröffentlicht: (2025)
von: Ge, Yubin, et al.
Veröffentlicht: (2025)
Beyond Overlap Metrics: Rewarding Reasoning and Preferences for Faithful Multi-Role Dialogue Summarization
von: Mei, Xiaoyong, et al.
Veröffentlicht: (2026)
von: Mei, Xiaoyong, et al.
Veröffentlicht: (2026)
Scalable and Accurate Graph Reasoning with LLM-based Multi-Agents
von: Hu, Yuwei, et al.
Veröffentlicht: (2024)
von: Hu, Yuwei, et al.
Veröffentlicht: (2024)
Sparks of Rationality: Do Reasoning LLMs Align with Human Judgment and Choice?
von: Tak, Ala N., et al.
Veröffentlicht: (2026)
von: Tak, Ala N., et al.
Veröffentlicht: (2026)
Injecting Salesperson's Dialogue Strategies in Large Language Models with Chain-of-Thought Reasoning
von: Chang, Wen-Yu, et al.
Veröffentlicht: (2024)
von: Chang, Wen-Yu, et al.
Veröffentlicht: (2024)
RuleArena: A Benchmark for Rule-Guided Reasoning with LLMs in Real-World Scenarios
von: Zhou, Ruiwen, et al.
Veröffentlicht: (2024)
von: Zhou, Ruiwen, et al.
Veröffentlicht: (2024)
METRO: Towards Strategy Induction from Expert Dialogue Transcripts for Non-collaborative Dialogues
von: Yang, Haofu, et al.
Veröffentlicht: (2026)
von: Yang, Haofu, et al.
Veröffentlicht: (2026)
STRIDE-ED: A Strategy-Grounded Stepwise Reasoning Framework for Empathetic Dialogue Systems
von: Ji, Hongru, et al.
Veröffentlicht: (2026)
von: Ji, Hongru, et al.
Veröffentlicht: (2026)
The Challenge of Teaching Reasoning to LLMs Without RL or Distillation
von: Du, Wei, et al.
Veröffentlicht: (2025)
von: Du, Wei, et al.
Veröffentlicht: (2025)
Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL
von: Liu, Che, et al.
Veröffentlicht: (2025)
von: Liu, Che, et al.
Veröffentlicht: (2025)
Reasoning Like a Doctor: Improving Medical Dialogue Systems via Diagnostic Reasoning Process Alignment
von: Xu, Kaishuai, et al.
Veröffentlicht: (2024)
von: Xu, Kaishuai, et al.
Veröffentlicht: (2024)
Thinking by Doing: Building Efficient World Model Reasoning in LLMs via Multi-turn Interaction
von: Shu, Bao, et al.
Veröffentlicht: (2025)
von: Shu, Bao, et al.
Veröffentlicht: (2025)
DialogueForge: LLM Simulation of Human-Chatbot Dialogue
von: Zhu, Ruizhe, et al.
Veröffentlicht: (2025)
von: Zhu, Ruizhe, et al.
Veröffentlicht: (2025)
Causal Discovery and Counterfactual Reasoning to Optimize Persuasive Dialogue Policies
von: Zeng, Donghuo, et al.
Veröffentlicht: (2025)
von: Zeng, Donghuo, et al.
Veröffentlicht: (2025)
LLM-Driven Multi-Turn Task-Oriented Dialogue Synthesis for Realistic Reasoning
von: Zhu, Yu, et al.
Veröffentlicht: (2026)
von: Zhu, Yu, et al.
Veröffentlicht: (2026)
Unearthing Gems from Stones: Policy Optimization with Negative Sample Augmentation for LLM Reasoning
von: Yang, Zhaohui, et al.
Veröffentlicht: (2025)
von: Yang, Zhaohui, et al.
Veröffentlicht: (2025)
"In Dialogues We Learn": Towards Personalized Dialogue Without Pre-defined Profiles through In-Dialogue Learning
von: Cheng, Chuanqi, et al.
Veröffentlicht: (2024)
von: Cheng, Chuanqi, et al.
Veröffentlicht: (2024)
Random Policy Valuation is Enough for LLM Reasoning with Verifiable Rewards
von: He, Haoran, et al.
Veröffentlicht: (2025)
von: He, Haoran, et al.
Veröffentlicht: (2025)
SpotAgent: Grounding Visual Geo-localization in Large Vision-Language Models through Agentic Reasoning
von: Jia, Furong, et al.
Veröffentlicht: (2026)
von: Jia, Furong, et al.
Veröffentlicht: (2026)
Dialogue is Better Than Monologue: Instructing Medical LLMs via Strategical Conversations
von: Liu, Zijie, et al.
Veröffentlicht: (2025)
von: Liu, Zijie, et al.
Veröffentlicht: (2025)
CrystalReasoner: Reasoning and RL for Property-Conditioned Crystal Structure Generation
von: Wu, Yuyang, et al.
Veröffentlicht: (2026)
von: Wu, Yuyang, et al.
Veröffentlicht: (2026)
Pre-Storage Reasoning for Episodic Memory: Shifting Inference Burden to Memory for Personalized Dialogue
von: Kim, Sangyeop, et al.
Veröffentlicht: (2025)
von: Kim, Sangyeop, et al.
Veröffentlicht: (2025)
Dialogue-based Explanations for Logical Reasoning using Structured Argumentation
von: Ho, Loan, et al.
Veröffentlicht: (2025)
von: Ho, Loan, et al.
Veröffentlicht: (2025)
RL Squeezes, SFT Expands: A Comparative Study of Reasoning LLMs
von: Matsutani, Kohsei, et al.
Veröffentlicht: (2025)
von: Matsutani, Kohsei, et al.
Veröffentlicht: (2025)
Towards Negotiative Dialogue for the Talkamatic Dialogue Manager
von: Larsson, Staffan, et al.
Veröffentlicht: (2024)
von: Larsson, Staffan, et al.
Veröffentlicht: (2024)
LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL
von: Peng, Yingzhe, et al.
Veröffentlicht: (2025)
von: Peng, Yingzhe, et al.
Veröffentlicht: (2025)
Recent Advances in Attack and Defense Approaches of Large Language Models
von: Cui, Jing, et al.
Veröffentlicht: (2024)
von: Cui, Jing, et al.
Veröffentlicht: (2024)
SE-Agent: Self-Evolution Trajectory Optimization in Multi-Step Reasoning with LLM-Based Agents
von: Lin, Jiaye, et al.
Veröffentlicht: (2025)
von: Lin, Jiaye, et al.
Veröffentlicht: (2025)
OnRL-RAG: Real-Time Personalized Mental Health Dialogue System
von: Bilal, Ahsan, et al.
Veröffentlicht: (2025)
von: Bilal, Ahsan, et al.
Veröffentlicht: (2025)
Can MLLMs Absorb Math Reasoning Abilities from LLMs as Free Lunch?
von: Hu, Yijie, et al.
Veröffentlicht: (2025)
von: Hu, Yijie, et al.
Veröffentlicht: (2025)
Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models
von: Zhou, Guanghao, et al.
Veröffentlicht: (2025)
von: Zhou, Guanghao, et al.
Veröffentlicht: (2025)
Evaluating Bias in Spoken Dialogue LLMs for Real-World Decisions and Recommendations
von: Wu, Yihao, et al.
Veröffentlicht: (2025)
von: Wu, Yihao, et al.
Veröffentlicht: (2025)
Improving Rule-based Reasoning in LLMs using Neurosymbolic Representations
von: Dhanraj, Varun, et al.
Veröffentlicht: (2025)
von: Dhanraj, Varun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Reasoning or Not? A Comprehensive Evaluation of Reasoning LLMs for Dialogue Summarization
von: Jin, Keyan, et al.
Veröffentlicht: (2025) -
Tailored Conversations beyond LLMs: A RL-Based Dialogue Manager
von: Galland, Lucie, et al.
Veröffentlicht: (2025) -
Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
von: Xie, Tian, et al.
Veröffentlicht: (2025) -
TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL
von: Yang, Tianze, et al.
Veröffentlicht: (2026) -
Towards Dialogues for Joint Human-AI Reasoning and Value Alignment
von: Bezou-Vrakatseli, Elfia, et al.
Veröffentlicht: (2024)