A General Highly Accurate Online Planning Method Integrating Large Language Models into Nested Rollout Policy Adaptation for Dialogue Tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Hui, Zhang, Fafa, Zhang, Xiaoyu, Mu, Chaoxu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
One for All: A General Framework of LLMs-based Multi-Criteria Decision Making on Human Expert Level
von: Wang, Hui, et al.
Veröffentlicht: (2025)
von: Wang, Hui, et al.
Veröffentlicht: (2025)
Planning of Heuristics: Strategic Planning on Large Language Models with Monte Carlo Tree Search for Automating Heuristic Optimization
von: Wang, Hui, et al.
Veröffentlicht: (2025)
von: Wang, Hui, et al.
Veröffentlicht: (2025)
CogMCTS: A Novel Cognitive-Guided Monte Carlo Tree Search Framework for Iterative Heuristic Evolution with Large Language Models
von: Wang, Hui, et al.
Veröffentlicht: (2025)
von: Wang, Hui, et al.
Veröffentlicht: (2025)
Generalized Nested Rollout Policy Adaptation with Limited Repetitions
von: Cazenave, Tristan
Veröffentlicht: (2024)
von: Cazenave, Tristan
Veröffentlicht: (2024)
Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Off-Policy Rollouts
von: Heuillet, Maxime, et al.
Veröffentlicht: (2025)
von: Heuillet, Maxime, et al.
Veröffentlicht: (2025)
DenseLoRA: Dense Low-Rank Adaptation of Large Language Models
von: Mu, Lin, et al.
Veröffentlicht: (2025)
von: Mu, Lin, et al.
Veröffentlicht: (2025)
Plug-and-Play Policy Planner for Large Language Model Powered Dialogue Agents
von: Deng, Yang, et al.
Veröffentlicht: (2023)
von: Deng, Yang, et al.
Veröffentlicht: (2023)
Knowledgeable Agents by Offline Reinforcement Learning from Large Language Model Rollouts
von: Pang, Jing-Cheng, et al.
Veröffentlicht: (2024)
von: Pang, Jing-Cheng, et al.
Veröffentlicht: (2024)
TLoRA: Task-aware Low Rank Adaptation of Large Language Models
von: Lin, Weicheng, et al.
Veröffentlicht: (2026)
von: Lin, Weicheng, et al.
Veröffentlicht: (2026)
Extracting Training Dialogue Data from Large Language Model based Task Bots
von: Zhang, Shuo, et al.
Veröffentlicht: (2026)
von: Zhang, Shuo, et al.
Veröffentlicht: (2026)
DuetSim: Building User Simulator with Dual Large Language Models for Task-Oriented Dialogues
von: Luo, Xiang, et al.
Veröffentlicht: (2024)
von: Luo, Xiang, et al.
Veröffentlicht: (2024)
Task as Context Prompting for Accurate Medical Symptom Coding Using Large Language Models
von: He, Chengyang, et al.
Veröffentlicht: (2025)
von: He, Chengyang, et al.
Veröffentlicht: (2025)
Anti-Overestimation Dialogue Policy Learning for Task-Completion Dialogue System
von: Tian, Chang, et al.
Veröffentlicht: (2022)
von: Tian, Chang, et al.
Veröffentlicht: (2022)
Large Language Models as User-Agents for Evaluating Task-Oriented-Dialogue Systems
von: Kazi, Taaha, et al.
Veröffentlicht: (2024)
von: Kazi, Taaha, et al.
Veröffentlicht: (2024)
General365: Benchmarking General Reasoning in Large Language Models Across Diverse and Challenging Tasks
von: Liu, Junlin, et al.
Veröffentlicht: (2026)
von: Liu, Junlin, et al.
Veröffentlicht: (2026)
TSS GAZ PTP: Towards Improving Gumbel AlphaZero with Two-stage Self-play for Multi-constrained Electric Vehicle Routing Problems
von: Wang, Hui, et al.
Veröffentlicht: (2025)
von: Wang, Hui, et al.
Veröffentlicht: (2025)
Driving Everywhere with Large Language Model Policy Adaptation
von: Li, Boyi, et al.
Veröffentlicht: (2024)
von: Li, Boyi, et al.
Veröffentlicht: (2024)
DFlow: Diverse Dialogue Flow Simulation with Large Language Models
von: Du, Wanyu, et al.
Veröffentlicht: (2024)
von: Du, Wanyu, et al.
Veröffentlicht: (2024)
Leveraging Graph Structures and Large Language Models for End-to-End Synthetic Task-Oriented Dialogues
von: Medjad, Maya, et al.
Veröffentlicht: (2025)
von: Medjad, Maya, et al.
Veröffentlicht: (2025)
A Prompt-driven Task Planning Method for Multi-drones based on Large Language Model
von: Liu, Yaohua
Veröffentlicht: (2024)
von: Liu, Yaohua
Veröffentlicht: (2024)
Knowledge Graph Fusion with Large Language Models for Accurate, Explainable Manufacturing Process Planning
von: Hoang, Danny, et al.
Veröffentlicht: (2025)
von: Hoang, Danny, et al.
Veröffentlicht: (2025)
Efficient Task Adaptation in Large Language Models via Selective Parameter Optimization
von: Wan, Weijie, et al.
Veröffentlicht: (2026)
von: Wan, Weijie, et al.
Veröffentlicht: (2026)
DND: Boosting Large Language Models with Dynamic Nested Depth
von: Chen, Tieyuan, et al.
Veröffentlicht: (2025)
von: Chen, Tieyuan, et al.
Veröffentlicht: (2025)
Multi-Party Supervised Fine-tuning of Language Models for Multi-Party Dialogue Generation
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
Task-Aligned Tool Recommendation for Large Language Models
von: Gao, Hang, et al.
Veröffentlicht: (2024)
von: Gao, Hang, et al.
Veröffentlicht: (2024)
Tree-Planner: Efficient Close-loop Task Planning with Large Language Models
von: Hu, Mengkang, et al.
Veröffentlicht: (2023)
von: Hu, Mengkang, et al.
Veröffentlicht: (2023)
SPEC-RL: Accelerating On-Policy Reinforcement Learning with Speculative Rollouts
von: Liu, Bingshuai, et al.
Veröffentlicht: (2025)
von: Liu, Bingshuai, et al.
Veröffentlicht: (2025)
Large Language Models as Planning Domain Generators
von: Oswald, James, et al.
Veröffentlicht: (2024)
von: Oswald, James, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models in Analysing Classroom Dialogue
von: Long, Yun, et al.
Veröffentlicht: (2024)
von: Long, Yun, et al.
Veröffentlicht: (2024)
Contextual Attention Modulation: Towards Efficient Multi-Task Adaptation in Large Language Models
von: Pan, Dayan, et al.
Veröffentlicht: (2025)
von: Pan, Dayan, et al.
Veröffentlicht: (2025)
Large Language Models as Generalizable Policies for Embodied Tasks
von: Szot, Andrew, et al.
Veröffentlicht: (2023)
von: Szot, Andrew, et al.
Veröffentlicht: (2023)
Sparsity Induction for Accurate Post-Training Pruning of Large Language Models
von: Jiang, Minhao, et al.
Veröffentlicht: (2026)
von: Jiang, Minhao, et al.
Veröffentlicht: (2026)
SELP: Generating Safe and Efficient Task Plans for Robot Agents with Large Language Models
von: Wu, Yi, et al.
Veröffentlicht: (2024)
von: Wu, Yi, et al.
Veröffentlicht: (2024)
Prompting Fairness: Integrating Causality to Debias Large Language Models
von: Li, Jingling, et al.
Veröffentlicht: (2024)
von: Li, Jingling, et al.
Veröffentlicht: (2024)
TaskBench: Benchmarking Large Language Models for Task Automation
von: Shen, Yongliang, et al.
Veröffentlicht: (2023)
von: Shen, Yongliang, et al.
Veröffentlicht: (2023)
A Survey of the Evolution of Language Model-Based Dialogue Systems: Data, Task and Models
von: Wang, Hongru, et al.
Veröffentlicht: (2023)
von: Wang, Hongru, et al.
Veröffentlicht: (2023)
DiffAgent: Fast and Accurate Text-to-Image API Selection with Large Language Model
von: Zhao, Lirui, et al.
Veröffentlicht: (2024)
von: Zhao, Lirui, et al.
Veröffentlicht: (2024)
Reasoning Pattern Matters: Learning to Reason without Human Rationales
von: Pang, Chaoxu, et al.
Veröffentlicht: (2025)
von: Pang, Chaoxu, et al.
Veröffentlicht: (2025)
Deploying Multi-task Online Server with Large Language Model
von: Qu, Yincen, et al.
Veröffentlicht: (2024)
von: Qu, Yincen, et al.
Veröffentlicht: (2024)
Aligning Large Language Models with Healthcare Stakeholders: A Pathway to Trustworthy AI Integration
von: Ding, Kexin, et al.
Veröffentlicht: (2025)
von: Ding, Kexin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
One for All: A General Framework of LLMs-based Multi-Criteria Decision Making on Human Expert Level
von: Wang, Hui, et al.
Veröffentlicht: (2025) -
Planning of Heuristics: Strategic Planning on Large Language Models with Monte Carlo Tree Search for Automating Heuristic Optimization
von: Wang, Hui, et al.
Veröffentlicht: (2025) -
CogMCTS: A Novel Cognitive-Guided Monte Carlo Tree Search Framework for Iterative Heuristic Evolution with Large Language Models
von: Wang, Hui, et al.
Veröffentlicht: (2025) -
Generalized Nested Rollout Policy Adaptation with Limited Repetitions
von: Cazenave, Tristan
Veröffentlicht: (2024) -
Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Off-Policy Rollouts
von: Heuillet, Maxime, et al.
Veröffentlicht: (2025)