Robust Planning with LLM-Modulo Framework: Case Study in Travel Planning
Fuente:
arXiv
Saved in:
| Main Authors: | Gundawar, Atharva, Verma, Mudit, Guan, Lin, Valmeekam, Karthik, Bhambri, Siddhant, Kambhampati, Subbarao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robust Planning with Compound LLM Architectures: An LLM-Modulo Approach
by: Gundawar, Atharva, et al.
Published: (2024)
by: Gundawar, Atharva, et al.
Published: (2024)
LLMs Can't Plan, But Can Help Planning in LLM-Modulo Frameworks
by: Kambhampati, Subbarao, et al.
Published: (2024)
by: Kambhampati, Subbarao, et al.
Published: (2024)
Planning in Strawberry Fields: Evaluating and Improving the Planning and Scheduling Capabilities of LRM o1
by: Valmeekam, Karthik, et al.
Published: (2024)
by: Valmeekam, Karthik, et al.
Published: (2024)
On the Brittle Foundations of ReAct Prompting for Agentic Large Language Models
by: Verma, Mudit, et al.
Published: (2024)
by: Verma, Mudit, et al.
Published: (2024)
Theory of Mind abilities of Large Language Models in Human-Robot Interaction : An Illusion?
by: Verma, Mudit, et al.
Published: (2024)
by: Verma, Mudit, et al.
Published: (2024)
Beyond Semantics: The Unreasonable Effectiveness of Reasonless Intermediate Tokens
by: Valmeekam, Karthik, et al.
Published: (2025)
by: Valmeekam, Karthik, et al.
Published: (2025)
Chain of Thoughtlessness? An Analysis of CoT in Planning
by: Stechly, Kaya, et al.
Published: (2024)
by: Stechly, Kaya, et al.
Published: (2024)
LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench
by: Valmeekam, Karthik, et al.
Published: (2024)
by: Valmeekam, Karthik, et al.
Published: (2024)
On the Self-Verification Limitations of Large Language Models on Reasoning and Planning Tasks
by: Stechly, Kaya, et al.
Published: (2024)
by: Stechly, Kaya, et al.
Published: (2024)
Do Cognitively Interpretable Reasoning Traces Improve LLM Performance?
by: Bhambri, Siddhant, et al.
Published: (2025)
by: Bhambri, Siddhant, et al.
Published: (2025)
Incorporating Human Flexibility through Reward Preferences in Human-AI Teaming
by: Bhambri, Siddhant, et al.
Published: (2023)
by: Bhambri, Siddhant, et al.
Published: (2023)
Interpretable Traces, Unexpected Outcomes: Investigating the Disconnect in Trace-Based Knowledge Distillation
by: Bhambri, Siddhant, et al.
Published: (2025)
by: Bhambri, Siddhant, et al.
Published: (2025)
Who is Helping Whom? Analyzing Inter-dependencies to Evaluate Cooperation in Human-AI Teaming
by: Biswas, Upasana, et al.
Published: (2025)
by: Biswas, Upasana, et al.
Published: (2025)
Performative Thinking? The Brittle Correlation Between CoT Length and Problem Complexity
by: Palod, Vardhan, et al.
Published: (2025)
by: Palod, Vardhan, et al.
Published: (2025)
Position: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces!
by: Kambhampati, Subbarao, et al.
Published: (2025)
by: Kambhampati, Subbarao, et al.
Published: (2025)
Extracting Heuristics from Large Language Models for Reward Shaping in Reinforcement Learning
by: Bhambri, Siddhant, et al.
Published: (2024)
by: Bhambri, Siddhant, et al.
Published: (2024)
Can Large Language Models Reason and Plan?
by: Kambhampati, Subbarao
Published: (2024)
by: Kambhampati, Subbarao
Published: (2024)
Local Coherence or Global Validity? Investigating RLVR Traces in Math Domains
by: Samineni, Soumya Rani, et al.
Published: (2025)
by: Samineni, Soumya Rani, et al.
Published: (2025)
RL in Name Only? Analyzing the Structural Assumptions in RL post-training for LLMs
by: Samineni, Soumya Rani, et al.
Published: (2025)
by: Samineni, Soumya Rani, et al.
Published: (2025)
Explainable Human-AI Interaction: A Planning Perspective
by: Sreedharan, Sarath, et al.
Published: (2024)
by: Sreedharan, Sarath, et al.
Published: (2024)
(How) Do reasoning models reason?
by: Subbarao Kambhampati, et al.
Published: (2025)
by: Subbarao Kambhampati, et al.
Published: (2025)
Mind The Gap: Quantifying Mechanistic Gaps in Algorithmic Reasoning via Neural Compilation
by: Saldyt, Lucas, et al.
Published: (2025)
by: Saldyt, Lucas, et al.
Published: (2025)
PAC Bench: Do Foundation Models Understand Prerequisites for Executing Manipulation Policies?
by: Gundawar, Atharva, et al.
Published: (2025)
by: Gundawar, Atharva, et al.
Published: (2025)
Algorithmic Language Models with Neurally Compiled Libraries
by: Saldyt, Lucas, et al.
Published: (2024)
by: Saldyt, Lucas, et al.
Published: (2024)
Superior Computer Chess with Model Predictive Control, Reinforcement Learning, and Rollout
by: Gundawar, Atharva, et al.
Published: (2024)
by: Gundawar, Atharva, et al.
Published: (2024)
TravelEval: A Comprehensive Benchmarking Framework for Evaluating LLM-Powered Travel Planning Agents
by: Chen, Weiyi, et al.
Published: (2026)
by: Chen, Weiyi, et al.
Published: (2026)
Task Success is not Enough: Investigating the Use of Video-Language Models as Behavior Critics for Catching Undesirable Agent Behaviors
by: Guan, Lin, et al.
Published: (2024)
by: Guan, Lin, et al.
Published: (2024)
DeepTravel: An End-to-End Agentic Reinforcement Learning Framework for Autonomous Travel Planning Agents
by: Ning, Yansong, et al.
Published: (2025)
by: Ning, Yansong, et al.
Published: (2025)
TravelAgent: An AI Assistant for Personalized Travel Planning
by: Chen, Aili, et al.
Published: (2024)
by: Chen, Aili, et al.
Published: (2024)
Bridging LLM Planning Agents and Formal Methods: A Case Study in Plan Verification
by: Ramani, Keshav, et al.
Published: (2025)
by: Ramani, Keshav, et al.
Published: (2025)
Hindsight PRIORs for Reward Learning from Human Preferences
by: Verma, Mudit, et al.
Published: (2024)
by: Verma, Mudit, et al.
Published: (2024)
NL2Plan: Robust LLM-Driven Planning from Minimal Text Descriptions
by: Gestrin, Elliot, et al.
Published: (2024)
by: Gestrin, Elliot, et al.
Published: (2024)
Revisiting the Travel Planning Capabilities of Large Language Models
by: Zhang, Bo-Wen, et al.
Published: (2026)
by: Zhang, Bo-Wen, et al.
Published: (2026)
PhyPlan: Generalizable and Rapid Physical Task Planning with Physics Informed Skill Networks for Robot Manipulators
by: Chopra, Mudit, et al.
Published: (2024)
by: Chopra, Mudit, et al.
Published: (2024)
Can large language models reason and plan?
by: Subbarao Kambhampati
Published: (2024)
by: Subbarao Kambhampati
Published: (2024)
SMaRT: Select, Mix, and ReinvenT -- A Strategy Fusion Framework for LLM-Driven Reasoning and Planning
by: Verma, Nikhil, et al.
Published: (2025)
by: Verma, Nikhil, et al.
Published: (2025)
HiMAP-Travel: Hierarchical Multi-Agent Planning for Long-Horizon Constrained Travel
by: Bui, The Viet, et al.
Published: (2026)
by: Bui, The Viet, et al.
Published: (2026)
RETAIL: Towards Real-world Travel Planning for Large Language Models
by: Deng, Bin, et al.
Published: (2025)
by: Deng, Bin, et al.
Published: (2025)
TripTailor: A Real-World Benchmark for Personalized Travel Planning
by: Shen, Yuanzhe, et al.
Published: (2025)
by: Shen, Yuanzhe, et al.
Published: (2025)
A Framework for Benchmarking and Aligning Task-Planning Safety in LLM-Based Embodied Agents
by: Huang, Yuting, et al.
Published: (2025)
by: Huang, Yuting, et al.
Published: (2025)
Similar Items
-
Robust Planning with Compound LLM Architectures: An LLM-Modulo Approach
by: Gundawar, Atharva, et al.
Published: (2024) -
LLMs Can't Plan, But Can Help Planning in LLM-Modulo Frameworks
by: Kambhampati, Subbarao, et al.
Published: (2024) -
Planning in Strawberry Fields: Evaluating and Improving the Planning and Scheduling Capabilities of LRM o1
by: Valmeekam, Karthik, et al.
Published: (2024) -
On the Brittle Foundations of ReAct Prompting for Agentic Large Language Models
by: Verma, Mudit, et al.
Published: (2024) -
Theory of Mind abilities of Large Language Models in Human-Robot Interaction : An Illusion?
by: Verma, Mudit, et al.
Published: (2024)