Can LLMs Fix Issues with Reasoning Models? Towards More Likely Models for AI Planning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Caglar, Turgay, Belhaj, Sirine, Chakraborti, Tathagata, Katz, Michael, Sreedharan, Sarath |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Seemingly Simple Planning Problems are Computationally Challenging: The Countdown Game
von: Katz, Michael, et al.
Veröffentlicht: (2025)
von: Katz, Michael, et al.
Veröffentlicht: (2025)
SPEAR: Security Posture Evaluation using AI Planner-Reasoning on Attack-Connectivity Hypergraphs
von: Podder, Rakesh, et al.
Veröffentlicht: (2025)
von: Podder, Rakesh, et al.
Veröffentlicht: (2025)
Explainable Human-AI Interaction: A Planning Perspective
von: Sreedharan, Sarath, et al.
Veröffentlicht: (2024)
von: Sreedharan, Sarath, et al.
Veröffentlicht: (2024)
Make Planning Research Rigorous Again!
von: Katz, Michael, et al.
Veröffentlicht: (2025)
von: Katz, Michael, et al.
Veröffentlicht: (2025)
Human-Modeling in Sequential Decision-Making: An Analysis through the Lens of Human-Aware AI
von: Tulli, Silvia, et al.
Veröffentlicht: (2024)
von: Tulli, Silvia, et al.
Veröffentlicht: (2024)
Expectation Alignment: Handling Reward Misspecification in the Presence of Expectation Mismatch
von: Mechergui, Malek, et al.
Veröffentlicht: (2024)
von: Mechergui, Malek, et al.
Veröffentlicht: (2024)
Inferring Implicit Goals Across Differing Task Models
von: Tulli, Silvia, et al.
Veröffentlicht: (2025)
von: Tulli, Silvia, et al.
Veröffentlicht: (2025)
Reducing Human-Robot Goal State Divergence with Environment Design
von: Sikes, Kelsey, et al.
Veröffentlicht: (2024)
von: Sikes, Kelsey, et al.
Veröffentlicht: (2024)
How Good Are LLMs at Processing Tool Outputs?
von: Kate, Kiran, et al.
Veröffentlicht: (2025)
von: Kate, Kiran, et al.
Veröffentlicht: (2025)
Reasoning in Large Language Models: A Geometric Perspective
von: Cosentino, Romain, et al.
Veröffentlicht: (2024)
von: Cosentino, Romain, et al.
Veröffentlicht: (2024)
ACPBench: Reasoning about Action, Change, and Planning
von: Kokel, Harsha, et al.
Veröffentlicht: (2024)
von: Kokel, Harsha, et al.
Veröffentlicht: (2024)
Model Space Reasoning as Search in Feedback Space for Planning Domain Generation
von: Oswald, James, et al.
Veröffentlicht: (2026)
von: Oswald, James, et al.
Veröffentlicht: (2026)
Can only LLMs do Reasoning?: Potential of Small Language Models in Task Planning
von: Choi, Gawon, et al.
Veröffentlicht: (2024)
von: Choi, Gawon, et al.
Veröffentlicht: (2024)
Giving AI Personalities Leads to More Human-Like Reasoning
von: Nighojkar, Animesh, et al.
Veröffentlicht: (2025)
von: Nighojkar, Animesh, et al.
Veröffentlicht: (2025)
ACPBench Hard: Unrestrained Reasoning about Action, Change, and Planning
von: Kokel, Harsha, et al.
Veröffentlicht: (2025)
von: Kokel, Harsha, et al.
Veröffentlicht: (2025)
LLMs-in-the-loop Part-1: Expert Small AI Models for Bio-Medical Text Translation
von: Keles, Bunyamin, et al.
Veröffentlicht: (2024)
von: Keles, Bunyamin, et al.
Veröffentlicht: (2024)
Thought of Search: Planning with Language Models Through The Lens of Efficiency
von: Katz, Michael, et al.
Veröffentlicht: (2024)
von: Katz, Michael, et al.
Veröffentlicht: (2024)
Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
Can Large Language Models Reason and Plan?
von: Kambhampati, Subbarao
Veröffentlicht: (2024)
von: Kambhampati, Subbarao
Veröffentlicht: (2024)
Can Agents Fix Agent Issues?
von: Rahardja, Alfin Wijaya, et al.
Veröffentlicht: (2025)
von: Rahardja, Alfin Wijaya, et al.
Veröffentlicht: (2025)
On the Role of Domain Experts in Creating Effective Tutoring Systems
von: Sreedharan, Sarath, et al.
Veröffentlicht: (2025)
von: Sreedharan, Sarath, et al.
Veröffentlicht: (2025)
AI Planning: A Primer and Survey (Preliminary Report)
von: Chen, Dillon Z., et al.
Veröffentlicht: (2024)
von: Chen, Dillon Z., et al.
Veröffentlicht: (2024)
Towards More Standardized AI Evaluation: From Models to Agents
von: Filali, Ali El, et al.
Veröffentlicht: (2026)
von: Filali, Ali El, et al.
Veröffentlicht: (2026)
Bridging the Reasoning Gap: Small LLMs Can Plan with Generalised Strategies
von: Borro, Andrey, et al.
Veröffentlicht: (2025)
von: Borro, Andrey, et al.
Veröffentlicht: (2025)
Towards Practical Tool Usage for Continually Learning LLMs
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2024)
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2024)
Can LLM-Reasoning Models Replace Classical Planning? A Benchmark Study
von: Goebel, Kai, et al.
Veröffentlicht: (2025)
von: Goebel, Kai, et al.
Veröffentlicht: (2025)
Conversational AI as a Coding Assistant: Understanding Programmers' Interactions with and Expectations from Large Language Models for Coding
von: Akhoroz, Mehmet, et al.
Veröffentlicht: (2025)
von: Akhoroz, Mehmet, et al.
Veröffentlicht: (2025)
LLMs Can't Plan, But Can Help Planning in LLM-Modulo Frameworks
von: Kambhampati, Subbarao, et al.
Veröffentlicht: (2024)
von: Kambhampati, Subbarao, et al.
Veröffentlicht: (2024)
Can LLMs Generate Human-Like Wayfinding Instructions? Towards Platform-Agnostic Embodied Instruction Synthesis
von: Dorbala, Vishnu Sashank, et al.
Veröffentlicht: (2024)
von: Dorbala, Vishnu Sashank, et al.
Veröffentlicht: (2024)
The Hidden AI Race: Tracking Environmental Costs of Innovation
von: Agarwal, Shyam, et al.
Veröffentlicht: (2025)
von: Agarwal, Shyam, et al.
Veröffentlicht: (2025)
Off-Trajectory Reasoning: Can LLMs Collaborate on Reasoning Trajectory?
von: Li, Aochong Oliver, et al.
Veröffentlicht: (2025)
von: Li, Aochong Oliver, et al.
Veröffentlicht: (2025)
Confidence-Aware Alignment Makes Reasoning LLMs More Reliable
von: Chen, Kejia, et al.
Veröffentlicht: (2026)
von: Chen, Kejia, et al.
Veröffentlicht: (2026)
On Reasoning Strength Planning in Large Reasoning Models
von: Sheng, Leheng, et al.
Veröffentlicht: (2025)
von: Sheng, Leheng, et al.
Veröffentlicht: (2025)
Unifying and Certifying Top-Quality Planning
von: Katz, Michael, et al.
Veröffentlicht: (2024)
von: Katz, Michael, et al.
Veröffentlicht: (2024)
Improved Generalized Planning with LLMs through Strategy Refinement and Reflection
von: Stein, Katharina, et al.
Veröffentlicht: (2025)
von: Stein, Katharina, et al.
Veröffentlicht: (2025)
When Models Know More Than They Say: Probing Analogical Reasoning in LLMs
von: McGovern, Hope, et al.
Veröffentlicht: (2026)
von: McGovern, Hope, et al.
Veröffentlicht: (2026)
Large Language Models as Planning Domain Generators
von: Oswald, James, et al.
Veröffentlicht: (2024)
von: Oswald, James, et al.
Veröffentlicht: (2024)
Do Thinking Tokens Help or Trap? Towards More Efficient Large Reasoning Model
von: Ding, Bowen, et al.
Veröffentlicht: (2025)
von: Ding, Bowen, et al.
Veröffentlicht: (2025)
Promises, Outlooks and Challenges of Diffusion Language Modeling
von: Deschenaux, Justin, et al.
Veröffentlicht: (2024)
von: Deschenaux, Justin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Seemingly Simple Planning Problems are Computationally Challenging: The Countdown Game
von: Katz, Michael, et al.
Veröffentlicht: (2025) -
SPEAR: Security Posture Evaluation using AI Planner-Reasoning on Attack-Connectivity Hypergraphs
von: Podder, Rakesh, et al.
Veröffentlicht: (2025) -
Explainable Human-AI Interaction: A Planning Perspective
von: Sreedharan, Sarath, et al.
Veröffentlicht: (2024) -
Make Planning Research Rigorous Again!
von: Katz, Michael, et al.
Veröffentlicht: (2025) -
Human-Modeling in Sequential Decision-Making: An Analysis through the Lens of Human-Aware AI
von: Tulli, Silvia, et al.
Veröffentlicht: (2024)