Make Planning Research Rigorous Again!
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Katz, Michael, Kokel, Harsha, Muise, Christian, Sohrabi, Shirin, Sreedharan, Sarath |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Seemingly Simple Planning Problems are Computationally Challenging: The Countdown Game
von: Katz, Michael, et al.
Veröffentlicht: (2025)
von: Katz, Michael, et al.
Veröffentlicht: (2025)
ACPBench: Reasoning about Action, Change, and Planning
von: Kokel, Harsha, et al.
Veröffentlicht: (2024)
von: Kokel, Harsha, et al.
Veröffentlicht: (2024)
ACPBench Hard: Unrestrained Reasoning about Action, Change, and Planning
von: Kokel, Harsha, et al.
Veröffentlicht: (2025)
von: Kokel, Harsha, et al.
Veröffentlicht: (2025)
Thought of Search: Planning with Language Models Through The Lens of Efficiency
von: Katz, Michael, et al.
Veröffentlicht: (2024)
von: Katz, Michael, et al.
Veröffentlicht: (2024)
Planning in the LLM Era: Building for Reliability and Efficiency
von: Katz, Michael, et al.
Veröffentlicht: (2026)
von: Katz, Michael, et al.
Veröffentlicht: (2026)
Learning and Reusing Policy Decompositions for Hierarchical Generalized Planning with LLM Agents
von: Sohrabi, Shirin, et al.
Veröffentlicht: (2026)
von: Sohrabi, Shirin, et al.
Veröffentlicht: (2026)
Automating Thought of Search: A Journey Towards Soundness and Completeness
von: Cao, Daniel, et al.
Veröffentlicht: (2024)
von: Cao, Daniel, et al.
Veröffentlicht: (2024)
Large Language Models as Planning Domain Generators
von: Oswald, James, et al.
Veröffentlicht: (2024)
von: Oswald, James, et al.
Veröffentlicht: (2024)
Model Space Reasoning as Search in Feedback Space for Planning Domain Generation
von: Oswald, James, et al.
Veröffentlicht: (2026)
von: Oswald, James, et al.
Veröffentlicht: (2026)
QueryGym: Step-by-Step Interaction with Relational Databases
von: Ananthakrishnan, Haritha, et al.
Veröffentlicht: (2025)
von: Ananthakrishnan, Haritha, et al.
Veröffentlicht: (2025)
Unifying and Certifying Top-Quality Planning
von: Katz, Michael, et al.
Veröffentlicht: (2024)
von: Katz, Michael, et al.
Veröffentlicht: (2024)
Some Orders Are Important: Partially Preserving Orders in Top-Quality Planning
von: Katz, Michael, et al.
Veröffentlicht: (2024)
von: Katz, Michael, et al.
Veröffentlicht: (2024)
Can LLMs Fix Issues with Reasoning Models? Towards More Likely Models for AI Planning
von: Caglar, Turgay, et al.
Veröffentlicht: (2023)
von: Caglar, Turgay, et al.
Veröffentlicht: (2023)
Explainable Human-AI Interaction: A Planning Perspective
von: Sreedharan, Sarath, et al.
Veröffentlicht: (2024)
von: Sreedharan, Sarath, et al.
Veröffentlicht: (2024)
Expectation Alignment: Handling Reward Misspecification in the Presence of Expectation Mismatch
von: Mechergui, Malek, et al.
Veröffentlicht: (2024)
von: Mechergui, Malek, et al.
Veröffentlicht: (2024)
Human-Modeling in Sequential Decision-Making: An Analysis through the Lens of Human-Aware AI
von: Tulli, Silvia, et al.
Veröffentlicht: (2024)
von: Tulli, Silvia, et al.
Veröffentlicht: (2024)
LLMs as Planning Formalizers: A Survey for Leveraging Large Language Models to Construct Automated Planning Models
von: Tantakoun, Marcus, et al.
Veröffentlicht: (2025)
von: Tantakoun, Marcus, et al.
Veröffentlicht: (2025)
Combining Planning and Reinforcement Learning for Solving Relational Multiagent Domains
von: Prabhakar, Nikhilesh, et al.
Veröffentlicht: (2025)
von: Prabhakar, Nikhilesh, et al.
Veröffentlicht: (2025)
Reducing Human-Robot Goal State Divergence with Environment Design
von: Sikes, Kelsey, et al.
Veröffentlicht: (2024)
von: Sikes, Kelsey, et al.
Veröffentlicht: (2024)
BPMN to PDDL: Translating Business Workflows for AI Planning
von: Nie, Jasper, et al.
Veröffentlicht: (2025)
von: Nie, Jasper, et al.
Veröffentlicht: (2025)
Inferring Implicit Goals Across Differing Task Models
von: Tulli, Silvia, et al.
Veröffentlicht: (2025)
von: Tulli, Silvia, et al.
Veröffentlicht: (2025)
Less is More: Learning Graph Tasks with Just LLMs
von: Shirai, Sola, et al.
Veröffentlicht: (2025)
von: Shirai, Sola, et al.
Veröffentlicht: (2025)
SPEAR: Security Posture Evaluation using AI Planner-Reasoning on Attack-Connectivity Hypergraphs
von: Podder, Rakesh, et al.
Veröffentlicht: (2025)
von: Podder, Rakesh, et al.
Veröffentlicht: (2025)
On the Role of Domain Experts in Creating Effective Tutoring Systems
von: Sreedharan, Sarath, et al.
Veröffentlicht: (2025)
von: Sreedharan, Sarath, et al.
Veröffentlicht: (2025)
Born Again Neural Networks
von: Furlanello, Tommaso, et al.
Veröffentlicht: (2018)
von: Furlanello, Tommaso, et al.
Veröffentlicht: (2018)
JSONSchemaBench: A Rigorous Benchmark of Structured Outputs for Language Models
von: Geng, Saibo, et al.
Veröffentlicht: (2025)
von: Geng, Saibo, et al.
Veröffentlicht: (2025)
Making Sigmoid-MSE Great Again: Output Reset Challenges Softmax Cross-Entropy in Neural Network Classification
von: Tyagi, Kanishka, et al.
Veröffentlicht: (2024)
von: Tyagi, Kanishka, et al.
Veröffentlicht: (2024)
LLM-Based World Models Can Make Decisions Solely, But Rigorous Evaluations are Needed
von: Yang, Chang, et al.
Veröffentlicht: (2024)
von: Yang, Chang, et al.
Veröffentlicht: (2024)
SAM-SP: Self-Prompting Makes SAM Great Again
von: Zhou, Chunpeng, et al.
Veröffentlicht: (2024)
von: Zhou, Chunpeng, et al.
Veröffentlicht: (2024)
Sparse Autoencoders, Again?
von: Lu, Yin, et al.
Veröffentlicht: (2025)
von: Lu, Yin, et al.
Veröffentlicht: (2025)
From Robustness to Explainability and Back Again
von: Huang, Xuanxiang, et al.
Veröffentlicht: (2023)
von: Huang, Xuanxiang, et al.
Veröffentlicht: (2023)
Planetarium: A Rigorous Benchmark for Translating Text to Structured Planning Languages
von: Zuo, Max, et al.
Veröffentlicht: (2024)
von: Zuo, Max, et al.
Veröffentlicht: (2024)
Scaling LLM Planning: NL2FLOW for Parametric Problem Generation and Rigorous Evaluation
von: Kang, Jungkoo
Veröffentlicht: (2025)
von: Kang, Jungkoo
Veröffentlicht: (2025)
Planning Anything with Rigor: General-Purpose Zero-Shot Planning with LLM-based Formalized Programming
von: Hao, Yilun, et al.
Veröffentlicht: (2024)
von: Hao, Yilun, et al.
Veröffentlicht: (2024)
Make Shuffling Great Again: A Side-Channel Resistant Fisher-Yates Algorithm for Protecting Neural Networks
von: Puškáč, Leonard, et al.
Veröffentlicht: (2025)
von: Puškáč, Leonard, et al.
Veröffentlicht: (2025)
Interaction Theater: A case of LLM Agents Interacting at Scale
von: Shekkizhar, Sarath, et al.
Veröffentlicht: (2026)
von: Shekkizhar, Sarath, et al.
Veröffentlicht: (2026)
TabSketchFM: Sketch-based Tabular Representation Learning for Data Discovery over Data Lakes
von: Khatiwada, Aamod, et al.
Veröffentlicht: (2024)
von: Khatiwada, Aamod, et al.
Veröffentlicht: (2024)
AI Planning: A Primer and Survey (Preliminary Report)
von: Chen, Dillon Z., et al.
Veröffentlicht: (2024)
von: Chen, Dillon Z., et al.
Veröffentlicht: (2024)
Rigor in AI: Doing Rigorous AI Work Requires a Broader, Responsible AI-Informed Conception of Rigor
von: Olteanu, Alexandra, et al.
Veröffentlicht: (2025)
von: Olteanu, Alexandra, et al.
Veröffentlicht: (2025)
Make Graph Neural Networks Great Again: A Generic Integration Paradigm of Topology-Free Patterns for Traffic Speed Prediction
von: Zhou, Yicheng, et al.
Veröffentlicht: (2024)
von: Zhou, Yicheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Seemingly Simple Planning Problems are Computationally Challenging: The Countdown Game
von: Katz, Michael, et al.
Veröffentlicht: (2025) -
ACPBench: Reasoning about Action, Change, and Planning
von: Kokel, Harsha, et al.
Veröffentlicht: (2024) -
ACPBench Hard: Unrestrained Reasoning about Action, Change, and Planning
von: Kokel, Harsha, et al.
Veröffentlicht: (2025) -
Thought of Search: Planning with Language Models Through The Lens of Efficiency
von: Katz, Michael, et al.
Veröffentlicht: (2024) -
Planning in the LLM Era: Building for Reliability and Efficiency
von: Katz, Michael, et al.
Veröffentlicht: (2026)