Learning to Search and Searching to Learn for Generalization in Planning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Aichmüller, Michael, Hesse, Yannik, Geffner, Hector |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Sketch Decompositions for Classical Planning via Deep Reinforcement Learning
par: Aichmüller, Michael, et autres
Publié: (2024)
par: Aichmüller, Michael, et autres
Publié: (2024)
Efficient Lookahead Encoding and Abstracted Width for Learning General Policies in Classical Planning
par: Aichmüller, Michael, et autres
Publié: (2026)
par: Aichmüller, Michael, et autres
Publié: (2026)
Learning Generalized Policies for Fully Observable Non-Deterministic Planning Domains
par: Hofmann, Till, et autres
Publié: (2024)
par: Hofmann, Till, et autres
Publié: (2024)
Learning General Policies From Examples
par: Bonet, Blai, et autres
Publié: (2025)
par: Bonet, Blai, et autres
Publié: (2025)
Learning More Expressive General Policies for Classical Planning Domains
par: Ståhlberg, Simon, et autres
Publié: (2024)
par: Ståhlberg, Simon, et autres
Publié: (2024)
Differentiable Learning of Lifted Action Schemas for Classical Planning
par: Reiter, Jonas, et autres
Publié: (2026)
par: Reiter, Jonas, et autres
Publié: (2026)
Learning General Policies with Policy Gradient Methods
par: Ståhlberg, Simon, et autres
Publié: (2025)
par: Ståhlberg, Simon, et autres
Publié: (2025)
Learning Lifted STRIPS Models from Action Traces Alone: A Simple, General, and Scalable Solution
par: Gösgens, Jonas, et autres
Publié: (2024)
par: Gösgens, Jonas, et autres
Publié: (2024)
Symmetries and Expressive Requirements for Learning General Policies
par: Drexler, Dominik, et autres
Publié: (2024)
par: Drexler, Dominik, et autres
Publié: (2024)
Learning to Ground Existentially Quantified Goals
par: Funkquist, Martin, et autres
Publié: (2024)
par: Funkquist, Martin, et autres
Publié: (2024)
Learning Lifted Action Models From Traces of Incomplete Actions and States
par: Jansen, Niklas, et autres
Publié: (2025)
par: Jansen, Niklas, et autres
Publié: (2025)
Learning Lifted Action Models from Traces with Minimal Information About Actions and States
par: Gösgens, Jonas, et autres
Publié: (2026)
par: Gösgens, Jonas, et autres
Publié: (2026)
First-Order Representation Languages for Goal-Conditioned RL
par: Ståhlberg, Simon, et autres
Publié: (2025)
par: Ståhlberg, Simon, et autres
Publié: (2025)
On Policy Reuse: An Expressive Language for Representing and Executing General Policies that Call Other Policies
par: Bonet, Blai, et autres
Publié: (2024)
par: Bonet, Blai, et autres
Publié: (2024)
Hybrid Reinforcement Learning and Search for Flight Trajectory Planning
par: Luise, Alberto, et autres
Publié: (2025)
par: Luise, Alberto, et autres
Publié: (2025)
Plan Before Search: Search Agents Need Plan
par: Qian, Zhipeng, et autres
Publié: (2026)
par: Qian, Zhipeng, et autres
Publié: (2026)
Model Space Reasoning as Search in Feedback Space for Planning Domain Generation
par: Oswald, James, et autres
Publié: (2026)
par: Oswald, James, et autres
Publié: (2026)
Plans for Evaluating Structured Generative Search Summaries
par: Sakai, Tetsuya, et autres
Publié: (2026)
par: Sakai, Tetsuya, et autres
Publié: (2026)
ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning
par: Chen, Mingyang, et autres
Publié: (2025)
par: Chen, Mingyang, et autres
Publié: (2025)
Subgoal-Guided Policy Heuristic Search with Learned Subgoals
par: Tuero, Jake, et autres
Publié: (2025)
par: Tuero, Jake, et autres
Publié: (2025)
Thought of Search: Planning with Language Models Through The Lens of Efficiency
par: Katz, Michael, et autres
Publié: (2024)
par: Katz, Michael, et autres
Publié: (2024)
Iterated Local Search with Linkage Learning
par: Tinós, Renato, et autres
Publié: (2024)
par: Tinós, Renato, et autres
Publié: (2024)
Learning to Reason via Program Generation, Emulation, and Search
par: Weir, Nathaniel, et autres
Publié: (2024)
par: Weir, Nathaniel, et autres
Publié: (2024)
Stream of Search (SoS): Learning to Search in Language
par: Gandhi, Kanishk, et autres
Publié: (2024)
par: Gandhi, Kanishk, et autres
Publié: (2024)
AutoSearch: Adaptive Search Depth for Efficient Agentic RAG via Reinforcement Learning
par: Sun, Jingbo, et autres
Publié: (2026)
par: Sun, Jingbo, et autres
Publié: (2026)
Comparative Analysis of Parameterized Action Actor-Critic Reinforcement Learning Algorithms for Web Search Match Plan Generation
par: Bapoo, Ubayd, et autres
Publié: (2025)
par: Bapoo, Ubayd, et autres
Publié: (2025)
Heuristic Search for Multi-Objective Probabilistic Planning
par: Chen, Dillon, et autres
Publié: (2023)
par: Chen, Dillon, et autres
Publié: (2023)
Enhancing Reinforcement Learning Through Guided Search
par: Arjonilla, Jérôme, et autres
Publié: (2024)
par: Arjonilla, Jérôme, et autres
Publié: (2024)
Multi-Objective Neural Architecture Search by Learning Search Space Partitions
par: Zhao, Yiyang, et autres
Publié: (2024)
par: Zhao, Yiyang, et autres
Publié: (2024)
Probe-then-Plan: Environment-Aware Planning for Industrial E-commerce Search
par: Chen, Mengxiang, et autres
Publié: (2026)
par: Chen, Mengxiang, et autres
Publié: (2026)
Beyond A*: Better Planning with Transformers via Search Dynamics Bootstrapping
par: Lehnert, Lucas, et autres
Publié: (2024)
par: Lehnert, Lucas, et autres
Publié: (2024)
Generative Active Learning for the Search of Small-molecule Protein Binders
par: Korablyov, Maksym, et autres
Publié: (2024)
par: Korablyov, Maksym, et autres
Publié: (2024)
AI Research Agents for Machine Learning: Search, Exploration, and Generalization in MLE-bench
par: Toledo, Edan, et autres
Publié: (2025)
par: Toledo, Edan, et autres
Publié: (2025)
Efficient Solution and Learning of Robust Factored MDPs
par: Schnitzer, Yannik, et autres
Publié: (2025)
par: Schnitzer, Yannik, et autres
Publié: (2025)
AI-SearchPlanner: Modular Agentic Search via Pareto-Optimal Multi-Objective Reinforcement Learning
par: Mei, Lang, et autres
Publié: (2025)
par: Mei, Lang, et autres
Publié: (2025)
Towards Agentic Self-Learning LLMs in Search Environment
par: Sun, Wangtao, et autres
Publié: (2025)
par: Sun, Wangtao, et autres
Publié: (2025)
Optimizing Generative Ranking Relevance via Reinforcement Learning in Xiaohongshu Search
par: Zeng, Ziyang, et autres
Publié: (2025)
par: Zeng, Ziyang, et autres
Publié: (2025)
DecoupleSearch: Decouple Planning and Search via Hierarchical Reward Modeling
par: Sun, Hao, et autres
Publié: (2025)
par: Sun, Hao, et autres
Publié: (2025)
Learning Type-Generalized Actions for Symbolic Planning
par: Tanneberg, Daniel, et autres
Publié: (2023)
par: Tanneberg, Daniel, et autres
Publié: (2023)
Planning In Natural Language Improves LLM Search For Code Generation
par: Wang, Evan, et autres
Publié: (2024)
par: Wang, Evan, et autres
Publié: (2024)
Documents similaires
-
Sketch Decompositions for Classical Planning via Deep Reinforcement Learning
par: Aichmüller, Michael, et autres
Publié: (2024) -
Efficient Lookahead Encoding and Abstracted Width for Learning General Policies in Classical Planning
par: Aichmüller, Michael, et autres
Publié: (2026) -
Learning Generalized Policies for Fully Observable Non-Deterministic Planning Domains
par: Hofmann, Till, et autres
Publié: (2024) -
Learning General Policies From Examples
par: Bonet, Blai, et autres
Publié: (2025) -
Learning More Expressive General Policies for Classical Planning Domains
par: Ståhlberg, Simon, et autres
Publié: (2024)