Planning with a Learned Policy Basis to Optimally Solve Complex Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Infante, Guillermo, Kuric, David, Jonsson, Anders, Gómez, Vicenç, van Hoof, Herke |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Globally Optimal Hierarchical Reinforcement Learning for Linearly-Solvable Markov Decision Processes
by: Infante, Guillermo, et al.
Published: (2021)
by: Infante, Guillermo, et al.
Published: (2021)
Hierarchical Average-Reward Linearly-solvable Markov Decision Processes
by: Infante, Guillermo, et al.
Published: (2024)
by: Infante, Guillermo, et al.
Published: (2024)
Making Universal Policies Universal
by: Höpner, Niklas, et al.
Published: (2025)
by: Höpner, Niklas, et al.
Published: (2025)
Data Augmentation for Instruction Following Policies via Trajectory Segmentation
by: Höpner, Niklas, et al.
Published: (2025)
by: Höpner, Niklas, et al.
Published: (2025)
Bridge the Inference Gaps of Neural Processes via Expectation Maximization
by: Wang, Qi, et al.
Published: (2025)
by: Wang, Qi, et al.
Published: (2025)
Gradient-Based Program Synthesis with Neurally Interpreted Languages
by: Macfarlane, Matthew V., et al.
Published: (2026)
by: Macfarlane, Matthew V., et al.
Published: (2026)
Uncoupled Learning of Differential Stackelberg Equilibria with Commitments
by: Loftin, Robert, et al.
Published: (2023)
by: Loftin, Robert, et al.
Published: (2023)
The Terminal Representation in Reinforcement Learning
by: Esterhuysen, Amir, et al.
Published: (2026)
by: Esterhuysen, Amir, et al.
Published: (2026)
Solving Deep Reinforcement Learning Tasks with Evolution Strategies and Linear Policy Networks
by: Wong, Annie, et al.
Published: (2024)
by: Wong, Annie, et al.
Published: (2024)
Learning The Minimum Action Distance
by: Steccanella, Lorenzo, et al.
Published: (2025)
by: Steccanella, Lorenzo, et al.
Published: (2025)
Scalable Constrained Multi-Agent Reinforcement Learning via State Augmentation and Consensus for Separable Dynamics
by: Amaya-Corredor, Santiago, et al.
Published: (2026)
by: Amaya-Corredor, Santiago, et al.
Published: (2026)
Accelerating Proximal Policy Optimization Learning Using Task Prediction for Solving Environments with Delayed Rewards
by: Ahmad, Ahmad, et al.
Published: (2024)
by: Ahmad, Ahmad, et al.
Published: (2024)
Improving Subgraph-GNNs via Edge-Level Ego-Network Encodings
by: Alvarez-Gonzalez, Nurudin, et al.
Published: (2023)
by: Alvarez-Gonzalez, Nurudin, et al.
Published: (2023)
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
by: Cherepanov, Egor, et al.
Published: (2025)
by: Cherepanov, Egor, et al.
Published: (2025)
Unlearning Clients, Features and Samples in Vertical Federated Learning
by: Varshney, Ayush K., et al.
Published: (2025)
by: Varshney, Ayush K., et al.
Published: (2025)
PeersimGym: An Environment for Solving the Task Offloading Problem with Reinforcement Learning
by: Metelo, Frederico, et al.
Published: (2024)
by: Metelo, Frederico, et al.
Published: (2024)
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance
by: He, Jinmin, et al.
Published: (2025)
by: He, Jinmin, et al.
Published: (2025)
Learning to Solve the Constrained Most Probable Explanation Task in Probabilistic Graphical Models
by: Arya, Shivvrat, et al.
Published: (2024)
by: Arya, Shivvrat, et al.
Published: (2024)
MLCopilot: Unleashing the Power of Large Language Models in Solving Machine Learning Tasks
by: Zhang, Lei, et al.
Published: (2023)
by: Zhang, Lei, et al.
Published: (2023)
Hierarchical Planning for Complex Tasks with Knowledge Graph-RAG and Symbolic Verification
by: Cornelio, Cristina, et al.
Published: (2025)
by: Cornelio, Cristina, et al.
Published: (2025)
On the Empirical Complexity of Reasoning and Planning in LLMs
by: Kang, Liwei, et al.
Published: (2024)
by: Kang, Liwei, et al.
Published: (2024)
Constructing an Optimal Behavior Basis for the Option Keyboard
by: Alegre, Lucas N., et al.
Published: (2025)
by: Alegre, Lucas N., et al.
Published: (2025)
Learning Optimal and Sample-Efficient Decision Policies with Guarantees
by: Shao, Daqian
Published: (2026)
by: Shao, Daqian
Published: (2026)
Autoregressive Policy Optimization for Constrained Allocation Tasks
by: Winkel, David, et al.
Published: (2024)
by: Winkel, David, et al.
Published: (2024)
Provably Efficient Exploration in Reward Machines with Low Regret
by: Bourel, Hippolyte, et al.
Published: (2024)
by: Bourel, Hippolyte, et al.
Published: (2024)
Boosting Hierarchical Reinforcement Learning with Meta-Learning for Complex Task Adaptation
by: Khajooeinejad, Arash, et al.
Published: (2024)
by: Khajooeinejad, Arash, et al.
Published: (2024)
Continual Task Learning through Adaptive Policy Self-Composition
by: Hu, Shengchao, et al.
Published: (2024)
by: Hu, Shengchao, et al.
Published: (2024)
Learning Policy Committees for Effective Personalization in MDPs with Diverse Tasks
by: Ge, Luise, et al.
Published: (2025)
by: Ge, Luise, et al.
Published: (2025)
Learning More Expressive General Policies for Classical Planning Domains
by: Ståhlberg, Simon, et al.
Published: (2024)
by: Ståhlberg, Simon, et al.
Published: (2024)
Inverse Entropic Optimal Transport Solves Semi-supervised Learning via Data Likelihood Maximization
by: Persiianov, Mikhail, et al.
Published: (2024)
by: Persiianov, Mikhail, et al.
Published: (2024)
Image-Based Deep Reinforcement Learning with Intrinsically Motivated Stimuli: On the Execution of Complex Robotic Tasks
by: Valencia, David, et al.
Published: (2024)
by: Valencia, David, et al.
Published: (2024)
Quantum Architecture Search for Solving Quantum Machine Learning Tasks
by: Kölle, Michael, et al.
Published: (2025)
by: Kölle, Michael, et al.
Published: (2025)
Combining Planning and Reinforcement Learning for Solving Relational Multiagent Domains
by: Prabhakar, Nikhilesh, et al.
Published: (2025)
by: Prabhakar, Nikhilesh, et al.
Published: (2025)
Learning Generalized Policies for Fully Observable Non-Deterministic Planning Domains
by: Hofmann, Till, et al.
Published: (2024)
by: Hofmann, Till, et al.
Published: (2024)
On the Global Optimality of Policy Gradient Methods in General Utility Reinforcement Learning
by: Barakat, Anas, et al.
Published: (2024)
by: Barakat, Anas, et al.
Published: (2024)
Optimal Policy Sparsification and Low Rank Decomposition for Deep Reinforcement Learning
by: Goddla, Vikram
Published: (2024)
by: Goddla, Vikram
Published: (2024)
COPR: Continual Human Preference Learning via Optimal Policy Regularization
by: Zhang, Han, et al.
Published: (2024)
by: Zhang, Han, et al.
Published: (2024)
PWM: Policy Learning with Multi-Task World Models
by: Georgiev, Ignat, et al.
Published: (2024)
by: Georgiev, Ignat, et al.
Published: (2024)
WorkArena: How Capable Are Web Agents at Solving Common Knowledge Work Tasks?
by: Drouin, Alexandre, et al.
Published: (2024)
by: Drouin, Alexandre, et al.
Published: (2024)
From Basis to Basis: Gaussian Particle Representation for Interpretable PDE Operators
by: Li, Zhihao, et al.
Published: (2026)
by: Li, Zhihao, et al.
Published: (2026)
Similar Items
-
Globally Optimal Hierarchical Reinforcement Learning for Linearly-Solvable Markov Decision Processes
by: Infante, Guillermo, et al.
Published: (2021) -
Hierarchical Average-Reward Linearly-solvable Markov Decision Processes
by: Infante, Guillermo, et al.
Published: (2024) -
Making Universal Policies Universal
by: Höpner, Niklas, et al.
Published: (2025) -
Data Augmentation for Instruction Following Policies via Trajectory Segmentation
by: Höpner, Niklas, et al.
Published: (2025) -
Bridge the Inference Gaps of Neural Processes via Expectation Maximization
by: Wang, Qi, et al.
Published: (2025)