A KL-regularization Framework for Learning to Plan with Adaptive Priors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Serra-Gomez, Álvaro, Ornia, Daniel Jarne, Tirumala, Dhruva, Moerland, Thomas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards a Practical Understanding of Lagrangian Methods in Safe Reinforcement Learning
von: Spoor, Lindsay, et al.
Veröffentlicht: (2025)
von: Spoor, Lindsay, et al.
Veröffentlicht: (2025)
Predictable Reinforcement Learning Dynamics through Entropy Rate Minimization
von: Ornia, Daniel Jarne, et al.
Veröffentlicht: (2023)
von: Ornia, Daniel Jarne, et al.
Veröffentlicht: (2023)
Sandbagging in a Simple Survival Bandit Problem
von: Dyer, Joel, et al.
Veröffentlicht: (2025)
von: Dyer, Joel, et al.
Veröffentlicht: (2025)
Learning to Plan, Planning to Learn: Adaptive Hierarchical RL-MPC for Sample-Efficient Decision Making
von: Hori, Toshiaki, et al.
Veröffentlicht: (2025)
von: Hori, Toshiaki, et al.
Veröffentlicht: (2025)
Bayesian Decision Making around Experts
von: Ornia, Daniel Jarne, et al.
Veröffentlicht: (2025)
von: Ornia, Daniel Jarne, et al.
Veröffentlicht: (2025)
Learning Agile Soccer Skills for a Bipedal Robot with Deep Reinforcement Learning
von: Haarnoja, Tuomas, et al.
Veröffentlicht: (2023)
von: Haarnoja, Tuomas, et al.
Veröffentlicht: (2023)
What Matters for Simulation to Online Reinforcement Learning on Real Robots
von: As, Yarden, et al.
Veröffentlicht: (2026)
von: As, Yarden, et al.
Veröffentlicht: (2026)
Adaptive Inverse Kinematics Framework for Learning Variable-Length Tool Manipulation in Robotics
von: Kothavale, Prathamesh, et al.
Veröffentlicht: (2025)
von: Kothavale, Prathamesh, et al.
Veröffentlicht: (2025)
CaRL: Learning Scalable Planning Policies with Simple Rewards
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2025)
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2025)
V-Max: A Reinforcement Learning Framework for Autonomous Driving
von: Charraut, Valentin, et al.
Veröffentlicht: (2025)
von: Charraut, Valentin, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Sustainable Energy: A Survey
von: Ponse, Koen, et al.
Veröffentlicht: (2024)
von: Ponse, Koen, et al.
Veröffentlicht: (2024)
Dynamic Obstacle Avoidance through Uncertainty-Based Adaptive Planning with Diffusion
von: Punyamoorty, Vineet, et al.
Veröffentlicht: (2024)
von: Punyamoorty, Vineet, et al.
Veröffentlicht: (2024)
WarmPrior: Straightening Flow-Matching Policies with Temporal Priors
von: Kang, Sinjae, et al.
Veröffentlicht: (2026)
von: Kang, Sinjae, et al.
Veröffentlicht: (2026)
Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own
von: Ye, Weirui, et al.
Veröffentlicht: (2023)
von: Ye, Weirui, et al.
Veröffentlicht: (2023)
Random Network Distillation Based Deep Reinforcement Learning for AGV Path Planning
von: Yin, Huilin, et al.
Veröffentlicht: (2024)
von: Yin, Huilin, et al.
Veröffentlicht: (2024)
Motion Planning Diffusion: Learning and Planning of Robot Motions with Diffusion Models
von: Carvalho, Joao, et al.
Veröffentlicht: (2023)
von: Carvalho, Joao, et al.
Veröffentlicht: (2023)
Intuitive Programming, Adaptive Task Planning, and Dynamic Role Allocation in Human-Robot Collaboration
von: Lagomarsino, Marta, et al.
Veröffentlicht: (2025)
von: Lagomarsino, Marta, et al.
Veröffentlicht: (2025)
Learn Once Plan Arbitrarily (LOPA): Attention-Enhanced Deep Reinforcement Learning Method for Global Path Planning
von: Huang, Guoming, et al.
Veröffentlicht: (2024)
von: Huang, Guoming, et al.
Veröffentlicht: (2024)
Emergent Risk Awareness in Rational Agents under Resource Constraints
von: Ornia, Daniel Jarne, et al.
Veröffentlicht: (2025)
von: Ornia, Daniel Jarne, et al.
Veröffentlicht: (2025)
UniConFlow: A Unified Constrained Flow-Matching Framework for Certified Motion Planning
von: Yang, Zewen, et al.
Veröffentlicht: (2025)
von: Yang, Zewen, et al.
Veröffentlicht: (2025)
Equivariant Action Sampling for Reinforcement Learning and Planning
von: Zhao, Linfeng, et al.
Veröffentlicht: (2024)
von: Zhao, Linfeng, et al.
Veröffentlicht: (2024)
RIZE: Adaptive Regularization for Imitation Learning
von: Karimi, Adib, et al.
Veröffentlicht: (2025)
von: Karimi, Adib, et al.
Veröffentlicht: (2025)
Safe Exploration via Policy Priors
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
Learning Social Heuristics for Human-Aware Path Planning
von: Eirale, Andrea, et al.
Veröffentlicht: (2025)
von: Eirale, Andrea, et al.
Veröffentlicht: (2025)
Planning the path with Reinforcement Learning: Optimal Robot Motion Planning in RoboCup Small Size League Environments
von: Machado, Mateus G., et al.
Veröffentlicht: (2024)
von: Machado, Mateus G., et al.
Veröffentlicht: (2024)
Adaptive Reinforcement Learning for Unobservable Random Delays
von: Wikman, John, et al.
Veröffentlicht: (2025)
von: Wikman, John, et al.
Veröffentlicht: (2025)
DHP: Discrete Hierarchical Planning for Hierarchical Reinforcement Learning Agents
von: Sharma, Shashank, et al.
Veröffentlicht: (2025)
von: Sharma, Shashank, et al.
Veröffentlicht: (2025)
HYPERmotion: Learning Hybrid Behavior Planning for Autonomous Loco-manipulation
von: Wang, Jin, et al.
Veröffentlicht: (2024)
von: Wang, Jin, et al.
Veröffentlicht: (2024)
Learning Adaptive Dexterous Grasping from Single Demonstrations
von: Shi, Liangzhi, et al.
Veröffentlicht: (2025)
von: Shi, Liangzhi, et al.
Veröffentlicht: (2025)
Adaptive Querying for Reward Learning from Human Feedback
von: Anand, Yashwanthi, et al.
Veröffentlicht: (2024)
von: Anand, Yashwanthi, et al.
Veröffentlicht: (2024)
ACL-QL: Adaptive Conservative Level in Q-Learning for Offline Reinforcement Learning
von: Wu, Kun, et al.
Veröffentlicht: (2024)
von: Wu, Kun, et al.
Veröffentlicht: (2024)
Revisiting Space Mission Planning: A Reinforcement Learning-Guided Approach for Multi-Debris Rendezvous
von: Bandyopadhyay, Agni, et al.
Veröffentlicht: (2024)
von: Bandyopadhyay, Agni, et al.
Veröffentlicht: (2024)
Hybrid Motion Planning with Deep Reinforcement Learning for Mobile Robot Navigation
von: Kolomeytsev, Yury, et al.
Veröffentlicht: (2025)
von: Kolomeytsev, Yury, et al.
Veröffentlicht: (2025)
Validity Learning on Failures: Mitigating the Distribution Shift in Autonomous Vehicle Planning
von: Arasteh, Fazel, et al.
Veröffentlicht: (2024)
von: Arasteh, Fazel, et al.
Veröffentlicht: (2024)
COLLAGE: Adaptive Fusion-based Retrieval for Augmented Policy Learning
von: Kumar, Sateesh, et al.
Veröffentlicht: (2025)
von: Kumar, Sateesh, et al.
Veröffentlicht: (2025)
Robot-Gated Interactive Imitation Learning with Adaptive Intervention Mechanism
von: Cai, Haoyuan, et al.
Veröffentlicht: (2025)
von: Cai, Haoyuan, et al.
Veröffentlicht: (2025)
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
von: Liu, Tenglong, et al.
Veröffentlicht: (2024)
von: Liu, Tenglong, et al.
Veröffentlicht: (2024)
Generalized Animal Imitator: Agile Locomotion with Versatile Motion Prior
von: Yang, Ruihan, et al.
Veröffentlicht: (2023)
von: Yang, Ruihan, et al.
Veröffentlicht: (2023)
A Goal-Oriented Reinforcement Learning-Based Path Planning Algorithm for Modular Self-Reconfigurable Satellites
von: Liu, Bofei, et al.
Veröffentlicht: (2025)
von: Liu, Bofei, et al.
Veröffentlicht: (2025)
robosuite: A Modular Simulation Framework and Benchmark for Robot Learning
von: Zhu, Yuke, et al.
Veröffentlicht: (2020)
von: Zhu, Yuke, et al.
Veröffentlicht: (2020)
Ähnliche Einträge
-
Towards a Practical Understanding of Lagrangian Methods in Safe Reinforcement Learning
von: Spoor, Lindsay, et al.
Veröffentlicht: (2025) -
Predictable Reinforcement Learning Dynamics through Entropy Rate Minimization
von: Ornia, Daniel Jarne, et al.
Veröffentlicht: (2023) -
Sandbagging in a Simple Survival Bandit Problem
von: Dyer, Joel, et al.
Veröffentlicht: (2025) -
Learning to Plan, Planning to Learn: Adaptive Hierarchical RL-MPC for Sample-Efficient Decision Making
von: Hori, Toshiaki, et al.
Veröffentlicht: (2025) -
Bayesian Decision Making around Experts
von: Ornia, Daniel Jarne, et al.
Veröffentlicht: (2025)