Safe Policy Exploration Improvement via Subgoals
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Angulo, Brian, Gorbov, Gregory, Panov, Aleksandr, Yakovlev, Konstantin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
M3PO: Massively Multi-Task Model-Based Policy Optimization
von: Narendra, Aditya, et al.
Veröffentlicht: (2025)
von: Narendra, Aditya, et al.
Veröffentlicht: (2025)
Safe Planning and Policy Optimization via World Model Learning
von: Latyshev, Artem, et al.
Veröffentlicht: (2025)
von: Latyshev, Artem, et al.
Veröffentlicht: (2025)
Safe Exploration via Policy Priors
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
Goal-Reaching Policy Learning from Non-Expert Observations via Effective Subgoal Guidance
von: Huang, RenMing, et al.
Veröffentlicht: (2024)
von: Huang, RenMing, et al.
Veröffentlicht: (2024)
CoRL-MPPI: Enhancing MPPI With Learnable Behaviours For Efficient And Provably-Safe Multi-Robot Collision Avoidance
von: Dergachev, Stepan, et al.
Veröffentlicht: (2025)
von: Dergachev, Stepan, et al.
Veröffentlicht: (2025)
Safe Interval Randomized Path Planning For Manipulators
von: Kerimov, Nuraddin, et al.
Veröffentlicht: (2024)
von: Kerimov, Nuraddin, et al.
Veröffentlicht: (2024)
Off-Policy Safe Reinforcement Learning with Constrained Optimistic Exploration
von: Li, Guopeng, et al.
Veröffentlicht: (2026)
von: Li, Guopeng, et al.
Veröffentlicht: (2026)
Object-Centric World Models Meet Monte Carlo Tree Search
von: Vakhitov, Rodion, et al.
Veröffentlicht: (2026)
von: Vakhitov, Rodion, et al.
Veröffentlicht: (2026)
SOE: Sample-Efficient Robot Policy Self-Improvement via On-Manifold Exploration
von: Jin, Yang, et al.
Veröffentlicht: (2025)
von: Jin, Yang, et al.
Veröffentlicht: (2025)
Relational Object-Centric Actor-Critic
von: Ugadiarov, Leonid, et al.
Veröffentlicht: (2023)
von: Ugadiarov, Leonid, et al.
Veröffentlicht: (2023)
SGN-CIRL: Scene Graph-based Navigation with Curriculum, Imitation, and Reinforcement Learning
von: Oskolkov, Nikita, et al.
Veröffentlicht: (2025)
von: Oskolkov, Nikita, et al.
Veröffentlicht: (2025)
SIME: Enhancing Policy Self-Improvement with Modal-level Exploration
von: Jin, Yang, et al.
Veröffentlicht: (2025)
von: Jin, Yang, et al.
Veröffentlicht: (2025)
Model-based Policy Optimization using Symbolic World Model
von: Gorodetskiy, Andrey, et al.
Veröffentlicht: (2024)
von: Gorodetskiy, Andrey, et al.
Veröffentlicht: (2024)
Anticipation-VLA: Solving Long-Horizon Embodied Tasks via Anticipation-based Subgoal Generation
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
MAPF-GPT: Imitation Learning for Multi-Agent Pathfinding at Scale
von: Andreychuk, Anton, et al.
Veröffentlicht: (2024)
von: Andreychuk, Anton, et al.
Veröffentlicht: (2024)
Advancing Learnable Multi-Agent Pathfinding Solvers with Active Fine-Tuning
von: Andreychuk, Anton, et al.
Veröffentlicht: (2025)
von: Andreychuk, Anton, et al.
Veröffentlicht: (2025)
LookPlanGraph: Embodied Instruction Following Method with VLM Graph Augmentation
von: Onishchenko, Anatoly O., et al.
Veröffentlicht: (2025)
von: Onishchenko, Anatoly O., et al.
Veröffentlicht: (2025)
Enhancing Sample Efficiency and Exploration in Reinforcement Learning through the Integration of Diffusion Models and Proximal Policy Optimization
von: Gao, Tianci, et al.
Veröffentlicht: (2024)
von: Gao, Tianci, et al.
Veröffentlicht: (2024)
Revisiting Safe Exploration in Safe Reinforcement learning
von: Eckel, David, et al.
Veröffentlicht: (2024)
von: Eckel, David, et al.
Veröffentlicht: (2024)
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
Knowledge-Guided Manipulation Using Multi-Task Reinforcement Learning
von: Narendra, Aditya, et al.
Veröffentlicht: (2026)
von: Narendra, Aditya, et al.
Veröffentlicht: (2026)
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning
von: As, Yarden, et al.
Veröffentlicht: (2024)
von: As, Yarden, et al.
Veröffentlicht: (2024)
Safe Multi-Agent Navigation guided by Goal-Conditioned Safe Reinforcement Learning
von: Feng, Meng, et al.
Veröffentlicht: (2025)
von: Feng, Meng, et al.
Veröffentlicht: (2025)
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation
von: Danesh, Mohamad H., et al.
Veröffentlicht: (2025)
von: Danesh, Mohamad H., et al.
Veröffentlicht: (2025)
A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
Don't Blind Your VLA: Aligning Visual Representations for OOD Generalization
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
Safe Deep Policy Adaptation
von: Xiao, Wenli, et al.
Veröffentlicht: (2023)
von: Xiao, Wenli, et al.
Veröffentlicht: (2023)
Hierarchical Entity-centric Reinforcement Learning with Factored Subgoal Diffusion
von: Haramati, Dan, et al.
Veröffentlicht: (2026)
von: Haramati, Dan, et al.
Veröffentlicht: (2026)
FOSP: Fine-tuning Offline Safe Policy through World Models
von: Cao, Chenyang, et al.
Veröffentlicht: (2024)
von: Cao, Chenyang, et al.
Veröffentlicht: (2024)
Learning Safe Autonomous Driving Policies Using Predictive Safety Representations
von: Keswani, Mahesh, et al.
Veröffentlicht: (2025)
von: Keswani, Mahesh, et al.
Veröffentlicht: (2025)
SATA: Safe and Adaptive Torque-Based Locomotion Policies Inspired by Animal Learning
von: Li, Peizhuo, et al.
Veröffentlicht: (2025)
von: Li, Peizhuo, et al.
Veröffentlicht: (2025)
DRARL: Disengagement-Reason-Augmented Reinforcement Learning for Efficient Improvement of Autonomous Driving Policy
von: Zhou, Weitao, et al.
Veröffentlicht: (2025)
von: Zhou, Weitao, et al.
Veröffentlicht: (2025)
HELP: Hierarchical Embodied Language Planner for Household Tasks
von: Korchemnyi, Alexandr V., et al.
Veröffentlicht: (2025)
von: Korchemnyi, Alexandr V., et al.
Veröffentlicht: (2025)
Can Tabular Foundation Models Guide Exploration in Robot Policy Learning?
von: Ou, Buqing, et al.
Veröffentlicht: (2026)
von: Ou, Buqing, et al.
Veröffentlicht: (2026)
Data-Efficient Policy Selection for Navigation in Partial Maps via Subgoal-Based Abstraction
von: Paudel, Abhishek, et al.
Veröffentlicht: (2023)
von: Paudel, Abhishek, et al.
Veröffentlicht: (2023)
AmbiK: Dataset of Ambiguous Tasks in Kitchen Environment
von: Ivanova, Anastasiia, et al.
Veröffentlicht: (2025)
von: Ivanova, Anastasiia, et al.
Veröffentlicht: (2025)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
SafeAuto: Knowledge-Enhanced Safe Autonomous Driving with Multimodal Foundation Models
von: Zhang, Jiawei, et al.
Veröffentlicht: (2025)
von: Zhang, Jiawei, et al.
Veröffentlicht: (2025)
GHIL-Glue: Hierarchical Control with Filtered Subgoal Images
von: Hatch, Kyle B., et al.
Veröffentlicht: (2024)
von: Hatch, Kyle B., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
M3PO: Massively Multi-Task Model-Based Policy Optimization
von: Narendra, Aditya, et al.
Veröffentlicht: (2025) -
Safe Planning and Policy Optimization via World Model Learning
von: Latyshev, Artem, et al.
Veröffentlicht: (2025) -
Safe Exploration via Policy Priors
von: Wendl, Manuel, et al.
Veröffentlicht: (2026) -
Goal-Reaching Policy Learning from Non-Expert Observations via Effective Subgoal Guidance
von: Huang, RenMing, et al.
Veröffentlicht: (2024) -
CoRL-MPPI: Enhancing MPPI With Learnable Behaviours For Efficient And Provably-Safe Multi-Robot Collision Avoidance
von: Dergachev, Stepan, et al.
Veröffentlicht: (2025)