Goal-Reaching Policy Learning from Non-Expert Observations via Effective Subgoal Guidance
Fuente:
arXiv
Guardado en:
| Autores principales: | Huang, RenMing, Liu, Shaochong, Pei, Yunqiang, Wang, Peng, Wang, Guoqing, Yang, Yang, Shen, Hengtao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Diffusion Models as Optimizers for Efficient Planning in Offline RL
por: Huang, Renming, et al.
Publicado: (2024)
por: Huang, Renming, et al.
Publicado: (2024)
Subgoal Diffuser: Coarse-to-fine Subgoal Generation to Guide Model Predictive Control for Robot Manipulation
por: Huang, Zixuan, et al.
Publicado: (2024)
por: Huang, Zixuan, et al.
Publicado: (2024)
Safe Policy Exploration Improvement via Subgoals
por: Angulo, Brian, et al.
Publicado: (2024)
por: Angulo, Brian, et al.
Publicado: (2024)
Hyper-GoalNet: Goal-Conditioned Manipulation Policy Learning with HyperNetworks
por: Zhou, Pei, et al.
Publicado: (2025)
por: Zhou, Pei, et al.
Publicado: (2025)
Planning with Learned Subgoals Selected by Temporal Information
por: Huang, Xi, et al.
Publicado: (2024)
por: Huang, Xi, et al.
Publicado: (2024)
Explicit-Implicit Subgoal Planning for Long-Horizon Tasks with Sparse Reward
por: Wang, Fangyuan, et al.
Publicado: (2023)
por: Wang, Fangyuan, et al.
Publicado: (2023)
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning
por: Vasan, Gautham, et al.
Publicado: (2024)
por: Vasan, Gautham, et al.
Publicado: (2024)
Data-Efficient Policy Selection for Navigation in Partial Maps via Subgoal-Based Abstraction
por: Paudel, Abhishek, et al.
Publicado: (2023)
por: Paudel, Abhishek, et al.
Publicado: (2023)
Score the Steps, Not Just the Goal: VLM-Based Subgoal Evaluation for Robotic Manipulation
por: ElMallah, Ramy, et al.
Publicado: (2025)
por: ElMallah, Ramy, et al.
Publicado: (2025)
Anticipation-VLA: Solving Long-Horizon Embodied Tasks via Anticipation-based Subgoal Generation
por: Zhang, Zhilong, et al.
Publicado: (2026)
por: Zhang, Zhilong, et al.
Publicado: (2026)
NavDP: Learning Sim-to-Real Navigation Diffusion Policy with Privileged Information Guidance
por: Cai, Wenzhe, et al.
Publicado: (2025)
por: Cai, Wenzhe, et al.
Publicado: (2025)
Reinforcement Learning with Data Bootstrapping for Dynamic Subgoal Pursuit in Humanoid Robot Navigation
por: Peng, Chengyang, et al.
Publicado: (2025)
por: Peng, Chengyang, et al.
Publicado: (2025)
SADP: Subgoal-Aware Diffusion Policy for Explainable Robots Learned from Foundation Model Generated Demonstrations
por: Hu, Site, et al.
Publicado: (2026)
por: Hu, Site, et al.
Publicado: (2026)
Subgoal-based Hierarchical Reinforcement Learning for Multi-Agent Collaboration
por: Xu, Cheng, et al.
Publicado: (2024)
por: Xu, Cheng, et al.
Publicado: (2024)
Reinforcement Learning Goal-Reaching Control with Guaranteed Lyapunov-Like Stabilizer for Mobile Robots
por: Shahna, Mehdi Heydari, et al.
Publicado: (2026)
por: Shahna, Mehdi Heydari, et al.
Publicado: (2026)
Goal-Reaching Trajectory Design Near Danger with Piecewise Affine Reach-avoid Computation
por: Chung, Long Kiu, et al.
Publicado: (2024)
por: Chung, Long Kiu, et al.
Publicado: (2024)
Goal Reaching with Eikonal-Constrained Hierarchical Quasimetric Reinforcement Learning
por: Giammarino, Vittorio, et al.
Publicado: (2025)
por: Giammarino, Vittorio, et al.
Publicado: (2025)
AION: Aerial Indoor Object-Goal Navigation Using Dual-Policy Reinforcement Learning
por: Yan, Zichen, et al.
Publicado: (2026)
por: Yan, Zichen, et al.
Publicado: (2026)
Vision-based Goal-Reaching Control for Mobile Robots Using a Hierarchical Learning Framework
por: Shahna, Mehdi Heydari, et al.
Publicado: (2026)
por: Shahna, Mehdi Heydari, et al.
Publicado: (2026)
Smoother Action Chunking Flow Policy via Prior-Corrected Orthogonal Trust-Region Guidance
por: Fang, Kai, et al.
Publicado: (2026)
por: Fang, Kai, et al.
Publicado: (2026)
Learning Human Reaching Optimality Principles from Minimal Observation Inverse Reinforcement Learning
por: Mehrdad, Sarmad, et al.
Publicado: (2025)
por: Mehrdad, Sarmad, et al.
Publicado: (2025)
Learning Cross-hand Policies for High-DOF Reaching and Grasping
por: She, Qijin, et al.
Publicado: (2024)
por: She, Qijin, et al.
Publicado: (2024)
DISC: Decoupling Instruction from State-Conditioned Control via Policy Generation
por: Ren, Hanxiang, et al.
Publicado: (2026)
por: Ren, Hanxiang, et al.
Publicado: (2026)
Expert Knowledge-driven Reinforcement Learning for Autonomous Racing via Trajectory Guidance and Dynamics Constraints
por: Leng, Bo, et al.
Publicado: (2026)
por: Leng, Bo, et al.
Publicado: (2026)
Shortcut Learning in Generalist Robot Policies: The Role of Dataset Diversity and Fragmentation
por: Xing, Youguang, et al.
Publicado: (2025)
por: Xing, Youguang, et al.
Publicado: (2025)
GR-MG: Leveraging Partially Annotated Data via Multi-Modal Goal-Conditioned Policy
por: Li, Peiyan, et al.
Publicado: (2024)
por: Li, Peiyan, et al.
Publicado: (2024)
GAMap: Zero-Shot Object Goal Navigation with Multi-Scale Geometric-Affordance Guidance
por: Yuan, Shuaihang, et al.
Publicado: (2024)
por: Yuan, Shuaihang, et al.
Publicado: (2024)
Act2Goal: From World Model To General Goal-conditioned Policy
por: Zhou, Pengfei, et al.
Publicado: (2025)
por: Zhou, Pengfei, et al.
Publicado: (2025)
ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models
por: Ye, Wencheng, et al.
Publicado: (2025)
por: Ye, Wencheng, et al.
Publicado: (2025)
Sample-Efficient Learning with Online Expert Correction for Autonomous Catheter Steering in Endovascular Bifurcation Navigation
por: Wang, Hao, et al.
Publicado: (2026)
por: Wang, Hao, et al.
Publicado: (2026)
TouchGuide: Inference-Time Steering of Visuomotor Policies via Touch Guidance
por: Zhang, Zhemeng, et al.
Publicado: (2026)
por: Zhang, Zhemeng, et al.
Publicado: (2026)
Hand-Eye Autonomous Delivery: Learning Humanoid Navigation, Locomotion and Reaching
por: Chen, Sirui, et al.
Publicado: (2025)
por: Chen, Sirui, et al.
Publicado: (2025)
GUIDES: Guidance Using Instructor-Distilled Embeddings for Pre-trained Robot Policy Enhancement
por: Gao, Minquan, et al.
Publicado: (2025)
por: Gao, Minquan, et al.
Publicado: (2025)
Heterogeneous Multi-Expert Reinforcement Learning for Long-Horizon Multi-Goal Tasks in Autonomous Forklifts
por: Chen, Yun, et al.
Publicado: (2026)
por: Chen, Yun, et al.
Publicado: (2026)
PPGuide: Steering Diffusion Policies with Performance Predictive Guidance
por: Wang, Zixing, et al.
Publicado: (2026)
por: Wang, Zixing, et al.
Publicado: (2026)
Implicit Drifting Policy: One-Step Action Generation via Conditional Expert Geometry
por: Yang, Zemin, et al.
Publicado: (2026)
por: Yang, Zemin, et al.
Publicado: (2026)
ACDC: Adaptive Curriculum Planning with Dynamic Contrastive Control for Goal-Conditioned Reinforcement Learning in Robotic Manipulation
por: Wang, Xuerui, et al.
Publicado: (2026)
por: Wang, Xuerui, et al.
Publicado: (2026)
H-GAR: A Hierarchical Interaction Framework via Goal-Driven Observation-Action Refinement for Robotic Manipulation
por: Zhu, Yijie, et al.
Publicado: (2025)
por: Zhu, Yijie, et al.
Publicado: (2025)
Hierarchical Entity-centric Reinforcement Learning with Factored Subgoal Diffusion
por: Haramati, Dan, et al.
Publicado: (2026)
por: Haramati, Dan, et al.
Publicado: (2026)
Learning Diffusion Policy from Primitive Skills for Robot Manipulation
por: Gu, Zhihao, et al.
Publicado: (2026)
por: Gu, Zhihao, et al.
Publicado: (2026)
Ejemplares similares
-
Diffusion Models as Optimizers for Efficient Planning in Offline RL
por: Huang, Renming, et al.
Publicado: (2024) -
Subgoal Diffuser: Coarse-to-fine Subgoal Generation to Guide Model Predictive Control for Robot Manipulation
por: Huang, Zixuan, et al.
Publicado: (2024) -
Safe Policy Exploration Improvement via Subgoals
por: Angulo, Brian, et al.
Publicado: (2024) -
Hyper-GoalNet: Goal-Conditioned Manipulation Policy Learning with HyperNetworks
por: Zhou, Pei, et al.
Publicado: (2025) -
Planning with Learned Subgoals Selected by Temporal Information
por: Huang, Xi, et al.
Publicado: (2024)