Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Hwang, Jaebak, Lee, Sanghyeon, Kim, Jeongmo, Han, Seungyul |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Wolfpack Adversarial Attack for Robust Multi-Agent Reinforcement Learning
by: Lee, Sunwoo, et al.
Published: (2025)
by: Lee, Sunwoo, et al.
Published: (2025)
Bridging Domain Gaps with Target-Aligned Generation for Offline Reinforcement Learning
by: Kim, Minung, et al.
Published: (2026)
by: Kim, Minung, et al.
Published: (2026)
Task-Aware Virtual Training: Enhancing Generalization in Meta-Reinforcement Learning for Out-of-Distribution Tasks
by: Kim, Jeongmo, et al.
Published: (2025)
by: Kim, Jeongmo, et al.
Published: (2025)
Exclusively Penalized Q-learning for Offline Reinforcement Learning
by: Yeom, Junghyuk, et al.
Published: (2024)
by: Yeom, Junghyuk, et al.
Published: (2024)
LLM-Guided Communication for Cooperative Multi-Agent Reinforcement Learning
by: Bae, Sangjun, et al.
Published: (2026)
by: Bae, Sangjun, et al.
Published: (2026)
Probabilistic Subgoal Representations for Hierarchical Reinforcement learning
by: Wang, Vivienne Huiling, et al.
Published: (2024)
by: Wang, Vivienne Huiling, et al.
Published: (2024)
Training toward significance with the decorrelated event classifier transformer neural network
by: Kim, Jaebak
Published: (2023)
by: Kim, Jaebak
Published: (2023)
Interaction-Breaking Adversarial Learning Framework for Robust Multi-Agent Reinforcement Learning
by: Lee, Sunwoo, et al.
Published: (2026)
by: Lee, Sunwoo, et al.
Published: (2026)
Refining Compositional Diffusion for Reliable Long-Horizon Planning
by: Lee, Kyowoon, et al.
Published: (2026)
by: Lee, Kyowoon, et al.
Published: (2026)
Subgoal Graph-Augmented Planning for LLM-Guided Open-World Reinforcement Learning
by: Fan, Shanwei, et al.
Published: (2025)
by: Fan, Shanwei, et al.
Published: (2025)
Flow Actor-Critic for Offline Reinforcement Learning
by: Chae, Jongseong, et al.
Published: (2026)
by: Chae, Jongseong, et al.
Published: (2026)
A Subgoal-driven Framework for Improving Long-Horizon LLM Agents
by: Wang, Taiyi, et al.
Published: (2026)
by: Wang, Taiyi, et al.
Published: (2026)
Goal-Space Planning with Subgoal Models
by: Lo, Chunlok, et al.
Published: (2022)
by: Lo, Chunlok, et al.
Published: (2022)
Subgoal-based Reward Shaping to Improve Efficiency in Reinforcement Learning
by: Okudo, Takato, et al.
Published: (2021)
by: Okudo, Takato, et al.
Published: (2021)
Fast and Precise: Adjusting Planning Horizon with Adaptive Subgoal Search
by: Zawalski, Michał, et al.
Published: (2022)
by: Zawalski, Michał, et al.
Published: (2022)
Planning Transformer: Long-Horizon Offline Reinforcement Learning with Planning Tokens
by: Clinton, Joseph, et al.
Published: (2024)
by: Clinton, Joseph, et al.
Published: (2024)
Transfer Reinforcement Learning in Heterogeneous Action Spaces using Subgoal Mapping
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning
by: Wu, Qingyuan, et al.
Published: (2025)
by: Wu, Qingyuan, et al.
Published: (2025)
Domain-Invariant Per-Frame Feature Extraction for Cross-Domain Imitation Learning with Visual Observations
by: Kim, Minung, et al.
Published: (2025)
by: Kim, Minung, et al.
Published: (2025)
Learning Hidden Subgoals under Temporal Ordering Constraints in Reinforcement Learning
by: Xu, Duo, et al.
Published: (2024)
by: Xu, Duo, et al.
Published: (2024)
Reinforcement Learning for Long-Horizon Interactive LLM Agents
by: Chen, Kevin, et al.
Published: (2025)
by: Chen, Kevin, et al.
Published: (2025)
FRAIN to Train: A Fast-and-Reliable Solution for Decentralized Federated Learning
by: Park, Sanghyeon, et al.
Published: (2025)
by: Park, Sanghyeon, et al.
Published: (2025)
Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL
by: Choi, Jinwoo, et al.
Published: (2026)
by: Choi, Jinwoo, et al.
Published: (2026)
Shaping Zero-Shot Coordination via State Blocking
by: Kang, Mingu, et al.
Published: (2026)
by: Kang, Mingu, et al.
Published: (2026)
Offline Reinforcement Learning with Universal Horizon Models
by: Chung, Hojun, et al.
Published: (2026)
by: Chung, Hojun, et al.
Published: (2026)
ARCLE: The Abstraction and Reasoning Corpus Learning Environment for Reinforcement Learning
by: Lee, Hosung, et al.
Published: (2024)
by: Lee, Hosung, et al.
Published: (2024)
HiMAC: Hierarchical Macro-Micro Learning for Long-Horizon LLM Agents
by: Jin, Hongbo, et al.
Published: (2026)
by: Jin, Hongbo, et al.
Published: (2026)
Learning for Long-Horizon Planning via Neuro-Symbolic Abductive Imitation
by: Shao, Jie-Jing, et al.
Published: (2024)
by: Shao, Jie-Jing, et al.
Published: (2024)
Subgoal Search For Complex Reasoning Tasks
by: Czechowski, Konrad, et al.
Published: (2021)
by: Czechowski, Konrad, et al.
Published: (2021)
SubgoalXL: Subgoal-based Expert Learning for Theorem Proving
by: Zhao, Xueliang, et al.
Published: (2024)
by: Zhao, Xueliang, et al.
Published: (2024)
DHP: Discrete Hierarchical Planning for Hierarchical Reinforcement Learning Agents
by: Sharma, Shashank, et al.
Published: (2025)
by: Sharma, Shashank, et al.
Published: (2025)
Generalized Gaussian Temporal Difference Error for Uncertainty-aware Reinforcement Learning
by: Kim, Seyeon, et al.
Published: (2024)
by: Kim, Seyeon, et al.
Published: (2024)
Prune-OPD: Efficient and Reliable On-Policy Distillation for Long-Horizon Reasoning
by: Yang, Zhicheng, et al.
Published: (2026)
by: Yang, Zhicheng, et al.
Published: (2026)
Distilling Reinforcement Learning Algorithms for In-Context Model-Based Planning
by: Son, Jaehyeon, et al.
Published: (2025)
by: Son, Jaehyeon, et al.
Published: (2025)
SPIRE: Synergistic Planning, Imitation, and Reinforcement Learning for Long-Horizon Manipulation
by: Zhou, Zihan, et al.
Published: (2024)
by: Zhou, Zihan, et al.
Published: (2024)
Horizon Generalization in Reinforcement Learning
by: Myers, Vivek, et al.
Published: (2025)
by: Myers, Vivek, et al.
Published: (2025)
GHIL-Glue: Hierarchical Control with Filtered Subgoal Images
by: Hatch, Kyle B., et al.
Published: (2024)
by: Hatch, Kyle B., et al.
Published: (2024)
RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents
by: Zhang, Zijing, et al.
Published: (2025)
by: Zhang, Zijing, et al.
Published: (2025)
Trust Region Masking for Long-Horizon LLM Reinforcement Learning
by: Li, Yingru, et al.
Published: (2025)
by: Li, Yingru, et al.
Published: (2025)
On the Effective Horizon of Inverse Reinforcement Learning
by: Xu, Yiqing, et al.
Published: (2023)
by: Xu, Yiqing, et al.
Published: (2023)
Similar Items
-
Wolfpack Adversarial Attack for Robust Multi-Agent Reinforcement Learning
by: Lee, Sunwoo, et al.
Published: (2025) -
Bridging Domain Gaps with Target-Aligned Generation for Offline Reinforcement Learning
by: Kim, Minung, et al.
Published: (2026) -
Task-Aware Virtual Training: Enhancing Generalization in Meta-Reinforcement Learning for Out-of-Distribution Tasks
by: Kim, Jeongmo, et al.
Published: (2025) -
Exclusively Penalized Q-learning for Offline Reinforcement Learning
by: Yeom, Junghyuk, et al.
Published: (2024) -
LLM-Guided Communication for Cooperative Multi-Agent Reinforcement Learning
by: Bae, Sangjun, et al.
Published: (2026)