LESSON: Learning to Integrate Exploration Strategies for Reinforcement Learning via an Option Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Woojun, Kim, Jeonghye, Sung, Youngchul |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive $Q$-Aid for Conditional Supervised Learning in Offline Reinforcement Learning
by: Kim, Jeonghye, et al.
Published: (2024)
by: Kim, Jeonghye, et al.
Published: (2024)
Decision ConvFormer: Local Filtering in MetaFormer is Sufficient for Decision Making
by: Kim, Jeonghye, et al.
Published: (2023)
by: Kim, Jeonghye, et al.
Published: (2023)
Penalizing Infeasible Actions and Reward Scaling in Reinforcement Learning with Offline Data
by: Kim, Jeonghye, et al.
Published: (2025)
by: Kim, Jeonghye, et al.
Published: (2025)
Online Pre-Training for Offline-to-Online Reinforcement Learning
by: Shin, Yongjae, et al.
Published: (2025)
by: Shin, Yongjae, et al.
Published: (2025)
ReflAct: World-Grounded Decision Making in LLM Agents via Goal-State Reflection
by: Kim, Jeonghye, et al.
Published: (2025)
by: Kim, Jeonghye, et al.
Published: (2025)
Reward Dimension Reduction for Scalable Multi-Objective Reinforcement Learning
by: Park, Giseung, et al.
Published: (2025)
by: Park, Giseung, et al.
Published: (2025)
B3C: A Minimalist Approach to Offline Multi-Agent Reinforcement Learning
by: Kim, Woojun, et al.
Published: (2025)
by: Kim, Woojun, et al.
Published: (2025)
Flow Actor-Critic for Offline Reinforcement Learning
by: Chae, Jongseong, et al.
Published: (2026)
by: Chae, Jongseong, et al.
Published: (2026)
Flow Matching with Injected Noise for Offline-to-Online Reinforcement Learning
by: Shin, Yongjae, et al.
Published: (2026)
by: Shin, Yongjae, et al.
Published: (2026)
The Max-Min Formulation of Multi-Objective Reinforcement Learning: From Theory to a Model-Free Algorithm
by: Park, Giseung, et al.
Published: (2024)
by: Park, Giseung, et al.
Published: (2024)
Constrained Multi-Objective Reinforcement Learning with Max-Min Criterion
by: Park, Giseung, et al.
Published: (2026)
by: Park, Giseung, et al.
Published: (2026)
Rebellious Student: Reversing Teacher Signals for Reasoning Exploration with Self-Distilled RLVR
by: Kim, Jeonghye, et al.
Published: (2026)
by: Kim, Jeonghye, et al.
Published: (2026)
Multi-Objective Reinforcement Learning with Max-Min Criterion: A Game-Theoretic Approach
by: Byeon, Woohyeon, et al.
Published: (2025)
by: Byeon, Woohyeon, et al.
Published: (2025)
Adaptive Action Chunking via Multi-Chunk Q Value Estimation
by: Shin, Yongjae, et al.
Published: (2026)
by: Shin, Yongjae, et al.
Published: (2026)
Align While Search: Belief-Guided Exploratory Inference for World-Grounded Embodied Agents
by: Bae, Seohui, et al.
Published: (2025)
by: Bae, Seohui, et al.
Published: (2025)
A Temporally Correlated Latent Exploration for Reinforcement Learning
by: Oh, SuMin, et al.
Published: (2024)
by: Oh, SuMin, et al.
Published: (2024)
An Autonomous Non-monolithic Agent with Multi-mode Exploration based on Options Framework
by: Kim, JaeYoon, et al.
Published: (2023)
by: Kim, JaeYoon, et al.
Published: (2023)
Adaptively Coordinating with Novel Partners via Learned Latent Strategies
by: Li, Benjamin, et al.
Published: (2025)
by: Li, Benjamin, et al.
Published: (2025)
Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy Optimization
by: Liu, Zeyuan, et al.
Published: (2026)
by: Liu, Zeyuan, et al.
Published: (2026)
Reinforcement Learning with Options and State Representation
by: Ghriss, Ayoub, et al.
Published: (2024)
by: Ghriss, Ayoub, et al.
Published: (2024)
Entropy-Aware Model Initialization for Effective Exploration in Deep Reinforcement Learning
by: Jang, Sooyoung, et al.
Published: (2021)
by: Jang, Sooyoung, et al.
Published: (2021)
Accelerating Reinforcement Learning with Value-Conditional State Entropy Exploration
by: Kim, Dongyoung, et al.
Published: (2023)
by: Kim, Dongyoung, et al.
Published: (2023)
Is Exploration All You Need? Effective Exploration Characteristics for Transfer in Reinforcement Learning
by: Balloch, Jonathan C., et al.
Published: (2024)
by: Balloch, Jonathan C., et al.
Published: (2024)
Learning-Driven Exploration for Reinforcement Learning
by: Usama, Muhammad, et al.
Published: (2019)
by: Usama, Muhammad, et al.
Published: (2019)
SCALAR: Learning and Composing Skills through LLM Guided Symbolic Planning and Deep RL Grounding
by: Zabounidis, Renos, et al.
Published: (2026)
by: Zabounidis, Renos, et al.
Published: (2026)
Understanding Reasoning in LLMs through Strategic Information Allocation under Uncertainty
by: Kim, Jeonghye, et al.
Published: (2026)
by: Kim, Jeonghye, et al.
Published: (2026)
Efficient Reinforcement Learning via Decoupling Exploration and Utilization
by: Yang, Jingpu, et al.
Published: (2023)
by: Yang, Jingpu, et al.
Published: (2023)
Distributional Reinforcement Learning on Path-dependent Options
by: Özsoy, Ahmet Umur
Published: (2025)
by: Özsoy, Ahmet Umur
Published: (2025)
EFRame: Deeper Reasoning via Exploration-Filter-Replay Reinforcement Learning Framework
by: Wang, Chen, et al.
Published: (2025)
by: Wang, Chen, et al.
Published: (2025)
IPCGRL: Language-Instructed Reinforcement Learning for Procedural Level Generation
by: Baek, In-Chang, et al.
Published: (2025)
by: Baek, In-Chang, et al.
Published: (2025)
Segment-driven Structural Induction and Semantic Alignment for Heterogeneous Tabular Representation
by: Jung, Woojun, et al.
Published: (2026)
by: Jung, Woojun, et al.
Published: (2026)
Mitigating Spurious Correlation via Distributionally Robust Learning with Hierarchical Ambiguity Sets
by: Jo, Sung Ho, et al.
Published: (2025)
by: Jo, Sung Ho, et al.
Published: (2025)
Adversarial Reinforcement Learning Framework for ESP Cheater Simulation
by: Park, Inkyu, et al.
Published: (2025)
by: Park, Inkyu, et al.
Published: (2025)
Swap-guided Preference Learning for Personalized Reinforcement Learning from Human Feedback
by: Kim, Gihoon, et al.
Published: (2026)
by: Kim, Gihoon, et al.
Published: (2026)
Search Inspired Exploration in Reinforcement Learning
by: Sotirchos, Georgios, et al.
Published: (2026)
by: Sotirchos, Georgios, et al.
Published: (2026)
Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs?
by: Kim, Jeonghye, et al.
Published: (2026)
by: Kim, Jeonghye, et al.
Published: (2026)
OptionZero: Planning with Learned Options
by: Huang, Po-Wei, et al.
Published: (2025)
by: Huang, Po-Wei, et al.
Published: (2025)
Finding Effective Security Strategies through Reinforcement Learning and Self-Play
by: Hammar, Kim, et al.
Published: (2020)
by: Hammar, Kim, et al.
Published: (2020)
Boosting deep Reinforcement Learning using pretraining with Logical Options
by: Ye, Zihan, et al.
Published: (2026)
by: Ye, Zihan, et al.
Published: (2026)
Unsupervised-to-Online Reinforcement Learning
by: Kim, Junsu, et al.
Published: (2024)
by: Kim, Junsu, et al.
Published: (2024)
Similar Items
-
Adaptive $Q$-Aid for Conditional Supervised Learning in Offline Reinforcement Learning
by: Kim, Jeonghye, et al.
Published: (2024) -
Decision ConvFormer: Local Filtering in MetaFormer is Sufficient for Decision Making
by: Kim, Jeonghye, et al.
Published: (2023) -
Penalizing Infeasible Actions and Reward Scaling in Reinforcement Learning with Offline Data
by: Kim, Jeonghye, et al.
Published: (2025) -
Online Pre-Training for Offline-to-Online Reinforcement Learning
by: Shin, Yongjae, et al.
Published: (2025) -
ReflAct: World-Grounded Decision Making in LLM Agents via Goal-State Reflection
by: Kim, Jeonghye, et al.
Published: (2025)