OptionZero: Planning with Learned Options
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Po-Wei, Peng, Pei-Chiun, Guei, Hung, Wu, Ti-Rong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MiniZero: Comparative Analysis of AlphaZero and MuZero on Go, Othello, and Atari Games
by: Wu, Ti-Rong, et al.
Published: (2023)
by: Wu, Ti-Rong, et al.
Published: (2023)
Demystifying MuZero Planning: Interpreting the Learned Model
by: Guei, Hung, et al.
Published: (2024)
by: Guei, Hung, et al.
Published: (2024)
MAPLE: Multi-State Aggregated Policy Evaluation for AlphaZero in Imperfect-Information Games
by: Li, Qian-Rong, et al.
Published: (2026)
by: Li, Qian-Rong, et al.
Published: (2026)
Learning Abstract World Model for Value-preserving Planning with Options
by: Rodriguez-Sanchez, Rafael, et al.
Published: (2024)
by: Rodriguez-Sanchez, Rafael, et al.
Published: (2024)
Evaluating Game Difficulty in Tetris Block Puzzle
by: Wang, Chun-Jui, et al.
Published: (2026)
by: Wang, Chun-Jui, et al.
Published: (2026)
Regret-Guided Search Control for Efficient Learning in AlphaZero
by: Tsai, Yun-Jui, et al.
Published: (2026)
by: Tsai, Yun-Jui, et al.
Published: (2026)
Scalable Option Learning in High-Throughput Environments
by: Henaff, Mikael, et al.
Published: (2025)
by: Henaff, Mikael, et al.
Published: (2025)
Diversity-Enriched Option-Critic
by: Kamat, Anand, et al.
Published: (2020)
by: Kamat, Anand, et al.
Published: (2020)
Unveiling Options with Neural Decomposition
by: Alikhasi, Mahdi, et al.
Published: (2024)
by: Alikhasi, Mahdi, et al.
Published: (2024)
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
by: Nair, Lakshmi, et al.
Published: (2025)
by: Nair, Lakshmi, et al.
Published: (2025)
Learning to Drive Safely with Hybrid Options
by: De Cooman, Bram, et al.
Published: (2025)
by: De Cooman, Bram, et al.
Published: (2025)
Joint Learning of Hierarchical Neural Options and Abstract World Model
by: Piriyakulkij, Wasu Top, et al.
Published: (2026)
by: Piriyakulkij, Wasu Top, et al.
Published: (2026)
Boosting deep Reinforcement Learning using pretraining with Logical Options
by: Ye, Zihan, et al.
Published: (2026)
by: Ye, Zihan, et al.
Published: (2026)
Option Discovery Using LLM-guided Semantic Hierarchical Reinforcement Learning
by: Shek, Chak Lam, et al.
Published: (2025)
by: Shek, Chak Lam, et al.
Published: (2025)
Enabling Option Learning in Sparse Rewards with Hindsight Experience Replay
by: Romio, Gabriel, et al.
Published: (2026)
by: Romio, Gabriel, et al.
Published: (2026)
Preference-based Reinforcement Learning beyond Pairwise Comparisons: Benefits of Multiple Options
by: Lee, Joongkyu, et al.
Published: (2025)
by: Lee, Joongkyu, et al.
Published: (2025)
Option-aware Temporally Abstracted Value for Offline Goal-Conditioned Reinforcement Learning
by: Ahn, Hongjoon, et al.
Published: (2025)
by: Ahn, Hongjoon, et al.
Published: (2025)
SOAP-RL: Sequential Option Advantage Propagation for Reinforcement Learning in POMDP Environments
by: Ishida, Shu, et al.
Published: (2024)
by: Ishida, Shu, et al.
Published: (2024)
Dynamic Sight Range Selection in Multi-Agent Reinforcement Learning
by: Liao, Wei-Chen, et al.
Published: (2025)
by: Liao, Wei-Chen, et al.
Published: (2025)
Learning Human-Like RL Agents Through Trajectory Optimization With Action Quantization
by: Guo, Jian-Ting, et al.
Published: (2025)
by: Guo, Jian-Ting, et al.
Published: (2025)
Constructing an Optimal Behavior Basis for the Option Keyboard
by: Alegre, Lucas N., et al.
Published: (2025)
by: Alegre, Lucas N., et al.
Published: (2025)
Accelerated Portfolio Optimization and Option Pricing with Reinforcement Learning
by: Keramati, Hadi, et al.
Published: (2025)
by: Keramati, Hadi, et al.
Published: (2025)
Not All Options Are Created Equal: Textual Option Weighting for Token-Efficient LLM-Based Knowledge Tracing
by: Kim, JongWoo, et al.
Published: (2024)
by: Kim, JongWoo, et al.
Published: (2024)
An Autonomous Non-monolithic Agent with Multi-mode Exploration based on Options Framework
by: Kim, JaeYoon, et al.
Published: (2023)
by: Kim, JaeYoon, et al.
Published: (2023)
Option-ID Based Elimination For Multiple Choice Questions
by: Zhu, Zhenhao, et al.
Published: (2025)
by: Zhu, Zhenhao, et al.
Published: (2025)
QOG:Question and Options Generation based on Language Model
by: Zhou, Jincheng
Published: (2024)
by: Zhou, Jincheng
Published: (2024)
HUGO -- Highlighting Unseen Grid Options: Combining Deep Reinforcement Learning with a Heuristic Target Topology Approach
by: Lehna, Malte, et al.
Published: (2024)
by: Lehna, Malte, et al.
Published: (2024)
Atomic Action Slicing: Planner-Aligned Options for Generalist VLA Agents
by: Tabakov, Stefan, et al.
Published: (2025)
by: Tabakov, Stefan, et al.
Published: (2025)
Machine Learning with a Reject Option: A survey
by: Hendrickx, Kilian, et al.
Published: (2021)
by: Hendrickx, Kilian, et al.
Published: (2021)
A Novel Approach to Solving Goal-Achieving Problems for Board Games
by: Shih, Chung-Chin, et al.
Published: (2021)
by: Shih, Chung-Chin, et al.
Published: (2021)
Diffusion Meets Options: Hierarchical Generative Skill Composition for Temporally-Extended Tasks
by: Feng, Zeyu, et al.
Published: (2024)
by: Feng, Zeyu, et al.
Published: (2024)
Concisely Explaining the Doubt: Minimum-Size Abductive Explanations for Linear Models with a Reject Option
by: Fernandes, Gleilson Pedro, et al.
Published: (2026)
by: Fernandes, Gleilson Pedro, et al.
Published: (2026)
QMOS: Enhancing LLMs for Telecommunication with Question Masked loss and Option Shuffling
by: Guda, Blessed, et al.
Published: (2024)
by: Guda, Blessed, et al.
Published: (2024)
Large Language Models Badly Generalize across Option Length, Problem Types, and Irrelevant Noun Replacements
by: Zhao, Guangxiang, et al.
Published: (2025)
by: Zhao, Guangxiang, et al.
Published: (2025)
Out-of-distribution Reject Option Method for Dataset Shift Problem in Early Disease Onset Prediction
by: Tosaki, Taisei, et al.
Published: (2024)
by: Tosaki, Taisei, et al.
Published: (2024)
Bridging Local and Global Knowledge via Transformer in Board Games
by: Ju, Yan-Ru, et al.
Published: (2024)
by: Ju, Yan-Ru, et al.
Published: (2024)
Conversational Orientation Reasoning: Egocentric-to-Allocentric Navigation with Multimodal Chain-of-Thought
by: Huang, Yu Ti
Published: (2025)
by: Huang, Yu Ti
Published: (2025)
Language Models as Zero-shot Lossless Gradient Compressors: Towards General Neural Parameter Prior Models
by: Wang, Hui-Po, et al.
Published: (2024)
by: Wang, Hui-Po, et al.
Published: (2024)
Strength Estimation and Human-Like Strength Adjustment in Games
by: Chen, Chun Jung, et al.
Published: (2025)
by: Chen, Chun Jung, et al.
Published: (2025)
A Study of Solving Life-and-Death Problems in Go Using Relevance-Zone Based Solvers
by: Shih, Chung-Chin, et al.
Published: (2025)
by: Shih, Chung-Chin, et al.
Published: (2025)
Similar Items
-
MiniZero: Comparative Analysis of AlphaZero and MuZero on Go, Othello, and Atari Games
by: Wu, Ti-Rong, et al.
Published: (2023) -
Demystifying MuZero Planning: Interpreting the Learned Model
by: Guei, Hung, et al.
Published: (2024) -
MAPLE: Multi-State Aggregated Policy Evaluation for AlphaZero in Imperfect-Information Games
by: Li, Qian-Rong, et al.
Published: (2026) -
Learning Abstract World Model for Value-preserving Planning with Options
by: Rodriguez-Sanchez, Rafael, et al.
Published: (2024) -
Evaluating Game Difficulty in Tetris Block Puzzle
by: Wang, Chun-Jui, et al.
Published: (2026)