Scale-Adaptive Balancing of Exploration and Exploitation in Classical Planning
Fuente:
arXiv
Saved in:
| Main Authors: | Wissow, Stephen, Asai, Masataro |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Extreme Value Monte Carlo Tree Search for Classical Planning
by: Asai, Masataro, et al.
Published: (2024)
by: Asai, Masataro, et al.
Published: (2024)
Bilevel MCTS for Amortized O(1) Node Selection in Classical Planning
by: Asai, Masataro
Published: (2025)
by: Asai, Masataro
Published: (2025)
"Don't Do That!": Guiding Embodied Systems through Large Language Model-based Constraint Generation
by: Seffo, Amin, et al.
Published: (2025)
by: Seffo, Amin, et al.
Published: (2025)
Query Decomposition for RAG: Balancing Exploration-Exploitation
by: Petcu, Roxana, et al.
Published: (2025)
by: Petcu, Roxana, et al.
Published: (2025)
Plan-MCTS: Plan Exploration for Action Exploitation in Web Navigation
by: Zhang, Weiming, et al.
Published: (2026)
by: Zhang, Weiming, et al.
Published: (2026)
Pass@k Training for Adaptively Balancing Exploration and Exploitation of Large Reasoning Models
by: Chen, Zhipeng, et al.
Published: (2025)
by: Chen, Zhipeng, et al.
Published: (2025)
$ϕ$-Decoding: Adaptive Foresight Sampling for Balanced Inference-Time Exploration and Exploitation
by: Xu, Fangzhi, et al.
Published: (2025)
by: Xu, Fangzhi, et al.
Published: (2025)
SELF-REDRAFT: Eliciting Intrinsic Exploration-Exploitation Balance in Test-Time Scaling for Code Generation
by: Chen, Yixiang, et al.
Published: (2025)
by: Chen, Yixiang, et al.
Published: (2025)
On Using Admissible Bounds for Learning Forward Search Heuristics
by: Núñez-Molina, Carlos, et al.
Published: (2023)
by: Núñez-Molina, Carlos, et al.
Published: (2023)
Balancing Exploration and Exploitation in LLM using Soft RLLF for Enhanced Negation Understanding
by: Nguyen, Ha-Thanh, et al.
Published: (2024)
by: Nguyen, Ha-Thanh, et al.
Published: (2024)
An Improved FOX Optimization Algorithm Using Adaptive Exploration and Exploitation for Global Optimization
by: Jumaah, Mahmood A., et al.
Published: (2025)
by: Jumaah, Mahmood A., et al.
Published: (2025)
Count-based Novelty Exploration in Classical Planning
by: Rosa, Giacomo, et al.
Published: (2024)
by: Rosa, Giacomo, et al.
Published: (2024)
B-STaR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners
by: Zeng, Weihao, et al.
Published: (2024)
by: Zeng, Weihao, et al.
Published: (2024)
Enhancing Adversarial Transferability by Balancing Exploration and Exploitation with Gradient-Guided Sampling
by: Niu, Zenghao, et al.
Published: (2025)
by: Niu, Zenghao, et al.
Published: (2025)
ETTRL: Balancing Exploration and Exploitation in LLM Test-Time Reinforcement Learning Via Entropy Mechanism
by: Liu, Jia, et al.
Published: (2025)
by: Liu, Jia, et al.
Published: (2025)
In-context Exploration-Exploitation for Reinforcement Learning
by: Dai, Zhenwen, et al.
Published: (2024)
by: Dai, Zhenwen, et al.
Published: (2024)
Exploitation Is All You Need... for Exploration
by: Rentschler, Micah, et al.
Published: (2025)
by: Rentschler, Micah, et al.
Published: (2025)
Exploration and Exploitation Errors Are Measurable for Language Model Agents
by: Park, Jaden, et al.
Published: (2026)
by: Park, Jaden, et al.
Published: (2026)
Navigating the Exploration-Exploitation Tradeoff in Inference-Time Scaling of Diffusion Models
by: Su, Xun, et al.
Published: (2025)
by: Su, Xun, et al.
Published: (2025)
Accelerating Large-Scale Dataset Distillation via Exploration-Exploitation Optimization
by: Alahmadi, Muhammad J., et al.
Published: (2026)
by: Alahmadi, Muhammad J., et al.
Published: (2026)
HTPO: Towards Exploration-Exploitation Balanced Policy Optimization via Hierarchical Token-level Objective Control
by: Yao, Xincheng, et al.
Published: (2026)
by: Yao, Xincheng, et al.
Published: (2026)
WESE: Weak Exploration to Strong Exploitation for LLM Agents
by: Huang, Xu, et al.
Published: (2024)
by: Huang, Xu, et al.
Published: (2024)
Structural Induced Exploration for Balanced and Scalable Multi-Robot Path Planning
by: Guo, Zikun, et al.
Published: (2025)
by: Guo, Zikun, et al.
Published: (2025)
MAGE: Meta-Reinforcement Learning for Language Agents toward Strategic Exploration and Exploitation
by: Yang, Lu, et al.
Published: (2026)
by: Yang, Lu, et al.
Published: (2026)
A Balanced Approach of Rapid Genetic Exploration and Surrogate Exploitation for Hyperparameter Optimization
by: Kim, Chul, et al.
Published: (2025)
by: Kim, Chul, et al.
Published: (2025)
Decoupling Exploration and Exploitation for Unsupervised Pre-training with Successor Features
by: Kim, JaeYoon, et al.
Published: (2024)
by: Kim, JaeYoon, et al.
Published: (2024)
ORBIT: On-policy Exploration-Exploitation for Controllable Multi-Budget Reasoning
by: Liang, Kun, et al.
Published: (2026)
by: Liang, Kun, et al.
Published: (2026)
Structured Exploration and Exploitation of Label Functions for Automated Data Annotation
by: Lam, Phong, et al.
Published: (2026)
by: Lam, Phong, et al.
Published: (2026)
Disentangling Exploration of Large Language Models by Optimal Exploitation
by: Grams, Tim, et al.
Published: (2025)
by: Grams, Tim, et al.
Published: (2025)
Adaptive Data Exploitation in Deep Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2025)
by: Yuan, Mingqi, et al.
Published: (2025)
Strategy-Augmented Planning for Large Language Models via Opponent Exploitation
by: Xu, Shuai, et al.
Published: (2025)
by: Xu, Shuai, et al.
Published: (2025)
Controlling Exploration-Exploitation in GFlowNets via Markov Chain Perspectives
by: Chen, Lin, et al.
Published: (2026)
by: Chen, Lin, et al.
Published: (2026)
HyPER: Bridging Exploration and Exploitation for Scalable LLM Reasoning with Hypothesis Path Expansion and Reduction
by: Qiu, Shengxuan, et al.
Published: (2026)
by: Qiu, Shengxuan, et al.
Published: (2026)
StratFormer: Adaptive Opponent Modeling and Exploitation in Imperfect-Information Games
by: Caen, Andy, et al.
Published: (2026)
by: Caen, Andy, et al.
Published: (2026)
Co-Exploration and Co-Exploitation via Shared Structure in Multi-Task Bandits
by: Mukherjee, Sumantrak, et al.
Published: (2025)
by: Mukherjee, Sumantrak, et al.
Published: (2025)
AdaSwitch: Balancing Exploration and Guidance in Knowledge Distillation via Adaptive Switching
by: Peng, Jingyu, et al.
Published: (2025)
by: Peng, Jingyu, et al.
Published: (2025)
Code Repair with LLMs gives an Exploration-Exploitation Tradeoff
by: Tang, Hao, et al.
Published: (2024)
by: Tang, Hao, et al.
Published: (2024)
First-Explore, then Exploit: Meta-Learning to Solve Hard Exploration-Exploitation Trade-Offs
by: Norman, Ben, et al.
Published: (2023)
by: Norman, Ben, et al.
Published: (2023)
DGRO: Enhancing LLM Reasoning via Exploration-Exploitation Control and Reward Variance Management
by: Su, Xuerui, et al.
Published: (2025)
by: Su, Xuerui, et al.
Published: (2025)
Dual-Dimensional Consistency: Balancing Budget and Quality in Adaptive Inference-Time Scaling
by: Xu, Rongman, et al.
Published: (2026)
by: Xu, Rongman, et al.
Published: (2026)
Similar Items
-
Extreme Value Monte Carlo Tree Search for Classical Planning
by: Asai, Masataro, et al.
Published: (2024) -
Bilevel MCTS for Amortized O(1) Node Selection in Classical Planning
by: Asai, Masataro
Published: (2025) -
"Don't Do That!": Guiding Embodied Systems through Large Language Model-based Constraint Generation
by: Seffo, Amin, et al.
Published: (2025) -
Query Decomposition for RAG: Balancing Exploration-Exploitation
by: Petcu, Roxana, et al.
Published: (2025) -
Plan-MCTS: Plan Exploration for Action Exploitation in Web Navigation
by: Zhang, Weiming, et al.
Published: (2026)