Amplifying Exploration in Monte-Carlo Tree Search by Focusing on the Unknown
Fuente:
arXiv
Saved in:
| Main Authors: | Derstroff, Cedric, Brugger, Jannis, Blüml, Jannis, Mezini, Mira, Kramer, Stefan, Kersting, Kristian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptable Hindsight Experience Replay for Search-Based Learning
by: Vazaios, Alexandros, et al.
Published: (2025)
by: Vazaios, Alexandros, et al.
Published: (2025)
Neural-Guided Equation Discovery
by: Brugger, Jannis, et al.
Published: (2025)
by: Brugger, Jannis, et al.
Published: (2025)
Deep Reinforcement Learning via Object-Centric Attention
by: Blüml, Jannis, et al.
Published: (2025)
by: Blüml, Jannis, et al.
Published: (2025)
Polynomial Regret Concentration of UCB for Non-Deterministic State Transitions
by: Cömer, Can, et al.
Published: (2025)
by: Cömer, Can, et al.
Published: (2025)
Peer Learning: Learning Complex Policies in Groups from Scratch via Action Recommendations
by: Derstroff, Cedric, et al.
Published: (2023)
by: Derstroff, Cedric, et al.
Published: (2023)
Boosting deep Reinforcement Learning using pretraining with Logical Options
by: Ye, Zihan, et al.
Published: (2026)
by: Ye, Zihan, et al.
Published: (2026)
Representation Matters for Mastering Chess: Improved Feature Representation in AlphaZero Outperforms Switching to Transformers
by: Czech, Johannes, et al.
Published: (2023)
by: Czech, Johannes, et al.
Published: (2023)
HackAtari: Atari Learning Environments for Robust and Continual Reinforcement Learning
by: Delfosse, Quentin, et al.
Published: (2024)
by: Delfosse, Quentin, et al.
Published: (2024)
OCAtari: Object-Centric Atari 2600 Reinforcement Learning Environments
by: Delfosse, Quentin, et al.
Published: (2023)
by: Delfosse, Quentin, et al.
Published: (2023)
Better Decisions through the Right Causal World Model
by: Dillies, Elisabeth, et al.
Published: (2025)
by: Dillies, Elisabeth, et al.
Published: (2025)
Prompting Neural-Guided Equation Discovery Based on Residuals
by: Brugger, Jannis, et al.
Published: (2025)
by: Brugger, Jannis, et al.
Published: (2025)
Automated Scientific Discovery: From Equation Discovery to Autonomous Discovery Systems
by: Kramer, Stefan, et al.
Published: (2023)
by: Kramer, Stefan, et al.
Published: (2023)
Deep Reinforcement Learning Agents are not even close to Human Intelligence
by: Delfosse, Quentin, et al.
Published: (2025)
by: Delfosse, Quentin, et al.
Published: (2025)
Monte Carlo Search Algorithms Discovering Monte Carlo Tree Search Exploration Terms
by: Cazenave, Tristan
Published: (2024)
by: Cazenave, Tristan
Published: (2024)
Checkmating One, by Using Many: Combining Mixture of Experts with MCTS to Improve in Chess
by: Helfenstein, Felix, et al.
Published: (2024)
by: Helfenstein, Felix, et al.
Published: (2024)
Monte Carlo Tree Search with Boltzmann Exploration
by: Painter, Michael, et al.
Published: (2024)
by: Painter, Michael, et al.
Published: (2024)
CodeSSM: Towards State Space Models for Code Understanding
by: Verma, Shweta, et al.
Published: (2025)
by: Verma, Shweta, et al.
Published: (2025)
Integrating Symbolic Execution into the Fine-Tuning of Code-Generating LLMs
by: Sakharova, Marina, et al.
Published: (2025)
by: Sakharova, Marina, et al.
Published: (2025)
OCALM: Object-Centric Assessment with Language Models
by: Kaufmann, Timo, et al.
Published: (2024)
by: Kaufmann, Timo, et al.
Published: (2024)
Monte Carlo Tree Search for Comprehensive Exploration in LLM-Based Automatic Heuristic Design
by: Zheng, Zhi, et al.
Published: (2025)
by: Zheng, Zhi, et al.
Published: (2025)
Towards Understanding What State Space Models Learn About Code
by: Wu, Jiali, et al.
Published: (2026)
by: Wu, Jiali, et al.
Published: (2026)
Efficient Post-training of LLMs for Code Generation With Offline Reinforcement Learning
by: Wu, Mingze, et al.
Published: (2026)
by: Wu, Mingze, et al.
Published: (2026)
Problem Solving Through Human-AI Preference-Based Cooperation
by: Dutta, Subhabrata, et al.
Published: (2024)
by: Dutta, Subhabrata, et al.
Published: (2024)
Evaluating and Mitigating Errors in LLM-Generated Web API Integrations
by: Maninger, Daniel, et al.
Published: (2025)
by: Maninger, Daniel, et al.
Published: (2025)
Analysis of Long Range Dependency Understanding in State Space Models
by: Ravikumar, Srividya, et al.
Published: (2026)
by: Ravikumar, Srividya, et al.
Published: (2026)
Focused Chain-of-Thought: Efficient LLM Reasoning via Structured Input Information
by: Struppek, Lukas, et al.
Published: (2025)
by: Struppek, Lukas, et al.
Published: (2025)
Epistemic Monte Carlo Tree Search
by: Oren, Yaniv, et al.
Published: (2022)
by: Oren, Yaniv, et al.
Published: (2022)
How Susceptible are LLMs to Influence in Prompts?
by: Anagnostidis, Sotiris, et al.
Published: (2024)
by: Anagnostidis, Sotiris, et al.
Published: (2024)
Generalized Proof-Number Monte-Carlo Tree Search
by: Kowalski, Jakub, et al.
Published: (2025)
by: Kowalski, Jakub, et al.
Published: (2025)
Proof Number Based Monte-Carlo Tree Search
by: Kowalski, Jakub, et al.
Published: (2023)
by: Kowalski, Jakub, et al.
Published: (2023)
A Critical Study of What Code-LLMs (Do Not) Learn
by: Anand, Abhinav, et al.
Published: (2024)
by: Anand, Abhinav, et al.
Published: (2024)
Towards a Characterisation of Monte-Carlo Tree Search Performance in Different Games
by: Soemers, Dennis J. N. J., et al.
Published: (2024)
by: Soemers, Dennis J. N. J., et al.
Published: (2024)
Interpretable end-to-end Neurosymbolic Reinforcement Learning agents
by: Grandien, Nils, et al.
Published: (2024)
by: Grandien, Nils, et al.
Published: (2024)
Learning from Less: Guiding Deep Reinforcement Learning with Differentiable Symbolic Planning
by: Ye, Zihan, et al.
Published: (2025)
by: Ye, Zihan, et al.
Published: (2025)
Array-Based Monte Carlo Tree Search
by: Ragan, James, et al.
Published: (2025)
by: Ragan, James, et al.
Published: (2025)
Lookahead Pathology in Monte-Carlo Tree Search
by: Nguyen, Khoi P. N., et al.
Published: (2022)
by: Nguyen, Khoi P. N., et al.
Published: (2022)
Twice Sequential Monte Carlo for Tree Search
by: Oren, Yaniv, et al.
Published: (2025)
by: Oren, Yaniv, et al.
Published: (2025)
Doubly Robust Monte Carlo Tree Search
by: Liu, Manqing, et al.
Published: (2025)
by: Liu, Manqing, et al.
Published: (2025)
Combining Monte Carlo Tree Search and Heuristic Search for Weighted Vertex Coloring
by: Grelier, Cyril, et al.
Published: (2023)
by: Grelier, Cyril, et al.
Published: (2023)
Surrogate Assisted Monte Carlo Tree Search in Combinatorial Optimization
by: Amiri, Saeid, et al.
Published: (2024)
by: Amiri, Saeid, et al.
Published: (2024)
Similar Items
-
Adaptable Hindsight Experience Replay for Search-Based Learning
by: Vazaios, Alexandros, et al.
Published: (2025) -
Neural-Guided Equation Discovery
by: Brugger, Jannis, et al.
Published: (2025) -
Deep Reinforcement Learning via Object-Centric Attention
by: Blüml, Jannis, et al.
Published: (2025) -
Polynomial Regret Concentration of UCB for Non-Deterministic State Transitions
by: Cömer, Can, et al.
Published: (2025) -
Peer Learning: Learning Complex Policies in Groups from Scratch via Action Recommendations
by: Derstroff, Cedric, et al.
Published: (2023)