Amplifying Exploration in Monte-Carlo Tree Search by Focusing on the Unknown
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Derstroff, Cedric, Brugger, Jannis, Blüml, Jannis, Mezini, Mira, Kramer, Stefan, Kersting, Kristian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptable Hindsight Experience Replay for Search-Based Learning
von: Vazaios, Alexandros, et al.
Veröffentlicht: (2025)
von: Vazaios, Alexandros, et al.
Veröffentlicht: (2025)
Neural-Guided Equation Discovery
von: Brugger, Jannis, et al.
Veröffentlicht: (2025)
von: Brugger, Jannis, et al.
Veröffentlicht: (2025)
Deep Reinforcement Learning via Object-Centric Attention
von: Blüml, Jannis, et al.
Veröffentlicht: (2025)
von: Blüml, Jannis, et al.
Veröffentlicht: (2025)
Polynomial Regret Concentration of UCB for Non-Deterministic State Transitions
von: Cömer, Can, et al.
Veröffentlicht: (2025)
von: Cömer, Can, et al.
Veröffentlicht: (2025)
Peer Learning: Learning Complex Policies in Groups from Scratch via Action Recommendations
von: Derstroff, Cedric, et al.
Veröffentlicht: (2023)
von: Derstroff, Cedric, et al.
Veröffentlicht: (2023)
Boosting deep Reinforcement Learning using pretraining with Logical Options
von: Ye, Zihan, et al.
Veröffentlicht: (2026)
von: Ye, Zihan, et al.
Veröffentlicht: (2026)
Representation Matters for Mastering Chess: Improved Feature Representation in AlphaZero Outperforms Switching to Transformers
von: Czech, Johannes, et al.
Veröffentlicht: (2023)
von: Czech, Johannes, et al.
Veröffentlicht: (2023)
HackAtari: Atari Learning Environments for Robust and Continual Reinforcement Learning
von: Delfosse, Quentin, et al.
Veröffentlicht: (2024)
von: Delfosse, Quentin, et al.
Veröffentlicht: (2024)
OCAtari: Object-Centric Atari 2600 Reinforcement Learning Environments
von: Delfosse, Quentin, et al.
Veröffentlicht: (2023)
von: Delfosse, Quentin, et al.
Veröffentlicht: (2023)
Better Decisions through the Right Causal World Model
von: Dillies, Elisabeth, et al.
Veröffentlicht: (2025)
von: Dillies, Elisabeth, et al.
Veröffentlicht: (2025)
Prompting Neural-Guided Equation Discovery Based on Residuals
von: Brugger, Jannis, et al.
Veröffentlicht: (2025)
von: Brugger, Jannis, et al.
Veröffentlicht: (2025)
Automated Scientific Discovery: From Equation Discovery to Autonomous Discovery Systems
von: Kramer, Stefan, et al.
Veröffentlicht: (2023)
von: Kramer, Stefan, et al.
Veröffentlicht: (2023)
Deep Reinforcement Learning Agents are not even close to Human Intelligence
von: Delfosse, Quentin, et al.
Veröffentlicht: (2025)
von: Delfosse, Quentin, et al.
Veröffentlicht: (2025)
Monte Carlo Search Algorithms Discovering Monte Carlo Tree Search Exploration Terms
von: Cazenave, Tristan
Veröffentlicht: (2024)
von: Cazenave, Tristan
Veröffentlicht: (2024)
Checkmating One, by Using Many: Combining Mixture of Experts with MCTS to Improve in Chess
von: Helfenstein, Felix, et al.
Veröffentlicht: (2024)
von: Helfenstein, Felix, et al.
Veröffentlicht: (2024)
Monte Carlo Tree Search with Boltzmann Exploration
von: Painter, Michael, et al.
Veröffentlicht: (2024)
von: Painter, Michael, et al.
Veröffentlicht: (2024)
CodeSSM: Towards State Space Models for Code Understanding
von: Verma, Shweta, et al.
Veröffentlicht: (2025)
von: Verma, Shweta, et al.
Veröffentlicht: (2025)
Integrating Symbolic Execution into the Fine-Tuning of Code-Generating LLMs
von: Sakharova, Marina, et al.
Veröffentlicht: (2025)
von: Sakharova, Marina, et al.
Veröffentlicht: (2025)
OCALM: Object-Centric Assessment with Language Models
von: Kaufmann, Timo, et al.
Veröffentlicht: (2024)
von: Kaufmann, Timo, et al.
Veröffentlicht: (2024)
Monte Carlo Tree Search for Comprehensive Exploration in LLM-Based Automatic Heuristic Design
von: Zheng, Zhi, et al.
Veröffentlicht: (2025)
von: Zheng, Zhi, et al.
Veröffentlicht: (2025)
Towards Understanding What State Space Models Learn About Code
von: Wu, Jiali, et al.
Veröffentlicht: (2026)
von: Wu, Jiali, et al.
Veröffentlicht: (2026)
Efficient Post-training of LLMs for Code Generation With Offline Reinforcement Learning
von: Wu, Mingze, et al.
Veröffentlicht: (2026)
von: Wu, Mingze, et al.
Veröffentlicht: (2026)
Problem Solving Through Human-AI Preference-Based Cooperation
von: Dutta, Subhabrata, et al.
Veröffentlicht: (2024)
von: Dutta, Subhabrata, et al.
Veröffentlicht: (2024)
Evaluating and Mitigating Errors in LLM-Generated Web API Integrations
von: Maninger, Daniel, et al.
Veröffentlicht: (2025)
von: Maninger, Daniel, et al.
Veröffentlicht: (2025)
Analysis of Long Range Dependency Understanding in State Space Models
von: Ravikumar, Srividya, et al.
Veröffentlicht: (2026)
von: Ravikumar, Srividya, et al.
Veröffentlicht: (2026)
Focused Chain-of-Thought: Efficient LLM Reasoning via Structured Input Information
von: Struppek, Lukas, et al.
Veröffentlicht: (2025)
von: Struppek, Lukas, et al.
Veröffentlicht: (2025)
Epistemic Monte Carlo Tree Search
von: Oren, Yaniv, et al.
Veröffentlicht: (2022)
von: Oren, Yaniv, et al.
Veröffentlicht: (2022)
How Susceptible are LLMs to Influence in Prompts?
von: Anagnostidis, Sotiris, et al.
Veröffentlicht: (2024)
von: Anagnostidis, Sotiris, et al.
Veröffentlicht: (2024)
Generalized Proof-Number Monte-Carlo Tree Search
von: Kowalski, Jakub, et al.
Veröffentlicht: (2025)
von: Kowalski, Jakub, et al.
Veröffentlicht: (2025)
Proof Number Based Monte-Carlo Tree Search
von: Kowalski, Jakub, et al.
Veröffentlicht: (2023)
von: Kowalski, Jakub, et al.
Veröffentlicht: (2023)
A Critical Study of What Code-LLMs (Do Not) Learn
von: Anand, Abhinav, et al.
Veröffentlicht: (2024)
von: Anand, Abhinav, et al.
Veröffentlicht: (2024)
Towards a Characterisation of Monte-Carlo Tree Search Performance in Different Games
von: Soemers, Dennis J. N. J., et al.
Veröffentlicht: (2024)
von: Soemers, Dennis J. N. J., et al.
Veröffentlicht: (2024)
Interpretable end-to-end Neurosymbolic Reinforcement Learning agents
von: Grandien, Nils, et al.
Veröffentlicht: (2024)
von: Grandien, Nils, et al.
Veröffentlicht: (2024)
Learning from Less: Guiding Deep Reinforcement Learning with Differentiable Symbolic Planning
von: Ye, Zihan, et al.
Veröffentlicht: (2025)
von: Ye, Zihan, et al.
Veröffentlicht: (2025)
Array-Based Monte Carlo Tree Search
von: Ragan, James, et al.
Veröffentlicht: (2025)
von: Ragan, James, et al.
Veröffentlicht: (2025)
Lookahead Pathology in Monte-Carlo Tree Search
von: Nguyen, Khoi P. N., et al.
Veröffentlicht: (2022)
von: Nguyen, Khoi P. N., et al.
Veröffentlicht: (2022)
Twice Sequential Monte Carlo for Tree Search
von: Oren, Yaniv, et al.
Veröffentlicht: (2025)
von: Oren, Yaniv, et al.
Veröffentlicht: (2025)
Doubly Robust Monte Carlo Tree Search
von: Liu, Manqing, et al.
Veröffentlicht: (2025)
von: Liu, Manqing, et al.
Veröffentlicht: (2025)
Combining Monte Carlo Tree Search and Heuristic Search for Weighted Vertex Coloring
von: Grelier, Cyril, et al.
Veröffentlicht: (2023)
von: Grelier, Cyril, et al.
Veröffentlicht: (2023)
Surrogate Assisted Monte Carlo Tree Search in Combinatorial Optimization
von: Amiri, Saeid, et al.
Veröffentlicht: (2024)
von: Amiri, Saeid, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Adaptable Hindsight Experience Replay for Search-Based Learning
von: Vazaios, Alexandros, et al.
Veröffentlicht: (2025) -
Neural-Guided Equation Discovery
von: Brugger, Jannis, et al.
Veröffentlicht: (2025) -
Deep Reinforcement Learning via Object-Centric Attention
von: Blüml, Jannis, et al.
Veröffentlicht: (2025) -
Polynomial Regret Concentration of UCB for Non-Deterministic State Transitions
von: Cömer, Can, et al.
Veröffentlicht: (2025) -
Peer Learning: Learning Complex Policies in Groups from Scratch via Action Recommendations
von: Derstroff, Cedric, et al.
Veröffentlicht: (2023)