Investigating Scale Independent UCT Exploration Factor Strategies
Fuente:
arXiv
Saved in:
| Main Authors: | Schmöcker, Robin, Schnell, Christoph, Dockhorn, Alexander |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Discovering State Equivalences in UCT Search Trees By Action Pruning
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
Grouping Nodes With Known Value Differences: A Lossless UCT-based Abstraction Algorithm
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
Investigating Intra-Abstraction Policies For Non-exact Abstraction Algorithms
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
Time-critical and confidence-based abstraction dropping methods
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
AUPO -- Abstracted Until Proven Otherwise: A Reward Distribution Based Abstraction Algorithm
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
Personalized Dynamic Difficulty Adjustment -- Imitation Learning Meets Reinforcement Learning
by: Fuchs, Ronja, et al.
Published: (2024)
by: Fuchs, Ronja, et al.
Published: (2024)
Match Point AI: A Novel AI Framework for Evaluating Data-Driven Tennis Strategies
by: Nübel, Carlo, et al.
Published: (2024)
by: Nübel, Carlo, et al.
Published: (2024)
Strategy Game-Playing with Size-Constrained State Abstraction
by: Xu, Linjie, et al.
Published: (2024)
by: Xu, Linjie, et al.
Published: (2024)
Markov Senior -- Learning Markov Junior Grammars to Generate User-specified Content
by: Oğuz, Mehmet Kayra, et al.
Published: (2024)
by: Oğuz, Mehmet Kayra, et al.
Published: (2024)
Generalization capabilities of MeshGraphNets to unseen geometries for fluid dynamics
by: Schmöcker, Robin, et al.
Published: (2024)
by: Schmöcker, Robin, et al.
Published: (2024)
Super-Exponential Regret for UCT, AlphaGo and Variants
by: Orseau, Laurent, et al.
Published: (2024)
by: Orseau, Laurent, et al.
Published: (2024)
Threshold UCT: Cost-Constrained Monte Carlo Tree Search with Pareto Curves
by: Kurečka, Martin, et al.
Published: (2024)
by: Kurečka, Martin, et al.
Published: (2024)
Online Optimization of Curriculum Learning Schedules using Evolutionary Optimization
by: Jiwatode, Mohit, et al.
Published: (2024)
by: Jiwatode, Mohit, et al.
Published: (2024)
From Gameplay Traces to Game Mechanics: Causal Induction with Large Language Models
by: Jiwatode, Mohit, et al.
Published: (2026)
by: Jiwatode, Mohit, et al.
Published: (2026)
Scaling Synthetic Task Generation for Agents via Exploration
by: Ramrakhya, Ram, et al.
Published: (2025)
by: Ramrakhya, Ram, et al.
Published: (2025)
Uncertainty-driven Exploration Strategies for Online Grasp Learning
by: Shi, Yitian, et al.
Published: (2023)
by: Shi, Yitian, et al.
Published: (2023)
AutoRL Hyperparameter Landscapes
by: Mohan, Aditya, et al.
Published: (2023)
by: Mohan, Aditya, et al.
Published: (2023)
Efficient Exploration at Scale
by: Asghari, Seyed Mohammad, et al.
Published: (2026)
by: Asghari, Seyed Mohammad, et al.
Published: (2026)
AutoDAN-Reasoning: Enhancing Strategies Exploration based Jailbreak Attacks with Test-Time Scaling
by: Liu, Xiaogeng, et al.
Published: (2025)
by: Liu, Xiaogeng, et al.
Published: (2025)
From Code to Play: Benchmarking Program Search for Games Using Large Language Models
by: Eberhardinger, Manuel, et al.
Published: (2024)
by: Eberhardinger, Manuel, et al.
Published: (2024)
Scale-Adaptive Balancing of Exploration and Exploitation in Classical Planning
by: Wissow, Stephen, et al.
Published: (2023)
by: Wissow, Stephen, et al.
Published: (2023)
Limits to AI Growth: The Ecological and Social Consequences of Scaling
by: Bhardwaj, Eshta, et al.
Published: (2025)
by: Bhardwaj, Eshta, et al.
Published: (2025)
Higher Replay Ratio Empowers Sample-Efficient Multi-Agent Reinforcement Learning
by: Xu, Linjie, et al.
Published: (2024)
by: Xu, Linjie, et al.
Published: (2024)
AI Alignment Strategies from a Risk Perspective: Independent Safety Mechanisms or Shared Failures?
by: Dung, Leonard, et al.
Published: (2025)
by: Dung, Leonard, et al.
Published: (2025)
Learning to Explore: Scaling Agentic Reasoning via Exploration-Aware Policy Optimization
by: Hua, Xingyuan, et al.
Published: (2026)
by: Hua, Xingyuan, et al.
Published: (2026)
A Comparison of Independent and Joint Fine-tuning Strategies for Retrieval-Augmented Generation
by: Lawton, Neal Gregory, et al.
Published: (2025)
by: Lawton, Neal Gregory, et al.
Published: (2025)
Season-Independent PV Disaggregation Using Multi-Scale Net Load Temporal Feature Extraction and Weather Factor Fusion
by: Chen, Xiaolu, et al.
Published: (2025)
by: Chen, Xiaolu, et al.
Published: (2025)
A Safe Exploration Strategy for Model-free Task Adaptation in Safety-constrained Grid Environments
by: Entezami, Erfan, et al.
Published: (2024)
by: Entezami, Erfan, et al.
Published: (2024)
Autonomous 3D Exploration in Large-Scale Environments with Dynamic Obstacles
by: Wiman, Emil, et al.
Published: (2023)
by: Wiman, Emil, et al.
Published: (2023)
AgentCPM-Explore: Realizing Long-Horizon Deep Exploration for Edge-Scale Agents
by: Chen, Haotian, et al.
Published: (2026)
by: Chen, Haotian, et al.
Published: (2026)
ExComm: Exploration-Stage Communication for Error-Resilient Agentic Test-Time Scaling
by: Song, Woomin, et al.
Published: (2026)
by: Song, Woomin, et al.
Published: (2026)
Personalization Increases Affective Alignment but Has Role-Dependent Effects on Epistemic Independence in LLMs
by: Kelley, Sean W., et al.
Published: (2026)
by: Kelley, Sean W., et al.
Published: (2026)
Nudging Beyond the Comfort Zone: Efficient Strategy-Guided Exploration for RLVR
by: Lee, Chanuk, et al.
Published: (2026)
by: Lee, Chanuk, et al.
Published: (2026)
Can Structured Templates Facilitate LLMs in Tackling Harder Tasks? : An Exploration of Scaling Laws by Difficulty
by: Yang, Zhichao, et al.
Published: (2025)
by: Yang, Zhichao, et al.
Published: (2025)
From Pixels to Factors: Learning Independently Controllable State Variables for Reinforcement Learning
by: Rodriguez-Sanchez, Rafael, et al.
Published: (2025)
by: Rodriguez-Sanchez, Rafael, et al.
Published: (2025)
Investigations into Proof Structures
by: Wernhard, Christoph, et al.
Published: (2023)
by: Wernhard, Christoph, et al.
Published: (2023)
The Initial Exploration Problem in Knowledge Graph Exploration
by: McNamara, Claire, et al.
Published: (2026)
by: McNamara, Claire, et al.
Published: (2026)
FastFT: Accelerating Reinforced Feature Transformation via Advanced Exploration Strategies
by: He, Tianqi, et al.
Published: (2025)
by: He, Tianqi, et al.
Published: (2025)
LASTIST: LArge-Scale Target-Independent STance dataset
by: Kim, DongJae, et al.
Published: (2025)
by: Kim, DongJae, et al.
Published: (2025)
Splits! Flexible Sociocultural Linguistic Investigation at Scale
by: Caplan, Eylon, et al.
Published: (2025)
by: Caplan, Eylon, et al.
Published: (2025)
Similar Items
-
Discovering State Equivalences in UCT Search Trees By Action Pruning
by: Schmöcker, Robin, et al.
Published: (2025) -
Grouping Nodes With Known Value Differences: A Lossless UCT-based Abstraction Algorithm
by: Schmöcker, Robin, et al.
Published: (2025) -
Investigating Intra-Abstraction Policies For Non-exact Abstraction Algorithms
by: Schmöcker, Robin, et al.
Published: (2025) -
Time-critical and confidence-based abstraction dropping methods
by: Schmöcker, Robin, et al.
Published: (2025) -
AUPO -- Abstracted Until Proven Otherwise: A Reward Distribution Based Abstraction Algorithm
by: Schmöcker, Robin, et al.
Published: (2025)