Discovering State Equivalences in UCT Search Trees By Action Pruning
Fuente:
arXiv
Saved in:
| Main Authors: | Schmöcker, Robin, Dockhorn, Alexander, Rosenhahn, Bodo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Grouping Nodes With Known Value Differences: A Lossless UCT-based Abstraction Algorithm
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
AUPO -- Abstracted Until Proven Otherwise: A Reward Distribution Based Abstraction Algorithm
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
Investigating Intra-Abstraction Policies For Non-exact Abstraction Algorithms
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
Investigating Scale Independent UCT Exploration Factor Strategies
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
Time-critical and confidence-based abstraction dropping methods
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
From Gameplay Traces to Game Mechanics: Causal Induction with Large Language Models
by: Jiwatode, Mohit, et al.
Published: (2026)
by: Jiwatode, Mohit, et al.
Published: (2026)
Naga: Vedic Encoding for Deep State Space Models
by: Schaller, Melanie, et al.
Published: (2025)
by: Schaller, Melanie, et al.
Published: (2025)
Personalized Dynamic Difficulty Adjustment -- Imitation Learning Meets Reinforcement Learning
by: Fuchs, Ronja, et al.
Published: (2024)
by: Fuchs, Ronja, et al.
Published: (2024)
Threshold UCT: Cost-Constrained Monte Carlo Tree Search with Pareto Curves
by: Kurečka, Martin, et al.
Published: (2024)
by: Kurečka, Martin, et al.
Published: (2024)
Interpretable Decision-Making for End-to-End Autonomous Driving
by: Mirzaie, Mona, et al.
Published: (2025)
by: Mirzaie, Mona, et al.
Published: (2025)
Strategy Game-Playing with Size-Constrained State Abstraction
by: Xu, Linjie, et al.
Published: (2024)
by: Xu, Linjie, et al.
Published: (2024)
Mastering Zero-Shot Interactions in Cooperative and Competitive Simultaneous Games
by: Mahlau, Yannik, et al.
Published: (2024)
by: Mahlau, Yannik, et al.
Published: (2024)
Markov Senior -- Learning Markov Junior Grammars to Generate User-specified Content
by: Oğuz, Mehmet Kayra, et al.
Published: (2024)
by: Oğuz, Mehmet Kayra, et al.
Published: (2024)
Generalization capabilities of MeshGraphNets to unseen geometries for fluid dynamics
by: Schmöcker, Robin, et al.
Published: (2024)
by: Schmöcker, Robin, et al.
Published: (2024)
Match Point AI: A Novel AI Framework for Evaluating Data-Driven Tennis Strategies
by: Nübel, Carlo, et al.
Published: (2024)
by: Nübel, Carlo, et al.
Published: (2024)
Super-Exponential Regret for UCT, AlphaGo and Variants
by: Orseau, Laurent, et al.
Published: (2024)
by: Orseau, Laurent, et al.
Published: (2024)
AutoML for Multi-Class Anomaly Compensation of Sensor Drift
by: Schaller, Melanie, et al.
Published: (2025)
by: Schaller, Melanie, et al.
Published: (2025)
Online Optimization of Curriculum Learning Schedules using Evolutionary Optimization
by: Jiwatode, Mohit, et al.
Published: (2024)
by: Jiwatode, Mohit, et al.
Published: (2024)
Multi-Agent Reinforcement Learning for Inverse Design in Photonic Integrated Circuits
by: Mahlau, Yannik, et al.
Published: (2025)
by: Mahlau, Yannik, et al.
Published: (2025)
Monte Carlo Search Algorithms Discovering Monte Carlo Tree Search Exploration Terms
by: Cazenave, Tristan
Published: (2024)
by: Cazenave, Tristan
Published: (2024)
Pruning by Block Benefit: Exploring the Properties of Vision Transformer Blocks during Domain Adaptation
by: Glandorf, Patrick, et al.
Published: (2025)
by: Glandorf, Patrick, et al.
Published: (2025)
From Code to Play: Benchmarking Program Search for Games Using Large Language Models
by: Eberhardinger, Manuel, et al.
Published: (2024)
by: Eberhardinger, Manuel, et al.
Published: (2024)
EquivPruner: Boosting Efficiency and Quality in LLM-Based Search via Action Pruning
by: Liu, Jiawei, et al.
Published: (2025)
by: Liu, Jiawei, et al.
Published: (2025)
QPM: Discrete Optimization for Globally Interpretable Image Classification
by: Norrenbrock, Thomas, et al.
Published: (2025)
by: Norrenbrock, Thomas, et al.
Published: (2025)
Segment Any Object Model (SAOM): Real-to-Simulation Fine-Tuning Strategy for Multi-Class Multi-Instance Segmentation
by: Khan, Mariia, et al.
Published: (2024)
by: Khan, Mariia, et al.
Published: (2024)
Policy-Space Search: Equivalences, Improvements, and Compression
by: Messa, Frederico, et al.
Published: (2024)
by: Messa, Frederico, et al.
Published: (2024)
Action-aware Dynamic Pruning for Efficient Vision-Language-Action Manipulation
by: Pei, Xiaohuan, et al.
Published: (2025)
by: Pei, Xiaohuan, et al.
Published: (2025)
Less Greedy Equivalence Search
by: Ejaz, Adiba, et al.
Published: (2025)
by: Ejaz, Adiba, et al.
Published: (2025)
Efficient Monte Carlo Tree Search via On-the-Fly State-Conditioned Action Abstraction
by: Kwak, Yunhyeok, et al.
Published: (2024)
by: Kwak, Yunhyeok, et al.
Published: (2024)
Discovering Symmetry Groups with Flow Matching
by: Chen, Yuxuan, et al.
Published: (2025)
by: Chen, Yuxuan, et al.
Published: (2025)
Structure-Aware Automatic Channel Pruning by Searching with Graph Embedding
by: Liu, Zifan, et al.
Published: (2025)
by: Liu, Zifan, et al.
Published: (2025)
Video Patch Pruning: Efficient Video Instance Segmentation via Early Token Reduction
by: Glandorf, Patrick, et al.
Published: (2026)
by: Glandorf, Patrick, et al.
Published: (2026)
AutoRL Hyperparameter Landscapes
by: Mohan, Aditya, et al.
Published: (2023)
by: Mohan, Aditya, et al.
Published: (2023)
Discovering Mathematical Formulas from Data via GPT-guided Monte Carlo Tree Search
by: Li, Yanjie, et al.
Published: (2024)
by: Li, Yanjie, et al.
Published: (2024)
ToolTree: Efficient LLM Agent Tool Planning via Dual-Feedback Monte Carlo Tree Search and Bidirectional Pruning
by: Yang, Shuo, et al.
Published: (2026)
by: Yang, Shuo, et al.
Published: (2026)
SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning
by: Wang, Hanzhen, et al.
Published: (2025)
by: Wang, Hanzhen, et al.
Published: (2025)
SLOPE: Search with Learned Optimal Pruning-based Expansion
by: Bokan, Davor, et al.
Published: (2024)
by: Bokan, Davor, et al.
Published: (2024)
Environment-Aware Adaptive Pruning with Interleaved Inference Orchestration for Vision-Language-Action Models
by: Huang, Yuting, et al.
Published: (2026)
by: Huang, Yuting, et al.
Published: (2026)
Discovering Spoofing Attempts on Language Model Watermarks
by: Gloaguen, Thibaud, et al.
Published: (2024)
by: Gloaguen, Thibaud, et al.
Published: (2024)
Discovering Symmetry Breaking in Physical Systems with Relaxed Group Convolution
by: Wang, Rui, et al.
Published: (2023)
by: Wang, Rui, et al.
Published: (2023)
Similar Items
-
Grouping Nodes With Known Value Differences: A Lossless UCT-based Abstraction Algorithm
by: Schmöcker, Robin, et al.
Published: (2025) -
AUPO -- Abstracted Until Proven Otherwise: A Reward Distribution Based Abstraction Algorithm
by: Schmöcker, Robin, et al.
Published: (2025) -
Investigating Intra-Abstraction Policies For Non-exact Abstraction Algorithms
by: Schmöcker, Robin, et al.
Published: (2025) -
Investigating Scale Independent UCT Exploration Factor Strategies
by: Schmöcker, Robin, et al.
Published: (2025) -
Time-critical and confidence-based abstraction dropping methods
by: Schmöcker, Robin, et al.
Published: (2025)