Tail-Risk-Safe Monte Carlo Tree Search under PAC-Level Guarantees
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Zuyuan, Ghosh, Arnob, Lan, Tian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lipschitz Lifelong Monte Carlo Tree Search for Mastering Non-Stationary Tasks
by: Zhang, Zuyuan, et al.
Published: (2025)
by: Zhang, Zuyuan, et al.
Published: (2025)
Optimistic Policy Learning under Pessimistic Adversaries with Regret and Violation Guarantees
by: Ganguly, Sourav, et al.
Published: (2026)
by: Ganguly, Sourav, et al.
Published: (2026)
Geometry of Drifting MDPs with Path-Integral Stability Certificates
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
Cochain Perspectives on Temporal-Difference Signals for Learning Beyond Markov Dynamics
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
C-MCTS: Safe Planning with Monte Carlo Tree Search
by: Parthasarathy, Dinesh, et al.
Published: (2023)
by: Parthasarathy, Dinesh, et al.
Published: (2023)
Operator-Guided Invariance Learning for Continuous Reinforcement Learning
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
Matrix-Space Reinforcement Learning for Reusing Local Transition Geometry
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
NonZero: Interaction-Guided Exploration for Multi-Agent Monte Carlo Tree Search
by: Tang, Sizhe, et al.
Published: (2026)
by: Tang, Sizhe, et al.
Published: (2026)
Metric-Gradient Projection for Stable Multi-Agent Policy Learning
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
Manifold-Constrained Energy-Based Transition Models for Offline Reinforcement Learning
by: Fang, Zeyu, et al.
Published: (2026)
by: Fang, Zeyu, et al.
Published: (2026)
Efficient Policy Optimization in Robust Constrained MDPs with Iteration Complexity Guarantees
by: Ganguly, Sourav, et al.
Published: (2025)
by: Ganguly, Sourav, et al.
Published: (2025)
Cooperative Backdoor Attack in Decentralized Reinforcement Learning with Theoretical Guarantee
by: Gao, Mengtong, et al.
Published: (2024)
by: Gao, Mengtong, et al.
Published: (2024)
Epistemic Monte Carlo Tree Search
by: Oren, Yaniv, et al.
Published: (2022)
by: Oren, Yaniv, et al.
Published: (2022)
HyPAC: Cost-Efficient LLMs-Human Hybrid Annotation with PAC Error Guarantees
by: Zeng, Hao, et al.
Published: (2026)
by: Zeng, Hao, et al.
Published: (2026)
Monte Carlo Tree Search in the Presence of Transition Uncertainty
by: Kohankhaki, Farnaz, et al.
Published: (2023)
by: Kohankhaki, Farnaz, et al.
Published: (2023)
Twice Sequential Monte Carlo for Tree Search
by: Oren, Yaniv, et al.
Published: (2025)
by: Oren, Yaniv, et al.
Published: (2025)
Doubly Robust Monte Carlo Tree Search
by: Liu, Manqing, et al.
Published: (2025)
by: Liu, Manqing, et al.
Published: (2025)
Monte Carlo Tree Search with Boltzmann Exploration
by: Painter, Michael, et al.
Published: (2024)
by: Painter, Michael, et al.
Published: (2024)
Spectral-Risk Safe Reinforcement Learning with Convergence Guarantees
by: Kim, Dohyeong, et al.
Published: (2024)
by: Kim, Dohyeong, et al.
Published: (2024)
Anytime Sequential Halving in Monte-Carlo Tree Search
by: Sagers, Dominic, et al.
Published: (2024)
by: Sagers, Dominic, et al.
Published: (2024)
Improving GFlowNets with Monte Carlo Tree Search
by: Morozov, Nikita, et al.
Published: (2024)
by: Morozov, Nikita, et al.
Published: (2024)
Certifiable Safe RLHF: Fixed-Penalty Constraint Optimization for Safer Language Models
by: Pandit, Kartik, et al.
Published: (2025)
by: Pandit, Kartik, et al.
Published: (2025)
Monte Carlo Permutation Search
by: Cazenave, Tristan
Published: (2025)
by: Cazenave, Tristan
Published: (2025)
Interactive Critique-Revision Training for Reliable Structured LLM Generation
by: Yu, Fei Xu, et al.
Published: (2026)
by: Yu, Fei Xu, et al.
Published: (2026)
Variance-Aware Prior-Based Tree Policies for Monte Carlo Tree Search
by: Weichart, Maximilian
Published: (2025)
by: Weichart, Maximilian
Published: (2025)
Efficient Monte Carlo Tree Search via On-the-Fly State-Conditioned Action Abstraction
by: Kwak, Yunhyeok, et al.
Published: (2024)
by: Kwak, Yunhyeok, et al.
Published: (2024)
Continuous Monte Carlo Graph Search
by: Kujanpää, Kalle, et al.
Published: (2022)
by: Kujanpää, Kalle, et al.
Published: (2022)
Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning
by: Xie, Yuxi, et al.
Published: (2024)
by: Xie, Yuxi, et al.
Published: (2024)
Optimizing Tensor Computation Graphs with Equality Saturation and Monte Carlo Tree Search
by: Hartmann, Jakob, et al.
Published: (2024)
by: Hartmann, Jakob, et al.
Published: (2024)
Monte Carlo Tree Search based Space Transfer for Black-box Optimization
by: Wang, Shukuan, et al.
Published: (2024)
by: Wang, Shukuan, et al.
Published: (2024)
Active Level Set Estimation for Continuous Search Space with Theoretical Guarantee
by: Ngo, Giang, et al.
Published: (2024)
by: Ngo, Giang, et al.
Published: (2024)
Object-Centric World Models Meet Monte Carlo Tree Search
by: Vakhitov, Rodion, et al.
Published: (2026)
by: Vakhitov, Rodion, et al.
Published: (2026)
Bayes Adaptive Monte Carlo Tree Search for Offline Model-based Reinforcement Learning
by: Chen, Jiayu, et al.
Published: (2024)
by: Chen, Jiayu, et al.
Published: (2024)
Explainable Graph Neural Architecture Search via Monte-Carlo Tree Search (Full version)
by: Sasaki, Yuya
Published: (2023)
by: Sasaki, Yuya
Published: (2023)
UNSAT Solver Synthesis via Monte Carlo Forest Search
by: Cameron, Chris, et al.
Published: (2022)
by: Cameron, Chris, et al.
Published: (2022)
Discovering Mathematical Formulas from Data via GPT-guided Monte Carlo Tree Search
by: Li, Yanjie, et al.
Published: (2024)
by: Li, Yanjie, et al.
Published: (2024)
Enhancing Reasoning through Process Supervision with Monte Carlo Tree Search
by: Li, Shuangtao, et al.
Published: (2025)
by: Li, Shuangtao, et al.
Published: (2025)
Monte Carlo Beam Search for Actor-Critic Reinforcement Learning in Continuous Control
by: Alzorgan, Hazim, et al.
Published: (2025)
by: Alzorgan, Hazim, et al.
Published: (2025)
Predicting and Publishing Accurate Imbalance Prices Using Monte Carlo Tree Search
by: Pavirani, Fabio, et al.
Published: (2024)
by: Pavirani, Fabio, et al.
Published: (2024)
Optimal Transport Perturbations for Safe Reinforcement Learning with Robustness Guarantees
by: Queeney, James, et al.
Published: (2023)
by: Queeney, James, et al.
Published: (2023)
Similar Items
-
Lipschitz Lifelong Monte Carlo Tree Search for Mastering Non-Stationary Tasks
by: Zhang, Zuyuan, et al.
Published: (2025) -
Optimistic Policy Learning under Pessimistic Adversaries with Regret and Violation Guarantees
by: Ganguly, Sourav, et al.
Published: (2026) -
Geometry of Drifting MDPs with Path-Integral Stability Certificates
by: Zhang, Zuyuan, et al.
Published: (2026) -
Cochain Perspectives on Temporal-Difference Signals for Learning Beyond Markov Dynamics
by: Zhang, Zuyuan, et al.
Published: (2026) -
C-MCTS: Safe Planning with Monte Carlo Tree Search
by: Parthasarathy, Dinesh, et al.
Published: (2023)