NonZero: Interaction-Guided Exploration for Multi-Agent Monte Carlo Tree Search
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Sizhe, Zhang, Zuyuan, Imani, Mahdi, Lan, Tian |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Metric-Gradient Projection for Stable Multi-Agent Policy Learning
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
Geometry of Drifting MDPs with Path-Integral Stability Certificates
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
Tail-Risk-Safe Monte Carlo Tree Search under PAC-Level Guarantees
by: Zhang, Zuyuan, et al.
Published: (2025)
by: Zhang, Zuyuan, et al.
Published: (2025)
Cochain Perspectives on Temporal-Difference Signals for Learning Beyond Markov Dynamics
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
Manifold-Constrained Energy-Based Transition Models for Offline Reinforcement Learning
by: Fang, Zeyu, et al.
Published: (2026)
by: Fang, Zeyu, et al.
Published: (2026)
Interactive Critique-Revision Training for Reliable Structured LLM Generation
by: Yu, Fei Xu, et al.
Published: (2026)
by: Yu, Fei Xu, et al.
Published: (2026)
Lipschitz Lifelong Monte Carlo Tree Search for Mastering Non-Stationary Tasks
by: Zhang, Zuyuan, et al.
Published: (2025)
by: Zhang, Zuyuan, et al.
Published: (2025)
Reason in Chains, Learn in Trees: Self-Rectification and Grafting for Multi-turn Agent Policy Optimization
by: Li, Yu, et al.
Published: (2026)
by: Li, Yu, et al.
Published: (2026)
ACDZero: MCTS Agent for Mastering Automated Cyber Defense
by: Li, Yu, et al.
Published: (2026)
by: Li, Yu, et al.
Published: (2026)
Operator-Guided Invariance Learning for Continuous Reinforcement Learning
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
Modeling Other Players with Bayesian Beliefs for Games with Incomplete Information
by: Zhang, Zuyuan, et al.
Published: (2024)
by: Zhang, Zuyuan, et al.
Published: (2024)
Monte Carlo Tree Search with Boltzmann Exploration
by: Painter, Michael, et al.
Published: (2024)
by: Painter, Michael, et al.
Published: (2024)
Bayesian Optimization through Gaussian Cox Process Models for Spatio-temporal Data
by: Mei, Yongsheng, et al.
Published: (2024)
by: Mei, Yongsheng, et al.
Published: (2024)
Optimizing Prompt Sequences using Monte Carlo Tree Search for LLM-Based Optimization
by: Yu, Fei Xu, et al.
Published: (2025)
by: Yu, Fei Xu, et al.
Published: (2025)
Agent Alpha: Tree Search Unifying Generation, Exploration and Evaluation for Computer-Use Agents
by: Tang, Sizhe, et al.
Published: (2026)
by: Tang, Sizhe, et al.
Published: (2026)
Matrix-Space Reinforcement Learning for Reusing Local Transition Geometry
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
A Contractive Feedback Semantics for Reinforcement Learning
by: Zhang, Zuyuan
Published: (2026)
by: Zhang, Zuyuan
Published: (2026)
MALinZero: Efficient Low-Dimensional Search for Mastering Complex Multi-Agent Planning
by: Tang, Sizhe, et al.
Published: (2025)
by: Tang, Sizhe, et al.
Published: (2025)
Global Optimization on Graph-Structured Data via Gaussian Processes with Spectral Representations
by: Hong, Shu, et al.
Published: (2025)
by: Hong, Shu, et al.
Published: (2025)
Epistemic Monte Carlo Tree Search
by: Oren, Yaniv, et al.
Published: (2022)
by: Oren, Yaniv, et al.
Published: (2022)
Provably Efficient Long-Horizon Exploration in Monte Carlo Tree Search through State Occupancy Regularization
by: Schramm, Liam, et al.
Published: (2024)
by: Schramm, Liam, et al.
Published: (2024)
Guided Exploration in Reinforcement Learning via Monte Carlo Critic Optimization
by: Kuznetsov, Igor
Published: (2022)
by: Kuznetsov, Igor
Published: (2022)
Provable and Practical: Efficient Exploration in Reinforcement Learning via Langevin Monte Carlo
by: Ishfaq, Haque, et al.
Published: (2023)
by: Ishfaq, Haque, et al.
Published: (2023)
Entropic Risk-Aware Monte Carlo Tree Search
by: Santos, Pedro P., et al.
Published: (2026)
by: Santos, Pedro P., et al.
Published: (2026)
An Efficient Algorithm for Thresholding Monte Carlo Tree Search
by: Nameki, Shoma, et al.
Published: (2026)
by: Nameki, Shoma, et al.
Published: (2026)
FedQHD: Closed-Form Function-Space Federated Reinforcement Learning
by: Hou, Yuchen, et al.
Published: (2026)
by: Hou, Yuchen, et al.
Published: (2026)
Improving Monte Carlo Tree Search for Symbolic Regression
by: Huang, Zhengyao, et al.
Published: (2025)
by: Huang, Zhengyao, et al.
Published: (2025)
Monte Carlo Tree Search in the Presence of Transition Uncertainty
by: Kohankhaki, Farnaz, et al.
Published: (2023)
by: Kohankhaki, Farnaz, et al.
Published: (2023)
Twice Sequential Monte Carlo for Tree Search
by: Oren, Yaniv, et al.
Published: (2025)
by: Oren, Yaniv, et al.
Published: (2025)
Doubly Robust Monte Carlo Tree Search
by: Liu, Manqing, et al.
Published: (2025)
by: Liu, Manqing, et al.
Published: (2025)
Preference Construction: A Bayesian Interactive Preference Elicitation Framework Based on Monte Carlo Tree Search
by: Wang, Yan, et al.
Published: (2025)
by: Wang, Yan, et al.
Published: (2025)
Decoupling Exploration and Policy Optimization: Uncertainty Guided Tree Search for Hard Exploration
by: Mhammedi, Zakaria, et al.
Published: (2026)
by: Mhammedi, Zakaria, et al.
Published: (2026)
Anytime Sequential Halving in Monte-Carlo Tree Search
by: Sagers, Dominic, et al.
Published: (2024)
by: Sagers, Dominic, et al.
Published: (2024)
Improving GFlowNets with Monte Carlo Tree Search
by: Morozov, Nikita, et al.
Published: (2024)
by: Morozov, Nikita, et al.
Published: (2024)
Monte Carlo Permutation Search
by: Cazenave, Tristan
Published: (2025)
by: Cazenave, Tristan
Published: (2025)
Enhancing Bayesian Network Structural Learning with Monte Carlo Tree Search
by: Laborda, Jorge D., et al.
Published: (2025)
by: Laborda, Jorge D., et al.
Published: (2025)
Tidiness Score-Guided Monte Carlo Tree Search for Visual Tabletop Rearrangement
by: Kee, Hogun, et al.
Published: (2025)
by: Kee, Hogun, et al.
Published: (2025)
Monte Carlo and quasi-Monte Carlo integration for likelihood functions
by: Tang, Yanbo
Published: (2025)
by: Tang, Yanbo
Published: (2025)
Collaborative AI Teaming in Unknown Environments via Active Goal Deduction
by: Zhang, Zuyuan, et al.
Published: (2024)
by: Zhang, Zuyuan, et al.
Published: (2024)
LiSFC-Search: Lifelong Search for Network SFC Optimization under Non-stationary Drifts
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
Similar Items
-
Metric-Gradient Projection for Stable Multi-Agent Policy Learning
by: Zhang, Zuyuan, et al.
Published: (2026) -
Geometry of Drifting MDPs with Path-Integral Stability Certificates
by: Zhang, Zuyuan, et al.
Published: (2026) -
Tail-Risk-Safe Monte Carlo Tree Search under PAC-Level Guarantees
by: Zhang, Zuyuan, et al.
Published: (2025) -
Cochain Perspectives on Temporal-Difference Signals for Learning Beyond Markov Dynamics
by: Zhang, Zuyuan, et al.
Published: (2026) -
Manifold-Constrained Energy-Based Transition Models for Offline Reinforcement Learning
by: Fang, Zeyu, et al.
Published: (2026)