Lookahead Pathology in Monte-Carlo Tree Search
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, Khoi P. N., Ramanujan, Raghuram |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improve Value Estimation of Q Function and Reshape Reward with Monte Carlo Tree Search
by: Li, Jiamian
Published: (2024)
by: Li, Jiamian
Published: (2024)
Tree Search for Simultaneous Move Games via Equilibrium Approximation
by: Yu, Ryan, et al.
Published: (2024)
by: Yu, Ryan, et al.
Published: (2024)
SMCEvolve: Principled Scientific Discovery via Sequential Monte Carlo Evolution
by: Jiang, Jiachen, et al.
Published: (2026)
by: Jiang, Jiachen, et al.
Published: (2026)
Conservative Equilibrium Discovery in Offline Game-Theoretic Multiagent Reinforcement Learning
by: Nguyen, Austin A., et al.
Published: (2026)
by: Nguyen, Austin A., et al.
Published: (2026)
United We Stand: Decentralized Multi-Agent Planning With Attrition
by: Nguyen, Nhat, et al.
Published: (2024)
by: Nguyen, Nhat, et al.
Published: (2024)
Boltzmann-based Exploration for Robust Decentralized Multi-Agent Planning (Extended Version)
by: Nguyen, Nhat D. A., et al.
Published: (2026)
by: Nguyen, Nhat D. A., et al.
Published: (2026)
GRASP: Graph Agentic Search over Propositions for Multi-hop Question Answering
by: Jenkins, Stockton, et al.
Published: (2026)
by: Jenkins, Stockton, et al.
Published: (2026)
Healthcare Mechanisms from Policy-as-Code Search under Strategic Provider Response
by: Wang, Zihan, et al.
Published: (2026)
by: Wang, Zihan, et al.
Published: (2026)
Optimizing Supply Chain Management using Gravitational Search Algorithm and Multi Agent System
by: Ojha, Muneendra
Published: (2012)
by: Ojha, Muneendra
Published: (2012)
Adaptive Anytime Multi-Agent Path Finding Using Bandit-Based Large Neighborhood Search
by: Phan, Thomy, et al.
Published: (2023)
by: Phan, Thomy, et al.
Published: (2023)
Contextual Knowledge Sharing in Multi-Agent Reinforcement Learning with Decentralized Communication and Coordination
by: Du, Hung, et al.
Published: (2025)
by: Du, Hung, et al.
Published: (2025)
Improving Learnt Local MAPF Policies with Heuristic Search
by: Veerapaneni, Rishi, et al.
Published: (2024)
by: Veerapaneni, Rishi, et al.
Published: (2024)
Improving Cooperation in Collaborative Embodied AI
by: Suprabha, Hima Jacob Leven, et al.
Published: (2025)
by: Suprabha, Hima Jacob Leven, et al.
Published: (2025)
Position: Multi-Agent Algorithmic Care Systems Demand Contestability for Trustworthy AI
by: Nguyen, Truong Thanh Hung, et al.
Published: (2026)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2026)
Accelerating Focal Search in Multi-Agent Path Finding with Tighter Lower Bounds
by: Tang, Yimin, et al.
Published: (2025)
by: Tang, Yimin, et al.
Published: (2025)
Large-Scale Multi-Robot Coverage Path Planning via Local Search
by: Tang, Jingtao, et al.
Published: (2023)
by: Tang, Jingtao, et al.
Published: (2023)
AgentSearchBench: A Benchmark for AI Agent Search in the Wild
by: Wu, Bin, et al.
Published: (2026)
by: Wu, Bin, et al.
Published: (2026)
The High Cost of Incivility: Quantifying Interaction Inefficiency via Multi-Agent Monte Carlo Simulations
by: Mangold, Benedikt
Published: (2025)
by: Mangold, Benedikt
Published: (2025)
Multi-Agent Motion Planning For Differential Drive Robots Through Stationary State Search
by: Yan, Jingtian, et al.
Published: (2024)
by: Yan, Jingtian, et al.
Published: (2024)
COMRES-VLM: Coordinated Multi-Robot Exploration and Search using Vision Language Models
by: Wang, Ruiyang, et al.
Published: (2025)
by: Wang, Ruiyang, et al.
Published: (2025)
Accelerating Search-Based Planning for Multi-Robot Manipulation by Leveraging Online-Generated Experiences
by: Shaoul, Yorai, et al.
Published: (2024)
by: Shaoul, Yorai, et al.
Published: (2024)
Learning to Recommend Multi-Agent Subgraphs from Calling Trees
by: Song, Xinyuan, et al.
Published: (2026)
by: Song, Xinyuan, et al.
Published: (2026)
I Can't Believe It's Corrupt: Evaluating Corruption in Multi-Agent Governance Systems
by: P, Vedanta S, et al.
Published: (2026)
by: P, Vedanta S, et al.
Published: (2026)
The Traitors: Deception and Trust in Multi-Agent Language Model Simulations
by: Curvo, Pedro M. P.
Published: (2025)
by: Curvo, Pedro M. P.
Published: (2025)
Grammar Search for Multi-Agent Systems
by: Singh, Mayank, et al.
Published: (2025)
by: Singh, Mayank, et al.
Published: (2025)
Cooperative Search and Track of Rogue Drones using Multiagent Reinforcement Learning
by: Valianti, Panayiota, et al.
Published: (2025)
by: Valianti, Panayiota, et al.
Published: (2025)
Efficient Mixture-of-Agents Serving via Tree-Structured Routing, Adaptive Pruning, and Dependency-Aware Prefill-Decode Overlap
by: Wang, Zijun, et al.
Published: (2025)
by: Wang, Zijun, et al.
Published: (2025)
Optimization Instability in Autonomous Agentic Workflows for Clinical Symptom Detection
by: Cagan, Cameron, et al.
Published: (2026)
by: Cagan, Cameron, et al.
Published: (2026)
Diverse Planning with Simulators via Linear Temporal Logic
by: Abdelwahed, Mustafa F., et al.
Published: (2025)
by: Abdelwahed, Mustafa F., et al.
Published: (2025)
PMAx: An Agentic Framework for AI-Driven Process Mining
by: Antonov, Anton, et al.
Published: (2026)
by: Antonov, Anton, et al.
Published: (2026)
Multicopy Reinforcement Learning Agents
by: Wolfe, Alicia P., et al.
Published: (2023)
by: Wolfe, Alicia P., et al.
Published: (2023)
SPIRAL: Symbolic LLM Planning via Grounded and Reflective Search
by: Zhang, Yifan, et al.
Published: (2025)
by: Zhang, Yifan, et al.
Published: (2025)
Discovering Sensorimotor Agency in Cellular Automata using Diversity Search
by: Hamon, Gautier, et al.
Published: (2024)
by: Hamon, Gautier, et al.
Published: (2024)
PlayBest: Professional Basketball Player Behavior Synthesis via Planning with Diffusion
by: Chen, Xiusi, et al.
Published: (2023)
by: Chen, Xiusi, et al.
Published: (2023)
An Agentic AI Workflow for Detecting Cognitive Concerns in Real-world Data
by: Tian, Jiazi, et al.
Published: (2025)
by: Tian, Jiazi, et al.
Published: (2025)
Adaptive Collaboration of Arena-Based Argumentative LLMs for Explainable and Contestable Legal Reasoning
by: Cao, Hoang-Loc, et al.
Published: (2026)
by: Cao, Hoang-Loc, et al.
Published: (2026)
Towards Natural Language Communication for Cooperative Autonomous Driving via Self-Play
by: Cui, Jiaxun, et al.
Published: (2025)
by: Cui, Jiaxun, et al.
Published: (2025)
MRBTP: Efficient Multi-Robot Behavior Tree Planning and Collaboration
by: Cai, Yishuai, et al.
Published: (2025)
by: Cai, Yishuai, et al.
Published: (2025)
HSCodeComp: A Realistic and Expert-level Benchmark for Deep Search Agents in Hierarchical Rule Application
by: Yang, Yiqian, et al.
Published: (2025)
by: Yang, Yiqian, et al.
Published: (2025)
MisoDICE: Multi-Agent Imitation from Unlabeled Mixed-Quality Demonstrations
by: Bui, The Viet, et al.
Published: (2025)
by: Bui, The Viet, et al.
Published: (2025)
Similar Items
-
Improve Value Estimation of Q Function and Reshape Reward with Monte Carlo Tree Search
by: Li, Jiamian
Published: (2024) -
Tree Search for Simultaneous Move Games via Equilibrium Approximation
by: Yu, Ryan, et al.
Published: (2024) -
SMCEvolve: Principled Scientific Discovery via Sequential Monte Carlo Evolution
by: Jiang, Jiachen, et al.
Published: (2026) -
Conservative Equilibrium Discovery in Offline Game-Theoretic Multiagent Reinforcement Learning
by: Nguyen, Austin A., et al.
Published: (2026) -
United We Stand: Decentralized Multi-Agent Planning With Attrition
by: Nguyen, Nhat, et al.
Published: (2024)