Saved in:
| Main Authors: | Zhang, Junyu, Yang, Feihong, Wang, Jian, Wang, Chao, Zhang, Xudong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.28273 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
COvolve: Adversarial Co-Evolution of Large-Language-Model-Generated Policies and Environments via Two-Player Zero-Sum Game
by: Sygkounas, Alkis, et al.
Published: (2026)
by: Sygkounas, Alkis, et al.
Published: (2026)
Policy Space Response Oracles: A Survey
by: Bighashdel, Ariyan, et al.
Published: (2024)
by: Bighashdel, Ariyan, et al.
Published: (2024)
Efficient Reinforcement Learning for Zero-Shot Coordination in Evolving Games
by: Hui, Bingyu, et al.
Published: (2025)
by: Hui, Bingyu, et al.
Published: (2025)
Computing Ex Ante Equilibrium in Heterogeneous Zero-Sum Team Games
by: Liu, Naming, et al.
Published: (2024)
by: Liu, Naming, et al.
Published: (2024)
Simulation-Free PSRO: Removing Game Simulation from Policy Space Response Oracles
by: Liu, Yingzhuo, et al.
Published: (2025)
by: Liu, Yingzhuo, et al.
Published: (2025)
Toward Optimal LLM Alignments Using Two-Player Games
by: Zheng, Rui, et al.
Published: (2024)
by: Zheng, Rui, et al.
Published: (2024)
Fusion-PSRO: Nash Policy Fusion for Policy Space Response Oracles
by: Lian, Jiesong, et al.
Published: (2024)
by: Lian, Jiesong, et al.
Published: (2024)
Are Large Language Models Strategic Decision Makers? A Study of Performance and Bias in Two-Player Non-Zero-Sum Games
by: Herr, Nathan, et al.
Published: (2024)
by: Herr, Nathan, et al.
Published: (2024)
Sample-Efficient Policy Space Response Oracles with Joint Experience Best Response
by: Bighashdel, Ariyan, et al.
Published: (2026)
by: Bighashdel, Ariyan, et al.
Published: (2026)
Revisiting Regularized Policy Optimization for Stable and Efficient Reinforcement Learning in Two-Player Games
by: Ota, Kazuki, et al.
Published: (2026)
by: Ota, Kazuki, et al.
Published: (2026)
Perturbing Best Responses in Zero-Sum Games
by: Dziwoki, Adam, et al.
Published: (2025)
by: Dziwoki, Adam, et al.
Published: (2025)
Learning to Play Two-Player Perfect-Information Games without Knowledge
by: Cohen-Solal, Quentin
Published: (2020)
by: Cohen-Solal, Quentin
Published: (2020)
Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization
by: Xu, Zelai, et al.
Published: (2025)
by: Xu, Zelai, et al.
Published: (2025)
Improving Sample Efficiency of Model-Free Algorithms for Zero-Sum Markov Games
by: Feng, Songtao, et al.
Published: (2023)
by: Feng, Songtao, et al.
Published: (2023)
Two-Player Zero-Sum Hybrid Games
by: Leudo, Santiago J., et al.
Published: (2024)
by: Leudo, Santiago J., et al.
Published: (2024)
Towards Real-world Deployment of NILM Systems: Challenges and Practices
by: Xue, Junyu, et al.
Published: (2024)
by: Xue, Junyu, et al.
Published: (2024)
Enhancing Two-Player Performance Through Single-Player Knowledge Transfer: An Empirical Study on Atari 2600 Games
by: Saadat, Kimiya, et al.
Published: (2024)
by: Saadat, Kimiya, et al.
Published: (2024)
Learning Global Hypothesis Space for Enhancing Synergistic Reasoning Chain
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Observer, Not Player: Simulating Theory of Mind in LLMs through Game Observation
by: Wang, Jerry, et al.
Published: (2025)
by: Wang, Jerry, et al.
Published: (2025)
Enhancing Player Enjoyment with a Two-Tier DRL and LLM-Based Agent System for Fighting Games
by: Wang, Shouren, et al.
Published: (2025)
by: Wang, Shouren, et al.
Published: (2025)
FM3Q: Factorized Multi-Agent MiniMax Q-Learning for Two-Team Zero-Sum Markov Game
by: Hu, Guangzheng, et al.
Published: (2024)
by: Hu, Guangzheng, et al.
Published: (2024)
Go-Oracle: Automated Test Oracle for Go Concurrency Bugs
by: Tsimpourlas, Foivos, et al.
Published: (2024)
by: Tsimpourlas, Foivos, et al.
Published: (2024)
Policy Newton Algorithm in Reproducing Kernel Hilbert Space
by: Zhang, Yixian, et al.
Published: (2025)
by: Zhang, Yixian, et al.
Published: (2025)
Code-Space Response Oracles: Generating Interpretable Multi-Agent Policies with Large Language Models
by: Hennes, Daniel, et al.
Published: (2026)
by: Hennes, Daniel, et al.
Published: (2026)
How Far Are LLMs from Professional Poker Players? Revisiting Game-Theoretic Reasoning with Agentic Tool Use
by: Lin, Minhua, et al.
Published: (2026)
by: Lin, Minhua, et al.
Published: (2026)
Two-Player Zero-Sum Games with Bandit Feedback
by: Yılmaz, Elif, et al.
Published: (2025)
by: Yılmaz, Elif, et al.
Published: (2025)
Leveraging MLLM Embeddings and Attribute Smoothing for Compositional Zero-Shot Learning
by: Yan, Xudong, et al.
Published: (2024)
by: Yan, Xudong, et al.
Published: (2024)
Bilevel Optimization over Saddle Points of Zero-Sum Markov Games
by: Zheng, Zihao, et al.
Published: (2026)
by: Zheng, Zihao, et al.
Published: (2026)
Are Large Vision Language Models Good Game Players?
by: Wang, Xinyu, et al.
Published: (2025)
by: Wang, Xinyu, et al.
Published: (2025)
ZeroS: Zero-Sum Linear Attention for Efficient Transformers
by: Lu, Jiecheng, et al.
Published: (2026)
by: Lu, Jiecheng, et al.
Published: (2026)
Offline Two-Player Zero-Sum Markov Games with KL Regularization
by: Chen, Claire, et al.
Published: (2026)
by: Chen, Claire, et al.
Published: (2026)
Active Policy Improvement from Multiple Black-box Oracles
by: Liu, Xuefeng, et al.
Published: (2023)
by: Liu, Xuefeng, et al.
Published: (2023)
Deep Reinforcement Learning Xiangqi Player with Monte Carlo Tree Search
by: Yilmaz, Berk, et al.
Published: (2025)
by: Yilmaz, Berk, et al.
Published: (2025)
Playing Large Games with Oracles and AI Debate
by: Chen, Xinyi, et al.
Published: (2023)
by: Chen, Xinyi, et al.
Published: (2023)
An Online Learning Approach for Two-Player Zero-Sum Linear Quadratic Games
by: Wang, Shanting, et al.
Published: (2026)
by: Wang, Shanting, et al.
Published: (2026)
Foundations of Global Consistency Checking with Noisy LLM Oracles
by: He, Paul, et al.
Published: (2026)
by: He, Paul, et al.
Published: (2026)
Player-Driven Emergence in LLM-Driven Game Narrative
by: Peng, Xiangyu, et al.
Published: (2024)
by: Peng, Xiangyu, et al.
Published: (2024)
Strategy Synthesis for Zero-Sum Neuro-Symbolic Concurrent Stochastic Games
by: Yan, Rui, et al.
Published: (2022)
by: Yan, Rui, et al.
Published: (2022)
Prompting Fairness: Artificial Intelligence as Game Players
by: Henry, Jazmia
Published: (2024)
by: Henry, Jazmia
Published: (2024)
ELA: Exploited Level Augmentation for Offline Learning in Zero-Sum Games
by: Lei, Shiqi, et al.
Published: (2024)
by: Lei, Shiqi, et al.
Published: (2024)
Similar Items
-
COvolve: Adversarial Co-Evolution of Large-Language-Model-Generated Policies and Environments via Two-Player Zero-Sum Game
by: Sygkounas, Alkis, et al.
Published: (2026) -
Policy Space Response Oracles: A Survey
by: Bighashdel, Ariyan, et al.
Published: (2024) -
Efficient Reinforcement Learning for Zero-Shot Coordination in Evolving Games
by: Hui, Bingyu, et al.
Published: (2025) -
Computing Ex Ante Equilibrium in Heterogeneous Zero-Sum Team Games
by: Liu, Naming, et al.
Published: (2024) -
Simulation-Free PSRO: Removing Game Simulation from Policy Space Response Oracles
by: Liu, Yingzhuo, et al.
Published: (2025)