Robust Deep Monte Carlo Counterfactual Regret Minimization: Addressing Theoretical Risks in Neural Fictitious Self-Play
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Jaafari, Zakaria El |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Accelerating Nash Equilibrium Convergence in Monte Carlo Settings Through Counterfactual Value Based Fictitious Play
von: Qi, Ju, et al.
Veröffentlicht: (2023)
von: Qi, Ju, et al.
Veröffentlicht: (2023)
Deep (Predictive) Discounted Counterfactual Regret Minimization
von: Xu, Hang, et al.
Veröffentlicht: (2025)
von: Xu, Hang, et al.
Veröffentlicht: (2025)
Real-Time Parallel Counterfactual Regret Minimization
von: Li, Boning, et al.
Veröffentlicht: (2026)
von: Li, Boning, et al.
Veröffentlicht: (2026)
Offline Fictitious Self-Play for Competitive Games
von: Chen, Jingxiao, et al.
Veröffentlicht: (2024)
von: Chen, Jingxiao, et al.
Veröffentlicht: (2024)
Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent
von: Xu, Hang, et al.
Veröffentlicht: (2024)
von: Xu, Hang, et al.
Veröffentlicht: (2024)
Solving Pasur Using GPU-Accelerated Counterfactual Regret Minimization
von: Baghal, Sina
Veröffentlicht: (2025)
von: Baghal, Sina
Veröffentlicht: (2025)
Parallelizing Counterfactual Regret Minimization
von: Kim, Juho, et al.
Veröffentlicht: (2026)
von: Kim, Juho, et al.
Veröffentlicht: (2026)
Meta-Learning in Self-Play Regret Minimization
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
Next-Token Prediction and Regret Minimization
von: Mohri, Mehryar, et al.
Veröffentlicht: (2026)
von: Mohri, Mehryar, et al.
Veröffentlicht: (2026)
GPU-Accelerated Counterfactual Regret Minimization
von: Kim, Juho
Veröffentlicht: (2024)
von: Kim, Juho
Veröffentlicht: (2024)
Regret Minimization and Convergence to Equilibria in General-sum Markov Games
von: Erez, Liad, et al.
Veröffentlicht: (2022)
von: Erez, Liad, et al.
Veröffentlicht: (2022)
The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play
von: La Malfa, Gabriele, et al.
Veröffentlicht: (2026)
von: La Malfa, Gabriele, et al.
Veröffentlicht: (2026)
Optimistic Thompson Sampling for No-Regret Learning in Unknown Games
von: Li, Yingru, et al.
Veröffentlicht: (2024)
von: Li, Yingru, et al.
Veröffentlicht: (2024)
Fast and Furious Symmetric Learning in Zero-Sum Games: Gradient Descent as Fictitious Play
von: Lazarsfeld, John, et al.
Veröffentlicht: (2025)
von: Lazarsfeld, John, et al.
Veröffentlicht: (2025)
Learning in Markov Games with Adaptive Adversaries: Policy Regret, Fundamental Barriers, and Efficient Algorithms
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
Do LLM Agents Have Regret? A Case Study in Online Learning and Games
von: Park, Chanwoo, et al.
Veröffentlicht: (2024)
von: Park, Chanwoo, et al.
Veröffentlicht: (2024)
From External to Swap Regret 2.0: An Efficient Reduction and Oblivious Adversary for Large Action Spaces
von: Dagan, Yuval, et al.
Veröffentlicht: (2023)
von: Dagan, Yuval, et al.
Veröffentlicht: (2023)
Easy as ABCs: Unifying Boltzmann Q-Learning and Counterfactual Regret Minimization
von: D'Amico-Wong, Luca, et al.
Veröffentlicht: (2024)
von: D'Amico-Wong, Luca, et al.
Veröffentlicht: (2024)
SpinGPT: A Large-Language-Model Approach to Playing Poker Correctly
von: Maugin, Narada, et al.
Veröffentlicht: (2025)
von: Maugin, Narada, et al.
Veröffentlicht: (2025)
Regret Minimization in Stackelberg Games with Side Information
von: Harris, Keegan, et al.
Veröffentlicht: (2024)
von: Harris, Keegan, et al.
Veröffentlicht: (2024)
Regret Minimization in Bilateral Trade With Perturbed Markets
von: Lunghi, Anna, et al.
Veröffentlicht: (2026)
von: Lunghi, Anna, et al.
Veröffentlicht: (2026)
Data-Augmented Game Starts for Accelerating Self-Play Exploration in Imperfect Information Games
von: Lanier, JB, et al.
Veröffentlicht: (2026)
von: Lanier, JB, et al.
Veröffentlicht: (2026)
Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
Selling Joint Ads: A Regret Minimization Perspective
von: Aggarwal, Gagan, et al.
Veröffentlicht: (2024)
von: Aggarwal, Gagan, et al.
Veröffentlicht: (2024)
Do LLMs Strategically Reveal, Conceal, and Infer Information? A Theoretical and Empirical Analysis in The Chameleon Game
von: Karabag, Mustafa O., et al.
Veröffentlicht: (2025)
von: Karabag, Mustafa O., et al.
Veröffentlicht: (2025)
Clone-Robust AI Alignment
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2025)
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2025)
Empirical Analysis of Fictitious Play for Nash Equilibrium Computation in Multiplayer Games
von: Ganzfried, Sam
Veröffentlicht: (2020)
von: Ganzfried, Sam
Veröffentlicht: (2020)
Regret Minimization for Piecewise Linear Rewards: Contracts, Auctions, and Beyond
von: Bacchiocchi, Francesco, et al.
Veröffentlicht: (2025)
von: Bacchiocchi, Francesco, et al.
Veröffentlicht: (2025)
Computational Lower Bounds for Regret Minimization in Normal-Form Games
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2024)
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning for Sequential Combinatorial Auctions
von: Ravindranath, Sai Srivatsa, et al.
Veröffentlicht: (2024)
von: Ravindranath, Sai Srivatsa, et al.
Veröffentlicht: (2024)
Decision Theory-Guided Deep Reinforcement Learning for Fast Learning
von: Wan, Zelin, et al.
Veröffentlicht: (2024)
von: Wan, Zelin, et al.
Veröffentlicht: (2024)
Aggregate Fictitious Play for Learning in Anonymous Polymatrix Games (Extended Version)
von: Kara, Semih, et al.
Veröffentlicht: (2025)
von: Kara, Semih, et al.
Veröffentlicht: (2025)
Efficiently Training Neural Networks for Imperfect Information Games by Sampling Information Sets
von: Bertram, Timo, et al.
Veröffentlicht: (2024)
von: Bertram, Timo, et al.
Veröffentlicht: (2024)
Self-optimization in distributed manufacturing systems using Modular State-based Stackelberg Games
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
Gradient-based Learning in State-based Potential Games for Self-Learning Production Systems
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
Regret Bounds for Competitive Resource Allocation with Endogenous Costs
von: Chai, Rui
Veröffentlicht: (2026)
von: Chai, Rui
Veröffentlicht: (2026)
GemNet: Menu-Based, Strategy-Proof Multi-Bidder Auctions Through Deep Learning
von: Wang, Tonghan, et al.
Veröffentlicht: (2024)
von: Wang, Tonghan, et al.
Veröffentlicht: (2024)
Comparing Uniform Price and Discriminatory Multi-Unit Auctions through Regret Minimization
von: Potfer, Marius, et al.
Veröffentlicht: (2025)
von: Potfer, Marius, et al.
Veröffentlicht: (2025)
Efficient Last-Iterate Convergence in Regret Minimization via Adaptive Reward Transformation
von: Ren, Hang, et al.
Veröffentlicht: (2025)
von: Ren, Hang, et al.
Veröffentlicht: (2025)
ElementaryNet: A Non-Strategic Neural Network for Predicting Human Behavior in Normal-Form Games
von: d'Eon, Greg, et al.
Veröffentlicht: (2025)
von: d'Eon, Greg, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Accelerating Nash Equilibrium Convergence in Monte Carlo Settings Through Counterfactual Value Based Fictitious Play
von: Qi, Ju, et al.
Veröffentlicht: (2023) -
Deep (Predictive) Discounted Counterfactual Regret Minimization
von: Xu, Hang, et al.
Veröffentlicht: (2025) -
Real-Time Parallel Counterfactual Regret Minimization
von: Li, Boning, et al.
Veröffentlicht: (2026) -
Offline Fictitious Self-Play for Competitive Games
von: Chen, Jingxiao, et al.
Veröffentlicht: (2024) -
Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent
von: Xu, Hang, et al.
Veröffentlicht: (2024)