Real-Time Parallel Counterfactual Regret Minimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Boning, Huang, Longbo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deep (Predictive) Discounted Counterfactual Regret Minimization
von: Xu, Hang, et al.
Veröffentlicht: (2025)
von: Xu, Hang, et al.
Veröffentlicht: (2025)
Parallelizing Counterfactual Regret Minimization
von: Kim, Juho, et al.
Veröffentlicht: (2026)
von: Kim, Juho, et al.
Veröffentlicht: (2026)
Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent
von: Xu, Hang, et al.
Veröffentlicht: (2024)
von: Xu, Hang, et al.
Veröffentlicht: (2024)
Solving Pasur Using GPU-Accelerated Counterfactual Regret Minimization
von: Baghal, Sina
Veröffentlicht: (2025)
von: Baghal, Sina
Veröffentlicht: (2025)
PokerSkill: LLMs Can Play Expert-Level Poker without Training or Solvers
von: Li, Boning, et al.
Veröffentlicht: (2026)
von: Li, Boning, et al.
Veröffentlicht: (2026)
Robust Deep Monte Carlo Counterfactual Regret Minimization: Addressing Theoretical Risks in Neural Fictitious Self-Play
von: Jaafari, Zakaria El
Veröffentlicht: (2025)
von: Jaafari, Zakaria El
Veröffentlicht: (2025)
RL-CFR: Improving Action Abstraction for Imperfect Information Extensive-Form Games with Reinforcement Learning
von: Li, Boning, et al.
Veröffentlicht: (2024)
von: Li, Boning, et al.
Veröffentlicht: (2024)
Next-Token Prediction and Regret Minimization
von: Mohri, Mehryar, et al.
Veröffentlicht: (2026)
von: Mohri, Mehryar, et al.
Veröffentlicht: (2026)
GPU-Accelerated Counterfactual Regret Minimization
von: Kim, Juho
Veröffentlicht: (2024)
von: Kim, Juho
Veröffentlicht: (2024)
Regret Minimization and Convergence to Equilibria in General-sum Markov Games
von: Erez, Liad, et al.
Veröffentlicht: (2022)
von: Erez, Liad, et al.
Veröffentlicht: (2022)
Effective, Efficient, and General Information Abstraction for Imperfect-Information Extensive-Form Games
von: Li, Boning, et al.
Veröffentlicht: (2026)
von: Li, Boning, et al.
Veröffentlicht: (2026)
Optimistic Thompson Sampling for No-Regret Learning in Unknown Games
von: Li, Yingru, et al.
Veröffentlicht: (2024)
von: Li, Yingru, et al.
Veröffentlicht: (2024)
Learning in Markov Games with Adaptive Adversaries: Policy Regret, Fundamental Barriers, and Efficient Algorithms
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
Do LLM Agents Have Regret? A Case Study in Online Learning and Games
von: Park, Chanwoo, et al.
Veröffentlicht: (2024)
von: Park, Chanwoo, et al.
Veröffentlicht: (2024)
From External to Swap Regret 2.0: An Efficient Reduction and Oblivious Adversary for Large Action Spaces
von: Dagan, Yuval, et al.
Veröffentlicht: (2023)
von: Dagan, Yuval, et al.
Veröffentlicht: (2023)
Easy as ABCs: Unifying Boltzmann Q-Learning and Counterfactual Regret Minimization
von: D'Amico-Wong, Luca, et al.
Veröffentlicht: (2024)
von: D'Amico-Wong, Luca, et al.
Veröffentlicht: (2024)
Accelerating Nash Equilibrium Convergence in Monte Carlo Settings Through Counterfactual Value Based Fictitious Play
von: Qi, Ju, et al.
Veröffentlicht: (2023)
von: Qi, Ju, et al.
Veröffentlicht: (2023)
Regret Minimization in Bilateral Trade With Perturbed Markets
von: Lunghi, Anna, et al.
Veröffentlicht: (2026)
von: Lunghi, Anna, et al.
Veröffentlicht: (2026)
Regret Minimization in Stackelberg Games with Side Information
von: Harris, Keegan, et al.
Veröffentlicht: (2024)
von: Harris, Keegan, et al.
Veröffentlicht: (2024)
Meta-Learning in Self-Play Regret Minimization
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
Selling Joint Ads: A Regret Minimization Perspective
von: Aggarwal, Gagan, et al.
Veröffentlicht: (2024)
von: Aggarwal, Gagan, et al.
Veröffentlicht: (2024)
Regret Minimization for Piecewise Linear Rewards: Contracts, Auctions, and Beyond
von: Bacchiocchi, Francesco, et al.
Veröffentlicht: (2025)
von: Bacchiocchi, Francesco, et al.
Veröffentlicht: (2025)
Computational Lower Bounds for Regret Minimization in Normal-Form Games
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2024)
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2024)
Regret Bounds for Competitive Resource Allocation with Endogenous Costs
von: Chai, Rui
Veröffentlicht: (2026)
von: Chai, Rui
Veröffentlicht: (2026)
Comparing Uniform Price and Discriminatory Multi-Unit Auctions through Regret Minimization
von: Potfer, Marius, et al.
Veröffentlicht: (2025)
von: Potfer, Marius, et al.
Veröffentlicht: (2025)
Efficient Last-Iterate Convergence in Regret Minimization via Adaptive Reward Transformation
von: Ren, Hang, et al.
Veröffentlicht: (2025)
von: Ren, Hang, et al.
Veröffentlicht: (2025)
GroupSegment-SHAP: Shapley Value Explanations with Group-Segment Players for Multivariate Time Series
von: Kim, Jinwoong, et al.
Veröffentlicht: (2026)
von: Kim, Jinwoong, et al.
Veröffentlicht: (2026)
Model-Based RL for Mean-Field Games is not Statistically Harder than Single-Agent RL
von: Huang, Jiawei, et al.
Veröffentlicht: (2024)
von: Huang, Jiawei, et al.
Veröffentlicht: (2024)
Governing AI Forgetting: Auditing for Machine Unlearning Compliance
von: Lin, Qinqi, et al.
Veröffentlicht: (2026)
von: Lin, Qinqi, et al.
Veröffentlicht: (2026)
ReLExS: Reinforcement Learning Explanations for Stackelberg No-Regret Learners
von: Huang, Xiangge, et al.
Veröffentlicht: (2024)
von: Huang, Xiangge, et al.
Veröffentlicht: (2024)
On the Limitations and Possibilities of Nash Regret Minimization in Zero-Sum Matrix Games under Noisy Feedback
von: Maiti, Arnab, et al.
Veröffentlicht: (2023)
von: Maiti, Arnab, et al.
Veröffentlicht: (2023)
Puzzle it Out: Local-to-Global World Model for Offline Multi-Agent Reinforcement Learning
von: Li, Sijia, et al.
Veröffentlicht: (2026)
von: Li, Sijia, et al.
Veröffentlicht: (2026)
Incentivizing Truthful Language Models via Peer Elicitation Games
von: Chen, Baiting, et al.
Veröffentlicht: (2025)
von: Chen, Baiting, et al.
Veröffentlicht: (2025)
Agent-oriented Joint Decision Support for Data Owners in Auction-based Federated Learning
von: Tang, Xiaoli, et al.
Veröffentlicht: (2024)
von: Tang, Xiaoli, et al.
Veröffentlicht: (2024)
Learning not to Regret
von: Sychrovský, David, et al.
Veröffentlicht: (2023)
von: Sychrovský, David, et al.
Veröffentlicht: (2023)
An Efficient Black-Box Reduction from Online Learning to Multicalibration, and a New Route to $Φ$-Regret Minimization
von: Farina, Gabriele, et al.
Veröffentlicht: (2026)
von: Farina, Gabriele, et al.
Veröffentlicht: (2026)
Strategic Candidacy in Generative AI Arenas
von: Hays, Chris, et al.
Veröffentlicht: (2026)
von: Hays, Chris, et al.
Veröffentlicht: (2026)
Code-Space Response Oracles: Generating Interpretable Multi-Agent Policies with Large Language Models
von: Hennes, Daniel, et al.
Veröffentlicht: (2026)
von: Hennes, Daniel, et al.
Veröffentlicht: (2026)
Test-Time Compute Games
von: Velasco, Ander Artola, et al.
Veröffentlicht: (2026)
von: Velasco, Ander Artola, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Deep (Predictive) Discounted Counterfactual Regret Minimization
von: Xu, Hang, et al.
Veröffentlicht: (2025) -
Parallelizing Counterfactual Regret Minimization
von: Kim, Juho, et al.
Veröffentlicht: (2026) -
Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent
von: Xu, Hang, et al.
Veröffentlicht: (2024) -
Solving Pasur Using GPU-Accelerated Counterfactual Regret Minimization
von: Baghal, Sina
Veröffentlicht: (2025) -
PokerSkill: LLMs Can Play Expert-Level Poker without Training or Solvers
von: Li, Boning, et al.
Veröffentlicht: (2026)