Heuristic Pathologies and Further Variance Reduction via Uncertainty Propagation in the AIVAT Family of Techniques
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Juho, Sandholm, Tuomas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Domain-Independent Game Abstraction using Word Embedding Techniques
von: Kim, Juho, et al.
Veröffentlicht: (2026)
von: Kim, Juho, et al.
Veröffentlicht: (2026)
Parallelizing Counterfactual Regret Minimization
von: Kim, Juho, et al.
Veröffentlicht: (2026)
von: Kim, Juho, et al.
Veröffentlicht: (2026)
Watermarking Game-Playing Agents in Perfect-Information Extensive-Form Games
von: Kim, Juho, et al.
Veröffentlicht: (2026)
von: Kim, Juho, et al.
Veröffentlicht: (2026)
General search techniques without common knowledge for imperfect-information games, and application to superhuman Fog of War chess
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2025)
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2025)
ApproxED: Approximate exploitability descent via learned best responses
von: Martin, Carlos, et al.
Veröffentlicht: (2023)
von: Martin, Carlos, et al.
Veröffentlicht: (2023)
Optimal Correlated Equilibria in General-Sum Extensive-Form Games: Fixed-Parameter Algorithms, Hardness, and Two-Sided Column-Generation
von: Zhang, Brian, et al.
Veröffentlicht: (2022)
von: Zhang, Brian, et al.
Veröffentlicht: (2022)
Convergence of $\text{log}(1/ε)$ for Gradient-Based Algorithms in Zero-Sum Games without the Condition Number: A Smoothed Analysis
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2024)
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2024)
On the Interplay between Social Welfare and Tractability of Equilibria
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2023)
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2023)
Solving Infinite-Player Games with Player-to-Strategy Networks
von: Martin, Carlos, et al.
Veröffentlicht: (2025)
von: Martin, Carlos, et al.
Veröffentlicht: (2025)
Simultaneous incremental support adjustment and metagame solving: An equilibrium-finding framework for continuous-action games
von: Martin, Carlos, et al.
Veröffentlicht: (2024)
von: Martin, Carlos, et al.
Veröffentlicht: (2024)
Exponential Lower Bounds on the Double Oracle Algorithm in Zero-Sum Games
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2024)
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2024)
On the Outcome Equivalence of Extensive-Form and Behavioral Correlated Equilibria
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2024)
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2024)
Faster Game Solving via Hyperparameter Schedules
von: Zhang, Naifeng, et al.
Veröffentlicht: (2024)
von: Zhang, Naifeng, et al.
Veröffentlicht: (2024)
Team Belief DAG: Generalizing the Sequence Form to Team Games for Fast Computation of Correlated Team Max-Min Equilibria via Regret Minimization
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2022)
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2022)
Equilibrium Refinements Improve Subgame Solving in Imperfect-Information Games
von: Kubicek, Ondrej, et al.
Veröffentlicht: (2026)
von: Kubicek, Ondrej, et al.
Veröffentlicht: (2026)
Joint-perturbation simultaneous pseudo-gradient
von: Martin, Carlos, et al.
Veröffentlicht: (2024)
von: Martin, Carlos, et al.
Veröffentlicht: (2024)
Faster Optimal Coalition Structure Generation via Offline Coalition Selection and Graph-Based Search
von: Taguelmimt, Redha, et al.
Veröffentlicht: (2024)
von: Taguelmimt, Redha, et al.
Veröffentlicht: (2024)
Scalable Mechanism Design for Multi-Agent Path Finding
von: Friedrich, Paul, et al.
Veröffentlicht: (2024)
von: Friedrich, Paul, et al.
Veröffentlicht: (2024)
Mediator Interpretation and Faster Learning Algorithms for Linear Correlated Equilibria in General Extensive-Form Games
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2023)
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2023)
LLMs as Strategic Agents: Beliefs, Best Response Behavior, and Emergent Heuristics
von: de Fortuny, Enric Junque, et al.
Veröffentlicht: (2025)
von: de Fortuny, Enric Junque, et al.
Veröffentlicht: (2025)
Verifying Approximate Equilibrium in Auctions
von: Pieroth, Fabian R., et al.
Veröffentlicht: (2024)
von: Pieroth, Fabian R., et al.
Veröffentlicht: (2024)
(Doubly) Exponential Lower Bounds for Follow the Regularized Leader in Potential Games
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2026)
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2026)
Computational Lower Bounds for Regret Minimization in Normal-Form Games
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2024)
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2024)
Barriers to Welfare Maximization with No-Regret Learning
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2024)
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2024)
Efficient $Φ$-Regret Minimization with Low-Degree Swap Deviations in Extensive-Form Games
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2024)
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2024)
Learning a Game by Paying the Agents
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2025)
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2025)
The Complexity of Symmetric Equilibria in Min-Max Optimization and Team Zero-Sum Games
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2025)
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2025)
Scale-Invariant Regret Matching and Online Learning with Optimal Convergence: Bridging Theory and Practice in Zero-Sum Games
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2025)
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2025)
Bicriteria Multidimensional Mechanism Design with Side Information
von: Balcan, Maria-Florina, et al.
Veröffentlicht: (2023)
von: Balcan, Maria-Florina, et al.
Veröffentlicht: (2023)
Revenue-Optimal Efficient Mechanism Design with General Type Spaces
von: Prasad, Siddharth, et al.
Veröffentlicht: (2025)
von: Prasad, Siddharth, et al.
Veröffentlicht: (2025)
Why AI Safety Requires Uncertainty, Incomplete Preferences, and Non-Archimedean Utilities
von: Benavoli, Alessio, et al.
Veröffentlicht: (2025)
von: Benavoli, Alessio, et al.
Veröffentlicht: (2025)
Hidden-Role Games: Equilibrium Concepts and Computation
von: Carminati, Luca, et al.
Veröffentlicht: (2023)
von: Carminati, Luca, et al.
Veröffentlicht: (2023)
A Lower Bound on Swap Regret in Extensive-Form Games
von: Daskalakis, Constantinos, et al.
Veröffentlicht: (2024)
von: Daskalakis, Constantinos, et al.
Veröffentlicht: (2024)
How Far Can LLMs Emulate Human Behavior?: A Strategic Analysis via the Buy-and-Sell Negotiation Game
von: Jeon, Mingyu, et al.
Veröffentlicht: (2025)
von: Jeon, Mingyu, et al.
Veröffentlicht: (2025)
The Complexity of Proper Equilibrium in Extensive-Form and Polytope Games
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2026)
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2026)
On the Complexity of Correlated Equilibria Beyond Normal-Form Games
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2026)
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2026)
The Complexity of Equilibrium Refinements in Potential Games
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2025)
von: Anagnostides, Ioannis, et al.
Veröffentlicht: (2025)
Maximin Share Guarantees via Limited Cost-Sensitive Sharing
von: Salavcova, Hana, et al.
Veröffentlicht: (2026)
von: Salavcova, Hana, et al.
Veröffentlicht: (2026)
Weakest Bidder Types and New Core-Selecting Combinatorial Auctions
von: Prasad, Siddharth, et al.
Veröffentlicht: (2025)
von: Prasad, Siddharth, et al.
Veröffentlicht: (2025)
Persona Vectors in Games: Measuring and Steering Strategies via Activation Vectors
von: Sun, Johnathan, et al.
Veröffentlicht: (2026)
von: Sun, Johnathan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Domain-Independent Game Abstraction using Word Embedding Techniques
von: Kim, Juho, et al.
Veröffentlicht: (2026) -
Parallelizing Counterfactual Regret Minimization
von: Kim, Juho, et al.
Veröffentlicht: (2026) -
Watermarking Game-Playing Agents in Perfect-Information Extensive-Form Games
von: Kim, Juho, et al.
Veröffentlicht: (2026) -
General search techniques without common knowledge for imperfect-information games, and application to superhuman Fog of War chess
von: Zhang, Brian Hu, et al.
Veröffentlicht: (2025) -
ApproxED: Approximate exploitability descent via learned best responses
von: Martin, Carlos, et al.
Veröffentlicht: (2023)