Adversarial Learning in Games with Bandit Feedback: Logarithmic Pure-Strategy Maximin Regret
Fuente:
arXiv
Saved in:
| Main Authors: | Ito, Shinji, Luo, Haipeng, Maiti, Arnab, Tsuchiya, Taira, Wu, Yue |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Instance-Dependent Regret Bounds for Learning Two-Player Zero-Sum Games with Bandit Feedback
by: Ito, Shinji, et al.
Published: (2025)
by: Ito, Shinji, et al.
Published: (2025)
Regret Bounds for Robust Online Decision Making
by: Appel, Alexander, et al.
Published: (2025)
by: Appel, Alexander, et al.
Published: (2025)
Reinforcement Learning for Reachability: Guaranteeing Asymptotic Optimality
by: Palasamudram, Amogh, et al.
Published: (2026)
by: Palasamudram, Amogh, et al.
Published: (2026)
Frequency Principle: Fourier Analysis Sheds Light on Deep Neural Networks
by: Xu, Zhi-Qin John, et al.
Published: (2019)
by: Xu, Zhi-Qin John, et al.
Published: (2019)
A geometric decomposition of finite games: Convergence vs. recurrence under exponential weights
by: Legacci, Davide, et al.
Published: (2024)
by: Legacci, Davide, et al.
Published: (2024)
A Parallelizable Approach for Characterizing NE in Zero-Sum Games After a Linear Number of Iterations of Gradient Descent
by: Kim, Taemin, et al.
Published: (2025)
by: Kim, Taemin, et al.
Published: (2025)
Backpropagation Through Time For Networks With Long-Term Dependencies
by: Bird, George, et al.
Published: (2021)
by: Bird, George, et al.
Published: (2021)
Robust equilibria in continuous games: From strategic to dynamic robustness
by: Lotidis, Kyriakos, et al.
Published: (2025)
by: Lotidis, Kyriakos, et al.
Published: (2025)
Accelerated regularized learning in finite N-person games
by: Lotidis, Kyriakos, et al.
Published: (2024)
by: Lotidis, Kyriakos, et al.
Published: (2024)
The Current and Future Perspectives of Zinc Oxide Nanoparticles in the Treatment of Diabetes Mellitus
by: Yousaf, Iqra
Published: (2024)
by: Yousaf, Iqra
Published: (2024)
Aligning Inductive Bias for Data-Efficient Generalization in State Space Models
by: Chen, Qiyu, et al.
Published: (2025)
by: Chen, Qiyu, et al.
Published: (2025)
Ambiguous Online Learning
by: Kosoy, Vanessa
Published: (2025)
by: Kosoy, Vanessa
Published: (2025)
Adaptive Discretization in Online Reinforcement Learning
by: Sinclair, Sean R., et al.
Published: (2021)
by: Sinclair, Sean R., et al.
Published: (2021)
Sparse Knowledge Distillation: A Mathematical Framework for Probability-Domain Temperature Scaling and Multi-Stage Compression
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
Online Decision Making with Generative Action Sets
by: Xu, Jianyu, et al.
Published: (2025)
by: Xu, Jianyu, et al.
Published: (2025)
Computing Game Symmetries and Equilibria That Respect Them
by: Tewolde, Emanuel, et al.
Published: (2025)
by: Tewolde, Emanuel, et al.
Published: (2025)
Breaking $1/ε$ Barrier in Quantum Zero-Sum Games: Generalizing Metric Subregularity for Spectraplexes
by: Su, Yiheng, et al.
Published: (2025)
by: Su, Yiheng, et al.
Published: (2025)
No-regret learning in harmonic games: Extrapolation in the face of conflicting interests
by: Legacci, Davide, et al.
Published: (2024)
by: Legacci, Davide, et al.
Published: (2024)
Understanding the Nature of Generative AI as Threshold Logic in High-Dimensional Space
by: Levin, Ilya
Published: (2026)
by: Levin, Ilya
Published: (2026)
Agnostic Learning under Targeted Poisoning: Optimal Rates and the Role of Randomness
by: Chornomaz, Bogdan, et al.
Published: (2025)
by: Chornomaz, Bogdan, et al.
Published: (2025)
Superior Scoring Rules for Probabilistic Evaluation of Single-Label Multi-Class Classification Tasks
by: Ahmadian, Rouhollah, et al.
Published: (2024)
by: Ahmadian, Rouhollah, et al.
Published: (2024)
Improved Approximation Ratio for Strategyproof Facility Location on a Cycle
by: Rogowski, Krzysztof, et al.
Published: (2025)
by: Rogowski, Krzysztof, et al.
Published: (2025)
Nested replicator dynamics, nested logit choice, and similarity-based learning
by: Mertikopoulos, Panayotis, et al.
Published: (2024)
by: Mertikopoulos, Panayotis, et al.
Published: (2024)
Golden Handcuffs make safer AI agents
by: Ebtekar, Aram, et al.
Published: (2026)
by: Ebtekar, Aram, et al.
Published: (2026)
Near-Optimal Consistency-Robustness Trade-Offs for Learning-Augmented Online Knapsack Problems
by: Daneshvaramoli, Mohammadreza, et al.
Published: (2024)
by: Daneshvaramoli, Mohammadreza, et al.
Published: (2024)
Margin in Abstract Spaces
by: Ashlagi, Yair, et al.
Published: (2026)
by: Ashlagi, Yair, et al.
Published: (2026)
Reinforcement Learning in MDPs with Information-Ordered Policies
by: Zhang, Zhongjun, et al.
Published: (2025)
by: Zhang, Zhongjun, et al.
Published: (2025)
Mitigating Catastrophic Forgetting in Streaming Generative and Predictive Learning via Stateful Replay
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Decision Making under Imperfect Recall: Algorithms and Benchmarks
by: Tewolde, Emanuel, et al.
Published: (2026)
by: Tewolde, Emanuel, et al.
Published: (2026)
MenuNet: A Strategy-Proof Mechanism for Matching Markets
by: Sun, Zhaohong, et al.
Published: (2026)
by: Sun, Zhaohong, et al.
Published: (2026)
A Quadratic Speedup in Finding Nash Equilibria of Quantum Zero-Sum Games
by: Vasconcelos, Francisca, et al.
Published: (2023)
by: Vasconcelos, Francisca, et al.
Published: (2023)
A Theoretical Study of (Hyper) Self-Attention through the Lens of Interactions: Representation, Training, Generalization
by: Ustaomeroglu, Muhammed, et al.
Published: (2025)
by: Ustaomeroglu, Muhammed, et al.
Published: (2025)
Think Thrice Before You Speak: Dual knowledge-enhanced Theory-of-Mind Reasoning for Persuasive Agents
by: Ma, Minghui, et al.
Published: (2026)
by: Ma, Minghui, et al.
Published: (2026)
Statistical Guarantees for Lifelong Reinforcement Learning using PAC-Bayes Theory
by: Zhang, Zhi, et al.
Published: (2024)
by: Zhang, Zhi, et al.
Published: (2024)
Integration of Deep Reinforcement Learning and Agent-based Simulation to Explore Strategies Counteracting Information Disorder
by: Lomasto, Luigi, et al.
Published: (2026)
by: Lomasto, Luigi, et al.
Published: (2026)
MMD-Balls as Credal Sets: A PAC-Bayesian Framework for Epistemic Uncertainty in Test-Time Adaptation
by: Ariq, Ahanaf Hasan
Published: (2026)
by: Ariq, Ahanaf Hasan
Published: (2026)
Inductive Venn-Abers and related regressors
by: Petej, Ivan, et al.
Published: (2026)
by: Petej, Ivan, et al.
Published: (2026)
Aggregation in conformal e-classification
by: Vovk, Vladimir
Published: (2026)
by: Vovk, Vladimir
Published: (2026)
Batched Nonparametric Bandits via k-Nearest Neighbor UCB
by: Arya, Sakshi
Published: (2025)
by: Arya, Sakshi
Published: (2025)
AI Agents for the Dhumbal Card Game: A Comparative Study
by: Malla, Sahaj Raj
Published: (2025)
by: Malla, Sahaj Raj
Published: (2025)
Similar Items
-
Instance-Dependent Regret Bounds for Learning Two-Player Zero-Sum Games with Bandit Feedback
by: Ito, Shinji, et al.
Published: (2025) -
Regret Bounds for Robust Online Decision Making
by: Appel, Alexander, et al.
Published: (2025) -
Reinforcement Learning for Reachability: Guaranteeing Asymptotic Optimality
by: Palasamudram, Amogh, et al.
Published: (2026) -
Frequency Principle: Fourier Analysis Sheds Light on Deep Neural Networks
by: Xu, Zhi-Qin John, et al.
Published: (2019) -
A geometric decomposition of finite games: Convergence vs. recurrence under exponential weights
by: Legacci, Davide, et al.
Published: (2024)