Adversarial Learning in Games with Bandit Feedback: Logarithmic Pure-Strategy Maximin Regret
Fuente:
arXiv
Guardado en:
| Autores principales: | Ito, Shinji, Luo, Haipeng, Maiti, Arnab, Tsuchiya, Taira, Wu, Yue |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Instance-Dependent Regret Bounds for Learning Two-Player Zero-Sum Games with Bandit Feedback
por: Ito, Shinji, et al.
Publicado: (2025)
por: Ito, Shinji, et al.
Publicado: (2025)
Regret Bounds for Robust Online Decision Making
por: Appel, Alexander, et al.
Publicado: (2025)
por: Appel, Alexander, et al.
Publicado: (2025)
Reinforcement Learning for Reachability: Guaranteeing Asymptotic Optimality
por: Palasamudram, Amogh, et al.
Publicado: (2026)
por: Palasamudram, Amogh, et al.
Publicado: (2026)
Frequency Principle: Fourier Analysis Sheds Light on Deep Neural Networks
por: Xu, Zhi-Qin John, et al.
Publicado: (2019)
por: Xu, Zhi-Qin John, et al.
Publicado: (2019)
A geometric decomposition of finite games: Convergence vs. recurrence under exponential weights
por: Legacci, Davide, et al.
Publicado: (2024)
por: Legacci, Davide, et al.
Publicado: (2024)
A Parallelizable Approach for Characterizing NE in Zero-Sum Games After a Linear Number of Iterations of Gradient Descent
por: Kim, Taemin, et al.
Publicado: (2025)
por: Kim, Taemin, et al.
Publicado: (2025)
Backpropagation Through Time For Networks With Long-Term Dependencies
por: Bird, George, et al.
Publicado: (2021)
por: Bird, George, et al.
Publicado: (2021)
Robust equilibria in continuous games: From strategic to dynamic robustness
por: Lotidis, Kyriakos, et al.
Publicado: (2025)
por: Lotidis, Kyriakos, et al.
Publicado: (2025)
Accelerated regularized learning in finite N-person games
por: Lotidis, Kyriakos, et al.
Publicado: (2024)
por: Lotidis, Kyriakos, et al.
Publicado: (2024)
The Current and Future Perspectives of Zinc Oxide Nanoparticles in the Treatment of Diabetes Mellitus
por: Yousaf, Iqra
Publicado: (2024)
por: Yousaf, Iqra
Publicado: (2024)
Aligning Inductive Bias for Data-Efficient Generalization in State Space Models
por: Chen, Qiyu, et al.
Publicado: (2025)
por: Chen, Qiyu, et al.
Publicado: (2025)
Ambiguous Online Learning
por: Kosoy, Vanessa
Publicado: (2025)
por: Kosoy, Vanessa
Publicado: (2025)
Adaptive Discretization in Online Reinforcement Learning
por: Sinclair, Sean R., et al.
Publicado: (2021)
por: Sinclair, Sean R., et al.
Publicado: (2021)
Sparse Knowledge Distillation: A Mathematical Framework for Probability-Domain Temperature Scaling and Multi-Stage Compression
por: Flouro, Aaron R., et al.
Publicado: (2026)
por: Flouro, Aaron R., et al.
Publicado: (2026)
Online Decision Making with Generative Action Sets
por: Xu, Jianyu, et al.
Publicado: (2025)
por: Xu, Jianyu, et al.
Publicado: (2025)
Computing Game Symmetries and Equilibria That Respect Them
por: Tewolde, Emanuel, et al.
Publicado: (2025)
por: Tewolde, Emanuel, et al.
Publicado: (2025)
Breaking $1/ε$ Barrier in Quantum Zero-Sum Games: Generalizing Metric Subregularity for Spectraplexes
por: Su, Yiheng, et al.
Publicado: (2025)
por: Su, Yiheng, et al.
Publicado: (2025)
No-regret learning in harmonic games: Extrapolation in the face of conflicting interests
por: Legacci, Davide, et al.
Publicado: (2024)
por: Legacci, Davide, et al.
Publicado: (2024)
Understanding the Nature of Generative AI as Threshold Logic in High-Dimensional Space
por: Levin, Ilya
Publicado: (2026)
por: Levin, Ilya
Publicado: (2026)
Agnostic Learning under Targeted Poisoning: Optimal Rates and the Role of Randomness
por: Chornomaz, Bogdan, et al.
Publicado: (2025)
por: Chornomaz, Bogdan, et al.
Publicado: (2025)
Superior Scoring Rules for Probabilistic Evaluation of Single-Label Multi-Class Classification Tasks
por: Ahmadian, Rouhollah, et al.
Publicado: (2024)
por: Ahmadian, Rouhollah, et al.
Publicado: (2024)
Improved Approximation Ratio for Strategyproof Facility Location on a Cycle
por: Rogowski, Krzysztof, et al.
Publicado: (2025)
por: Rogowski, Krzysztof, et al.
Publicado: (2025)
Nested replicator dynamics, nested logit choice, and similarity-based learning
por: Mertikopoulos, Panayotis, et al.
Publicado: (2024)
por: Mertikopoulos, Panayotis, et al.
Publicado: (2024)
Golden Handcuffs make safer AI agents
por: Ebtekar, Aram, et al.
Publicado: (2026)
por: Ebtekar, Aram, et al.
Publicado: (2026)
Near-Optimal Consistency-Robustness Trade-Offs for Learning-Augmented Online Knapsack Problems
por: Daneshvaramoli, Mohammadreza, et al.
Publicado: (2024)
por: Daneshvaramoli, Mohammadreza, et al.
Publicado: (2024)
Margin in Abstract Spaces
por: Ashlagi, Yair, et al.
Publicado: (2026)
por: Ashlagi, Yair, et al.
Publicado: (2026)
Reinforcement Learning in MDPs with Information-Ordered Policies
por: Zhang, Zhongjun, et al.
Publicado: (2025)
por: Zhang, Zhongjun, et al.
Publicado: (2025)
Mitigating Catastrophic Forgetting in Streaming Generative and Predictive Learning via Stateful Replay
por: Du, Wenzhang
Publicado: (2025)
por: Du, Wenzhang
Publicado: (2025)
Decision Making under Imperfect Recall: Algorithms and Benchmarks
por: Tewolde, Emanuel, et al.
Publicado: (2026)
por: Tewolde, Emanuel, et al.
Publicado: (2026)
MenuNet: A Strategy-Proof Mechanism for Matching Markets
por: Sun, Zhaohong, et al.
Publicado: (2026)
por: Sun, Zhaohong, et al.
Publicado: (2026)
A Quadratic Speedup in Finding Nash Equilibria of Quantum Zero-Sum Games
por: Vasconcelos, Francisca, et al.
Publicado: (2023)
por: Vasconcelos, Francisca, et al.
Publicado: (2023)
A Theoretical Study of (Hyper) Self-Attention through the Lens of Interactions: Representation, Training, Generalization
por: Ustaomeroglu, Muhammed, et al.
Publicado: (2025)
por: Ustaomeroglu, Muhammed, et al.
Publicado: (2025)
Think Thrice Before You Speak: Dual knowledge-enhanced Theory-of-Mind Reasoning for Persuasive Agents
por: Ma, Minghui, et al.
Publicado: (2026)
por: Ma, Minghui, et al.
Publicado: (2026)
Statistical Guarantees for Lifelong Reinforcement Learning using PAC-Bayes Theory
por: Zhang, Zhi, et al.
Publicado: (2024)
por: Zhang, Zhi, et al.
Publicado: (2024)
Integration of Deep Reinforcement Learning and Agent-based Simulation to Explore Strategies Counteracting Information Disorder
por: Lomasto, Luigi, et al.
Publicado: (2026)
por: Lomasto, Luigi, et al.
Publicado: (2026)
MMD-Balls as Credal Sets: A PAC-Bayesian Framework for Epistemic Uncertainty in Test-Time Adaptation
por: Ariq, Ahanaf Hasan
Publicado: (2026)
por: Ariq, Ahanaf Hasan
Publicado: (2026)
Inductive Venn-Abers and related regressors
por: Petej, Ivan, et al.
Publicado: (2026)
por: Petej, Ivan, et al.
Publicado: (2026)
Aggregation in conformal e-classification
por: Vovk, Vladimir
Publicado: (2026)
por: Vovk, Vladimir
Publicado: (2026)
Batched Nonparametric Bandits via k-Nearest Neighbor UCB
por: Arya, Sakshi
Publicado: (2025)
por: Arya, Sakshi
Publicado: (2025)
AI Agents for the Dhumbal Card Game: A Comparative Study
por: Malla, Sahaj Raj
Publicado: (2025)
por: Malla, Sahaj Raj
Publicado: (2025)
Ejemplares similares
-
Instance-Dependent Regret Bounds for Learning Two-Player Zero-Sum Games with Bandit Feedback
por: Ito, Shinji, et al.
Publicado: (2025) -
Regret Bounds for Robust Online Decision Making
por: Appel, Alexander, et al.
Publicado: (2025) -
Reinforcement Learning for Reachability: Guaranteeing Asymptotic Optimality
por: Palasamudram, Amogh, et al.
Publicado: (2026) -
Frequency Principle: Fourier Analysis Sheds Light on Deep Neural Networks
por: Xu, Zhi-Qin John, et al.
Publicado: (2019) -
A geometric decomposition of finite games: Convergence vs. recurrence under exponential weights
por: Legacci, Davide, et al.
Publicado: (2024)