Responding to Promises: No-regret learning against followers with memory
Fuente:
arXiv
Guardado en:
| Autores principales: | Hebbar, Vijeth, Langbort, Cédric |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On Network Congestion Reduction Using Public Signals Under Boundedly Rational User Equilibria (Full Version)
por: Massicot, Olivier, et al.
Publicado: (2024)
por: Massicot, Olivier, et al.
Publicado: (2024)
Almost-Bayesian Quadratic Persuasion (Extended Version)
por: Massicot, Olivier, et al.
Publicado: (2022)
por: Massicot, Olivier, et al.
Publicado: (2022)
Misinformation Regulation in the Presence of Competition between Social Media Platforms (Extended Version)
por: Sasaki, So, et al.
Publicado: (2024)
por: Sasaki, So, et al.
Publicado: (2024)
Strategizing against No-regret Learners
por: Deng, Yuan, et al.
Publicado: (2019)
por: Deng, Yuan, et al.
Publicado: (2019)
Revisiting Regret Benchmarks in Online Non-Stochastic Control
por: Hebbar, Vijeth, et al.
Publicado: (2025)
por: Hebbar, Vijeth, et al.
Publicado: (2025)
Hierarchies of No-regret Algorithms
por: Xu, R., et al.
Publicado: (2026)
por: Xu, R., et al.
Publicado: (2026)
No-regret incentive-compatible online learning under exact truthfulness with non-myopic experts
por: Komiyama, Junpei, et al.
Publicado: (2025)
por: Komiyama, Junpei, et al.
Publicado: (2025)
Last iterate convergence in no-regret learning: constrained min-max optimization for convex-concave landscapes
por: Lei, Qi, et al.
Publicado: (2020)
por: Lei, Qi, et al.
Publicado: (2020)
On the price of exact truthfulness in incentive-compatible online learning with bandit feedback: A regret lower bound for WSU-UX
por: Mortazavi, Ali, et al.
Publicado: (2024)
por: Mortazavi, Ali, et al.
Publicado: (2024)
Convergence to Nash Equilibrium and No-regret Guarantee in (Markov) Potential Games
por: Dong, Jing, et al.
Publicado: (2024)
por: Dong, Jing, et al.
Publicado: (2024)
The Economics of No-regret Learning Algorithms
por: Hartline, Jason
Publicado: (2026)
por: Hartline, Jason
Publicado: (2026)
Bayes correlated equilibria, no-regret dynamics in Bayesian games, and the price of anarchy
por: Fujii, Kaito
Publicado: (2023)
por: Fujii, Kaito
Publicado: (2023)
Promises Made, Promises Kept: Safe Pareto Improvements via Ex Post Verifiable Commitments
por: Sauerberg, Nathaniel, et al.
Publicado: (2025)
por: Sauerberg, Nathaniel, et al.
Publicado: (2025)
Optimal No-regret Learning in Repeated First-price Auctions
por: Han, Yanjun, et al.
Publicado: (2020)
por: Han, Yanjun, et al.
Publicado: (2020)
The Computational Intractability of Not Worst Responding
por: Ahunbay, Mete Şeref, et al.
Publicado: (2026)
por: Ahunbay, Mete Şeref, et al.
Publicado: (2026)
An $α$-regret analysis of Adversarial Bilateral Trade
por: Azar, Yossi, et al.
Publicado: (2022)
por: Azar, Yossi, et al.
Publicado: (2022)
Playing against a stationary opponent
por: Grand-Clément, Julien, et al.
Publicado: (2025)
por: Grand-Clément, Julien, et al.
Publicado: (2025)
A memory-based spatial evolutionary game with the dynamic interaction between learners and profiteers
por: Pi, Bin, et al.
Publicado: (2024)
por: Pi, Bin, et al.
Publicado: (2024)
Conditional cooperation with longer memory
por: Glynatsi, Nikoleta E., et al.
Publicado: (2024)
por: Glynatsi, Nikoleta E., et al.
Publicado: (2024)
Randomized learning-augmented auctions with revenue guarantees
por: Caragiannis, Ioannis, et al.
Publicado: (2024)
por: Caragiannis, Ioannis, et al.
Publicado: (2024)
Computing stable limit cycles of learning in games
por: Biggar, Oliver, et al.
Publicado: (2026)
por: Biggar, Oliver, et al.
Publicado: (2026)
Evolutionary hypergame dynamics: Introspection reasoning and social learning
por: Zhang, Feipeng, et al.
Publicado: (2025)
por: Zhang, Feipeng, et al.
Publicado: (2025)
A Renegotiable contract-theoretic incentive mechanism for Federated learning
por: Tan, Xavier, et al.
Publicado: (2025)
por: Tan, Xavier, et al.
Publicado: (2025)
Fairness in Multi-Proposer-Multi-Responder Ultimatum Game
por: Krakovská, Hana, et al.
Publicado: (2024)
por: Krakovská, Hana, et al.
Publicado: (2024)
Persuading a Behavioral Agent: Approximately Best Responding and Learning
por: Chen, Yiling, et al.
Publicado: (2023)
por: Chen, Yiling, et al.
Publicado: (2023)
Steady-state Based Approach to Online Non-stochastic Control
por: Hebbar, Vijeth, et al.
Publicado: (2026)
por: Hebbar, Vijeth, et al.
Publicado: (2026)
"What are my options?": Explaining RL Agents with Diverse Near-Optimal Alternatives (Extended)
por: Brindise, Noel, et al.
Publicado: (2025)
por: Brindise, Noel, et al.
Publicado: (2025)
Mobility operator service capacity sharing contract design to risk-pool against network disruptions
por: Pantelidis, Theodoros P., et al.
Publicado: (2020)
por: Pantelidis, Theodoros P., et al.
Publicado: (2020)
Profit Maximization in Bilateral Trade against a Smooth Adversary
por: Di Gregorio, Simone, et al.
Publicado: (2026)
por: Di Gregorio, Simone, et al.
Publicado: (2026)
Game-Theoretic Modeling of Stealthy Intrusion Defense against MDP-Based Attackers
por: Kouam, Willie, et al.
Publicado: (2026)
por: Kouam, Willie, et al.
Publicado: (2026)
Misspecified learning and evolutionary stability
por: He, Kevin, et al.
Publicado: (2025)
por: He, Kevin, et al.
Publicado: (2025)
Sink equilibria and the attractors of learning in games
por: Biggar, Oliver, et al.
Publicado: (2025)
por: Biggar, Oliver, et al.
Publicado: (2025)
Coarse Q-learning: Indifference, Indeterminacy, and Instability
por: Jehiel, Philippe, et al.
Publicado: (2024)
por: Jehiel, Philippe, et al.
Publicado: (2024)
Generalized Individual Q-learning for Polymatrix Games with Partial Observations
por: Donmez, Ahmed Said, et al.
Publicado: (2024)
por: Donmez, Ahmed Said, et al.
Publicado: (2024)
Improved learning rates in multi-unit uniform price auctions
por: Potfer, Marius, et al.
Publicado: (2025)
por: Potfer, Marius, et al.
Publicado: (2025)
Disentangling trust from cooperation: Evolution of trust as reduced monitoring in social dilemmas
por: Perret, Cedric, et al.
Publicado: (2025)
por: Perret, Cedric, et al.
Publicado: (2025)
Personalized Dynamic Pricing Policy for Electric Vehicles: Reinforcement learning approach
por: Bae, Sangjun, et al.
Publicado: (2024)
por: Bae, Sangjun, et al.
Publicado: (2024)
Group-wise oracle-efficient algorithms for online multi-group learning
por: Deng, Samuel, et al.
Publicado: (2024)
por: Deng, Samuel, et al.
Publicado: (2024)
A potentialization algorithm for games with applications to multi-agent learning in repeated games
por: Lakheshar, Philipp, et al.
Publicado: (2026)
por: Lakheshar, Philipp, et al.
Publicado: (2026)
Zero-sum turn games using Q-learning: finite computation with security guarantees
por: Anderson, Sean, et al.
Publicado: (2025)
por: Anderson, Sean, et al.
Publicado: (2025)
Ejemplares similares
-
On Network Congestion Reduction Using Public Signals Under Boundedly Rational User Equilibria (Full Version)
por: Massicot, Olivier, et al.
Publicado: (2024) -
Almost-Bayesian Quadratic Persuasion (Extended Version)
por: Massicot, Olivier, et al.
Publicado: (2022) -
Misinformation Regulation in the Presence of Competition between Social Media Platforms (Extended Version)
por: Sasaki, So, et al.
Publicado: (2024) -
Strategizing against No-regret Learners
por: Deng, Yuan, et al.
Publicado: (2019) -
Revisiting Regret Benchmarks in Online Non-Stochastic Control
por: Hebbar, Vijeth, et al.
Publicado: (2025)