Stochastic contextual bandits with graph feedback: from independence number to MAS number
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wen, Yuxiao, Han, Yanjun, Zhou, Zhengyuan |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Joint Value Estimation and Bidding in Repeated First-Price Auctions
par: Wen, Yuxiao, et autres
Publié: (2025)
par: Wen, Yuxiao, et autres
Publié: (2025)
The (Marginal) Value of a Search Ad: An Online Causal Framework for Repeated Second-price Auctions
par: Wen, Yuxiao, et autres
Publié: (2026)
par: Wen, Yuxiao, et autres
Publié: (2026)
Optimal No-regret Learning in Repeated First-price Auctions
par: Han, Yanjun, et autres
Publié: (2020)
par: Han, Yanjun, et autres
Publié: (2020)
Fairness in two-player zero-sum games with bandit feedback
par: Akash, S, et autres
Publié: (2026)
par: Akash, S, et autres
Publié: (2026)
Learning to Bid Optimally and Efficiently in Adversarial First-price Auctions
par: Han, Yanjun, et autres
Publié: (2020)
par: Han, Yanjun, et autres
Publié: (2020)
On the price of exact truthfulness in incentive-compatible online learning with bandit feedback: A regret lower bound for WSU-UX
par: Mortazavi, Ali, et autres
Publié: (2024)
par: Mortazavi, Ali, et autres
Publié: (2024)
A survey on multi-player bandits
par: Boursier, Etienne, et autres
Publié: (2022)
par: Boursier, Etienne, et autres
Publié: (2022)
Bayesian Algorithms for Adversarial Online Learning: from Finite to Infinite Action Spaces
par: Terenin, Alexander, et autres
Publié: (2025)
par: Terenin, Alexander, et autres
Publié: (2025)
Truthful mechanisms for linear bandit games with private contexts
par: Hu, Yiting, et autres
Publié: (2025)
par: Hu, Yiting, et autres
Publié: (2025)
Optimal Scoring Rule Design under Partial Knowledge
par: Chen, Yiling, et autres
Publié: (2021)
par: Chen, Yiling, et autres
Publié: (2021)
Pure Exploration via Frank-Wolfe Self-Play
par: Liu, Xinyu, et autres
Publié: (2025)
par: Liu, Xinyu, et autres
Publié: (2025)
Model-Based Reinforcement Learning for Offline Zero-Sum Markov Games
par: Yan, Yuling, et autres
Publié: (2022)
par: Yan, Yuling, et autres
Publié: (2022)
Learning to Bid in Non-Stationary Repeated First-Price Auctions
par: Hu, Zihao, et autres
Publié: (2025)
par: Hu, Zihao, et autres
Publié: (2025)
Sharp Results for Hypothesis Testing with Risk-Sensitive Agents
par: Shi, Flora C., et autres
Publié: (2024)
par: Shi, Flora C., et autres
Publié: (2024)
Instance-Adaptive Hypothesis Tests with Heterogeneous Agents
par: Shi, Flora C., et autres
Publié: (2025)
par: Shi, Flora C., et autres
Publié: (2025)
Statistical Impossibility and Possibility of Aligning LLMs with Human Preferences: From Condorcet Paradox to Nash Equilibrium
par: Liu, Kaizhao, et autres
Publié: (2025)
par: Liu, Kaizhao, et autres
Publié: (2025)
Adaptive, Doubly Optimal No-Regret Learning in Strongly Monotone and Exp-Concave Games with Gradient Feedback
par: Jordan, Michael I., et autres
Publié: (2023)
par: Jordan, Michael I., et autres
Publié: (2023)
Computing Optimal Regularizers for Online Linear Optimization
par: Gatmiry, Khashayar, et autres
Publié: (2024)
par: Gatmiry, Khashayar, et autres
Publié: (2024)
Isotonic Mechanism for Exponential Family Estimation in Machine Learning Peer Review
par: Yan, Yuling, et autres
Publié: (2023)
par: Yan, Yuling, et autres
Publié: (2023)
Principal-Agent Hypothesis Testing
par: Bates, Stephen, et autres
Publié: (2022)
par: Bates, Stephen, et autres
Publié: (2022)
You Are the Best Reviewer of Your Own Papers: The Isotonic Mechanism
par: Su, Weijie
Publié: (2022)
par: Su, Weijie
Publié: (2022)
Doubly Optimal No-Regret Online Learning in Strongly Monotone Games with Bandit Feedback
par: Ba, Wenjia, et autres
Publié: (2021)
par: Ba, Wenjia, et autres
Publié: (2021)
Adaptive Discretization against an Adversary: Lipschitz bandits, Dynamic Pricing, and Auction Tuning
par: Podimata, Chara, et autres
Publié: (2020)
par: Podimata, Chara, et autres
Publié: (2020)
p-Mean Regret for Stochastic Bandits
par: Krishna, Anand, et autres
Publié: (2024)
par: Krishna, Anand, et autres
Publié: (2024)
Nash Equilibria via Stochastic Eigendecomposition
par: Gemp, Ian
Publié: (2024)
par: Gemp, Ian
Publié: (2024)
Learning-Augmented Online Bidding in Stochastic Settings
par: Angelopoulos, Spyros, et autres
Publié: (2025)
par: Angelopoulos, Spyros, et autres
Publié: (2025)
Partially Observable Stochastic Games with Neural Perception Mechanisms
par: Yan, Rui, et autres
Publié: (2023)
par: Yan, Rui, et autres
Publié: (2023)
Learning and Collusion in Multi-unit Auctions
par: Brânzei, Simina, et autres
Publié: (2023)
par: Brânzei, Simina, et autres
Publié: (2023)
Learning the Expected Core of Strictly Convex Stochastic Cooperative Games
par: Tran, Nam Phuong, et autres
Publié: (2024)
par: Tran, Nam Phuong, et autres
Publié: (2024)
Robust Adversarial Reinforcement Learning in Stochastic Games via Sequence Modeling
par: Tang, Xiaohang, et autres
Publié: (2025)
par: Tang, Xiaohang, et autres
Publié: (2025)
Decentralized Multi-Agent Reinforcement Learning for Continuous-Space Stochastic Games
par: Altabaa, Awni, et autres
Publié: (2023)
par: Altabaa, Awni, et autres
Publié: (2023)
Actor-Dual-Critic Dynamics for Zero-sum and Identical-Interest Stochastic Games
par: Donmez, Ahmed Said, et autres
Publié: (2026)
par: Donmez, Ahmed Said, et autres
Publié: (2026)
Last-Iterate Convergence of Payoff-Based Independent Learning in Zero-Sum Stochastic Games
par: Chen, Zaiwei, et autres
Publié: (2024)
par: Chen, Zaiwei, et autres
Publié: (2024)
A Bayesian Learning Algorithm for Unknown Zero-sum Stochastic Games with an Arbitrary Opponent
par: Jafarnia-Jahromi, Mehdi, et autres
Publié: (2021)
par: Jafarnia-Jahromi, Mehdi, et autres
Publié: (2021)
Gradient Manipulation in Distributed Stochastic Gradient Descent with Strategic Agents: Truthful Incentives with Convergence Guarantees
par: Chen, Ziqin, et autres
Publié: (2026)
par: Chen, Ziqin, et autres
Publié: (2026)
Neural Mean-Field Games: Extending Mean-Field Game Theory with Neural Stochastic Differential Equations
par: Thöni, Anna C. M., et autres
Publié: (2025)
par: Thöni, Anna C. M., et autres
Publié: (2025)
Peer Expectation in Robust Forecast Aggregation: The Possibility/Impossibility
par: Kong, Yuqing
Publié: (2024)
par: Kong, Yuqing
Publié: (2024)
Low-Rank Graphon Estimation: Theory and Applications to Graphon Games
par: Klopp, Olga, et autres
Publié: (2025)
par: Klopp, Olga, et autres
Publié: (2025)
Linear Functions to the Extended Reals
par: Waggoner, Bo
Publié: (2021)
par: Waggoner, Bo
Publié: (2021)
A conversion theorem and minimax optimality for continuum contextual bandits
par: Akhavan, Arya, et autres
Publié: (2024)
par: Akhavan, Arya, et autres
Publié: (2024)
Documents similaires
-
Joint Value Estimation and Bidding in Repeated First-Price Auctions
par: Wen, Yuxiao, et autres
Publié: (2025) -
The (Marginal) Value of a Search Ad: An Online Causal Framework for Repeated Second-price Auctions
par: Wen, Yuxiao, et autres
Publié: (2026) -
Optimal No-regret Learning in Repeated First-price Auctions
par: Han, Yanjun, et autres
Publié: (2020) -
Fairness in two-player zero-sum games with bandit feedback
par: Akash, S, et autres
Publié: (2026) -
Learning to Bid Optimally and Efficiently in Adversarial First-price Auctions
par: Han, Yanjun, et autres
Publié: (2020)