Learning Equilibria in Matching Games with Bandit Feedback
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Athanasopoulos, Andreas, Dimitrakakis, Christos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Probably Correct Optimal Stable Matching for Two-Sided Markets Under Uncertainty
von: Athanasopoulos, Andreas, et al.
Veröffentlicht: (2025)
von: Athanasopoulos, Andreas, et al.
Veröffentlicht: (2025)
Two-Player Zero-Sum Games with Bandit Feedback
von: Yılmaz, Elif, et al.
Veröffentlicht: (2025)
von: Yılmaz, Elif, et al.
Veröffentlicht: (2025)
Rawlsian many-to-one matching with non-linear utility
von: Nana, Hortence, et al.
Veröffentlicht: (2025)
von: Nana, Hortence, et al.
Veröffentlicht: (2025)
Strategic Linear Contextual Bandits
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2024)
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2024)
Fair Contracts in Principal-Agent Games with Heterogeneous Types
von: Tłuczek, Jakub, et al.
Veröffentlicht: (2025)
von: Tłuczek, Jakub, et al.
Veröffentlicht: (2025)
Isoperimetry is All We Need: Langevin Posterior Sampling for RL with Sublinear Regret
von: Jorge, Emilio, et al.
Veröffentlicht: (2024)
von: Jorge, Emilio, et al.
Veröffentlicht: (2024)
Environment Design for Inverse Reinforcement Learning
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2022)
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2022)
Queueing Matching Bandits with Preference Feedback
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024)
von: Kim, Jung-hun, et al.
Veröffentlicht: (2024)
Sequential Cohort Selection
von: Nana, Hortence Phalonne, et al.
Veröffentlicht: (2025)
von: Nana, Hortence Phalonne, et al.
Veröffentlicht: (2025)
Bandits with Preference Feedback: A Stackelberg Game Perspective
von: Pásztor, Barna, et al.
Veröffentlicht: (2024)
von: Pásztor, Barna, et al.
Veröffentlicht: (2024)
The Harder Path: Last Iterate Convergence for Uncoupled Learning in Zero-Sum Games with Bandit Feedback
von: Fiegel, Côme, et al.
Veröffentlicht: (2026)
von: Fiegel, Côme, et al.
Veröffentlicht: (2026)
Performative Prediction with Bandit Feedback: Learning through Reparameterization
von: Chen, Yatong, et al.
Veröffentlicht: (2023)
von: Chen, Yatong, et al.
Veröffentlicht: (2023)
Doubly Optimal No-Regret Online Learning in Strongly Monotone Games with Bandit Feedback
von: Ba, Wenjia, et al.
Veröffentlicht: (2021)
von: Ba, Wenjia, et al.
Veröffentlicht: (2021)
Near-Optimal Last-Iterate Convergence for Zero-Sum Games with Bandit Feedback and Opponent Actions
von: Hait, Soumita, et al.
Veröffentlicht: (2026)
von: Hait, Soumita, et al.
Veröffentlicht: (2026)
Nearest Neighbour with Bandit Feedback
von: Pasteris, Stephen, et al.
Veröffentlicht: (2023)
von: Pasteris, Stephen, et al.
Veröffentlicht: (2023)
Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback
von: Goyal, Tanmay, et al.
Veröffentlicht: (2025)
von: Goyal, Tanmay, et al.
Veröffentlicht: (2025)
Deep Learning Methods for Detecting Thermal Runaway Events in Battery Production Lines
von: Athanasopoulos, Athanasios, et al.
Veröffentlicht: (2025)
von: Athanasopoulos, Athanasios, et al.
Veröffentlicht: (2025)
Constrained Feedback Learning for Non-Stationary Multi-Armed Bandits
von: Li, Shaoang, et al.
Veröffentlicht: (2025)
von: Li, Shaoang, et al.
Veröffentlicht: (2025)
Learning Markov Decision Processes under Fully Bandit Feedback
von: Zhuo, Zhengjia, et al.
Veröffentlicht: (2026)
von: Zhuo, Zhengjia, et al.
Veröffentlicht: (2026)
A General Framework for Clustering and Distribution Matching with Bandit Feedback
von: Yavas, Recep Can, et al.
Veröffentlicht: (2024)
von: Yavas, Recep Can, et al.
Veröffentlicht: (2024)
Lipschitz Bandits with Stochastic Delayed Feedback
von: Liu, Zhongxuan, et al.
Veröffentlicht: (2025)
von: Liu, Zhongxuan, et al.
Veröffentlicht: (2025)
Graph Feedback Bandits with Similar Arms
von: Qi, Han, et al.
Veröffentlicht: (2024)
von: Qi, Han, et al.
Veröffentlicht: (2024)
Nonparametric Kernel Clustering with Bandit Feedback
von: Thuot, Victor, et al.
Veröffentlicht: (2026)
von: Thuot, Victor, et al.
Veröffentlicht: (2026)
Optimal Analysis for Bandit Learning in Matching Markets with Serial Dictatorship
von: Wang, Zilong, et al.
Veröffentlicht: (2025)
von: Wang, Zilong, et al.
Veröffentlicht: (2025)
Adaptive Client Sampling in Federated Learning via Online Learning with Bandit Feedback
von: Zhao, Boxin, et al.
Veröffentlicht: (2021)
von: Zhao, Boxin, et al.
Veröffentlicht: (2021)
Last-Iterate Convergence of No-Regret Learning for Equilibria in Bargaining Games
von: Kamp, Serafina, et al.
Veröffentlicht: (2025)
von: Kamp, Serafina, et al.
Veröffentlicht: (2025)
Nearly Tight Bounds for Cross-Learning Contextual Bandits with Graphical Feedback
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2025)
von: Huang, Ruiyuan, et al.
Veröffentlicht: (2025)
Pure Exploration for a Good Policy in Reinforcement Learning with Bandit Feedback
von: Li, Zitian, et al.
Veröffentlicht: (2026)
von: Li, Zitian, et al.
Veröffentlicht: (2026)
Optimal Clustering with Bandit Feedback
von: Yang, Junwen, et al.
Veröffentlicht: (2022)
von: Yang, Junwen, et al.
Veröffentlicht: (2022)
Operator Splitting for Learning to Predict Equilibria in Convex Games
von: McKenzie, Daniel, et al.
Veröffentlicht: (2021)
von: McKenzie, Daniel, et al.
Veröffentlicht: (2021)
Learning to Schedule Online Tasks with Bandit Feedback
von: Xu, Yongxin, et al.
Veröffentlicht: (2024)
von: Xu, Yongxin, et al.
Veröffentlicht: (2024)
Cascading Bandits With Feedback
von: Prakash, R Sri, et al.
Veröffentlicht: (2025)
von: Prakash, R Sri, et al.
Veröffentlicht: (2025)
Stochastic $k$-Submodular Bandits with Full Bandit Feedback
von: Nie, Guanyu, et al.
Veröffentlicht: (2024)
von: Nie, Guanyu, et al.
Veröffentlicht: (2024)
Constrained Pareto Set Identification with Bandit Feedback
von: Kone, Cyrille, et al.
Veröffentlicht: (2025)
von: Kone, Cyrille, et al.
Veröffentlicht: (2025)
Does Feedback Help in Bandits with Arm Erasures?
von: Karakas, Merve, et al.
Veröffentlicht: (2025)
von: Karakas, Merve, et al.
Veröffentlicht: (2025)
Bandit and Delayed Feedback in Online Structured Prediction
von: Shibukawa, Yuki, et al.
Veröffentlicht: (2025)
von: Shibukawa, Yuki, et al.
Veröffentlicht: (2025)
Beyond Bandit Feedback in Online Multiclass Classification
von: van der Hoeven, Dirk, et al.
Veröffentlicht: (2021)
von: van der Hoeven, Dirk, et al.
Veröffentlicht: (2021)
Efficient Contextual Bandits with Uninformed Feedback Graphs
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2024)
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2024)
Multiclass Online Learnability under Bandit Feedback
von: Raman, Ananth, et al.
Veröffentlicht: (2023)
von: Raman, Ananth, et al.
Veröffentlicht: (2023)
Biased Dueling Bandits with Stochastic Delayed Feedback
von: Yi, Bongsoo, et al.
Veröffentlicht: (2024)
von: Yi, Bongsoo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Probably Correct Optimal Stable Matching for Two-Sided Markets Under Uncertainty
von: Athanasopoulos, Andreas, et al.
Veröffentlicht: (2025) -
Two-Player Zero-Sum Games with Bandit Feedback
von: Yılmaz, Elif, et al.
Veröffentlicht: (2025) -
Rawlsian many-to-one matching with non-linear utility
von: Nana, Hortence, et al.
Veröffentlicht: (2025) -
Strategic Linear Contextual Bandits
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2024) -
Fair Contracts in Principal-Agent Games with Heterogeneous Types
von: Tłuczek, Jakub, et al.
Veröffentlicht: (2025)