Principal-Agent Bandit Games with Self-Interested and Exploratory Learning Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Junyan, Ratliff, Lillian J. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals
von: Liu, Junyan, et al.
Veröffentlicht: (2025)
von: Liu, Junyan, et al.
Veröffentlicht: (2025)
Online Learning for Uninformed Markov Games: Empirical Nash-Value Regret and Non-Stationarity Adaptation
von: Liu, Junyan, et al.
Veröffentlicht: (2026)
von: Liu, Junyan, et al.
Veröffentlicht: (2026)
Adaptive Calibration in Non-Stationary Environments
von: Liu, Junyan, et al.
Veröffentlicht: (2026)
von: Liu, Junyan, et al.
Veröffentlicht: (2026)
Incentivized Learning in Principal-Agent Bandit Games
von: Scheid, Antoine, et al.
Veröffentlicht: (2024)
von: Scheid, Antoine, et al.
Veröffentlicht: (2024)
Convergence of Learning Dynamics in Stackelberg Games
von: Fiez, Tanner, et al.
Veröffentlicht: (2019)
von: Fiez, Tanner, et al.
Veröffentlicht: (2019)
Strategically Robust Multi-Agent Reinforcement Learning with Linear Function Approximation
von: Gonzales, Jake, et al.
Veröffentlicht: (2026)
von: Gonzales, Jake, et al.
Veröffentlicht: (2026)
Online SuBmodular + SuPermodular (BP) Maximization with Bandit Feedback
von: Narang, Adhyyan, et al.
Veröffentlicht: (2022)
von: Narang, Adhyyan, et al.
Veröffentlicht: (2022)
On the Limitations and Possibilities of Nash Regret Minimization in Zero-Sum Matrix Games under Noisy Feedback
von: Maiti, Arnab, et al.
Veröffentlicht: (2023)
von: Maiti, Arnab, et al.
Veröffentlicht: (2023)
A Learning Algorithm That Attains the Human Optimum in a Repeated Human-Machine Interaction Game
von: Isa, Jason T., et al.
Veröffentlicht: (2025)
von: Isa, Jason T., et al.
Veröffentlicht: (2025)
Convergence Analysis of Gradient-Based Learning with Non-Uniform Learning Rates in Non-Cooperative Multi-Agent Settings
von: Chasnov, Benjamin, et al.
Veröffentlicht: (2019)
von: Chasnov, Benjamin, et al.
Veröffentlicht: (2019)
Improved Regret and Contextual Linear Extension for Pandora's Box and Prophet Inequality
von: Liu, Junyan, et al.
Veröffentlicht: (2025)
von: Liu, Junyan, et al.
Veröffentlicht: (2025)
Efficient Near-Optimal Algorithm for Online Shortest Paths in Directed Acyclic Graphs with Bandit Feedback Against Adaptive Adversaries
von: Maiti, Arnab, et al.
Veröffentlicht: (2025)
von: Maiti, Arnab, et al.
Veröffentlicht: (2025)
Dynamics of Learning under User Choice: Overspecialization and Peer-Model Probing
von: Narang, Adhyyan, et al.
Veröffentlicht: (2026)
von: Narang, Adhyyan, et al.
Veröffentlicht: (2026)
dUltra: Ultra-Fast Diffusion Language Models via Reinforcement Learning
von: Chen, Shirui, et al.
Veröffentlicht: (2025)
von: Chen, Shirui, et al.
Veröffentlicht: (2025)
Sample Complexity Reduction via Policy Difference Estimation in Tabular Reinforcement Learning
von: Narang, Adhyyan, et al.
Veröffentlicht: (2024)
von: Narang, Adhyyan, et al.
Veröffentlicht: (2024)
Uniform Last-Iterate Guarantee for Bandits and Reinforcement Learning
von: Liu, Junyan, et al.
Veröffentlicht: (2024)
von: Liu, Junyan, et al.
Veröffentlicht: (2024)
Self-Interested Agents in Collaborative Machine Learning: An Incentivized Adaptive Data-Centric Framework
von: Vijayan, Nithia, et al.
Veröffentlicht: (2024)
von: Vijayan, Nithia, et al.
Veröffentlicht: (2024)
Online Learning with Improving Agents: Multiclass, Budgeted Agents and Bandit Learners
von: Ashkezari, Sajad, et al.
Veröffentlicht: (2026)
von: Ashkezari, Sajad, et al.
Veröffentlicht: (2026)
Efficient Uncoupled Learning Dynamics with $\tilde{O}\!\left(T^{-1/4}\right)$ Last-Iterate Convergence in Bilinear Saddle-Point Problems over Convex Sets under Bandit Feedback
von: Maiti, Arnab, et al.
Veröffentlicht: (2026)
von: Maiti, Arnab, et al.
Veröffentlicht: (2026)
Fair Contracts in Principal-Agent Games with Heterogeneous Types
von: Tłuczek, Jakub, et al.
Veröffentlicht: (2025)
von: Tłuczek, Jakub, et al.
Veröffentlicht: (2025)
Multi-Agent Lipschitz Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2026)
Principal-Agent Reinforcement Learning: Orchestrating AI Agents with Contracts
von: Ivanov, Dima, et al.
Veröffentlicht: (2024)
von: Ivanov, Dima, et al.
Veröffentlicht: (2024)
Generalized Principal-Agent Problem with a Learning Agent
von: Lin, Tao, et al.
Veröffentlicht: (2024)
von: Lin, Tao, et al.
Veröffentlicht: (2024)
Learning Equilibria in Matching Games with Bandit Feedback
von: Athanasopoulos, Andreas, et al.
Veröffentlicht: (2025)
von: Athanasopoulos, Andreas, et al.
Veröffentlicht: (2025)
Collaborating in Multi-Armed Bandits with Strategic Agents
von: Barnea, Idan, et al.
Veröffentlicht: (2026)
von: Barnea, Idan, et al.
Veröffentlicht: (2026)
Multi-Objective Multi-Agent Bandits: From Learning Efficiency to Fairness Optimization
von: Wang, John, et al.
Veröffentlicht: (2026)
von: Wang, John, et al.
Veröffentlicht: (2026)
Distributed Multi-Agent Bandits Over Erdős-Rényi Random Networks
von: Liu, Jingyuan, et al.
Veröffentlicht: (2025)
von: Liu, Jingyuan, et al.
Veröffentlicht: (2025)
Learning in Online Principal-Agent Interactions: The Power of Menus
von: Han, Minbiao, et al.
Veröffentlicht: (2023)
von: Han, Minbiao, et al.
Veröffentlicht: (2023)
Multi-Agent Stochastic Bandits Robust to Adversarial Corruptions
von: Ghaffari, Fatemeh, et al.
Veröffentlicht: (2024)
von: Ghaffari, Fatemeh, et al.
Veröffentlicht: (2024)
Adaptive Sample Sharing for Multi Agent Linear Bandits
von: Cherkaoui, Hamza, et al.
Veröffentlicht: (2023)
von: Cherkaoui, Hamza, et al.
Veröffentlicht: (2023)
A Comprehensive Review of Multi-Agent Reinforcement Learning in Video Games
von: Li, Zhengyang, et al.
Veröffentlicht: (2025)
von: Li, Zhengyang, et al.
Veröffentlicht: (2025)
Contextual Agent Security: A Policy for Every Purpose
von: Tsai, Lillian, et al.
Veröffentlicht: (2025)
von: Tsai, Lillian, et al.
Veröffentlicht: (2025)
S4S: Solving for a Diffusion Model Solver
von: Frankel, Eric, et al.
Veröffentlicht: (2025)
von: Frankel, Eric, et al.
Veröffentlicht: (2025)
On the Universal Near Optimality of Hedge in Combinatorial Settings
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2025)
Emergent specialization from participation dynamics and multi-learner retraining
von: Dean, Sarah, et al.
Veröffentlicht: (2022)
von: Dean, Sarah, et al.
Veröffentlicht: (2022)
Fair Algorithms with Probing for Multi-Agent Multi-Armed Bandits
von: Xu, Tianyi, et al.
Veröffentlicht: (2025)
von: Xu, Tianyi, et al.
Veröffentlicht: (2025)
Distributed Algorithms for Multi-Agent Multi-Armed Bandits with Collision
von: Zhou, Daoyuan, et al.
Veröffentlicht: (2025)
von: Zhou, Daoyuan, et al.
Veröffentlicht: (2025)
Stochastic Principal-Agent Problems: Efficient Computation and Learning
von: Gan, Jiarui, et al.
Veröffentlicht: (2023)
von: Gan, Jiarui, et al.
Veröffentlicht: (2023)
Improving Generalization in Game Agents with Data Augmentation in Imitation Learning
von: Yadgaroff, Derek, et al.
Veröffentlicht: (2023)
von: Yadgaroff, Derek, et al.
Veröffentlicht: (2023)
Scalable Principal-Agent Contract Design via Gradient-Based Optimization
von: Galanti, Tomer, et al.
Veröffentlicht: (2025)
von: Galanti, Tomer, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent Arrivals
von: Liu, Junyan, et al.
Veröffentlicht: (2025) -
Online Learning for Uninformed Markov Games: Empirical Nash-Value Regret and Non-Stationarity Adaptation
von: Liu, Junyan, et al.
Veröffentlicht: (2026) -
Adaptive Calibration in Non-Stationary Environments
von: Liu, Junyan, et al.
Veröffentlicht: (2026) -
Incentivized Learning in Principal-Agent Bandit Games
von: Scheid, Antoine, et al.
Veröffentlicht: (2024) -
Convergence of Learning Dynamics in Stackelberg Games
von: Fiez, Tanner, et al.
Veröffentlicht: (2019)