Contractual Reinforcement Learning: Pulling Arms with Invisible Hands
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Jibang, Chen, Siyu, Wang, Mengdi, Wang, Huazheng, Xu, Haifeng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Generalized Principal-Agency: Contracts, Information, Games and Beyond
di: Gan, Jiarui, et al.
Pubblicazione: (2022)
di: Gan, Jiarui, et al.
Pubblicazione: (2022)
An Isotonic Mechanism for Overlapping Ownership
di: Wu, Jibang, et al.
Pubblicazione: (2023)
di: Wu, Jibang, et al.
Pubblicazione: (2023)
Paid with Models: Optimal Contract Design for Collaborative Machine Learning
di: Wang, Bingchen, et al.
Pubblicazione: (2024)
di: Wang, Bingchen, et al.
Pubblicazione: (2024)
Robust Stackelberg Equilibria
di: Gan, Jiarui, et al.
Pubblicazione: (2023)
di: Gan, Jiarui, et al.
Pubblicazione: (2023)
Contracting with a Learning Agent
di: Guruganesh, Guru, et al.
Pubblicazione: (2024)
di: Guruganesh, Guru, et al.
Pubblicazione: (2024)
Generalized Principal-Agent Problem with a Learning Agent
di: Lin, Tao, et al.
Pubblicazione: (2024)
di: Lin, Tao, et al.
Pubblicazione: (2024)
Efficient Inverse Multiagent Learning
di: Goktas, Denizalp, et al.
Pubblicazione: (2025)
di: Goktas, Denizalp, et al.
Pubblicazione: (2025)
Are Bounded Contracts Learnable and Approximately Optimal?
di: Chen, Yurong, et al.
Pubblicazione: (2024)
di: Chen, Yurong, et al.
Pubblicazione: (2024)
Calibeating Made Simple
di: Chen, Yurong, et al.
Pubblicazione: (2026)
di: Chen, Yurong, et al.
Pubblicazione: (2026)
Computational Performance of Deep Reinforcement Learning to find Nash Equilibria
di: Graf, Christoph, et al.
Pubblicazione: (2021)
di: Graf, Christoph, et al.
Pubblicazione: (2021)
The Pseudo-Dimension of Contracts
di: Duetting, Paul, et al.
Pubblicazione: (2025)
di: Duetting, Paul, et al.
Pubblicazione: (2025)
Fraud-Proof Revenue Division on Subscription Platforms
di: Ghosh, Abheek, et al.
Pubblicazione: (2025)
di: Ghosh, Abheek, et al.
Pubblicazione: (2025)
Persuasive Calibration
di: Feng, Yiding, et al.
Pubblicazione: (2025)
di: Feng, Yiding, et al.
Pubblicazione: (2025)
No Screening is More Efficient with Multiple Objects
di: Noda, Shunya, et al.
Pubblicazione: (2024)
di: Noda, Shunya, et al.
Pubblicazione: (2024)
Human strategic decision making in parametrized games
di: Ganzfried, Sam
Pubblicazione: (2021)
di: Ganzfried, Sam
Pubblicazione: (2021)
A New Lower Bound for the Random Offerer Mechanism in Bilateral Trade using AI-Guided Evolutionary Search
di: Cai, Yang, et al.
Pubblicazione: (2026)
di: Cai, Yang, et al.
Pubblicazione: (2026)
Two-Sided Time-Independent Regret for Matching Markets with Limited Interviews
di: Mirfakhar, Amirmahdi, et al.
Pubblicazione: (2026)
di: Mirfakhar, Amirmahdi, et al.
Pubblicazione: (2026)
Nash Convergence of Mean-Based Learning Algorithms in First-Price Auctions
di: Deng, Xiaotie, et al.
Pubblicazione: (2021)
di: Deng, Xiaotie, et al.
Pubblicazione: (2021)
Persuading a Behavioral Agent: Approximately Best Responding and Learning
di: Chen, Yiling, et al.
Pubblicazione: (2023)
di: Chen, Yiling, et al.
Pubblicazione: (2023)
Learned Collusion
di: Compte, Olivier
Pubblicazione: (2023)
di: Compte, Olivier
Pubblicazione: (2023)
Misspecified Explore-then-Exploit Leads to Supra-Competitive Prices
di: Baek, Jackie, et al.
Pubblicazione: (2026)
di: Baek, Jackie, et al.
Pubblicazione: (2026)
Dueling Over Dessert, Mastering the Art of Repeated Cake Cutting
di: Brânzei, Simina, et al.
Pubblicazione: (2024)
di: Brânzei, Simina, et al.
Pubblicazione: (2024)
Deep Learning for Double Auction
di: Liu, Jiayin, et al.
Pubblicazione: (2025)
di: Liu, Jiayin, et al.
Pubblicazione: (2025)
Tell Me Why: Incentivizing Explanations
di: Srinivasan, Siddarth, et al.
Pubblicazione: (2025)
di: Srinivasan, Siddarth, et al.
Pubblicazione: (2025)
Randomized Truthful Auctions with Learning Agents
di: Aggarwal, Gagan, et al.
Pubblicazione: (2024)
di: Aggarwal, Gagan, et al.
Pubblicazione: (2024)
How Humans Help LLMs: Assessing and Incentivizing Human Preference Annotators
di: Liu, Shang, et al.
Pubblicazione: (2025)
di: Liu, Shang, et al.
Pubblicazione: (2025)
Networked Information Aggregation via Machine Learning
di: Kearns, Michael, et al.
Pubblicazione: (2025)
di: Kearns, Michael, et al.
Pubblicazione: (2025)
Multiplayer Bandit Learning, from Competition to Cooperation
di: Brânzei, Simina, et al.
Pubblicazione: (2019)
di: Brânzei, Simina, et al.
Pubblicazione: (2019)
Learning to Coordinate Bidders in Non-Truthful Auctions
di: Fu, Hu, et al.
Pubblicazione: (2025)
di: Fu, Hu, et al.
Pubblicazione: (2025)
From No-Regret to Strategically Robust Learning in Repeated Auctions
di: Zhao, Junyao
Pubblicazione: (2026)
di: Zhao, Junyao
Pubblicazione: (2026)
Learning to Play Multi-Follower Bayesian Stackelberg Games
di: Personnat, Gerson, et al.
Pubblicazione: (2025)
di: Personnat, Gerson, et al.
Pubblicazione: (2025)
Prediction-sharing During Training and Inference
di: Gafni, Yotam, et al.
Pubblicazione: (2024)
di: Gafni, Yotam, et al.
Pubblicazione: (2024)
Computing Equilibrium beyond Unilateral Deviation
di: Liu, Mingyang, et al.
Pubblicazione: (2026)
di: Liu, Mingyang, et al.
Pubblicazione: (2026)
Fairness in Repeated Matching: A Maximin Perspective
di: Lim, Eugene, et al.
Pubblicazione: (2025)
di: Lim, Eugene, et al.
Pubblicazione: (2025)
The Price of Proportional Representation in Temporal Voting
di: Teh, Nicholas
Pubblicazione: (2026)
di: Teh, Nicholas
Pubblicazione: (2026)
Multi-agent Adaptive Mechanism Design
di: Han, Qiushi, et al.
Pubblicazione: (2025)
di: Han, Qiushi, et al.
Pubblicazione: (2025)
The Bounds of Algorithmic Collusion; $Q$-learning, Gradient Learning, and the Folk Theorem
di: Askenazi-Golan, Galit, et al.
Pubblicazione: (2024)
di: Askenazi-Golan, Galit, et al.
Pubblicazione: (2024)
Competing Bandits: The Perils of Exploration Under Competition
di: Aridor, Guy, et al.
Pubblicazione: (2020)
di: Aridor, Guy, et al.
Pubblicazione: (2020)
Rational Adversaries and the Maintenance of Fragility: A Game-Theoretic Theory of Rational Stagnation
di: Hirota, Daisuke
Pubblicazione: (2025)
di: Hirota, Daisuke
Pubblicazione: (2025)
Human-AI Productivity Paradoxes: Modeling the Interplay of Skill, Effort, and AI Assistance
di: Aouad, Ali, et al.
Pubblicazione: (2026)
di: Aouad, Ali, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Generalized Principal-Agency: Contracts, Information, Games and Beyond
di: Gan, Jiarui, et al.
Pubblicazione: (2022) -
An Isotonic Mechanism for Overlapping Ownership
di: Wu, Jibang, et al.
Pubblicazione: (2023) -
Paid with Models: Optimal Contract Design for Collaborative Machine Learning
di: Wang, Bingchen, et al.
Pubblicazione: (2024) -
Robust Stackelberg Equilibria
di: Gan, Jiarui, et al.
Pubblicazione: (2023) -
Contracting with a Learning Agent
di: Guruganesh, Guru, et al.
Pubblicazione: (2024)