Jackpot! Alignment as a Maximal Lottery
Fuente:
arXiv
Saved in:
| Main Authors: | Maura-Rivero, Roberto-Rafael, Lanctot, Marc, Visin, Francesco, Larson, Kate |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models
by: Maura-Rivero, Roberto-Rafael, et al.
Published: (2025)
by: Maura-Rivero, Roberto-Rafael, et al.
Published: (2025)
Structural Reinforcement Learning for Heterogeneous Agent Macroeconomics
by: Yang, Yucheng, et al.
Published: (2025)
by: Yang, Yucheng, et al.
Published: (2025)
Preference Learning with Response Time: Robust Losses and Guarantees
by: Sawarni, Ayush, et al.
Published: (2025)
by: Sawarni, Ayush, et al.
Published: (2025)
Prospect-Theory Behavior from Bellman Optimality in MDPs with Catastrophic States
by: Chen, Yujiao
Published: (2026)
by: Chen, Yujiao
Published: (2026)
The Burden of Interactive Alignment with Inconsistent Preferences
by: Shirali, Ali
Published: (2025)
by: Shirali, Ali
Published: (2025)
Continuous Classification Aggregation
by: Meng, Zijun
Published: (2025)
by: Meng, Zijun
Published: (2025)
Contracting with a Learning Agent
by: Guruganesh, Guru, et al.
Published: (2024)
by: Guruganesh, Guru, et al.
Published: (2024)
Generalized Principal-Agent Problem with a Learning Agent
by: Lin, Tao, et al.
Published: (2024)
by: Lin, Tao, et al.
Published: (2024)
A Regret Analysis of Bilateral Trade
by: Cesa-Bianchi, Nicolò, et al.
Published: (2021)
by: Cesa-Bianchi, Nicolò, et al.
Published: (2021)
Group Selection as a Safeguard Against AI Substitution
by: Zhong, Qiankun, et al.
Published: (2026)
by: Zhong, Qiankun, et al.
Published: (2026)
The Pseudo-Dimension of Contracts
by: Duetting, Paul, et al.
Published: (2025)
by: Duetting, Paul, et al.
Published: (2025)
Fraud-Proof Revenue Division on Subscription Platforms
by: Ghosh, Abheek, et al.
Published: (2025)
by: Ghosh, Abheek, et al.
Published: (2025)
Persuasive Calibration
by: Feng, Yiding, et al.
Published: (2025)
by: Feng, Yiding, et al.
Published: (2025)
Efficient Inverse Multiagent Learning
by: Goktas, Denizalp, et al.
Published: (2025)
by: Goktas, Denizalp, et al.
Published: (2025)
Are Bounded Contracts Learnable and Approximately Optimal?
by: Chen, Yurong, et al.
Published: (2024)
by: Chen, Yurong, et al.
Published: (2024)
Contractual Reinforcement Learning: Pulling Arms with Invisible Hands
by: Wu, Jibang, et al.
Published: (2024)
by: Wu, Jibang, et al.
Published: (2024)
No Screening is More Efficient with Multiple Objects
by: Noda, Shunya, et al.
Published: (2024)
by: Noda, Shunya, et al.
Published: (2024)
Paid with Models: Optimal Contract Design for Collaborative Machine Learning
by: Wang, Bingchen, et al.
Published: (2024)
by: Wang, Bingchen, et al.
Published: (2024)
Human strategic decision making in parametrized games
by: Ganzfried, Sam
Published: (2021)
by: Ganzfried, Sam
Published: (2021)
Calibeating Made Simple
by: Chen, Yurong, et al.
Published: (2026)
by: Chen, Yurong, et al.
Published: (2026)
A New Lower Bound for the Random Offerer Mechanism in Bilateral Trade using AI-Guided Evolutionary Search
by: Cai, Yang, et al.
Published: (2026)
by: Cai, Yang, et al.
Published: (2026)
Model-Based Soft Maximization of Suitable Metrics of Long-Term Human Power
by: Heitzig, Jobst, et al.
Published: (2025)
by: Heitzig, Jobst, et al.
Published: (2025)
Pricing AI Model Accuracy
by: Kumar, Nikhil
Published: (2025)
by: Kumar, Nikhil
Published: (2025)
NDAI Agreements
by: Stephenson, Matthew, et al.
Published: (2025)
by: Stephenson, Matthew, et al.
Published: (2025)
Token Is All You Price
by: Zhong, Weijie
Published: (2025)
by: Zhong, Weijie
Published: (2025)
The Economics of AI Foundation Models: Openness, Competition, and Governance
by: Xu, Fasheng, et al.
Published: (2025)
by: Xu, Fasheng, et al.
Published: (2025)
Market-based Architectures in RL and Beyond
by: Sudhir, Abhimanyu Pallavi, et al.
Published: (2025)
by: Sudhir, Abhimanyu Pallavi, et al.
Published: (2025)
Optimal Use of Preferences in Artificial Intelligence Algorithms
by: Gans, Joshua S.
Published: (2026)
by: Gans, Joshua S.
Published: (2026)
Failures of Contingent Thinking
by: Piermont, Evan, et al.
Published: (2020)
by: Piermont, Evan, et al.
Published: (2020)
Generative AI and Copyright: A Dynamic Perspective
by: Yang, S. Alex, et al.
Published: (2024)
by: Yang, S. Alex, et al.
Published: (2024)
Learning what they think vs. learning what they do: The micro-foundations of vicarious learning
by: Park, Sanghyun, et al.
Published: (2020)
by: Park, Sanghyun, et al.
Published: (2020)
Cheap Talking Algorithms
by: Condorelli, Daniele, et al.
Published: (2023)
by: Condorelli, Daniele, et al.
Published: (2023)
AI-powered mechanisms as judges: Breaking ties in chess
by: Anbarci, Nejat, et al.
Published: (2022)
by: Anbarci, Nejat, et al.
Published: (2022)
Learning Macroeconomic Policies through Dynamic Stackelberg Mean-Field Games
by: Mi, Qirui, et al.
Published: (2024)
by: Mi, Qirui, et al.
Published: (2024)
A Model of Artificial Jagged Intelligence
by: Gans, Joshua
Published: (2026)
by: Gans, Joshua
Published: (2026)
Decision-Making Behavior Evaluation Framework for LLMs under Uncertain Context
by: Jia, Jingru, et al.
Published: (2024)
by: Jia, Jingru, et al.
Published: (2024)
A General Framework for Estimating Preferences Using Response Time Data
by: Echenique, Federico, et al.
Published: (2025)
by: Echenique, Federico, et al.
Published: (2025)
Markets for Models
by: Dasaratha, Krishna, et al.
Published: (2025)
by: Dasaratha, Krishna, et al.
Published: (2025)
The Hidden Cost of Waiting for Accurate Predictions
by: Shirali, Ali, et al.
Published: (2025)
by: Shirali, Ali, et al.
Published: (2025)
Beyond Softmax: A New Perspective on Gradient Bandits
by: Melo, Emerson, et al.
Published: (2025)
by: Melo, Emerson, et al.
Published: (2025)
Similar Items
-
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models
by: Maura-Rivero, Roberto-Rafael, et al.
Published: (2025) -
Structural Reinforcement Learning for Heterogeneous Agent Macroeconomics
by: Yang, Yucheng, et al.
Published: (2025) -
Preference Learning with Response Time: Robust Losses and Guarantees
by: Sawarni, Ayush, et al.
Published: (2025) -
Prospect-Theory Behavior from Bellman Optimality in MDPs with Catastrophic States
by: Chen, Yujiao
Published: (2026) -
The Burden of Interactive Alignment with Inconsistent Preferences
by: Shirali, Ali
Published: (2025)