VickreyFeedback: Cost-efficient Data Construction for Reinforcement Learning from Human Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Guoxi, Duan, Jiuding |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Generalized Model for Multidimensional Intransitivity
by: Duan, Jiuding, et al.
Published: (2024)
by: Duan, Jiuding, et al.
Published: (2024)
Instructing LLMs to Negotiate using Reinforcement Learning with Verifiable Rewards
by: Liu, Shuze Daniel, et al.
Published: (2026)
by: Liu, Shuze Daniel, et al.
Published: (2026)
Explainable Graph Neural Networks via Structural Externalities
by: Wu, Lijun, et al.
Published: (2025)
by: Wu, Lijun, et al.
Published: (2025)
Modelling bounded rational decision-making through Wasserstein constraints
by: Evans, Benjamin Patrick, et al.
Published: (2025)
by: Evans, Benjamin Patrick, et al.
Published: (2025)
Capturing the Complexity of Human Strategic Decision-Making with Machine Learning
by: Zhu, Jian-Qiao, et al.
Published: (2024)
by: Zhu, Jian-Qiao, et al.
Published: (2024)
Modeling the Feedback of AI Price Estimations on Actual Market Values
by: Silaghi, Viorel, et al.
Published: (2024)
by: Silaghi, Viorel, et al.
Published: (2024)
DeepVoting: Learning and Fine-Tuning Voting Rules with Canonical Embeddings
by: Matone, Leonardo, et al.
Published: (2024)
by: Matone, Leonardo, et al.
Published: (2024)
Learning to Charge More: A Theoretical Study of Collusion by Q-Learning Agents
by: Chica, Cristian, et al.
Published: (2025)
by: Chica, Cristian, et al.
Published: (2025)
Optimal Trade and Industrial Policies in the Global Economy: A Deep Learning Framework
by: Wang, Zi, et al.
Published: (2024)
by: Wang, Zi, et al.
Published: (2024)
Algorithmic Collusion of Pricing and Advertising on E-commerce Platforms
by: Zhao, Hangcheng, et al.
Published: (2025)
by: Zhao, Hangcheng, et al.
Published: (2025)
Bridging the Gap Between Estimated and True Regret Towards Reliable Regret Estimation in Deep Learning based Mechanism Design
by: You, Shuyuan, et al.
Published: (2026)
by: You, Shuyuan, et al.
Published: (2026)
Tacit Bidder-Side Collusion: Artificial Intelligence in Dynamic Auctions
by: Tolety, Sriram
Published: (2025)
by: Tolety, Sriram
Published: (2025)
An AI Capability Threshold for Rent-Funded Universal Basic Income in an AI-Automated Economy
by: Nayebi, Aran
Published: (2025)
by: Nayebi, Aran
Published: (2025)
Manipulation and Peer Mechanisms: A Survey
by: Olckers, Matthew, et al.
Published: (2022)
by: Olckers, Matthew, et al.
Published: (2022)
Algorithmic Collusion by Large Language Models
by: Fish, Sara, et al.
Published: (2024)
by: Fish, Sara, et al.
Published: (2024)
Artificial Intelligence and Algorithmic Price Collusion in Two-sided Markets
by: Chica, Cristian, et al.
Published: (2024)
by: Chica, Cristian, et al.
Published: (2024)
Information Aggregation with AI Agents
by: Galanis, Spyros
Published: (2026)
by: Galanis, Spyros
Published: (2026)
Algorithmic pricing with independent learners and relative experience replay
by: Han, Bingyan
Published: (2021)
by: Han, Bingyan
Published: (2021)
Explore-then-Commit Algorithms for Decentralized Two-Sided Matching Markets
by: Pagare, Tejas, et al.
Published: (2024)
by: Pagare, Tejas, et al.
Published: (2024)
Advancing Ad Auction Realism: Practical Insights & Modeling Implications
by: Chen, Ming, et al.
Published: (2023)
by: Chen, Ming, et al.
Published: (2023)
Integrative Experiments Identify How Punishment Impacts Welfare in Public Goods Games
by: Alsobay, Mohammed, et al.
Published: (2025)
by: Alsobay, Mohammed, et al.
Published: (2025)
On Benchmark Hacking in ML Contests: Modeling, Insights and Design
by: Qiu, Xiaoyun, et al.
Published: (2026)
by: Qiu, Xiaoyun, et al.
Published: (2026)
Not Yet: Humans Outperform LLMs in a Colonel Blotto Tournament
by: Dagaev, Dmitry, et al.
Published: (2026)
by: Dagaev, Dmitry, et al.
Published: (2026)
Normative Equivalence in Human-AI Cooperation: Behaviour, Not Identity, Drives Cooperation in Mixed-Agent Groups
by: Mutzner, Nico, et al.
Published: (2026)
by: Mutzner, Nico, et al.
Published: (2026)
Auction-Based Regulation for Artificial Intelligence
by: Bornstein, Marco, et al.
Published: (2024)
by: Bornstein, Marco, et al.
Published: (2024)
AI Cap-and-Trade: Efficiency Incentives for Accessibility and Sustainability
by: Bornstein, Marco, et al.
Published: (2026)
by: Bornstein, Marco, et al.
Published: (2026)
Assessing Large Language Models' ability to predict how humans balance self-interest and the interest of others
by: Capraro, Valerio, et al.
Published: (2023)
by: Capraro, Valerio, et al.
Published: (2023)
Bandit Profit-maximization for Targeted Marketing
by: Huh, Joon Suk, et al.
Published: (2024)
by: Huh, Joon Suk, et al.
Published: (2024)
Contractual Reinforcement Learning: Pulling Arms with Invisible Hands
by: Wu, Jibang, et al.
Published: (2024)
by: Wu, Jibang, et al.
Published: (2024)
Learning and Calibrating Heterogeneous Bounded Rational Market Behaviour with Multi-Agent Reinforcement Learning
by: Evans, Benjamin Patrick, et al.
Published: (2024)
by: Evans, Benjamin Patrick, et al.
Published: (2024)
Human behaviour through a LENS: How Linguistic content triggers Emotions and Norms and determines Strategy choices
by: Capraro, Valerio
Published: (2024)
by: Capraro, Valerio
Published: (2024)
A General Framework for Optimizing and Learning Nash Equilibrium
by: Zhang, Di, et al.
Published: (2024)
by: Zhang, Di, et al.
Published: (2024)
Overcoming the Machine Penalty with Imperfectly Fair AI Agents
by: Wang, Zhen, et al.
Published: (2024)
by: Wang, Zhen, et al.
Published: (2024)
Can AI with High Reasoning Ability Replicate Human-like Decision Making in Economic Experiments?
by: Kitadai, Ayato, et al.
Published: (2024)
by: Kitadai, Ayato, et al.
Published: (2024)
Data-Driven Mechanism Design using Multi-Agent Revealed Preferences
by: Snow, Luke, et al.
Published: (2024)
by: Snow, Luke, et al.
Published: (2024)
Human strategic decision making in parametrized games
by: Ganzfried, Sam
Published: (2021)
by: Ganzfried, Sam
Published: (2021)
Training Language Models for Bilateral Trade with Private Information
by: Bergemann, Dirk, et al.
Published: (2026)
by: Bergemann, Dirk, et al.
Published: (2026)
Contracting with a Learning Agent
by: Guruganesh, Guru, et al.
Published: (2024)
by: Guruganesh, Guru, et al.
Published: (2024)
Efficient Inverse Multiagent Learning
by: Goktas, Denizalp, et al.
Published: (2025)
by: Goktas, Denizalp, et al.
Published: (2025)
Wasserstein Markets for Differentially-Private Data
by: Chhachhi, Saurab, et al.
Published: (2024)
by: Chhachhi, Saurab, et al.
Published: (2024)
Similar Items
-
A Generalized Model for Multidimensional Intransitivity
by: Duan, Jiuding, et al.
Published: (2024) -
Instructing LLMs to Negotiate using Reinforcement Learning with Verifiable Rewards
by: Liu, Shuze Daniel, et al.
Published: (2026) -
Explainable Graph Neural Networks via Structural Externalities
by: Wu, Lijun, et al.
Published: (2025) -
Modelling bounded rational decision-making through Wasserstein constraints
by: Evans, Benjamin Patrick, et al.
Published: (2025) -
Capturing the Complexity of Human Strategic Decision-Making with Machine Learning
by: Zhu, Jian-Qiao, et al.
Published: (2024)