Human Misperception of Generative-AI Alignment: A Laboratory Experiment
Fuente:
arXiv
Saved in:
| Main Authors: | He, Kevin, Shorrer, Ran, Xia, Mengjia |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Strategic Algorithmic Monoculture: Experimental Evidence from Coordination Games
by: Ballestero, Gonzalo, et al.
Published: (2026)
by: Ballestero, Gonzalo, et al.
Published: (2026)
A Revealed Preference Framework for AI Alignment
by: Suleymanov, Elchin
Published: (2026)
by: Suleymanov, Elchin
Published: (2026)
Human-AI Productivity Paradoxes: Modeling the Interplay of Skill, Effort, and AI Assistance
by: Aouad, Ali, et al.
Published: (2026)
by: Aouad, Ali, et al.
Published: (2026)
Will AI Trade? A Computational Inversion of the No-Trade Theorem
by: Li, Hanyu, et al.
Published: (2025)
by: Li, Hanyu, et al.
Published: (2025)
Creative Ownership in the Age of AI
by: Liang, Annie, et al.
Published: (2026)
by: Liang, Annie, et al.
Published: (2026)
Fine-Tuning Games: Bargaining and Adaptation for General-Purpose Models
by: Laufer, Benjamin, et al.
Published: (2023)
by: Laufer, Benjamin, et al.
Published: (2023)
Rational Adversaries and the Maintenance of Fragility: A Game-Theoretic Theory of Rational Stagnation
by: Hirota, Daisuke
Published: (2025)
by: Hirota, Daisuke
Published: (2025)
Friend or Foe: Delegating to an AI Whose Alignment is Unknown
by: Fudenberg, Drew, et al.
Published: (2025)
by: Fudenberg, Drew, et al.
Published: (2025)
Tell Me Why: Incentivizing Explanations
by: Srinivasan, Siddarth, et al.
Published: (2025)
by: Srinivasan, Siddarth, et al.
Published: (2025)
Understanding EFX Allocations: Counting and Variants
by: Neoh, Tzeh Yuan, et al.
Published: (2025)
by: Neoh, Tzeh Yuan, et al.
Published: (2025)
Collective dynamics of strategic classification
by: Couto, Marta C., et al.
Published: (2025)
by: Couto, Marta C., et al.
Published: (2025)
Misspecified Explore-then-Exploit Leads to Supra-Competitive Prices
by: Baek, Jackie, et al.
Published: (2026)
by: Baek, Jackie, et al.
Published: (2026)
Learned Collusion
by: Compte, Olivier
Published: (2023)
by: Compte, Olivier
Published: (2023)
Extrapolating Volition with Recursive Information Markets
by: Sudhir, Abhimanyu Pallavi, et al.
Published: (2026)
by: Sudhir, Abhimanyu Pallavi, et al.
Published: (2026)
Artificial Intelligence for Multi-Unit Auction design
by: Khezr, Peyman, et al.
Published: (2024)
by: Khezr, Peyman, et al.
Published: (2024)
Post-AGI Economies: Autonomy and the First Fundamental Theorem of Welfare Economics
by: Perrier, Elija
Published: (2026)
by: Perrier, Elija
Published: (2026)
Algorithmic Persuasion Through Simulation
by: Harris, Keegan, et al.
Published: (2023)
by: Harris, Keegan, et al.
Published: (2023)
Dueling Over Dessert, Mastering the Art of Repeated Cake Cutting
by: Brânzei, Simina, et al.
Published: (2024)
by: Brânzei, Simina, et al.
Published: (2024)
Two-Sided Time-Independent Regret for Matching Markets with Limited Interviews
by: Mirfakhar, Amirmahdi, et al.
Published: (2026)
by: Mirfakhar, Amirmahdi, et al.
Published: (2026)
Delegation and Verification Under AI
by: Huang, Lingxiao, et al.
Published: (2026)
by: Huang, Lingxiao, et al.
Published: (2026)
Misspecified learning and evolutionary stability
by: He, Kevin, et al.
Published: (2025)
by: He, Kevin, et al.
Published: (2025)
The Backfiring Effect of Weak AI Safety Regulation
by: Laufer, Benjamin, et al.
Published: (2025)
by: Laufer, Benjamin, et al.
Published: (2025)
Human strategic decision making in parametrized games
by: Ganzfried, Sam
Published: (2021)
by: Ganzfried, Sam
Published: (2021)
Opponent Modeling in Multiplayer Imperfect-Information Games
by: Ganzfried, Sam, et al.
Published: (2022)
by: Ganzfried, Sam, et al.
Published: (2022)
A New Lower Bound for the Random Offerer Mechanism in Bilateral Trade using AI-Guided Evolutionary Search
by: Cai, Yang, et al.
Published: (2026)
by: Cai, Yang, et al.
Published: (2026)
Algorithmic Collusion by Large Language Models
by: Fish, Sara, et al.
Published: (2024)
by: Fish, Sara, et al.
Published: (2024)
Facility Location Games with Scaling Effects
by: He, Yu, et al.
Published: (2024)
by: He, Yu, et al.
Published: (2024)
Generalized Principal-Agent Problem with a Learning Agent
by: Lin, Tao, et al.
Published: (2024)
by: Lin, Tao, et al.
Published: (2024)
Computing Most Equitable Voting Rules
by: Xia, Lirong
Published: (2024)
by: Xia, Lirong
Published: (2024)
Feedback in Dynamic Contests: Theory and Experiment
by: Goel, Sumit, et al.
Published: (2025)
by: Goel, Sumit, et al.
Published: (2025)
Monotonic Mechanisms for Selling Multiple Goods
by: Ben-Moshe, Ran, et al.
Published: (2022)
by: Ben-Moshe, Ran, et al.
Published: (2022)
The Cost of EFX: Generalized-Mean Welfare and Complexity Dichotomies with Few Surplus Items
by: Lim, Eugene, et al.
Published: (2026)
by: Lim, Eugene, et al.
Published: (2026)
Language-based game theory in the age of artificial intelligence
by: Capraro, Valerio, et al.
Published: (2024)
by: Capraro, Valerio, et al.
Published: (2024)
Private Private Information
by: He, Kevin, et al.
Published: (2021)
by: He, Kevin, et al.
Published: (2021)
The Value of Context: Human versus Black Box Evaluators
by: Iakovlev, Andrei, et al.
Published: (2024)
by: Iakovlev, Andrei, et al.
Published: (2024)
A General Framework for a Class of Quarrels: The Quarrelling Paradox Revisited
by: Abizadeh, Arash, et al.
Published: (2022)
by: Abizadeh, Arash, et al.
Published: (2022)
Rank-Guaranteed Auctions
by: He, Wei, et al.
Published: (2024)
by: He, Wei, et al.
Published: (2024)
Generalizing Instant Runoff Voting to Allow Indifferences
by: Delemazure, Théo, et al.
Published: (2024)
by: Delemazure, Théo, et al.
Published: (2024)
Approximate Revenue Maximization for Diffusion Auctions
by: Huang, Yifan, et al.
Published: (2025)
by: Huang, Yifan, et al.
Published: (2025)
Markets with Heterogeneous Agents: Dynamics and Survival of Bayesian vs. No-Regret Learners
by: Easley, David, et al.
Published: (2025)
by: Easley, David, et al.
Published: (2025)
Similar Items
-
Strategic Algorithmic Monoculture: Experimental Evidence from Coordination Games
by: Ballestero, Gonzalo, et al.
Published: (2026) -
A Revealed Preference Framework for AI Alignment
by: Suleymanov, Elchin
Published: (2026) -
Human-AI Productivity Paradoxes: Modeling the Interplay of Skill, Effort, and AI Assistance
by: Aouad, Ali, et al.
Published: (2026) -
Will AI Trade? A Computational Inversion of the No-Trade Theorem
by: Li, Hanyu, et al.
Published: (2025) -
Creative Ownership in the Age of AI
by: Liang, Annie, et al.
Published: (2026)