A Decision-Theoretic Approach for Managing Misalignment
Fuente:
arXiv
Guardado en:
| Autores principales: | Herrmann, Daniel A., Chari, Abinav, Qian, Isabelle, Sharvesh, Sree, Levinstein, B. A. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Pay for The Second-Best Service: A Game-Theoretic Approach Against Dishonest LLM Providers
por: Cao, Yuhan, et al.
Publicado: (2025)
por: Cao, Yuhan, et al.
Publicado: (2025)
Incentives, Equilibria, and the Limits of Healthcare AI: A Game-Theoretic Perspective
por: Ercole, Ari
Publicado: (2026)
por: Ercole, Ari
Publicado: (2026)
A Learning Framework for Distribution-Based Game-Theoretic Solution Concepts
por: Jha, Tushant, et al.
Publicado: (2019)
por: Jha, Tushant, et al.
Publicado: (2019)
Linear Social Choice with Few Queries: A Moment-Based Approach
por: Ge, Luise, et al.
Publicado: (2026)
por: Ge, Luise, et al.
Publicado: (2026)
Responsibility Gap in Collective Decision Making
por: Naumov, Pavel, et al.
Publicado: (2025)
por: Naumov, Pavel, et al.
Publicado: (2025)
The End Justifies the Mean: A Linear Ranking Rule for Proportional Sequential Decisions
por: Baharav, Carmel, et al.
Publicado: (2026)
por: Baharav, Carmel, et al.
Publicado: (2026)
Algorithmic Decision-Making under Agents with Persistent Improvement
por: Xie, Tian, et al.
Publicado: (2024)
por: Xie, Tian, et al.
Publicado: (2024)
Hypergame Rationalisability: Solving Agent Misalignment In Strategic Play
por: Trencsenyi, Vince
Publicado: (2025)
por: Trencsenyi, Vince
Publicado: (2025)
Bidding Games on Markov Decision Processes with Quantitative Reachability Objectives
por: Avni, Guy, et al.
Publicado: (2024)
por: Avni, Guy, et al.
Publicado: (2024)
Cooperation Dynamics in Multi-Agent Systems: Exploring Game-Theoretic Scenarios with Mean-Field Equilibria
por: Sathi, Vaigarai, et al.
Publicado: (2023)
por: Sathi, Vaigarai, et al.
Publicado: (2023)
Beyond Nash Equilibrium: Bounded Rationality of LLMs and humans in Strategic Decision-making
por: Zheng, Kehan, et al.
Publicado: (2025)
por: Zheng, Kehan, et al.
Publicado: (2025)
MR-LDM -- The Merge-Reactive Longitudinal Decision Model: Game Theoretic Human Decision Modeling for Interactive Sim Agents
por: Holley, Dustin, et al.
Publicado: (2025)
por: Holley, Dustin, et al.
Publicado: (2025)
An Empirical Game-Theoretic Analysis of Autonomous Cyber-Defence Agents
por: Palmer, Gregory, et al.
Publicado: (2025)
por: Palmer, Gregory, et al.
Publicado: (2025)
A Constraint Programming Approach to Fair High School Course Scheduling
por: Kiyohara, Mitsuka, et al.
Publicado: (2024)
por: Kiyohara, Mitsuka, et al.
Publicado: (2024)
EconEvals: Benchmarks and Litmus Tests for Economic Decision-Making by LLM Agents
por: Fish, Sara, et al.
Publicado: (2025)
por: Fish, Sara, et al.
Publicado: (2025)
A Game-Theoretic Negotiation Framework for Cross-Cultural Consensus in LLMs
por: Zhang, Guoxi, et al.
Publicado: (2025)
por: Zhang, Guoxi, et al.
Publicado: (2025)
Steering Language Models with Game-Theoretic Solvers
por: Gemp, Ian, et al.
Publicado: (2024)
por: Gemp, Ian, et al.
Publicado: (2024)
Beyond Right to be Forgotten: Managing Heterogeneity Side Effects Through Strategic Incentives
por: Shao, Jiaqi, et al.
Publicado: (2024)
por: Shao, Jiaqi, et al.
Publicado: (2024)
Automated Approach for Solving Infinite-state Polynomial Reachability Games
por: Chatterjee, Krishnendu, et al.
Publicado: (2026)
por: Chatterjee, Krishnendu, et al.
Publicado: (2026)
Tiny Multi-Agent DRL for Twins Migration in UAV Metaverses: A Multi-Leader Multi-Follower Stackelberg Game Approach
por: Kang, Jiawen, et al.
Publicado: (2024)
por: Kang, Jiawen, et al.
Publicado: (2024)
Is Four Enough? Automated Reasoning Approaches and Dual Bounds for Condorcet Dimensions of Elections
por: Zilberstein, Itai, et al.
Publicado: (2026)
por: Zilberstein, Itai, et al.
Publicado: (2026)
CERN for AI: A Theoretical Framework for Autonomous Simulation-Based Artificial Intelligence Testing and Alignment
por: Bojic, Ljubisa, et al.
Publicado: (2023)
por: Bojic, Ljubisa, et al.
Publicado: (2023)
Categorical Approach to Conflict Resolution: Integrating Category Theory into the Graph Model for Conflict Resolution
por: Kato, Yukiko
Publicado: (2023)
por: Kato, Yukiko
Publicado: (2023)
AI of the People, by the People, for the People: A Social Choice Approach to Collective Control of Artificial Intelligence
por: Bachmann, Paul Anton, et al.
Publicado: (2026)
por: Bachmann, Paul Anton, et al.
Publicado: (2026)
Probabilistic Neuro-Symbolic Reasoning for Sparse Historical Data: A Framework Integrating Bayesian Inference, Causal Models, and Game-Theoretic Allocation
por: Kublashvili, Saba
Publicado: (2025)
por: Kublashvili, Saba
Publicado: (2025)
A Graph-Theoretical Perspective on Law Design for Multiagent Systems
por: Shi, Qi, et al.
Publicado: (2025)
por: Shi, Qi, et al.
Publicado: (2025)
Framing the Game: How Context Shapes LLM Decision-Making
por: Robinson, Isaac, et al.
Publicado: (2025)
por: Robinson, Isaac, et al.
Publicado: (2025)
Decision-Making on Timing and Route Selection: A Game-Theoretic Approach
por: Wang, Chenlan, et al.
Publicado: (2025)
por: Wang, Chenlan, et al.
Publicado: (2025)
A Computable Game-Theoretic Framework for Multi-Agent Theory of Mind
por: Zhu, Fengming, et al.
Publicado: (2025)
por: Zhu, Fengming, et al.
Publicado: (2025)
Computing Voting Rules with Elicited Incomplete Votes
por: Halpern, Daniel, et al.
Publicado: (2024)
por: Halpern, Daniel, et al.
Publicado: (2024)
Rational Adversaries and the Maintenance of Fragility: A Game-Theoretic Theory of Rational Stagnation
por: Hirota, Daisuke
Publicado: (2025)
por: Hirota, Daisuke
Publicado: (2025)
Diffusion of Responsibility in Collective Decision Making
por: Naumov, Pavel, et al.
Publicado: (2025)
por: Naumov, Pavel, et al.
Publicado: (2025)
Are Large Language Models Strategic Decision Makers? A Study of Performance and Bias in Two-Player Non-Zero-Sum Games
por: Herr, Nathan, et al.
Publicado: (2024)
por: Herr, Nathan, et al.
Publicado: (2024)
A Framework for Adversarial Analysis of Decision Support Systems Prior to Deployment
por: Bissey, Brett, et al.
Publicado: (2025)
por: Bissey, Brett, et al.
Publicado: (2025)
ALYMPICS: LLM Agents Meet Game Theory -- Exploring Strategic Decision-Making with AI Agents
por: Mao, Shaoguang, et al.
Publicado: (2023)
por: Mao, Shaoguang, et al.
Publicado: (2023)
Robust Reward Design for Markov Decision Processes
por: Wu, Shuo, et al.
Publicado: (2024)
por: Wu, Shuo, et al.
Publicado: (2024)
ADAPT: A Game-Theoretic and Neuro-Symbolic Framework for Automated Distributed Adaptive Penetration Testing
por: Lei, Haozhe, et al.
Publicado: (2024)
por: Lei, Haozhe, et al.
Publicado: (2024)
Can Media Act as a Soft Regulator of Safe AI Development? A Game Theoretical Analysis
por: da Fonseca, Henrique Correia, et al.
Publicado: (2025)
por: da Fonseca, Henrique Correia, et al.
Publicado: (2025)
Do LLMs Strategically Reveal, Conceal, and Infer Information? A Theoretical and Empirical Analysis in The Chameleon Game
por: Karabag, Mustafa O., et al.
Publicado: (2025)
por: Karabag, Mustafa O., et al.
Publicado: (2025)
Information Bargaining: Bilateral Commitment in Bayesian Persuasion
por: Lin, Yue, et al.
Publicado: (2025)
por: Lin, Yue, et al.
Publicado: (2025)
Ejemplares similares
-
Pay for The Second-Best Service: A Game-Theoretic Approach Against Dishonest LLM Providers
por: Cao, Yuhan, et al.
Publicado: (2025) -
Incentives, Equilibria, and the Limits of Healthcare AI: A Game-Theoretic Perspective
por: Ercole, Ari
Publicado: (2026) -
A Learning Framework for Distribution-Based Game-Theoretic Solution Concepts
por: Jha, Tushant, et al.
Publicado: (2019) -
Linear Social Choice with Few Queries: A Moment-Based Approach
por: Ge, Luise, et al.
Publicado: (2026) -
Responsibility Gap in Collective Decision Making
por: Naumov, Pavel, et al.
Publicado: (2025)