When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Backmann, Steffen, Piedrahita, David Guzman, Tewolde, Emanuel, Mihalcea, Rada, Schölkopf, Bernhard, Jin, Zhijing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas
von: Tewolde, Emanuel, et al.
Veröffentlicht: (2026)
von: Tewolde, Emanuel, et al.
Veröffentlicht: (2026)
Are Language Models Consequentialist or Deontological Moral Reasoners?
von: Samway, Keenan, et al.
Veröffentlicht: (2025)
von: Samway, Keenan, et al.
Veröffentlicht: (2025)
Democratic or Authoritarian? Probing a New Dimension of Political Biases in Large Language Models
von: Piedrahita, David Guzman, et al.
Veröffentlicht: (2025)
von: Piedrahita, David Guzman, et al.
Veröffentlicht: (2025)
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
von: Borah, Angana, et al.
Veröffentlicht: (2024)
von: Borah, Angana, et al.
Veröffentlicht: (2024)
Preserving Historical Truth: Detecting Historical Revisionism in Large Language Models
von: Ortu, Francesco, et al.
Veröffentlicht: (2026)
von: Ortu, Francesco, et al.
Veröffentlicht: (2026)
Cooperate or Collapse: Emergence of Sustainable Cooperation in a Society of LLM Agents
von: Piatti, Giorgio, et al.
Veröffentlicht: (2024)
von: Piatti, Giorgio, et al.
Veröffentlicht: (2024)
Implicit Personalization in Language Models: A Systematic Study
von: Jin, Zhijing, et al.
Veröffentlicht: (2024)
von: Jin, Zhijing, et al.
Veröffentlicht: (2024)
When Do Language Models Endorse Limitations on Human Rights Principles?
von: Samway, Keenan, et al.
Veröffentlicht: (2026)
von: Samway, Keenan, et al.
Veröffentlicht: (2026)
Do LLMs Think Fast and Slow? A Causal Study on Sentiment Analysis
von: Lyu, Zhiheng, et al.
Veröffentlicht: (2024)
von: Lyu, Zhiheng, et al.
Veröffentlicht: (2024)
The Staircase of Ethics: Probing LLM Value Priorities through Multi-Step Induction to Complex Moral Dilemmas
von: Wu, Ya, et al.
Veröffentlicht: (2025)
von: Wu, Ya, et al.
Veröffentlicht: (2025)
Are LLMs Good Safety Agents or a Propaganda Engine?
von: Yadav, Neemesh, et al.
Veröffentlicht: (2025)
von: Yadav, Neemesh, et al.
Veröffentlicht: (2025)
The Curious Case of Curiosity across Human Cultures and LLMs
von: Borah, Angana, et al.
Veröffentlicht: (2025)
von: Borah, Angana, et al.
Veröffentlicht: (2025)
Exploring the Jungle of Bias: Political Bias Attribution in Language Models via Dependency Analysis
von: Jenny, David F., et al.
Veröffentlicht: (2023)
von: Jenny, David F., et al.
Veröffentlicht: (2023)
One Word at a Time: Incremental Completion Decomposition Breaks LLM Safety
von: Arif, Samee, et al.
Veröffentlicht: (2026)
von: Arif, Samee, et al.
Veröffentlicht: (2026)
Why AI Is WEIRD and Should Not Be This Way: Towards AI For Everyone, With Everyone, By Everyone
von: Mihalcea, Rada, et al.
Veröffentlicht: (2024)
von: Mihalcea, Rada, et al.
Veröffentlicht: (2024)
Simulating Ethics: Using LLM Debate Panels to Model Deliberation on Medical Dilemmas
von: Zohny, Hazem
Veröffentlicht: (2025)
von: Zohny, Hazem
Veröffentlicht: (2025)
SocialHarmBench: Revealing LLM Vulnerabilities to Socially Harmful Requests
von: Pandey, Punya Syon, et al.
Veröffentlicht: (2025)
von: Pandey, Punya Syon, et al.
Veröffentlicht: (2025)
Corrupted by Reasoning: Reasoning Language Models Become Free-Riders in Public Goods Games
von: Piedrahita, David Guzman, et al.
Veröffentlicht: (2025)
von: Piedrahita, David Guzman, et al.
Veröffentlicht: (2025)
Towards Algorithmic Fidelity: Mental Health Representation across Demographics in Synthetic vs. Human-generated Data
von: Mori, Shinka, et al.
Veröffentlicht: (2024)
von: Mori, Shinka, et al.
Veröffentlicht: (2024)
How Robust Are Router-LLMs? Analysis of the Fragility of LLM Routing Capabilities
von: Kassem, Aly M., et al.
Veröffentlicht: (2025)
von: Kassem, Aly M., et al.
Veröffentlicht: (2025)
Which Humans? Inclusivity and Representation in Human-Centered AI
von: Mihalcea, Rada, et al.
Veröffentlicht: (2025)
von: Mihalcea, Rada, et al.
Veröffentlicht: (2025)
Future of Pandemic Prevention and Response CCC Workshop Report
von: Danks, David, et al.
Veröffentlicht: (2024)
von: Danks, David, et al.
Veröffentlicht: (2024)
Human Action Co-occurrence in Lifestyle Vlogs using Graph Link Prediction
von: Ignat, Oana, et al.
Veröffentlicht: (2023)
von: Ignat, Oana, et al.
Veröffentlicht: (2023)
Social Catalysts, Not Moral Agents: The Illusion of Alignment in LLM Societies
von: Hu, Yueqing, et al.
Veröffentlicht: (2026)
von: Hu, Yueqing, et al.
Veröffentlicht: (2026)
Can Large Language Models Infer Causation from Correlation?
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
von: Jin, Zhijing, et al.
Veröffentlicht: (2023)
CausalCite: A Causal Formulation of Paper Citations
von: Kumar, Ishan, et al.
Veröffentlicht: (2023)
von: Kumar, Ishan, et al.
Veröffentlicht: (2023)
Uplifting Lower-Income Data: Strategies for Socioeconomic Perspective Shifts in Large Multi-modal Models
von: Nwatu, Joan, et al.
Veröffentlicht: (2024)
von: Nwatu, Joan, et al.
Veröffentlicht: (2024)
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
von: Jiao, Junfeng, et al.
Veröffentlicht: (2025)
Causality for Natural Language Processing
von: Jin, Zhijing
Veröffentlicht: (2025)
von: Jin, Zhijing
Veröffentlicht: (2025)
Culture Affordance Atlas: Reconciling Object Diversity Through Functional Mapping
von: Nwatu, Joan, et al.
Veröffentlicht: (2025)
von: Nwatu, Joan, et al.
Veröffentlicht: (2025)
From Hallucination to Scheming: A Unified Taxonomy and Benchmark Analysis for LLM Deception
von: Shi, Jerick, et al.
Veröffentlicht: (2026)
von: Shi, Jerick, et al.
Veröffentlicht: (2026)
Social Welfare Function Leaderboard: When LLM Agents Allocate Social Welfare
von: Shi, Zhengliang, et al.
Veröffentlicht: (2025)
von: Shi, Zhengliang, et al.
Veröffentlicht: (2025)
Causal Responsibility Attribution for Human-AI Collaboration
von: Qi, Yahang, et al.
Veröffentlicht: (2024)
von: Qi, Yahang, et al.
Veröffentlicht: (2024)
Evaluating Cooperation in LLM Social Groups through Elected Leadership
von: Faulkner, Ryan, et al.
Veröffentlicht: (2026)
von: Faulkner, Ryan, et al.
Veröffentlicht: (2026)
Are Human Interactions Replicable by Generative Agents? A Case Study on Pronoun Usage in Hierarchical Interactions
von: Deng, Naihao, et al.
Veröffentlicht: (2025)
von: Deng, Naihao, et al.
Veröffentlicht: (2025)
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models
von: Choi, Younwoo, et al.
Veröffentlicht: (2025)
von: Choi, Younwoo, et al.
Veröffentlicht: (2025)
How Do AI Agents Spend Your Money? Analyzing and Predicting Token Consumption in Agentic Coding Tasks
von: Bai, Longju, et al.
Veröffentlicht: (2026)
von: Bai, Longju, et al.
Veröffentlicht: (2026)
Quriosity: Analyzing Human Questioning Behavior and Causal Inquiry through Curiosity-Driven Queries
von: Ceraolo, Roberto, et al.
Veröffentlicht: (2024)
von: Ceraolo, Roberto, et al.
Veröffentlicht: (2024)
The Parrot Dilemma: Human-Labeled vs. LLM-augmented Data in Classification Tasks
von: Møller, Anders Giovanni, et al.
Veröffentlicht: (2023)
von: Møller, Anders Giovanni, et al.
Veröffentlicht: (2023)
Normative Evaluation of Large Language Models with Everyday Moral Dilemmas
von: Sachdeva, Pratik S., et al.
Veröffentlicht: (2025)
von: Sachdeva, Pratik S., et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas
von: Tewolde, Emanuel, et al.
Veröffentlicht: (2026) -
Are Language Models Consequentialist or Deontological Moral Reasoners?
von: Samway, Keenan, et al.
Veröffentlicht: (2025) -
Democratic or Authoritarian? Probing a New Dimension of Political Biases in Large Language Models
von: Piedrahita, David Guzman, et al.
Veröffentlicht: (2025) -
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
von: Borah, Angana, et al.
Veröffentlicht: (2024) -
Preserving Historical Truth: Detecting Historical Revisionism in Large Language Models
von: Ortu, Francesco, et al.
Veröffentlicht: (2026)