When Is It Acceptable to Break the Rules? Knowledge Representation of Moral Judgement Based on Empirical Data
Fuente:
arXiv
Saved in:
| Main Authors: | Awad, Edmond, Levine, Sydney, Loreggia, Andrea, Mattei, Nicholas, Rahwan, Iyad, Rossi, Francesca, Talamadupula, Kartik, Tenenbaum, Joshua, Kleiman-Weiner, Max |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sycophantic Chatbots Cause Delusional Spiraling, Even in Ideal Bayesians
by: Chandra, Kartik, et al.
Published: (2026)
by: Chandra, Kartik, et al.
Published: (2026)
When Empowerment Disempowers
by: Yang, Claire, et al.
Published: (2025)
by: Yang, Claire, et al.
Published: (2025)
"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer?
by: Reichman, Benjamin, et al.
Published: (2025)
by: Reichman, Benjamin, et al.
Published: (2025)
Die Business Judgement Rule
by: Willen, Max
Published: (2020)
by: Willen, Max
Published: (2020)
Value Internalization: Learning and Generalizing from Social Reward
by: Rong, Frieda, et al.
Published: (2024)
by: Rong, Frieda, et al.
Published: (2024)
Evaluating LLMs in Open-Source Games
by: Sistla, Swadesh, et al.
Published: (2025)
by: Sistla, Swadesh, et al.
Published: (2025)
Resource‐Rational Virtual Bargaining for Moral Judgment: Toward a Probabilistic Cognitive Model
by: Diego Trujillo, et al.
Published: (2025)
by: Diego Trujillo, et al.
Published: (2025)
Mutual benefits of social learning and algorithmic mediation for cumulative culture
by: Czaplicka, Agnieszka, et al.
Published: (2024)
by: Czaplicka, Agnieszka, et al.
Published: (2024)
Quantum Transfer Learning for Acceptability Judgements
by: Buonaiuto, Giuseppe, et al.
Published: (2024)
by: Buonaiuto, Giuseppe, et al.
Published: (2024)
Boundedly Rational Meta-Learning in Sequential Consumer Choice
by: Khosravi, Mehrzad, et al.
Published: (2026)
by: Khosravi, Mehrzad, et al.
Published: (2026)
Estimating the Empowerment of Language Model Agents
by: Song, Jinyeop, et al.
Published: (2025)
by: Song, Jinyeop, et al.
Published: (2025)
Are Language Models Consequentialist or Deontological Moral Reasoners?
by: Samway, Keenan, et al.
Published: (2025)
by: Samway, Keenan, et al.
Published: (2025)
False consensus biases AI against vulnerable stakeholders
by: Dong, Mengchen, et al.
Published: (2024)
by: Dong, Mengchen, et al.
Published: (2024)
The Science Fiction Science Method
by: Rahwan, Iyad, et al.
Published: (2025)
by: Rahwan, Iyad, et al.
Published: (2025)
Concise Reasoning via Reinforcement Learning
by: Fatemi, Mehdi, et al.
Published: (2025)
by: Fatemi, Mehdi, et al.
Published: (2025)
Dynamics of Algorithmic Content Amplification on TikTok
by: Baumann, Fabian, et al.
Published: (2025)
by: Baumann, Fabian, et al.
Published: (2025)
Preserving Sense of Agency: User Preferences for Robot Autonomy and User Control across Household Tasks
by: Yang, Claire, et al.
Published: (2025)
by: Yang, Claire, et al.
Published: (2025)
Empirical evidence of Large Language Model's influence on human spoken communication
by: Yakura, Hiromu, et al.
Published: (2024)
by: Yakura, Hiromu, et al.
Published: (2024)
Group Selection as a Safeguard Against AI Substitution
by: Zhong, Qiankun, et al.
Published: (2026)
by: Zhong, Qiankun, et al.
Published: (2026)
The Lock-in Hypothesis: Stagnation by Algorithm
by: Qiu, Tianyi Alex, et al.
Published: (2025)
by: Qiu, Tianyi Alex, et al.
Published: (2025)
Are Human Conversations Special? A Large Language Model Perspective
by: Jawale, Toshish, et al.
Published: (2024)
by: Jawale, Toshish, et al.
Published: (2024)
Emotional RAG LLMs: Reading Comprehension for the Open Internet
by: Reichman, Benjamin, et al.
Published: (2024)
by: Reichman, Benjamin, et al.
Published: (2024)
The Role of Social Learning and Collective Norm Formation in Fostering Cooperation in LLM Multi-Agent Systems
by: Gupta, Prateek, et al.
Published: (2025)
by: Gupta, Prateek, et al.
Published: (2025)
From Human Judgements to Predictive Models: Unravelling Acceptability in Code-Mixed Sentences
by: Kodali, Prashant, et al.
Published: (2024)
by: Kodali, Prashant, et al.
Published: (2024)
Experimental Evidence That Conversational Artificial Intelligence Can Steer Consumer Behavior Without Detection
by: Werner, Tobias, et al.
Published: (2024)
by: Werner, Tobias, et al.
Published: (2024)
Recognising, Anticipating, and Mitigating LLM Pollution of Online Behavioural Research
by: Rilla, Raluca, et al.
Published: (2025)
by: Rilla, Raluca, et al.
Published: (2025)
Modeling Others' Minds as Code
by: Jha, Kunal, et al.
Published: (2025)
by: Jha, Kunal, et al.
Published: (2025)
Evaluating Gender Bias of LLMs in Making Morality Judgements
by: Bajaj, Divij, et al.
Published: (2024)
by: Bajaj, Divij, et al.
Published: (2024)
A Research in China Based on the Moral Judgement Test
by: Shaogang Yang
Published: (2012)
by: Shaogang Yang
Published: (2012)
Language Model Alignment in Multilingual Trolley Problems
by: Jin, Zhijing, et al.
Published: (2024)
by: Jin, Zhijing, et al.
Published: (2024)
SafetyAnalyst: Interpretable, Transparent, and Steerable Safety Moderation for AI Behavior
by: Li, Jing-Jing, et al.
Published: (2024)
by: Li, Jing-Jing, et al.
Published: (2024)
EXPLORER: Exploration-guided Reasoning for Textual Reinforcement Learning
by: Basu, Kinjal, et al.
Published: (2024)
by: Basu, Kinjal, et al.
Published: (2024)
Imagining and building wise machines: The centrality of AI metacognition
by: Johnson, Samuel G. B., et al.
Published: (2024)
by: Johnson, Samuel G. B., et al.
Published: (2024)
Cross-environment Cooperation Enables Zero-shot Multi-agent Coordination
by: Jha, Kunal, et al.
Published: (2025)
by: Jha, Kunal, et al.
Published: (2025)
Cooperate or Collapse: Emergence of Sustainable Cooperation in a Society of LLM Agents
by: Piatti, Giorgio, et al.
Published: (2024)
by: Piatti, Giorgio, et al.
Published: (2024)
Generative Value Conflicts Reveal LLM Priorities
by: Liu, Andy, et al.
Published: (2025)
by: Liu, Andy, et al.
Published: (2025)
Understanding the Progression of Educational Topics via Semantic Matching
by: Alkhidir, Tamador, et al.
Published: (2024)
by: Alkhidir, Tamador, et al.
Published: (2024)
Moral Judgement on the Use of Market System Ideology in Social Domains
by: Kenji Noguchi, et al.
Published: (2025)
by: Kenji Noguchi, et al.
Published: (2025)
Conniving With Continuations: Representing Goals in a Domain‐Specific Language of Thought
by: Kartik Chandra, et al.
Published: (2026)
by: Kartik Chandra, et al.
Published: (2026)
When Autonomy Breaks: The Hidden Existential Risk of AI
by: Krook, Joshua
Published: (2025)
by: Krook, Joshua
Published: (2025)
Similar Items
-
Sycophantic Chatbots Cause Delusional Spiraling, Even in Ideal Bayesians
by: Chandra, Kartik, et al.
Published: (2026) -
When Empowerment Disempowers
by: Yang, Claire, et al.
Published: (2025) -
"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer?
by: Reichman, Benjamin, et al.
Published: (2025) -
Die Business Judgement Rule
by: Willen, Max
Published: (2020) -
Value Internalization: Learning and Generalizing from Social Reward
by: Rong, Frieda, et al.
Published: (2024)