One Model, Many Morals: Uncovering Cross-Linguistic Misalignments in Computational Moral Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Farid, Sualeha, Lin, Jayden, Chen, Zean, Kumar, Shivani, Jurgens, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Are Rules Meant to be Broken? Understanding Multilingual Moral Reasoning as a Computational Pipeline with UniMoral
von: Kumar, Shivani, et al.
Veröffentlicht: (2025)
von: Kumar, Shivani, et al.
Veröffentlicht: (2025)
The Fellowship of the LLMs: Multi-Model Workflows for Synthetic Preference Optimization Dataset Generation
von: Arif, Samee, et al.
Veröffentlicht: (2024)
von: Arif, Samee, et al.
Veröffentlicht: (2024)
ProMoral-Bench: Evaluating Prompting Strategies for Moral Reasoning and Safety in LLMs
von: Thomas, Rohan Subramanian, et al.
Veröffentlicht: (2026)
von: Thomas, Rohan Subramanian, et al.
Veröffentlicht: (2026)
MoralBench: Moral Evaluation of LLMs
von: Ji, Jianchao, et al.
Veröffentlicht: (2024)
von: Ji, Jianchao, et al.
Veröffentlicht: (2024)
Big Reasoning with Small Models: Instruction Retrieval at Inference Time
von: Alkiek, Kenan, et al.
Veröffentlicht: (2025)
von: Alkiek, Kenan, et al.
Veröffentlicht: (2025)
Uncovering Cross-Linguistic Disparities in LLMs using Sparse Autoencoders
von: Xuan, Richmond Sin Jing, et al.
Veröffentlicht: (2025)
von: Xuan, Richmond Sin Jing, et al.
Veröffentlicht: (2025)
UQA: Corpus for Urdu Question Answering
von: Arif, Samee, et al.
Veröffentlicht: (2024)
von: Arif, Samee, et al.
Veröffentlicht: (2024)
Exploring the psychology of LLMs' Moral and Legal Reasoning
von: Almeida, Guilherme F. C. F., et al.
Veröffentlicht: (2023)
von: Almeida, Guilherme F. C. F., et al.
Veröffentlicht: (2023)
Ethical Reasoning and Moral Value Alignment of LLMs Depend on the Language we Prompt them in
von: Agarwal, Utkarsh, et al.
Veröffentlicht: (2024)
von: Agarwal, Utkarsh, et al.
Veröffentlicht: (2024)
Are Large Language Models Moral Hypocrites? A Study Based on Moral Foundations
von: Nunes, José Luiz, et al.
Veröffentlicht: (2024)
von: Nunes, José Luiz, et al.
Veröffentlicht: (2024)
Histoires Morales: A French Dataset for Assessing Moral Alignment
von: Leteno, Thibaud, et al.
Veröffentlicht: (2025)
von: Leteno, Thibaud, et al.
Veröffentlicht: (2025)
Untangling Input Language from Reasoning Language: A Diagnostic Framework for Cross-Lingual Moral Alignment in LLMs
von: Li, Nan, et al.
Veröffentlicht: (2026)
von: Li, Nan, et al.
Veröffentlicht: (2026)
Residual Connections and the Causal Shift: Uncovering a Structural Misalignment in Transformers
von: Lys, Jonathan, et al.
Veröffentlicht: (2026)
von: Lys, Jonathan, et al.
Veröffentlicht: (2026)
Probabilistic Aggregation and Targeted Embedding Optimization for Collective Moral Reasoning in Large Language Models
von: Yuan, Chenchen, et al.
Veröffentlicht: (2025)
von: Yuan, Chenchen, et al.
Veröffentlicht: (2025)
Understanding Moral Reasoning Trajectories in Large Language Models: Toward Probing-Based Explainability
von: Huang, Fan, et al.
Veröffentlicht: (2026)
von: Huang, Fan, et al.
Veröffentlicht: (2026)
Literary Narrative as Moral Probe : A Cross-System Framework for Evaluating AI Ethical Reasoning and Refusal Behavior
von: Flynn, David C.
Veröffentlicht: (2026)
von: Flynn, David C.
Veröffentlicht: (2026)
Moral Lenses, Political Coordinates: Towards Ideological Positioning of Morally Conditioned LLMs
von: Yuan, Chenchen, et al.
Veröffentlicht: (2026)
von: Yuan, Chenchen, et al.
Veröffentlicht: (2026)
GPT-4's One-Dimensional Mapping of Morality: How the Accuracy of Country-Estimates Depends on Moral Domain
von: Strimling, Pontus, et al.
Veröffentlicht: (2024)
von: Strimling, Pontus, et al.
Veröffentlicht: (2024)
Tracing Moral Foundations in Large Language Models
von: Yu, Chenxiao, et al.
Veröffentlicht: (2026)
von: Yu, Chenxiao, et al.
Veröffentlicht: (2026)
Mechanistic Origin of Moral Indifference in Language Models
von: Li, Lingyu, et al.
Veröffentlicht: (2026)
von: Li, Lingyu, et al.
Veröffentlicht: (2026)
Do Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues Test
von: Khandelwal, Aditi, et al.
Veröffentlicht: (2024)
von: Khandelwal, Aditi, et al.
Veröffentlicht: (2024)
Does Cross-Cultural Alignment Change the Commonsense Morality of Language Models?
von: Jinnai, Yuu
Veröffentlicht: (2024)
von: Jinnai, Yuu
Veröffentlicht: (2024)
Metaphors are a Source of Cross-Domain Misalignment of Large Reasoning Models
von: Hu, Zhibo, et al.
Veröffentlicht: (2026)
von: Hu, Zhibo, et al.
Veröffentlicht: (2026)
Do Language Models Understand Morality? Towards a Robust Detection of Moral Content
von: Bulla, Luana, et al.
Veröffentlicht: (2024)
von: Bulla, Luana, et al.
Veröffentlicht: (2024)
Morality is Non-Binary: Building a Pluralist Moral Sentence Embedding Space using Contrastive Learning
von: Park, Jeongwoo, et al.
Veröffentlicht: (2024)
von: Park, Jeongwoo, et al.
Veröffentlicht: (2024)
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment
von: Vida, Karina, et al.
Veröffentlicht: (2024)
von: Vida, Karina, et al.
Veröffentlicht: (2024)
Do LLMs Understand Social Knowledge? Evaluating the Sociability of Large Language Models with SocKET Benchmark
von: Choi, Minje, et al.
Veröffentlicht: (2023)
von: Choi, Minje, et al.
Veröffentlicht: (2023)
Exploring Cultural Variations in Moral Judgments with Large Language Models
von: Mohammadi, Hadi, et al.
Veröffentlicht: (2025)
von: Mohammadi, Hadi, et al.
Veröffentlicht: (2025)
SaGE: Evaluating Moral Consistency in Large Language Models
von: Bonagiri, Vamshi Krishna, et al.
Veröffentlicht: (2024)
von: Bonagiri, Vamshi Krishna, et al.
Veröffentlicht: (2024)
Moral Mazes in the Era of LLMs
von: Nguyen, Dang, et al.
Veröffentlicht: (2026)
von: Nguyen, Dang, et al.
Veröffentlicht: (2026)
Large Language Models as Mirrors of Societal Moral Standards
von: Papadopoulou, Evi, et al.
Veröffentlicht: (2024)
von: Papadopoulou, Evi, et al.
Veröffentlicht: (2024)
Inductive Linguistic Reasoning with Large Language Models
von: Ramji, Raghav, et al.
Veröffentlicht: (2024)
von: Ramji, Raghav, et al.
Veröffentlicht: (2024)
The Moral Consistency Pipeline: Continuous Ethical Evaluation for Large Language Models
von: Jamshidi, Saeid, et al.
Veröffentlicht: (2025)
von: Jamshidi, Saeid, et al.
Veröffentlicht: (2025)
Thought Crime: Backdoors and Emergent Misalignment in Reasoning Models
von: Chua, James, et al.
Veröffentlicht: (2025)
von: Chua, James, et al.
Veröffentlicht: (2025)
Moral Persuasion in Large Language Models: Evaluating Susceptibility and Ethical Alignment
von: Huang, Allison, et al.
Veröffentlicht: (2024)
von: Huang, Allison, et al.
Veröffentlicht: (2024)
Measuring Moral LLM Responses in Multilingual Capacities
von: Basu, Kimaya, et al.
Veröffentlicht: (2025)
von: Basu, Kimaya, et al.
Veröffentlicht: (2025)
GreedLlama: Performance of Financial Value-Aligned Large Language Models in Moral Reasoning
von: Yu, Jeffy, et al.
Veröffentlicht: (2024)
von: Yu, Jeffy, et al.
Veröffentlicht: (2024)
Morality is Contextual: Learning Interpretable Moral Contexts from Human Data with Probabilistic Clustering and Large Language Models
von: Morlat, Geoffroy, et al.
Veröffentlicht: (2025)
von: Morlat, Geoffroy, et al.
Veröffentlicht: (2025)
Parallel Scaling Law: Unveiling Reasoning Generalization through A Cross-Linguistic Perspective
von: Yang, Wen, et al.
Veröffentlicht: (2025)
von: Yang, Wen, et al.
Veröffentlicht: (2025)
Modeling the One-to-Many Property in Open-Domain Dialogue with LLMs
von: Lee, Jing Yang, et al.
Veröffentlicht: (2025)
von: Lee, Jing Yang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Are Rules Meant to be Broken? Understanding Multilingual Moral Reasoning as a Computational Pipeline with UniMoral
von: Kumar, Shivani, et al.
Veröffentlicht: (2025) -
The Fellowship of the LLMs: Multi-Model Workflows for Synthetic Preference Optimization Dataset Generation
von: Arif, Samee, et al.
Veröffentlicht: (2024) -
ProMoral-Bench: Evaluating Prompting Strategies for Moral Reasoning and Safety in LLMs
von: Thomas, Rohan Subramanian, et al.
Veröffentlicht: (2026) -
MoralBench: Moral Evaluation of LLMs
von: Ji, Jianchao, et al.
Veröffentlicht: (2024) -
Big Reasoning with Small Models: Instruction Retrieval at Inference Time
von: Alkiek, Kenan, et al.
Veröffentlicht: (2025)