Moral Mazes in the Era of LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nguyen, Dang, Fu, Harvey Yiyun, West, Peter, Holtzman, Ari, Tan, Chenhao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Know Thyself? On the Incapability and Implications of AI Self-Recognition
von: Bai, Xiaoyan, et al.
Veröffentlicht: (2025)
von: Bai, Xiaoyan, et al.
Veröffentlicht: (2025)
The Story is Not the Science: Execution-Grounded Evaluation of Mechanistic Interpretability Research
von: Bai, Xiaoyan, et al.
Veröffentlicht: (2026)
von: Bai, Xiaoyan, et al.
Veröffentlicht: (2026)
On the Effectiveness and Generalization of Race Representations for Debiasing High-Stakes Decisions
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
AbsenceBench: Language Models Can't Tell What's Missing
von: Fu, Harvey Yiyun, et al.
Veröffentlicht: (2025)
von: Fu, Harvey Yiyun, et al.
Veröffentlicht: (2025)
GPT-4V Cannot Generate Radiology Reports Yet
von: Jiang, Yuyang, et al.
Veröffentlicht: (2024)
von: Jiang, Yuyang, et al.
Veröffentlicht: (2024)
The Text Uncanny Valley: Non-Monotonic Performance Degradation in LLM Information Retrieval
von: Tong, Zekai, et al.
Veröffentlicht: (2026)
von: Tong, Zekai, et al.
Veröffentlicht: (2026)
AI as Entertainment
von: Kommers, Cody, et al.
Veröffentlicht: (2026)
von: Kommers, Cody, et al.
Veröffentlicht: (2026)
Predicting vs. Acting: A Trade-off Between World Modeling & Agent Modeling
von: Li, Margaret, et al.
Veröffentlicht: (2024)
von: Li, Margaret, et al.
Veröffentlicht: (2024)
Widespread Gender and Pronoun Bias in Moral Judgments Across LLMs
von: Fernandes, Gustavo Lúcius, et al.
Veröffentlicht: (2026)
von: Fernandes, Gustavo Lúcius, et al.
Veröffentlicht: (2026)
Linearly Decoding Refused Knowledge in Aligned Language Models
von: Shrivastava, Aryan, et al.
Veröffentlicht: (2025)
von: Shrivastava, Aryan, et al.
Veröffentlicht: (2025)
Why Slop Matters
von: Kommers, Cody, et al.
Veröffentlicht: (2025)
von: Kommers, Cody, et al.
Veröffentlicht: (2025)
Prompting as Scientific Inquiry
von: Holtzman, Ari, et al.
Veröffentlicht: (2025)
von: Holtzman, Ari, et al.
Veröffentlicht: (2025)
Wikipedia in the Era of LLMs: Evolution and Risks
von: Huang, Siming, et al.
Veröffentlicht: (2025)
von: Huang, Siming, et al.
Veröffentlicht: (2025)
Language of Bargaining
von: Heddaya, Mourad, et al.
Veröffentlicht: (2023)
von: Heddaya, Mourad, et al.
Veröffentlicht: (2023)
On Wednesdays, We Ask Questions: Optimizing "Active Listening" in Automated Legal Triage and Referral
von: Steenhuis, Quinten, et al.
Veröffentlicht: (2026)
von: Steenhuis, Quinten, et al.
Veröffentlicht: (2026)
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment
von: Vida, Karina, et al.
Veröffentlicht: (2024)
von: Vida, Karina, et al.
Veröffentlicht: (2024)
HypoBench: Towards Systematic and Principled Benchmarking for Hypothesis Generation
von: Liu, Haokun, et al.
Veröffentlicht: (2025)
von: Liu, Haokun, et al.
Veröffentlicht: (2025)
Literature Meets Data: A Synergistic Approach to Hypothesis Generation
von: Liu, Haokun, et al.
Veröffentlicht: (2024)
von: Liu, Haokun, et al.
Veröffentlicht: (2024)
Hypothesis Generation with Large Language Models
von: Zhou, Yangqiaoyu, et al.
Veröffentlicht: (2024)
von: Zhou, Yangqiaoyu, et al.
Veröffentlicht: (2024)
Writing in Symbiosis: Mapping Human Creative Agency in the AI Era
von: Doshi, Vivan, et al.
Veröffentlicht: (2025)
von: Doshi, Vivan, et al.
Veröffentlicht: (2025)
Gender Bias in Machine Translation and The Era of Large Language Models
von: Vanmassenhove, Eva
Veröffentlicht: (2024)
von: Vanmassenhove, Eva
Veröffentlicht: (2024)
Dropouts in Confidence: Moral Uncertainty in Human-LLM Alignment
von: Kwon, Jea, et al.
Veröffentlicht: (2025)
von: Kwon, Jea, et al.
Veröffentlicht: (2025)
CPsyCoun: A Report-based Multi-turn Dialogue Reconstruction and Evaluation Framework for Chinese Psychological Counseling
von: Zhang, Chenhao, et al.
Veröffentlicht: (2024)
von: Zhang, Chenhao, et al.
Veröffentlicht: (2024)
CaseSumm: A Large-Scale Dataset for Long-Context Summarization from U.S. Supreme Court Opinions
von: Heddaya, Mourad, et al.
Veröffentlicht: (2024)
von: Heddaya, Mourad, et al.
Veröffentlicht: (2024)
Right to be Forgotten in the Era of Large Language Models: Implications, Challenges, and Solutions
von: Zhang, Dawen, et al.
Veröffentlicht: (2023)
von: Zhang, Dawen, et al.
Veröffentlicht: (2023)
Exploring Persona-dependent LLM Alignment for the Moral Machine Experiment
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
Attributions toward Artificial Agents in a modified Moral Turing Test
von: Aharoni, Eyal, et al.
Veröffentlicht: (2024)
von: Aharoni, Eyal, et al.
Veröffentlicht: (2024)
The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era
von: Jadhav, Rudra, et al.
Veröffentlicht: (2026)
von: Jadhav, Rudra, et al.
Veröffentlicht: (2026)
Moral Susceptibility and Robustness under Persona Role-Play in Large Language Models
von: Costa, Davi Bastos, et al.
Veröffentlicht: (2025)
von: Costa, Davi Bastos, et al.
Veröffentlicht: (2025)
When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
von: Backmann, Steffen, et al.
Veröffentlicht: (2025)
von: Backmann, Steffen, et al.
Veröffentlicht: (2025)
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
von: Xu, Shanshan, et al.
Veröffentlicht: (2025)
von: Xu, Shanshan, et al.
Veröffentlicht: (2025)
"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
Assessing the Performance of Human-Capable LLMs -- Are LLMs Coming for Your Job?
von: Mavi, John, et al.
Veröffentlicht: (2024)
von: Mavi, John, et al.
Veröffentlicht: (2024)
LLM Probability Concentration: How Alignment Shrinks the Generative Horizon
von: Yang, Chenghao, et al.
Veröffentlicht: (2025)
von: Yang, Chenghao, et al.
Veröffentlicht: (2025)
CLEAR: A Clinically-Grounded Tabular Framework for Radiology Report Evaluation
von: Jiang, Yuyang, et al.
Veröffentlicht: (2025)
von: Jiang, Yuyang, et al.
Veröffentlicht: (2025)
The simulation of judgment in LLMs
von: Loru, Edoardo, et al.
Veröffentlicht: (2025)
von: Loru, Edoardo, et al.
Veröffentlicht: (2025)
Measuring Teaching with LLMs
von: Hardy, Michael
Veröffentlicht: (2025)
von: Hardy, Michael
Veröffentlicht: (2025)
The Political Preferences of LLMs
von: Rozado, David
Veröffentlicht: (2024)
von: Rozado, David
Veröffentlicht: (2024)
Why are all LLMs Obsessed with Japanese Culture? On the Hidden Cultural and Regional Biases of LLMs
von: de Landa, Joseba Fernandez, et al.
Veröffentlicht: (2026)
von: de Landa, Joseba Fernandez, et al.
Veröffentlicht: (2026)
Are LLMs (Really) Ideological? An IRT-based Analysis and Alignment Tool for Perceived Socio-Economic Bias in LLMs
von: Wachter, Jasmin, et al.
Veröffentlicht: (2025)
von: Wachter, Jasmin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Know Thyself? On the Incapability and Implications of AI Self-Recognition
von: Bai, Xiaoyan, et al.
Veröffentlicht: (2025) -
The Story is Not the Science: Execution-Grounded Evaluation of Mechanistic Interpretability Research
von: Bai, Xiaoyan, et al.
Veröffentlicht: (2026) -
On the Effectiveness and Generalization of Race Representations for Debiasing High-Stakes Decisions
von: Nguyen, Dang, et al.
Veröffentlicht: (2025) -
AbsenceBench: Language Models Can't Tell What's Missing
von: Fu, Harvey Yiyun, et al.
Veröffentlicht: (2025) -
GPT-4V Cannot Generate Radiology Reports Yet
von: Jiang, Yuyang, et al.
Veröffentlicht: (2024)