The Moral Gap of Large Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Skorski, Maciej, Landowska, Alina |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Beyond Human Judgment: A Bayesian Evaluation of LLMs' Moral Values Understanding
par: Skorski, Maciej, et autres
Publié: (2025)
par: Skorski, Maciej, et autres
Publié: (2025)
Mapping Technological Futures: Anticipatory Discourse Through Text Mining
par: Skorski, Maciej, et autres
Publié: (2025)
par: Skorski, Maciej, et autres
Publié: (2025)
Toward Cultural Interpretability: A Linguistic Anthropological Framework for Describing and Evaluating Large Language Models (LLMs)
par: Jones, Graham M., et autres
Publié: (2024)
par: Jones, Graham M., et autres
Publié: (2024)
The Moral Machine Experiment on Large Language Models
par: Takemoto, Kazuhiro
Publié: (2023)
par: Takemoto, Kazuhiro
Publié: (2023)
MoReBench: Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes
par: Chiu, Yu Ying, et autres
Publié: (2025)
par: Chiu, Yu Ying, et autres
Publié: (2025)
Does Writing with Language Models Reduce Content Diversity?
par: Padmakumar, Vishakh, et autres
Publié: (2023)
par: Padmakumar, Vishakh, et autres
Publié: (2023)
Prediction-Powered Ranking of Large Language Models
par: Chatzi, Ivi, et autres
Publié: (2024)
par: Chatzi, Ivi, et autres
Publié: (2024)
Linear Representations of Political Perspective Emerge in Large Language Models
par: Kim, Junsol, et autres
Publié: (2025)
par: Kim, Junsol, et autres
Publié: (2025)
Can Large Language Models Unlock Novel Scientific Research Ideas?
par: Kumar, Sandeep, et autres
Publié: (2024)
par: Kumar, Sandeep, et autres
Publié: (2024)
Computational Approaches to Understanding Large Language Model Impact on Writing and Information Ecosystems
par: Liang, Weixin
Publié: (2025)
par: Liang, Weixin
Publié: (2025)
ProgressGym: Alignment with a Millennium of Moral Progress
par: Qiu, Tianyi, et autres
Publié: (2024)
par: Qiu, Tianyi, et autres
Publié: (2024)
Large Language Models Can Infer Personality from Free-Form User Interactions
par: Peters, Heinrich, et autres
Publié: (2024)
par: Peters, Heinrich, et autres
Publié: (2024)
LLM4PM: A case study on using Large Language Models for Process Modeling in Enterprise Organizations
par: Ziche, Clara, et autres
Publié: (2024)
par: Ziche, Clara, et autres
Publié: (2024)
The "Colonial Impulse" of Natural Language Processing: An Audit of Bengali Sentiment Analysis Tools and Their Identity-based Biases
par: Das, Dipto, et autres
Publié: (2024)
par: Das, Dipto, et autres
Publié: (2024)
LLM-based Cognitive Models of Students with Misconceptions
par: Sonkar, Shashank, et autres
Publié: (2024)
par: Sonkar, Shashank, et autres
Publié: (2024)
When "A Helpful Assistant" Is Not Really Helpful: Personas in System Prompts Do Not Improve Performances of Large Language Models
par: Zheng, Mingqian, et autres
Publié: (2023)
par: Zheng, Mingqian, et autres
Publié: (2023)
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
par: Si, Chenglei, et autres
Publié: (2025)
par: Si, Chenglei, et autres
Publié: (2025)
Causal Reasoning and Large Language Models: Opening a New Frontier for Causality
par: Kıcıman, Emre, et autres
Publié: (2023)
par: Kıcıman, Emre, et autres
Publié: (2023)
Implicit Personalization in Language Models: A Systematic Study
par: Jin, Zhijing, et autres
Publié: (2024)
par: Jin, Zhijing, et autres
Publié: (2024)
Multidimensional Human Activity Recognition With Large Language Model: A Conceptual Framework
par: Hasan, Syed Mhamudul
Publié: (2024)
par: Hasan, Syed Mhamudul
Publié: (2024)
On the Pros and Cons of Active Learning for Moral Preference Elicitation
par: Keswani, Vijay, et autres
Publié: (2024)
par: Keswani, Vijay, et autres
Publié: (2024)
Evaluating Multimodal Language Models as Visual Assistants for Visually Impaired Users
par: Karamolegkou, Antonia, et autres
Publié: (2025)
par: Karamolegkou, Antonia, et autres
Publié: (2025)
Humanizing LLMs: A Survey of Psychological Measurements with Tools, Datasets, and Human-Agent Applications
par: Dong, Wenhan, et autres
Publié: (2025)
par: Dong, Wenhan, et autres
Publié: (2025)
Attention to Non-Adopters
par: Zhou, Kaitlyn, et autres
Publié: (2025)
par: Zhou, Kaitlyn, et autres
Publié: (2025)
Watching the Watchers: A Comparative Fairness Audit of Cloud-based Content Moderation Services
par: Hartmann, David, et autres
Publié: (2024)
par: Hartmann, David, et autres
Publié: (2024)
Activation Steering via Generative Causal Mediation
par: Sankaranarayanan, Aruna, et autres
Publié: (2026)
par: Sankaranarayanan, Aruna, et autres
Publié: (2026)
Story Ribbons: Reimagining Storyline Visualizations with Large Language Models
par: Yeh, Catherine, et autres
Publié: (2025)
par: Yeh, Catherine, et autres
Publié: (2025)
Evaluating Large Language Models in Theory of Mind Tasks
par: Kosinski, Michal
Publié: (2023)
par: Kosinski, Michal
Publié: (2023)
The Perils & Promises of Fact-checking with Large Language Models
par: Quelle, Dorian, et autres
Publié: (2023)
par: Quelle, Dorian, et autres
Publié: (2023)
Self-reflecting Large Language Models: A Hegelian Dialectical Approach
par: Abdali, Sara, et autres
Publié: (2025)
par: Abdali, Sara, et autres
Publié: (2025)
'Simulacrum of Stories': Examining Large Language Models as Qualitative Research Participants
par: Kapania, Shivani, et autres
Publié: (2024)
par: Kapania, Shivani, et autres
Publié: (2024)
Multi-turn Evaluation of Anthropomorphic Behaviours in Large Language Models
par: Ibrahim, Lujain, et autres
Publié: (2025)
par: Ibrahim, Lujain, et autres
Publié: (2025)
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
par: Si, Chenglei, et autres
Publié: (2024)
par: Si, Chenglei, et autres
Publié: (2024)
Multi-Level Feedback Generation with Large Language Models for Empowering Novice Peer Counselors
par: Chaszczewicz, Alicja, et autres
Publié: (2024)
par: Chaszczewicz, Alicja, et autres
Publié: (2024)
Sample-Efficient Human Evaluation of Large Language Models via Maximum Discrepancy Competition
par: Feng, Kehua, et autres
Publié: (2024)
par: Feng, Kehua, et autres
Publié: (2024)
Communication Bias in Large Language Models: A Regulatory Perspective
par: Kuenzler, Adrian, et autres
Publié: (2025)
par: Kuenzler, Adrian, et autres
Publié: (2025)
A Systematic Review on the Evaluation of Large Language Models in Theory of Mind Tasks
par: Sarıtaş, Karahan, et autres
Publié: (2025)
par: Sarıtaş, Karahan, et autres
Publié: (2025)
SELF-PERCEPT: Introspection Improves Large Language Models' Detection of Multi-Person Mental Manipulation in Conversations
par: Khanna, Danush, et autres
Publié: (2025)
par: Khanna, Danush, et autres
Publié: (2025)
Scaling Laws for Moral Machine Judgment in Large Language Models
par: Takemoto, Kazuhiro
Publié: (2026)
par: Takemoto, Kazuhiro
Publié: (2026)
Large Language Models Can Infer Psychological Dispositions of Social Media Users
par: Peters, Heinrich, et autres
Publié: (2023)
par: Peters, Heinrich, et autres
Publié: (2023)
Documents similaires
-
Beyond Human Judgment: A Bayesian Evaluation of LLMs' Moral Values Understanding
par: Skorski, Maciej, et autres
Publié: (2025) -
Mapping Technological Futures: Anticipatory Discourse Through Text Mining
par: Skorski, Maciej, et autres
Publié: (2025) -
Toward Cultural Interpretability: A Linguistic Anthropological Framework for Describing and Evaluating Large Language Models (LLMs)
par: Jones, Graham M., et autres
Publié: (2024) -
The Moral Machine Experiment on Large Language Models
par: Takemoto, Kazuhiro
Publié: (2023) -
MoReBench: Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes
par: Chiu, Yu Ying, et autres
Publié: (2025)