Beyond Human Judgment: A Bayesian Evaluation of LLMs' Moral Values Understanding
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Skorski, Maciej, Landowska, Alina |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
EQUITRIAGE: A Fairness Audit of Gender Bias in LLM-Based Emergency Department Triage
par: Young, Richard J., et autres
Publié: (2026)
par: Young, Richard J., et autres
Publié: (2026)
MEDLEY-BENCH: Scale Buys Evaluation but Not Control in AI Metacognition
par: Abtahi, Farhad, et autres
Publié: (2026)
par: Abtahi, Farhad, et autres
Publié: (2026)
Automated Quality Assessment for LLM-Based Complex Qualitative Coding: A Confidence-Diversity Framework
par: Zhao, Zhilong, et autres
Publié: (2025)
par: Zhao, Zhilong, et autres
Publié: (2025)
AVEC: Bootstrapping Privacy for Local LLMs
par: Gaikwad, Madhava
Publié: (2025)
par: Gaikwad, Madhava
Publié: (2025)
Computable Gap Assessment of Artificial Intelligence Governance in Children's Centres: Evidence-Mechanism-Governance-Indicator Modelling of UNICEF's Guidance on AI and Children 3.0 Based on the Graph-GAP Framework
par: Meng, Wei
Publié: (2025)
par: Meng, Wei
Publié: (2025)
The Likelihood Ratio Wall: Structural Limits on Accurate Risk Assessment for Rare Violence
par: Pollanen, Marco
Publié: (2026)
par: Pollanen, Marco
Publié: (2026)
Conformal Prediction Sets for Next-Token Prediction in Large Language Models: Balancing Coverage Guarantees with Set Efficiency
par: Kotla, Yoshith Roy, et autres
Publié: (2025)
par: Kotla, Yoshith Roy, et autres
Publié: (2025)
ChatGPT Based Data Augmentation for Improved Parameter-Efficient Debiasing of LLMs
par: Han, Pengrui, et autres
Publié: (2024)
par: Han, Pengrui, et autres
Publié: (2024)
AETAS: Analysis of Evolving Temporal Affect and Semantics for Legal History
par: Wang, Qizhi
Publié: (2025)
par: Wang, Qizhi
Publié: (2025)
On-Device Generative AI for GDPR-Compliant Visual Monitoring: Natural Language Alerts from Local Object Detection
par: Schappacher-Tilp, Gudrun, et autres
Publié: (2026)
par: Schappacher-Tilp, Gudrun, et autres
Publié: (2026)
On the Validity of Traditional Vulnerability Scoring Systems for Adversarial Attacks against LLMs
par: Bahar, Atmane Ayoub Mansour, et autres
Publié: (2024)
par: Bahar, Atmane Ayoub Mansour, et autres
Publié: (2024)
Moral Semantics Survive Machine Translation: Cross-Lingual Evidence from Moral Foundations Corpora
par: Skorski, Maciej
Publié: (2026)
par: Skorski, Maciej
Publié: (2026)
A Novel Kuhnian Ontology for Epistemic Classification of STM Scholarly Articles
par: Saqr, Khalid M.
Publié: (2020)
par: Saqr, Khalid M.
Publié: (2020)
The Ethics Engine: A Modular Pipeline for Accessible Psychometric Assessment of Large Language Models
par: Van Clief, Jake, et autres
Publié: (2025)
par: Van Clief, Jake, et autres
Publié: (2025)
2025 Southeast Asia Eleven Nations Influence Index Report
par: Meng, Wei
Publié: (2025)
par: Meng, Wei
Publié: (2025)
Recipient Profiling: Predicting Characteristics from Messages
par: Borquez, Martin, et autres
Publié: (2024)
par: Borquez, Martin, et autres
Publié: (2024)
When No Benchmark Exists: Validating Comparative LLM Safety Scoring Without Ground-Truth Labels
par: Gautam, Sushant, et autres
Publié: (2026)
par: Gautam, Sushant, et autres
Publié: (2026)
Can LLMs Understand What We Cannot Say? Measuring Multilevel Alignment Through Abortion Stigma Across Cognitive, Interpersonal, and Structural Levels
par: Sharma, Anika, et autres
Publié: (2025)
par: Sharma, Anika, et autres
Publié: (2025)
Talk, Walk, and Market Response: Multimodal Measurement of AI Washing and Its Capital Market Consequences in China
par: Zhanjie, Wen, et autres
Publié: (2026)
par: Zhanjie, Wen, et autres
Publié: (2026)
How Few-shot Demonstrations Affect Prompt-based Defenses Against LLM Jailbreak Attacks
par: Wang, Yanshu, et autres
Publié: (2026)
par: Wang, Yanshu, et autres
Publié: (2026)
The Fragility Of Moral Judgment In Large Language Models
par: van Nuenen, Tom, et autres
Publié: (2026)
par: van Nuenen, Tom, et autres
Publié: (2026)
Modeling and Visualization Reasoning for Stakeholders in Education and Industry Integration Systems: Research on Structured Synthetic Dialogue Data Generation Based on NIST Standards
par: Meng, Wei
Publié: (2025)
par: Meng, Wei
Publié: (2025)
Enhancing Mental Health Support through Human-AI Collaboration: Toward Secure and Empathetic AI-enabled chatbots
par: AlMakinah, Rawan, et autres
Publié: (2024)
par: AlMakinah, Rawan, et autres
Publié: (2024)
Machine Learning and Public Health: Identifying and Mitigating Algorithmic Bias through a Systematic Review
par: Altamirano, Sara, et autres
Publié: (2025)
par: Altamirano, Sara, et autres
Publié: (2025)
Claim Automation using Large Language Model
par: Mo, Zhengda, et autres
Publié: (2026)
par: Mo, Zhengda, et autres
Publié: (2026)
The AI Fiction Paradox
par: Elkins, Katherine
Publié: (2026)
par: Elkins, Katherine
Publié: (2026)
Beyond Static Knowledge Messengers: Towards Adaptive, Fair, and Scalable Federated Learning for Medical AI
par: Arafat, Jahidul, et autres
Publié: (2025)
par: Arafat, Jahidul, et autres
Publié: (2025)
Mapping the Web of Science, a large-scale graph and text-based dataset with LLM embeddings
par: Kunt, Tim, et autres
Publié: (2026)
par: Kunt, Tim, et autres
Publié: (2026)
Transforming Business with Generative AI: Research, Innovation, Market Deployment and Future Shifts in Business Models
par: Singh, Narotam, et autres
Publié: (2024)
par: Singh, Narotam, et autres
Publié: (2024)
Conformal Path Reasoning: Trustworthy Knowledge Graph Question Answering via Path-Level Calibration
par: Lin, Shuhang, et autres
Publié: (2026)
par: Lin, Shuhang, et autres
Publié: (2026)
AI Literacy and LLM Engagement in Higher Education: A Cross-National Quantitative Study
par: Hossain, Shahin, et autres
Publié: (2025)
par: Hossain, Shahin, et autres
Publié: (2025)
EMPATHIA: Multi-Faceted Human-AI Collaboration for Refugee Integration
par: Barhdadi, Mohamed Rayan, et autres
Publié: (2025)
par: Barhdadi, Mohamed Rayan, et autres
Publié: (2025)
IndianBailJudgments-1200: A Multi-Attribute Dataset for Legal NLP on Indian Bail Orders
par: Deshmukh, Sneha, et autres
Publié: (2025)
par: Deshmukh, Sneha, et autres
Publié: (2025)
Scalable AI-Driven Analytics for User Engagement and Stance Detection on Social Media
par: Seneviratne, Thammitage Piyumi Wathsala, et autres
Publié: (2026)
par: Seneviratne, Thammitage Piyumi Wathsala, et autres
Publié: (2026)
AI-Powered Citation Auditing: A Zero-Assumption Protocol for Systematic Reference Verification in Academic Research
par: van Rensburg, L. J. Janse
Publié: (2025)
par: van Rensburg, L. J. Janse
Publié: (2025)
When Names Change Verdicts: Intervention Consistency Reveals Systematic Bias in LLM Decision-Making
par: Basu, Abhinaba, et autres
Publié: (2026)
par: Basu, Abhinaba, et autres
Publié: (2026)
Intelligence Without Integrity: Why Capable LLMs May Undermine Reliability
par: Allen, Ryan, et autres
Publié: (2026)
par: Allen, Ryan, et autres
Publié: (2026)
Democratizing LLMs: An Exploration of Cost-Performance Trade-offs in Self-Refined Open-Source Models
par: Shashidhar, Sumuk, et autres
Publié: (2023)
par: Shashidhar, Sumuk, et autres
Publié: (2023)
Can AI Make Conflicts Worse? An Alignment Failure in LLM Deployment Across Conflict Contexts
par: Kryshtal, Andrii
Publié: (2026)
par: Kryshtal, Andrii
Publié: (2026)
Estimating Exam Item Difficulty with LLMs: A Benchmark on Brazil's ENEM Corpus
par: Brant, Thiago, et autres
Publié: (2026)
par: Brant, Thiago, et autres
Publié: (2026)
Documents similaires
-
EQUITRIAGE: A Fairness Audit of Gender Bias in LLM-Based Emergency Department Triage
par: Young, Richard J., et autres
Publié: (2026) -
MEDLEY-BENCH: Scale Buys Evaluation but Not Control in AI Metacognition
par: Abtahi, Farhad, et autres
Publié: (2026) -
Automated Quality Assessment for LLM-Based Complex Qualitative Coding: A Confidence-Diversity Framework
par: Zhao, Zhilong, et autres
Publié: (2025) -
AVEC: Bootstrapping Privacy for Local LLMs
par: Gaikwad, Madhava
Publié: (2025) -
Computable Gap Assessment of Artificial Intelligence Governance in Children's Centres: Evidence-Mechanism-Governance-Indicator Modelling of UNICEF's Guidance on AI and Children 3.0 Based on the Graph-GAP Framework
par: Meng, Wei
Publié: (2025)