Culturally-Attuned Moral Machines: Implicit Learning of Human Value Systems by AI through Inverse Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Oliveira, Nigini, Li, Jasmine, Khalvati, Koosha, Barragan, Rodolfo Cortes, Reinecke, Katharina, Meltzoff, Andrew N., Rao, Rajesh P. N. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Framing an AI with Values Reduces AI Reliance in AI-supported Writing Tasks
by: Gao, Alice, et al.
Published: (2026)
by: Gao, Alice, et al.
Published: (2026)
NormAd: A Framework for Measuring the Cultural Adaptability of Large Language Models
by: Rao, Abhinav, et al.
Published: (2024)
by: Rao, Abhinav, et al.
Published: (2024)
Predicting Star Scientists in the Field of Artificial Intelligence: A Machine Learning Approach
by: Shirouyeh, Koosha, et al.
Published: (2024)
by: Shirouyeh, Koosha, et al.
Published: (2024)
Reinforcement Learning from Human Feedback: Whose Culture, Whose Values, Whose Perspectives?
by: Barman, Kristian González, et al.
Published: (2024)
by: Barman, Kristian González, et al.
Published: (2024)
Recursive Neural Programs: Variational Learning of Image Grammars and Part-Whole Hierarchies
by: Fisher, Ares, et al.
Published: (2022)
by: Fisher, Ares, et al.
Published: (2022)
SAGE: Sustainable Agent-Guided Expert-tuning for Culturally Attuned Translation in Low-Resource Southeast Asia
by: Lu, Zhixiang, et al.
Published: (2026)
by: Lu, Zhixiang, et al.
Published: (2026)
Exploring Radiologists' Expectations of Explainable Machine Learning Models in Medical Image Analysis
by: Ketabi, Sara, et al.
Published: (2026)
by: Ketabi, Sara, et al.
Published: (2026)
Meta-Representational Predictive Coding: Biomimetic Self-Supervised Learning
by: Ororbia, Alexander, et al.
Published: (2025)
by: Ororbia, Alexander, et al.
Published: (2025)
Active Predictive Coding: A Unified Neural Framework for Learning Hierarchical World Models for Perception and Planning
by: Rao, Rajesh P. N., et al.
Published: (2022)
by: Rao, Rajesh P. N., et al.
Published: (2022)
Implicit Humanization in Everyday LLM Moral Judgments
by: Ayad, Hoda, et al.
Published: (2026)
by: Ayad, Hoda, et al.
Published: (2026)
MoVa: Towards Generalizable Classification of Human Morals and Values
by: Chen, Ziyu, et al.
Published: (2025)
by: Chen, Ziyu, et al.
Published: (2025)
The Computational Complexity of Variational Inequalities and Applications in Game Theory
by: Kapron, Bruce M., et al.
Published: (2024)
by: Kapron, Bruce M., et al.
Published: (2024)
Agentic Reinforcement Learning with Implicit Step Rewards
by: Liu, Xiaoqian, et al.
Published: (2025)
by: Liu, Xiaoqian, et al.
Published: (2025)
Learning the Value Systems of Agents with Preference-based and Inverse Reinforcement Learning
by: Holgado-Sánchez, Andrés, et al.
Published: (2026)
by: Holgado-Sánchez, Andrés, et al.
Published: (2026)
Open-radiomics: A Collection of Standardized Datasets and a Technical Protocol for Reproducible Radiomics Machine Learning Pipelines
by: Namdar, Khashayar, et al.
Published: (2022)
by: Namdar, Khashayar, et al.
Published: (2022)
An End-to-End System for Culturally-Attuned Driving Feedback using a Dual-Component NLG Engine
by: Thompson, Iniakpokeikiye Peter, et al.
Published: (2025)
by: Thompson, Iniakpokeikiye Peter, et al.
Published: (2025)
Every Question Has Its Own Value: Reinforcement Learning with Explicit Human Values
by: Yu, Dian, et al.
Published: (2025)
by: Yu, Dian, et al.
Published: (2025)
Can LLMs Grasp Implicit Cultural Values? Benchmarking LLMs' Cultural Intelligence with CQ-Bench
by: Liu, Ziyi, et al.
Published: (2025)
by: Liu, Ziyi, et al.
Published: (2025)
Language Scent: Exploring Cross-Language Information Navigation
by: Zhu, Jiawen Stefanie, et al.
Published: (2026)
by: Zhu, Jiawen Stefanie, et al.
Published: (2026)
Enhancing Image Caption Generation Using Reinforcement Learning with Human Feedback
by: L, Adarsh N, et al.
Published: (2024)
by: L, Adarsh N, et al.
Published: (2024)
Unleashing Implicit Rewards: Prefix-Value Learning for Distribution-Level Optimization
by: Gao, Shiping, et al.
Published: (2026)
by: Gao, Shiping, et al.
Published: (2026)
Transferable Post-training via Inverse Value Learning
by: Lu, Xinyu, et al.
Published: (2024)
by: Lu, Xinyu, et al.
Published: (2024)
KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts
by: Hwang, Taebaek, et al.
Published: (2025)
by: Hwang, Taebaek, et al.
Published: (2025)
ClinicalFMamba: Advancing Clinical Assessment using Mamba-based Multimodal Neuroimaging Fusion
by: Zhou, Meng, et al.
Published: (2025)
by: Zhou, Meng, et al.
Published: (2025)
BLIP: Facilitating the Exploration of Undesirable Consequences of Digital Technologies
by: Pang, Rock Yuren, et al.
Published: (2024)
by: Pang, Rock Yuren, et al.
Published: (2024)
Insights from the Inverse: Reconstructing LLM Training Goals Through Inverse Reinforcement Learning
by: Joselowitz, Jared, et al.
Published: (2024)
by: Joselowitz, Jared, et al.
Published: (2024)
Implicit In-context Learning
by: Li, Zhuowei, et al.
Published: (2024)
by: Li, Zhuowei, et al.
Published: (2024)
That is Unacceptable: the Moral Foundations of Canceling
by: Lo, Soda Marem, et al.
Published: (2025)
by: Lo, Soda Marem, et al.
Published: (2025)
AI-Driven MRI-based Brain Tumour Segmentation Benchmarking
by: Ludwig, Connor, et al.
Published: (2025)
by: Ludwig, Connor, et al.
Published: (2025)
Hierarchical Active Inference using Successor Representations
by: Rangarajan, Prashant, et al.
Published: (2026)
by: Rangarajan, Prashant, et al.
Published: (2026)
Process Reinforcement through Implicit Rewards
by: Cui, Ganqu, et al.
Published: (2025)
by: Cui, Ganqu, et al.
Published: (2025)
Question Answering with Texts and Tables through Deep Reinforcement Learning
by: José, Marcos M., et al.
Published: (2024)
by: José, Marcos M., et al.
Published: (2024)
Bridging Cultural Nuances in Dialogue Agents through Cultural Value Surveys
by: Cao, Yong, et al.
Published: (2024)
by: Cao, Yong, et al.
Published: (2024)
Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions
by: Kim, Jiseon, et al.
Published: (2026)
by: Kim, Jiseon, et al.
Published: (2026)
Supervised Fine-Tuning as Inverse Reinforcement Learning
by: Sun, Hao
Published: (2024)
by: Sun, Hao
Published: (2024)
Privacy-Preserving and Incentive-Driven Relay-Based Framework for Cross-Domain Blockchain Interoperability
by: Moradi, Saeed, et al.
Published: (2025)
by: Moradi, Saeed, et al.
Published: (2025)
Rethinking Machine Ethics -- Can LLMs Perform Moral Reasoning through the Lens of Moral Theories?
by: Zhou, Jingyan, et al.
Published: (2023)
by: Zhou, Jingyan, et al.
Published: (2023)
Know Your Audience: The benefits and pitfalls of generating plain language summaries beyond the "general" audience
by: August, Tal, et al.
Published: (2024)
by: August, Tal, et al.
Published: (2024)
Preference Distillation via Value based Reinforcement Learning
by: Kwon, Minchan, et al.
Published: (2025)
by: Kwon, Minchan, et al.
Published: (2025)
Efficient Inverse Design Optimization through Multi-fidelity Simulations, Machine Learning, and Search Space Reduction Strategies
by: Grbcic, Luka, et al.
Published: (2023)
by: Grbcic, Luka, et al.
Published: (2023)
Similar Items
-
Framing an AI with Values Reduces AI Reliance in AI-supported Writing Tasks
by: Gao, Alice, et al.
Published: (2026) -
NormAd: A Framework for Measuring the Cultural Adaptability of Large Language Models
by: Rao, Abhinav, et al.
Published: (2024) -
Predicting Star Scientists in the Field of Artificial Intelligence: A Machine Learning Approach
by: Shirouyeh, Koosha, et al.
Published: (2024) -
Reinforcement Learning from Human Feedback: Whose Culture, Whose Values, Whose Perspectives?
by: Barman, Kristian González, et al.
Published: (2024) -
Recursive Neural Programs: Variational Learning of Image Grammars and Part-Whole Hierarchies
by: Fisher, Ares, et al.
Published: (2022)