Learning the Value Systems of Societies with Preference-based Multi-objective Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Holgado-Sánchez, Andrés, Vamplew, Peter, Dazeley, Richard, Ossowski, Sascha, Billhardt, Holger |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning the Value Systems of Agents with Preference-based and Inverse Reinforcement Learning
by: Holgado-Sánchez, Andrés, et al.
Published: (2026)
by: Holgado-Sánchez, Andrés, et al.
Published: (2026)
Learning from Preferences and Mixed Demonstrations in General Settings
by: Brown, Jason R, et al.
Published: (2025)
by: Brown, Jason R, et al.
Published: (2025)
Measure what Matters: Psychometric Evaluation of AI with Situational Judgment Tests
by: Yost, Alexandra, et al.
Published: (2025)
by: Yost, Alexandra, et al.
Published: (2025)
Survey Transfer Learning: Recycling Data with Silicon Responses
by: Amini, Ali
Published: (2025)
by: Amini, Ali
Published: (2025)
ES-C51: Expected Sarsa Based C51 Distributional Reinforcement Learning Algorithm
by: Tandon, Rijul, et al.
Published: (2025)
by: Tandon, Rijul, et al.
Published: (2025)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
by: Bachar, Or, et al.
Published: (2026)
by: Bachar, Or, et al.
Published: (2026)
Misaligned from Within: Large Language Models Reproduce Our Double-Loop Learning Blindness
by: Rogers, Tim, et al.
Published: (2025)
by: Rogers, Tim, et al.
Published: (2025)
Fusing Rewards and Preferences in Reinforcement Learning
by: Khorasani, Sadegh, et al.
Published: (2025)
by: Khorasani, Sadegh, et al.
Published: (2025)
When Can Human-AI Teams Outperform Individuals? Tight Bounds with Impossibility Guarantees
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
An Aircraft Upset Recovery System with Reinforcement Learning
by: Demir, Mahir, et al.
Published: (2026)
by: Demir, Mahir, et al.
Published: (2026)
LUCAS-MEGA: A Large-Scale Multimodal Dataset for Representation Learning in Soil-Environment Systems
by: Leng, Kuangdai, et al.
Published: (2026)
by: Leng, Kuangdai, et al.
Published: (2026)
Feature Selection and Regularization in Multi-Class Classification: An Empirical Study of One-vs-Rest Logistic Regression with Gradient Descent Optimization and L1 Sparsity Constraints
by: Arafat, Jahidul, et al.
Published: (2025)
by: Arafat, Jahidul, et al.
Published: (2025)
Uncertainty Quantification for Scientific Machine Learning using Sparse Variational Gaussian Process Kolmogorov-Arnold Networks (SVGP KAN)
by: Ju, Y. Sungtaek
Published: (2025)
by: Ju, Y. Sungtaek
Published: (2025)
Synheart Emotion: Privacy-Preserving On-Device Emotion Recognition from Biosignals
by: Ademtew, Henok, et al.
Published: (2025)
by: Ademtew, Henok, et al.
Published: (2025)
An Automatic Ground Collision Avoidance System with Reinforcement Learning
by: Sevgili, Seyyid Osman, et al.
Published: (2026)
by: Sevgili, Seyyid Osman, et al.
Published: (2026)
BrainForm: a Serious Game for BCI Training and Data Collection
by: Romani, Michele, et al.
Published: (2025)
by: Romani, Michele, et al.
Published: (2025)
Adaptive XAI in High Stakes Environments: Modeling Swift Trust with Multimodal Feedback in Human AI Teams
by: Fernando, Nishani, et al.
Published: (2025)
by: Fernando, Nishani, et al.
Published: (2025)
Explicit modelling of subject dependency in BCI decoding
by: Romani, Michele, et al.
Published: (2025)
by: Romani, Michele, et al.
Published: (2025)
SHAPoint: Task-Agnostic, Efficient, and Interpretable Point-Based Risk Scoring via Shapley Values
by: Meirman, Tomer D., et al.
Published: (2025)
by: Meirman, Tomer D., et al.
Published: (2025)
Designing AI for Prosecutorial Governance: Case Prioritization and Statutory Oversight in Mexico
by: Sobrino, Fernanda, et al.
Published: (2026)
by: Sobrino, Fernanda, et al.
Published: (2026)
Deep Learning for Solving and Estimating Dynamic Models in Economics and Finance
by: Scheidegger, Simon
Published: (2026)
by: Scheidegger, Simon
Published: (2026)
Differentiating Viral and Bacterial Infections: A Machine Learning Model Based on Routine Blood Test Values
by: Gunčar, Gregor, et al.
Published: (2023)
by: Gunčar, Gregor, et al.
Published: (2023)
BioAlchemy: Distilling Biological Literature into Reasoning-Ready Reinforcement Learning Training Data
by: Hsu, Brian, et al.
Published: (2026)
by: Hsu, Brian, et al.
Published: (2026)
Scalable and Interpretable Scientific Discovery via Sparse Variational Gaussian Process Kolmogorov-Arnold Networks (SVGP KAN)
by: Ju, Y. Sungtaek
Published: (2025)
by: Ju, Y. Sungtaek
Published: (2025)
Automated Database Indexing using Model-free Reinforcement Learning
by: Licks, Gabriel Paludo, et al.
Published: (2020)
by: Licks, Gabriel Paludo, et al.
Published: (2020)
Explainable Artificial Intelligence (XAI) 2.0: A Manifesto of Open Challenges and Interdisciplinary Research Directions
by: Longo, Luca, et al.
Published: (2023)
by: Longo, Luca, et al.
Published: (2023)
Widening the Role of Group Recommender Systems with CAJO
by: Ricci, Francesco, et al.
Published: (2025)
by: Ricci, Francesco, et al.
Published: (2025)
Analysing Factorizations of Action-Value Networks for Cooperative Multi-Agent Reinforcement Learning
by: Castellini, Jacopo, et al.
Published: (2019)
by: Castellini, Jacopo, et al.
Published: (2019)
Efficiency Without Cognitive Change: Evidence from Human Interaction with Narrow AI Systems
by: Benítez, María Angélica, et al.
Published: (2025)
by: Benítez, María Angélica, et al.
Published: (2025)
Benchmarking the Discovery Engine
by: Foxabbott, Jack, et al.
Published: (2025)
by: Foxabbott, Jack, et al.
Published: (2025)
Multi-Hypothesis Prediction for Portfolio Optimization: A Structured Ensemble Learning Approach to Risk Diversification
by: Dominguez, Alejandro Rodriguez, et al.
Published: (2025)
by: Dominguez, Alejandro Rodriguez, et al.
Published: (2025)
Geoinformatics-Guided Machine Learning for Power Plant Classification
by: Austin-Gabriel, Blessing, et al.
Published: (2025)
by: Austin-Gabriel, Blessing, et al.
Published: (2025)
EyeLayer: Integrating Human Attention Patterns into LLM-Based Code Summarization
by: Zhang, Jiahao, et al.
Published: (2026)
by: Zhang, Jiahao, et al.
Published: (2026)
Conversational No-code, Multi-agentic Disease Module Identification and Drug Repurposing Prediction with ChatDRex
by: Süwer, Simon, et al.
Published: (2025)
by: Süwer, Simon, et al.
Published: (2025)
Between Knowledge and Care: Evaluating Generative AI-Based IUI in Type 2 Diabetes Management Through Patient and Physician Perspectives
by: Meng, Yibo, et al.
Published: (2025)
by: Meng, Yibo, et al.
Published: (2025)
Smart Recommendations for Renting Bikes in Bike Sharing Systems
by: Billhardt, Holger, et al.
Published: (2024)
by: Billhardt, Holger, et al.
Published: (2024)
Multi-modal Integration Analysis of Alzheimer's Disease Using Large Language Models and Knowledge Graphs
by: Kiguchi, Kanan, et al.
Published: (2025)
by: Kiguchi, Kanan, et al.
Published: (2025)
Active Inference for an Intelligent Agent in Autonomous Reconnaissance Missions
by: Schubert, Johan, et al.
Published: (2025)
by: Schubert, Johan, et al.
Published: (2025)
Interactive Groupwise Comparison for Reinforcement Learning from Human Feedback
by: Kompatscher, Jan, et al.
Published: (2025)
by: Kompatscher, Jan, et al.
Published: (2025)
Digging deeper: deep joint species distribution modeling reveals environmental drivers of Earthworm Communities
by: Si-moussi, Sara, et al.
Published: (2025)
by: Si-moussi, Sara, et al.
Published: (2025)
Similar Items
-
Learning the Value Systems of Agents with Preference-based and Inverse Reinforcement Learning
by: Holgado-Sánchez, Andrés, et al.
Published: (2026) -
Learning from Preferences and Mixed Demonstrations in General Settings
by: Brown, Jason R, et al.
Published: (2025) -
Measure what Matters: Psychometric Evaluation of AI with Situational Judgment Tests
by: Yost, Alexandra, et al.
Published: (2025) -
Survey Transfer Learning: Recycling Data with Silicon Responses
by: Amini, Ali
Published: (2025) -
ES-C51: Expected Sarsa Based C51 Distributional Reinforcement Learning Algorithm
by: Tandon, Rijul, et al.
Published: (2025)