Learning the Value Systems of Societies from Preferences
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Holgado-Sánchez, Andrés, Billhardt, Holger, Ossowski, Sascha, Degli-Esposti, Sara |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Learning the Value Systems of Societies with Preference-based Multi-objective Reinforcement Learning
par: Holgado-Sánchez, Andrés, et autres
Publié: (2026)
par: Holgado-Sánchez, Andrés, et autres
Publié: (2026)
Learning the Value Systems of Agents with Preference-based and Inverse Reinforcement Learning
par: Holgado-Sánchez, Andrés, et autres
Publié: (2026)
par: Holgado-Sánchez, Andrés, et autres
Publié: (2026)
Algorithms for learning value-aligned policies considering admissibility relaxation
par: Holgado-Sánchez, Andrés, et autres
Publié: (2024)
par: Holgado-Sánchez, Andrés, et autres
Publié: (2024)
Smart Recommendations for Renting Bikes in Bike Sharing Systems
par: Billhardt, Holger, et autres
Publié: (2024)
par: Billhardt, Holger, et autres
Publié: (2024)
Towards a prioritised use of transportation infrastructures: the case of vehicle-specific dynamic access restrictions to city centres
par: Billhardt, Holger, et autres
Publié: (2024)
par: Billhardt, Holger, et autres
Publié: (2024)
Bike3S: A Tool for Bike Sharing Systems Simulation
par: Fernández, Alberto, et autres
Publié: (2024)
par: Fernández, Alberto, et autres
Publié: (2024)
Legal and ethical implications of applications based on agreement technologies: the case of auction-based road intersections
par: Santos, José-Antonio, et autres
Publié: (2024)
par: Santos, José-Antonio, et autres
Publié: (2024)
Agreement Technologies for Coordination in Smart Cities
par: Billhardt, Holger, et autres
Publié: (2024)
par: Billhardt, Holger, et autres
Publié: (2024)
Introduction to AI Safety, Ethics, and Society
par: Hendrycks, Dan
Publié: (2024)
par: Hendrycks, Dan
Publié: (2024)
Revolutionising Distance Learning: A Comparative Study of Learning Progress with AI-Driven Tutoring
par: Möller, Moritz, et autres
Publié: (2024)
par: Möller, Moritz, et autres
Publié: (2024)
Culturally-Attuned Moral Machines: Implicit Learning of Human Value Systems by AI through Inverse Reinforcement Learning
par: Oliveira, Nigini, et autres
Publié: (2023)
par: Oliveira, Nigini, et autres
Publié: (2023)
Whose Preferences? Differences in Fairness Preferences and Their Impact on the Fairness of AI Utilizing Human Feedback
par: Lerner, Emilia Agis, et autres
Publié: (2024)
par: Lerner, Emilia Agis, et autres
Publié: (2024)
The LLM Has Left The Chat: Evidence of Bail Preferences in Large Language Models
par: Ensign, Danielle, et autres
Publié: (2025)
par: Ensign, Danielle, et autres
Publié: (2025)
I Prefer not to Say: Protecting User Consent in Models with Optional Personal Data
par: Leemann, Tobias, et autres
Publié: (2022)
par: Leemann, Tobias, et autres
Publié: (2022)
On The Fairness Impacts of Hardware Selection in Machine Learning
par: Nelaturu, Sree Harsha, et autres
Publié: (2023)
par: Nelaturu, Sree Harsha, et autres
Publié: (2023)
Taxi dispatching strategies with compensations
par: Billhardt, Holger, et autres
Publié: (2024)
par: Billhardt, Holger, et autres
Publié: (2024)
On-Time Delivery in Crowdshipping Systems: An Agent-Based Approach Using Streaming Data
par: Dötterl, Jeremias, et autres
Publié: (2024)
par: Dötterl, Jeremias, et autres
Publié: (2024)
Safe and Certifiable AI Systems: Concepts, Challenges, and Lessons Learned
par: Schweighofer, Kajetan, et autres
Publié: (2025)
par: Schweighofer, Kajetan, et autres
Publié: (2025)
BiasGuard: Guardrailing Fairness in Machine Learning Production Systems
par: Cohen-Inger, Nurit, et autres
Publié: (2025)
par: Cohen-Inger, Nurit, et autres
Publié: (2025)
Safeguarding Autonomy: a Focus on Machine Learning Decision Systems
par: Subías-Beltrán, Paula, et autres
Publié: (2025)
par: Subías-Beltrán, Paula, et autres
Publié: (2025)
Values in the Wild: Discovering and Analyzing Values in Real-World Language Model Interactions
par: Huang, Saffron, et autres
Publié: (2025)
par: Huang, Saffron, et autres
Publié: (2025)
Reward Models Inherit Value Biases from Pretraining
par: Christian, Brian, et autres
Publié: (2026)
par: Christian, Brian, et autres
Publié: (2026)
Still More Shades of Null: An Evaluation Suite for Responsible Missing Value Imputation
par: Khan, Falaah Arif, et autres
Publié: (2024)
par: Khan, Falaah Arif, et autres
Publié: (2024)
Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook
par: Lee, Jaehyeok, et autres
Publié: (2026)
par: Lee, Jaehyeok, et autres
Publié: (2026)
A Deep Learning Approach Towards Student Performance Prediction in Online Courses: Challenges Based on a Global Perspective
par: Moubayed, Abdallah, et autres
Publié: (2024)
par: Moubayed, Abdallah, et autres
Publié: (2024)
LearnLM: Improving Gemini for Learning
par: LearnLM Team, et autres
Publié: (2024)
par: LearnLM Team, et autres
Publié: (2024)
Value Lens: Using Large Language Models to Understand Human Values
par: Fernández, Eduardo de la Cruz, et autres
Publié: (2025)
par: Fernández, Eduardo de la Cruz, et autres
Publié: (2025)
Enhancing Equitable Access to AI in Housing and Homelessness System of Care through Federated Learning
par: Taib, Musa, et autres
Publié: (2024)
par: Taib, Musa, et autres
Publié: (2024)
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
par: Zhou, Han, et autres
Publié: (2024)
par: Zhou, Han, et autres
Publié: (2024)
Non-myopic Matching and Rebalancing in Large-Scale On-Demand Ride-Pooling Systems Using Simulation-Informed Reinforcement Learning
par: Namdarpour, Farnoosh, et autres
Publié: (2025)
par: Namdarpour, Farnoosh, et autres
Publié: (2025)
Influence of Recommender Systems on Users: A Dynamical Systems Analysis
par: Lankireddy, Prabhat, et autres
Publié: (2024)
par: Lankireddy, Prabhat, et autres
Publié: (2024)
EigenBench: A Comparative Behavioral Measure of Value Alignment
par: Chang, Jonathn, et autres
Publié: (2025)
par: Chang, Jonathn, et autres
Publié: (2025)
Unintended Harms of Value-Aligned LLMs: Psychological and Empirical Insights
par: Choi, Sooyung, et autres
Publié: (2025)
par: Choi, Sooyung, et autres
Publié: (2025)
Learning Recourse Costs from Pairwise Feature Comparisons
par: Rawal, Kaivalya, et autres
Publié: (2024)
par: Rawal, Kaivalya, et autres
Publié: (2024)
The Wolf Within: Covert Injection of Malice into MLLM Societies via an MLLM Operative
par: Tan, Zhen, et autres
Publié: (2024)
par: Tan, Zhen, et autres
Publié: (2024)
Putnam's Critical and Explanatory Tendencies Interpreted from a Machine Learning Perspective
par: Soudin, Sheldon Z.
Publié: (2025)
par: Soudin, Sheldon Z.
Publié: (2025)
GreedLlama: Performance of Financial Value-Aligned Large Language Models in Moral Reasoning
par: Yu, Jeffy, et autres
Publié: (2024)
par: Yu, Jeffy, et autres
Publié: (2024)
Prerequisite Structure Discovery in Intelligent Tutoring Systems
par: Annabi, Louis, et autres
Publié: (2024)
par: Annabi, Louis, et autres
Publié: (2024)
Investigating Thematic Patterns and User Preferences in LLM Interactions using BERTopic
par: Bhandarkar, Abhay, et autres
Publié: (2025)
par: Bhandarkar, Abhay, et autres
Publié: (2025)
A Trustworthiness-based Metaphysics of Artificial Intelligence Systems
par: Ferrario, Andrea
Publié: (2025)
par: Ferrario, Andrea
Publié: (2025)
Documents similaires
-
Learning the Value Systems of Societies with Preference-based Multi-objective Reinforcement Learning
par: Holgado-Sánchez, Andrés, et autres
Publié: (2026) -
Learning the Value Systems of Agents with Preference-based and Inverse Reinforcement Learning
par: Holgado-Sánchez, Andrés, et autres
Publié: (2026) -
Algorithms for learning value-aligned policies considering admissibility relaxation
par: Holgado-Sánchez, Andrés, et autres
Publié: (2024) -
Smart Recommendations for Renting Bikes in Bike Sharing Systems
par: Billhardt, Holger, et autres
Publié: (2024) -
Towards a prioritised use of transportation infrastructures: the case of vehicle-specific dynamic access restrictions to city centres
par: Billhardt, Holger, et autres
Publié: (2024)