Hybrid Approaches for Moral Value Alignment in AI Agents: a Manifesto
Fuente:
arXiv
Saved in:
| Main Authors: | Tennant, Elizaveta, Hailes, Stephen, Musolesi, Mirco |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dynamics of Moral Behavior in Heterogeneous Populations of Learning Agents
by: Tennant, Elizaveta, et al.
Published: (2024)
by: Tennant, Elizaveta, et al.
Published: (2024)
Moral Alignment for LLM Agents
by: Tennant, Elizaveta, et al.
Published: (2024)
by: Tennant, Elizaveta, et al.
Published: (2024)
Opponent Shaping in LLM Agents
by: Segura, Marta Emili Garcia, et al.
Published: (2025)
by: Segura, Marta Emili Garcia, et al.
Published: (2025)
Trust-based Consensus in Multi-Agent Reinforcement Learning Systems
by: Fung, Ho Long, et al.
Published: (2022)
by: Fung, Ho Long, et al.
Published: (2022)
(Ir)rationality in AI: State of the Art, Research Challenges and Open Questions
by: Macmillan-Scott, Olivia, et al.
Published: (2023)
by: Macmillan-Scott, Olivia, et al.
Published: (2023)
Investigating the Impact of Direct Punishment on the Emergence of Cooperation in Multi-Agent Reinforcement Learning Systems
by: Dasgupta, Nayana, et al.
Published: (2023)
by: Dasgupta, Nayana, et al.
Published: (2023)
CoMIX: A Multi-agent Reinforcement Learning Training Architecture for Efficient Decentralized Coordination and Independent Decision-Making
by: Minelli, Giovanni, et al.
Published: (2023)
by: Minelli, Giovanni, et al.
Published: (2023)
Reinforcement Learning Discovers Efficient Decentralized Graph Path Search Strategies
by: Pisacane, Alexei, et al.
Published: (2024)
by: Pisacane, Alexei, et al.
Published: (2024)
Multi-Agent Risks from Advanced AI
by: Hammond, Lewis, et al.
Published: (2025)
by: Hammond, Lewis, et al.
Published: (2025)
Social Catalysts, Not Moral Agents: The Illusion of Alignment in LLM Societies
by: Hu, Yueqing, et al.
Published: (2026)
by: Hu, Yueqing, et al.
Published: (2026)
When Reasoning Models Hurt Behavioral Simulation: A Solver-Sampler Mismatch in Multi-Agent LLM Negotiation
by: Andric, Sandro
Published: (2026)
by: Andric, Sandro
Published: (2026)
Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem
by: LaCroix, Travis
Published: (2026)
by: LaCroix, Travis
Published: (2026)
Multi-Agent Reinforcement Learning Simulation for Environmental Policy Synthesis
by: Rudd-Jones, James, et al.
Published: (2025)
by: Rudd-Jones, James, et al.
Published: (2025)
Enhancing Clinical Decision-Making: Integrating Multi-Agent Systems with Ethical AI Governance
by: Chen, Ying-Jung, et al.
Published: (2025)
by: Chen, Ying-Jung, et al.
Published: (2025)
MAEBE: Multi-Agent Emergent Behavior Framework
by: Erisken, Sinem, et al.
Published: (2025)
by: Erisken, Sinem, et al.
Published: (2025)
Toward Super Agent System with Hybrid AI Routers
by: Yao, Yuhang, et al.
Published: (2025)
by: Yao, Yuhang, et al.
Published: (2025)
Toward Inclusive Educational AI: Auditing Frontier LLMs through a Multiplexity Lens
by: Mushtaq, Abdullah, et al.
Published: (2025)
by: Mushtaq, Abdullah, et al.
Published: (2025)
Exploring Collaboration Mechanisms for LLM Agents: A Social Psychology View
by: Zhang, Jintian, et al.
Published: (2023)
by: Zhang, Jintian, et al.
Published: (2023)
Thinking Outside the (Gray) Box: A Context-Based Score for Assessing Value and Originality in Neural Text Generation
by: Franceschelli, Giorgio, et al.
Published: (2025)
by: Franceschelli, Giorgio, et al.
Published: (2025)
A Hierarchical Hybrid AI Approach: Integrating Deep Reinforcement Learning and Scripted Agents in Combat Simulations
by: Black, Scotty, et al.
Published: (2025)
by: Black, Scotty, et al.
Published: (2025)
How Far Are We From True Auto-Research?
by: Zhang, Zhengxin, et al.
Published: (2026)
by: Zhang, Zhengxin, et al.
Published: (2026)
Improving Health Information Access in the World's Largest Maternal Mobile Health Program via Bandit Algorithms
by: Lalan, Arshika, et al.
Published: (2024)
by: Lalan, Arshika, et al.
Published: (2024)
Epistemic diversity across language models mitigates knowledge collapse
by: Hodel, Damian, et al.
Published: (2025)
by: Hodel, Damian, et al.
Published: (2025)
Simultaneously Achieving Group Exposure Fairness and Within-Group Meritocracy in Stochastic Bandits
by: Pokhriyal, Subham, et al.
Published: (2024)
by: Pokhriyal, Subham, et al.
Published: (2024)
Hybrid Agentic AI and Multi-Agent Systems in Smart Manufacturing
by: Farahani, Mojtaba A., et al.
Published: (2025)
by: Farahani, Mojtaba A., et al.
Published: (2025)
AI Agents as Policymakers in Simulated Epidemics
by: Aoki, Goshi, et al.
Published: (2026)
by: Aoki, Goshi, et al.
Published: (2026)
AI Agent for Education: von Neumann Multi-Agent System Framework
by: Jiang, Yuan-Hao, et al.
Published: (2024)
by: Jiang, Yuan-Hao, et al.
Published: (2024)
Bit-politeia: An AI Agent Community in Blockchain
by: Yang, Xing
Published: (2026)
by: Yang, Xing
Published: (2026)
Generative AI collective behavior needs an interactionist paradigm
by: Ferrarotti, Laura, et al.
Published: (2026)
by: Ferrarotti, Laura, et al.
Published: (2026)
Toward Adaptive Categories: Dimensional Governance for Agentic AI
by: Engin, Zeynep, et al.
Published: (2025)
by: Engin, Zeynep, et al.
Published: (2025)
Personalized Constitutionally-Aligned Agentic Superego: Secure AI Behavior Aligned to Diverse Human Values
by: Watson, Nell, et al.
Published: (2025)
by: Watson, Nell, et al.
Published: (2025)
Towards Computational Social Dynamics of Semi-Autonomous AI Agents
by: Lidarity, S. O., et al.
Published: (2026)
by: Lidarity, S. O., et al.
Published: (2026)
Copyright in Generative Deep Learning
by: Franceschelli, Giorgio, et al.
Published: (2021)
by: Franceschelli, Giorgio, et al.
Published: (2021)
Creativity and Machine Learning: A Survey
by: Franceschelli, Giorgio, et al.
Published: (2021)
by: Franceschelli, Giorgio, et al.
Published: (2021)
DeepCreativity: Measuring Creativity with Deep Learning Techniques
by: Franceschelli, Giorgio, et al.
Published: (2022)
by: Franceschelli, Giorgio, et al.
Published: (2022)
Governing What the EU AI Act Excludes: Accountability for Autonomous AI Agents in Smart City Critical Infrastructure
by: Butt, Talal Ashraf, et al.
Published: (2026)
by: Butt, Talal Ashraf, et al.
Published: (2026)
LOKA Protocol: A Decentralized Framework for Trustworthy and Ethical AI Agent Ecosystems
by: Ranjan, Rajesh, et al.
Published: (2025)
by: Ranjan, Rajesh, et al.
Published: (2025)
Fairness in Agentic AI: A Unified Framework for Ethical and Equitable Multi-Agent System
by: Ranjan, Rajesh, et al.
Published: (2025)
by: Ranjan, Rajesh, et al.
Published: (2025)
GGBond: Growing Graph-Based AI-Agent Society for Socially-Aware Recommender Simulation
by: Zhong, Hailin, et al.
Published: (2025)
by: Zhong, Hailin, et al.
Published: (2025)
CogniPair: From LLM Chatbots to Conscious AI Agents -- GNWT-Based Multi-Agent Digital Twins for Social Pairing -- Dating & Hiring Applications
by: Ye, Wanghao, et al.
Published: (2025)
by: Ye, Wanghao, et al.
Published: (2025)
Similar Items
-
Dynamics of Moral Behavior in Heterogeneous Populations of Learning Agents
by: Tennant, Elizaveta, et al.
Published: (2024) -
Moral Alignment for LLM Agents
by: Tennant, Elizaveta, et al.
Published: (2024) -
Opponent Shaping in LLM Agents
by: Segura, Marta Emili Garcia, et al.
Published: (2025) -
Trust-based Consensus in Multi-Agent Reinforcement Learning Systems
by: Fung, Ho Long, et al.
Published: (2022) -
(Ir)rationality in AI: State of the Art, Research Challenges and Open Questions
by: Macmillan-Scott, Olivia, et al.
Published: (2023)