Personalized Constitutionally-Aligned Agentic Superego: Secure AI Behavior Aligned to Diverse Human Values
Fuente:
arXiv
Saved in:
| Main Authors: | Watson, Nell, Amer, Ahmed, Harris, Evan, Ravindra, Preeti, Zhang, Shujun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Is Decentralized AI Governable? From Regulative Policy to Constitutive Protocol
by: Hu, Botao Amber, et al.
Published: (2026)
by: Hu, Botao Amber, et al.
Published: (2026)
Socio-technical aspects of Agentic AI
by: Donta, Praveen Kumar, et al.
Published: (2025)
by: Donta, Praveen Kumar, et al.
Published: (2025)
Mapping Human Anti-collusion Mechanisms to Multi-agent AI Systems
by: Idowu, Jamiu, et al.
Published: (2026)
by: Idowu, Jamiu, et al.
Published: (2026)
PAARS: Persona Aligned Agentic Retail Shoppers
by: Mansour, Saab, et al.
Published: (2025)
by: Mansour, Saab, et al.
Published: (2025)
Advancing Responsible Innovation in Agentic AI: A study of Ethical Frameworks for Household Automation
by: Chandra, Joydeep, et al.
Published: (2025)
by: Chandra, Joydeep, et al.
Published: (2025)
Fairness in Agentic AI: A Unified Framework for Ethical and Equitable Multi-Agent System
by: Ranjan, Rajesh, et al.
Published: (2025)
by: Ranjan, Rajesh, et al.
Published: (2025)
AgentCity: Constitutional Governance for Autonomous Agent Economies via Separation of Power
by: Ruan, Anbang, et al.
Published: (2026)
by: Ruan, Anbang, et al.
Published: (2026)
Exploring Network-Knowledge Graph Duality: A Case Study in Agentic Supply Chain Risk Analysis
by: Heus, Evan, et al.
Published: (2025)
by: Heus, Evan, et al.
Published: (2025)
Robustness of Agentic AI Systems via Adversarially-Aligned Jacobian Regularization
by: Mumcu, Furkan, et al.
Published: (2026)
by: Mumcu, Furkan, et al.
Published: (2026)
Sentinel Agents for Secure and Trustworthy Agentic AI in Multi-Agent Systems
by: Gosmar, Diego, et al.
Published: (2025)
by: Gosmar, Diego, et al.
Published: (2025)
FinAgent: An Agentic AI Framework Integrating Personal Finance and Nutrition Planning
by: Syed, Toqeer Ali, et al.
Published: (2025)
by: Syed, Toqeer Ali, et al.
Published: (2025)
CODE-GEN: A Human-in-the-Loop RAG-Based Agentic AI System for Multiple-Choice Question Generation
by: Duan, Xiaojing, et al.
Published: (2026)
by: Duan, Xiaojing, et al.
Published: (2026)
Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem
by: LaCroix, Travis
Published: (2026)
by: LaCroix, Travis
Published: (2026)
Aligning Individual and Collective Objectives in Multi-Agent Cooperation
by: Li, Yang, et al.
Published: (2024)
by: Li, Yang, et al.
Published: (2024)
Security Threats in Agentic AI System
by: Khan, Raihan, et al.
Published: (2024)
by: Khan, Raihan, et al.
Published: (2024)
Toward Personalizing Quantum Computing Education: An Evolutionary LLM-Powered Approach
by: Elhaimeur, Iizalaarab, et al.
Published: (2025)
by: Elhaimeur, Iizalaarab, et al.
Published: (2025)
Scaling Behavior of Single LLM-Driven Multi-Agent Systems
by: Li, Jialing, et al.
Published: (2026)
by: Li, Jialing, et al.
Published: (2026)
Aligning Artificial Superintelligence via a Multi-Box Protocol
by: Negozio, Avraham Yair
Published: (2025)
by: Negozio, Avraham Yair
Published: (2025)
Hybrid Approaches for Moral Value Alignment in AI Agents: a Manifesto
by: Tennant, Elizaveta, et al.
Published: (2023)
by: Tennant, Elizaveta, et al.
Published: (2023)
AgentTutor: Empowering Personalized Learning with Multi-Turn Interactive Teaching in Intelligent Education Systems
by: Liu, Yuxin, et al.
Published: (2025)
by: Liu, Yuxin, et al.
Published: (2025)
Aligning Compound AI Systems via System-level DPO
by: Wang, Xiangwen, et al.
Published: (2025)
by: Wang, Xiangwen, et al.
Published: (2025)
Personality-Driven Student Agent-Based Modeling in Mathematics Education: How Well Do Student Agents Align with Human Learners?
by: Xiao, Bushi, et al.
Published: (2026)
by: Xiao, Bushi, et al.
Published: (2026)
Governed Reasoning for Institutional AI
by: Seck, Mamadou
Published: (2026)
by: Seck, Mamadou
Published: (2026)
Agentifying Agentic AI
by: Dignum, Virginia, et al.
Published: (2025)
by: Dignum, Virginia, et al.
Published: (2025)
AI Agents as Policymakers in Simulated Epidemics
by: Aoki, Goshi, et al.
Published: (2026)
by: Aoki, Goshi, et al.
Published: (2026)
AI-Mediated Explainable Regulation for Justice
by: Hofweber, Thomas, et al.
Published: (2026)
by: Hofweber, Thomas, et al.
Published: (2026)
Governing What the EU AI Act Excludes: Accountability for Autonomous AI Agents in Smart City Critical Infrastructure
by: Butt, Talal Ashraf, et al.
Published: (2026)
by: Butt, Talal Ashraf, et al.
Published: (2026)
Bit-politeia: An AI Agent Community in Blockchain
by: Yang, Xing
Published: (2026)
by: Yang, Xing
Published: (2026)
Invisible Orchestrators Suppress Protective Behavior and Dissociate Power-Holders: Safety Risks in Multi-Agent LLM Systems
by: Fukui, Hiroki
Published: (2026)
by: Fukui, Hiroki
Published: (2026)
Aligned Agents, Biased Swarm: Measuring Bias Amplification in Multi-Agent Systems
by: Li, Keyu, et al.
Published: (2026)
by: Li, Keyu, et al.
Published: (2026)
Towards Computational Social Dynamics of Semi-Autonomous AI Agents
by: Lidarity, S. O., et al.
Published: (2026)
by: Lidarity, S. O., et al.
Published: (2026)
From Narrative to Action: A Hierarchical LLM-Agent Framework for Human Mobility Generation
by: Li, Qiumeng, et al.
Published: (2025)
by: Li, Qiumeng, et al.
Published: (2025)
Urban-MAS: Human-Centered Urban Prediction with LLM-Based Multi-Agent System
by: Lou, Shangyu
Published: (2025)
by: Lou, Shangyu
Published: (2025)
AI Agent for Education: von Neumann Multi-Agent System Framework
by: Jiang, Yuan-Hao, et al.
Published: (2024)
by: Jiang, Yuan-Hao, et al.
Published: (2024)
Harm in AI-Driven Societies: An Audit of Toxicity Adoption on Chirper.ai
by: Coppolillo, Erica, et al.
Published: (2026)
by: Coppolillo, Erica, et al.
Published: (2026)
LOKA Protocol: A Decentralized Framework for Trustworthy and Ethical AI Agent Ecosystems
by: Ranjan, Rajesh, et al.
Published: (2025)
by: Ranjan, Rajesh, et al.
Published: (2025)
The potential role of AI agents in transforming nuclear medicine research and cancer management in India
by: Vashistha, Rajat, et al.
Published: (2025)
by: Vashistha, Rajat, et al.
Published: (2025)
GGBond: Growing Graph-Based AI-Agent Society for Socially-Aware Recommender Simulation
by: Zhong, Hailin, et al.
Published: (2025)
by: Zhong, Hailin, et al.
Published: (2025)
Too Human to Model:The Uncanny Valley of LLMs in Social Simulation -- When Generative Language Agents Misalign with Modelling Principles
by: Zeng, Yongchao, et al.
Published: (2025)
by: Zeng, Yongchao, et al.
Published: (2025)
Formalizing the Safety, Security, and Functional Properties of Agentic AI Systems
by: Allegrini, Edoardo, et al.
Published: (2025)
by: Allegrini, Edoardo, et al.
Published: (2025)
Similar Items
-
Is Decentralized AI Governable? From Regulative Policy to Constitutive Protocol
by: Hu, Botao Amber, et al.
Published: (2026) -
Socio-technical aspects of Agentic AI
by: Donta, Praveen Kumar, et al.
Published: (2025) -
Mapping Human Anti-collusion Mechanisms to Multi-agent AI Systems
by: Idowu, Jamiu, et al.
Published: (2026) -
PAARS: Persona Aligned Agentic Retail Shoppers
by: Mansour, Saab, et al.
Published: (2025) -
Advancing Responsible Innovation in Agentic AI: A study of Ethical Frameworks for Household Automation
by: Chandra, Joydeep, et al.
Published: (2025)