Characterizing AI Agents for Alignment and Governance
Fuente:
arXiv
Saved in:
| Main Authors: | Kasirzadeh, Atoosa, Gabriel, Iason |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Two Types of AI Existential Risk: Decisive and Accumulative
by: Kasirzadeh, Atoosa
Published: (2024)
by: Kasirzadeh, Atoosa
Published: (2024)
Bridging the Gap in the Responsible AI Divides
by: Gyevnár, Bálint, et al.
Published: (2026)
by: Gyevnár, Bálint, et al.
Published: (2026)
Explanation Hacking: The perils of algorithmic recourse
by: Sullivan, Emily, et al.
Published: (2024)
by: Sullivan, Emily, et al.
Published: (2024)
Measurement challenges in AI catastrophic risk governance and safety frameworks
by: Kasirzadeh, Atoosa
Published: (2024)
by: Kasirzadeh, Atoosa
Published: (2024)
The Algorithmic State Architecture (ASA): An Integrated Framework for AI-Enabled Government
by: Engin, Zeynep, et al.
Published: (2025)
by: Engin, Zeynep, et al.
Published: (2025)
Structured AI Decision-Making in Disaster Management
by: Dcruz, Julian Gerald, et al.
Published: (2025)
by: Dcruz, Julian Gerald, et al.
Published: (2025)
Systematic Hazard Analysis for Frontier AI using STPA
by: Mylius, Simon
Published: (2025)
by: Mylius, Simon
Published: (2025)
Navigating the sociotechnical labyrinth: Dynamic certification for responsible embodied AI
by: Bakirtzis, Georgios, et al.
Published: (2024)
by: Bakirtzis, Georgios, et al.
Published: (2024)
AI Safety for Everyone
by: Gyevnar, Balint, et al.
Published: (2025)
by: Gyevnar, Balint, et al.
Published: (2025)
Metacognition Should Be the Scientific Framework for Bounded and Effective Self-Governance in Generative AI
by: Ji, Eugene Yu, et al.
Published: (2026)
by: Ji, Eugene Yu, et al.
Published: (2026)
Human-AI Safety: A Descendant of Generative AI and Control Systems Safety
by: Bajcsy, Andrea, et al.
Published: (2024)
by: Bajcsy, Andrea, et al.
Published: (2024)
The SMART+ Framework for AI Systems
by: Kandikatla, Laxmiraju, et al.
Published: (2025)
by: Kandikatla, Laxmiraju, et al.
Published: (2025)
Technical Risks of (Lethal) Autonomous Weapons Systems
by: Podar, Heramb, et al.
Published: (2025)
by: Podar, Heramb, et al.
Published: (2025)
Towards More Efficient Shared Autonomous Mobility: A Learning-Based Fleet Repositioning Approach
by: Filipovska, Monika, et al.
Published: (2022)
by: Filipovska, Monika, et al.
Published: (2022)
A Taxonomy and Review of Algorithms for Modeling and Predicting Human Driver Behavior
by: Bhattacharyya, Raunak P., et al.
Published: (2020)
by: Bhattacharyya, Raunak P., et al.
Published: (2020)
Governing What the EU AI Act Excludes: Accountability for Autonomous AI Agents in Smart City Critical Infrastructure
by: Butt, Talal Ashraf, et al.
Published: (2026)
by: Butt, Talal Ashraf, et al.
Published: (2026)
Leveraging High-Fidelity Digital Models and Reinforcement Learning for Mission Engineering: A Case Study of Aerial Firefighting Under Perfect Information
by: Çetinkaya, İbrahim Oğuz, et al.
Published: (2025)
by: Çetinkaya, İbrahim Oğuz, et al.
Published: (2025)
Operationalizing Reconstructive Authority: Runtime Construction, Dependency Resolution, and Execution Gating in Autonomous Agent Systems
by: TraslaIA, Marcelo Fernandez -
Published: (2026)
by: TraslaIA, Marcelo Fernandez -
Published: (2026)
Governed Reasoning for Institutional AI
by: Seck, Mamadou
Published: (2026)
by: Seck, Mamadou
Published: (2026)
Adapting Probabilistic Risk Assessment for AI
by: Wisakanto, Anna Katariina, et al.
Published: (2025)
by: Wisakanto, Anna Katariina, et al.
Published: (2025)
Multi-Agent Risks from Advanced AI
by: Hammond, Lewis, et al.
Published: (2025)
by: Hammond, Lewis, et al.
Published: (2025)
AgentCity: Constitutional Governance for Autonomous Agent Economies via Separation of Power
by: Ruan, Anbang, et al.
Published: (2026)
by: Ruan, Anbang, et al.
Published: (2026)
Consent Chain Degradation in Embodied Multi-Agent Systems: Bridging the Gap Between AI Agent Governance and Robot Ethics
by: Haklidir, Mehmet
Published: (2026)
by: Haklidir, Mehmet
Published: (2026)
AI, Digital Platforms, and the New Systemic Risk
by: Hacker, Philipp, et al.
Published: (2025)
by: Hacker, Philipp, et al.
Published: (2025)
AI-Powered CPS-Enabled Vulnerable-User-Aware Urban Transportation Digital Twin: Methods and Applications
by: Fu, Yongjie, et al.
Published: (2024)
by: Fu, Yongjie, et al.
Published: (2024)
Soft-Label Governance for Distributional Safety in Multi-Agent Systems
by: Aiersilan, Aizierjiang, et al.
Published: (2026)
by: Aiersilan, Aizierjiang, et al.
Published: (2026)
Hybrid Approaches for Moral Value Alignment in AI Agents: a Manifesto
by: Tennant, Elizaveta, et al.
Published: (2023)
by: Tennant, Elizaveta, et al.
Published: (2023)
Is Decentralized AI Governable? From Regulative Policy to Constitutive Protocol
by: Hu, Botao Amber, et al.
Published: (2026)
by: Hu, Botao Amber, et al.
Published: (2026)
Governance by Design: A Parsonian Institutional Architecture for Internet-Wide Agent Societies
by: Ruan, Anbang
Published: (2026)
by: Ruan, Anbang
Published: (2026)
Social Catalysts, Not Moral Agents: The Illusion of Alignment in LLM Societies
by: Hu, Yueqing, et al.
Published: (2026)
by: Hu, Yueqing, et al.
Published: (2026)
Empowering Cognitive Digital Twins with Generative Foundation Models: Developing a Low-Carbon Integrated Freight Transportation System
by: Li, Xueping, et al.
Published: (2024)
by: Li, Xueping, et al.
Published: (2024)
Improving Operational Efficiency In EV Ridepooling Fleets By Predictive Exploitation of Idle Times
by: Provoost, Jesper C., et al.
Published: (2022)
by: Provoost, Jesper C., et al.
Published: (2022)
Reinforcement Learning for Sustainable Energy: A Survey
by: Ponse, Koen, et al.
Published: (2024)
by: Ponse, Koen, et al.
Published: (2024)
AI Agents as Policymakers in Simulated Epidemics
by: Aoki, Goshi, et al.
Published: (2026)
by: Aoki, Goshi, et al.
Published: (2026)
Genocide by Algorithm in Gaza: Artificial Intelligence, Countervailing Responsibility, and the Corruption of Public Discourse
by: Radeljic, Branislav
Published: (2026)
by: Radeljic, Branislav
Published: (2026)
AI Agent for Education: von Neumann Multi-Agent System Framework
by: Jiang, Yuan-Hao, et al.
Published: (2024)
by: Jiang, Yuan-Hao, et al.
Published: (2024)
Epistemic Injustice in Generative AI
by: Kay, Jackie, et al.
Published: (2024)
by: Kay, Jackie, et al.
Published: (2024)
Bit-politeia: An AI Agent Community in Blockchain
by: Yang, Xing
Published: (2026)
by: Yang, Xing
Published: (2026)
Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice via Social Determinants
by: Tang, Zeyu, et al.
Published: (2025)
by: Tang, Zeyu, et al.
Published: (2025)
Towards Computational Social Dynamics of Semi-Autonomous AI Agents
by: Lidarity, S. O., et al.
Published: (2026)
by: Lidarity, S. O., et al.
Published: (2026)
Similar Items
-
Two Types of AI Existential Risk: Decisive and Accumulative
by: Kasirzadeh, Atoosa
Published: (2024) -
Bridging the Gap in the Responsible AI Divides
by: Gyevnár, Bálint, et al.
Published: (2026) -
Explanation Hacking: The perils of algorithmic recourse
by: Sullivan, Emily, et al.
Published: (2024) -
Measurement challenges in AI catastrophic risk governance and safety frameworks
by: Kasirzadeh, Atoosa
Published: (2024) -
The Algorithmic State Architecture (ASA): An Integrated Framework for AI-Enabled Government
by: Engin, Zeynep, et al.
Published: (2025)