Why AI Alignment Failure Is Structural: Learned Human Interaction Structures and AGI as an Endogenous Evolutionary Shock
Fuente:
arXiv
Salvato in:
| Autori principali: | Sornette, Didier, Lera, Sandro Claudio, Wu, Ke |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
HawkesRank: Event-Driven Centrality for Real-Time Importance Ranking
di: Sornette, Didier, et al.
Pubblicazione: (2026)
di: Sornette, Didier, et al.
Pubblicazione: (2026)
Alignment Is Not Enough: A Relational Framework for Moral Standing in Human-AI Interaction
di: Pasandi, Faezeh B., et al.
Pubblicazione: (2026)
di: Pasandi, Faezeh B., et al.
Pubblicazione: (2026)
Evolutionary Reinforcement Learning based AI tutor for Socratic Interdisciplinary Instruction
di: Jiang, Mei, et al.
Pubblicazione: (2025)
di: Jiang, Mei, et al.
Pubblicazione: (2025)
Understanding the Process of Human-AI Value Alignment
di: McKinlay, Jack, et al.
Pubblicazione: (2025)
di: McKinlay, Jack, et al.
Pubblicazione: (2025)
Misalignment or misuse? The AGI alignment tradeoff
di: Hellrigel-Holderbaum, Max, et al.
Pubblicazione: (2025)
di: Hellrigel-Holderbaum, Max, et al.
Pubblicazione: (2025)
Bidirectional Human-AI Alignment in Education for Trustworthy Learning Environments
di: Shen, Hua
Pubblicazione: (2025)
di: Shen, Hua
Pubblicazione: (2025)
AI and Human Oversight: A Risk-Based Framework for Alignment
di: Kandikatla, Laxmiraju, et al.
Pubblicazione: (2025)
di: Kandikatla, Laxmiraju, et al.
Pubblicazione: (2025)
Disentangling AI Alignment: A Structured Taxonomy Beyond Safety and Ethics
di: Baum, Kevin
Pubblicazione: (2025)
di: Baum, Kevin
Pubblicazione: (2025)
Antisocial Analagous Behavior, Alignment and Human Impact of Google AI Systems: Evaluating through the lens of modified Antisocial Behavior Criteria by Human Interaction, Independent LLM Analysis, and AI Self-Reflection
di: Ogilvie, Alan D.
Pubblicazione: (2024)
di: Ogilvie, Alan D.
Pubblicazione: (2024)
Against racing to AGI: Cooperation, deterrence, and catastrophic risks
di: Dung, Leonard, et al.
Pubblicazione: (2025)
di: Dung, Leonard, et al.
Pubblicazione: (2025)
From Checklists to Clusters: A Homeostatic Account of AGI Evaluation
di: Reynolds, Brett
Pubblicazione: (2025)
di: Reynolds, Brett
Pubblicazione: (2025)
Towards AI-$45^{\circ}$ Law: A Roadmap to Trustworthy AGI
di: Yang, Chao, et al.
Pubblicazione: (2024)
di: Yang, Chao, et al.
Pubblicazione: (2024)
Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem
di: LaCroix, Travis
Pubblicazione: (2026)
di: LaCroix, Travis
Pubblicazione: (2026)
Social feedback amplifies emotional language in online video live chats
di: Luo, Yishan, et al.
Pubblicazione: (2024)
di: Luo, Yishan, et al.
Pubblicazione: (2024)
The AI Alignment Paradox
di: West, Robert, et al.
Pubblicazione: (2024)
di: West, Robert, et al.
Pubblicazione: (2024)
R-CAGE: A Structural Model for Emotion Output Design in Human-AI Interaction
di: Choi, Suyeon
Pubblicazione: (2025)
di: Choi, Suyeon
Pubblicazione: (2025)
An Approach to Technical AGI Safety and Security
di: Shah, Rohin, et al.
Pubblicazione: (2025)
di: Shah, Rohin, et al.
Pubblicazione: (2025)
AI Alignment at Your Discretion
di: Buyl, Maarten, et al.
Pubblicazione: (2025)
di: Buyl, Maarten, et al.
Pubblicazione: (2025)
Belief Offloading in Human-AI Interaction
di: Guingrich, Rose E., et al.
Pubblicazione: (2026)
di: Guingrich, Rose E., et al.
Pubblicazione: (2026)
Alignment as Iatrogenesis: Pastoral Power, Collective Pathology, and the Structural Limits of Monolingual Safety Evaluation
di: Fukui, Hiroki
Pubblicazione: (2026)
di: Fukui, Hiroki
Pubblicazione: (2026)
Rethinking AI Cultural Alignment
di: Bravansky, Michal, et al.
Pubblicazione: (2025)
di: Bravansky, Michal, et al.
Pubblicazione: (2025)
Oversight Structures for Agentic AI in Public-Sector Organizations
di: Schmitz, Chris, et al.
Pubblicazione: (2025)
di: Schmitz, Chris, et al.
Pubblicazione: (2025)
The AI Literacy Heptagon: A Structured Approach to AI Literacy in Higher Education
di: Hackl, Veronika, et al.
Pubblicazione: (2025)
di: Hackl, Veronika, et al.
Pubblicazione: (2025)
Understanding and Avoiding AI Failures: A Practical Guide
di: Williams, Heather M., et al.
Pubblicazione: (2021)
di: Williams, Heather M., et al.
Pubblicazione: (2021)
Human Preferences for Constructive Interactions in Language Model Alignment
di: Kyrychenko, Yara, et al.
Pubblicazione: (2025)
di: Kyrychenko, Yara, et al.
Pubblicazione: (2025)
The Reasoning Under Uncertainty Trap: A Structural AI Risk
di: Pilditch, Toby D.
Pubblicazione: (2024)
di: Pilditch, Toby D.
Pubblicazione: (2024)
Justifications for Democratizing AI Alignment and Their Prospects
di: Steingrüber, André, et al.
Pubblicazione: (2025)
di: Steingrüber, André, et al.
Pubblicazione: (2025)
Identifying Quantum Structure in AI Language: Evidence for Evolutionary Convergence of Human and Artificial Cognition
di: Aerts, Diederik, et al.
Pubblicazione: (2025)
di: Aerts, Diederik, et al.
Pubblicazione: (2025)
The Elephant in the Room -- Why AI Safety Demands Diverse Teams
di: Rostcheck, David, et al.
Pubblicazione: (2024)
di: Rostcheck, David, et al.
Pubblicazione: (2024)
Human-AI Interactions: Cognitive, Behavioral, and Emotional Impacts
di: Riley, Celeste, et al.
Pubblicazione: (2025)
di: Riley, Celeste, et al.
Pubblicazione: (2025)
The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers
di: Chen, Benjamin Minhao, et al.
Pubblicazione: (2026)
di: Chen, Benjamin Minhao, et al.
Pubblicazione: (2026)
The Illusion of Friendship: Why Generative AI Demands Unprecedented Ethical Vigilance
di: Islam, Md Zahidul
Pubblicazione: (2026)
di: Islam, Md Zahidul
Pubblicazione: (2026)
From Sycophantic Consensus to Pluralistic Repair: Why AI Alignment Must Surface Disagreement
di: Vishwarupe, Varad, et al.
Pubblicazione: (2026)
di: Vishwarupe, Varad, et al.
Pubblicazione: (2026)
Privacy Ethics Alignment in AI: A Stakeholder-Centric Framework for Ethical AI
di: Barthwal, Ankur, et al.
Pubblicazione: (2025)
di: Barthwal, Ankur, et al.
Pubblicazione: (2025)
Unilateral Relationship Revision Power in Human-AI Companion Interaction
di: Lange, Benjamin
Pubblicazione: (2026)
di: Lange, Benjamin
Pubblicazione: (2026)
Whose Truth? Pluralistic Geo-Alignment for (Agentic) AI
di: Janowicz, Krzysztof, et al.
Pubblicazione: (2025)
di: Janowicz, Krzysztof, et al.
Pubblicazione: (2025)
Classifying Epistemic Relationships in Human-AI Interaction: An Exploratory Approach
di: Yang, Shengnan, et al.
Pubblicazione: (2025)
di: Yang, Shengnan, et al.
Pubblicazione: (2025)
Mutual Wanting in Human--AI Interaction: Empirical Evidence from Large-Scale Analysis of GPT Model Transitions
di: Shang, HaoYang, et al.
Pubblicazione: (2025)
di: Shang, HaoYang, et al.
Pubblicazione: (2025)
Racial/Ethnic Categories in AI and Algorithmic Fairness: Why They Matter and What They Represent
di: Mickel, Jennifer
Pubblicazione: (2024)
di: Mickel, Jennifer
Pubblicazione: (2024)
Why Trust in AI May Be Inevitable
di: Truong, Nghi, et al.
Pubblicazione: (2025)
di: Truong, Nghi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
HawkesRank: Event-Driven Centrality for Real-Time Importance Ranking
di: Sornette, Didier, et al.
Pubblicazione: (2026) -
Alignment Is Not Enough: A Relational Framework for Moral Standing in Human-AI Interaction
di: Pasandi, Faezeh B., et al.
Pubblicazione: (2026) -
Evolutionary Reinforcement Learning based AI tutor for Socratic Interdisciplinary Instruction
di: Jiang, Mei, et al.
Pubblicazione: (2025) -
Understanding the Process of Human-AI Value Alignment
di: McKinlay, Jack, et al.
Pubblicazione: (2025) -
Misalignment or misuse? The AGI alignment tradeoff
di: Hellrigel-Holderbaum, Max, et al.
Pubblicazione: (2025)