Is Power-Seeking AI an Existential Risk?
Fuente:
arXiv
Saved in:
| Main Author: | Carlsmith, Joseph |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Two Types of AI Existential Risk: Decisive and Accumulative
by: Kasirzadeh, Atoosa
Published: (2024)
by: Kasirzadeh, Atoosa
Published: (2024)
Existential Conversations with Large Language Models: Content, Community, and Culture
by: Shanahan, Murray, et al.
Published: (2024)
by: Shanahan, Murray, et al.
Published: (2024)
AI Consciousness and Existential Risk
by: VanRullen, Rufin
Published: (2025)
by: VanRullen, Rufin
Published: (2025)
AI-Powered Autonomous Weapons Risk Geopolitical Instability and Threaten AI Research
by: Simmons-Edler, Riley, et al.
Published: (2024)
by: Simmons-Edler, Riley, et al.
Published: (2024)
How Effective Is Constitutional AI in Small LLMs? A Study on DeepSeek-R1 and Its Peers
by: Menke, Antonio-Gabriel Chacón, et al.
Published: (2025)
by: Menke, Antonio-Gabriel Chacón, et al.
Published: (2025)
Risk Sources and Risk Management Measures in Support of Standards for General-Purpose AI Systems
by: Gipiškis, Rokas, et al.
Published: (2024)
by: Gipiškis, Rokas, et al.
Published: (2024)
AssurAI: Experience with Constructing Korean Socio-cultural Datasets to Discover Potential Risks of Generative AI
by: Lim, Chae-Gyun, et al.
Published: (2025)
by: Lim, Chae-Gyun, et al.
Published: (2025)
Dimensional Characterization and Pathway Modeling for Catastrophic AI Risks
by: Chin, Ze Shen
Published: (2025)
by: Chin, Ze Shen
Published: (2025)
The AI Risk Spectrum: From Dangerous Capabilities to Existential Threats
by: Grey, Markov, et al.
Published: (2025)
by: Grey, Markov, et al.
Published: (2025)
Humanity in the Age of AI: Reassessing 2025's Existential-Risk Narratives
by: Louadi, Mohamed El
Published: (2025)
by: Louadi, Mohamed El
Published: (2025)
Liability and Insurance for Catastrophic Losses: the Nuclear Power Precedent and Lessons for AI
by: Trout, Cristian
Published: (2024)
by: Trout, Cristian
Published: (2024)
The Case for ESM3 as a General-Purpose AI Model with Systemic Risk Under the EU AI Act
by: Qureshi, Taro, et al.
Published: (2026)
by: Qureshi, Taro, et al.
Published: (2026)
Strategic Polysemy in AI Discourse: A Philosophical Analysis of Language, Hype, and Power
by: LaCroix, Travis, et al.
Published: (2026)
by: LaCroix, Travis, et al.
Published: (2026)
Diagnosing Hallucination Risk in AI Surgical Decision-Support: A Sequential Framework for Sequential Validation
by: Chen, Dong, et al.
Published: (2025)
by: Chen, Dong, et al.
Published: (2025)
Understanding and Mitigating Risks of Generative AI in Financial Services
by: Gehrmann, Sebastian, et al.
Published: (2025)
by: Gehrmann, Sebastian, et al.
Published: (2025)
Risks of AI Scientists: Prioritizing Safeguarding Over Autonomy
by: Tang, Xiangru, et al.
Published: (2024)
by: Tang, Xiangru, et al.
Published: (2024)
When Autonomy Breaks: The Hidden Existential Risk of AI
by: Krook, Joshua
Published: (2025)
by: Krook, Joshua
Published: (2025)
Open Problems in Frontier AI Risk Management
by: Ziosi, Marta, et al.
Published: (2026)
by: Ziosi, Marta, et al.
Published: (2026)
Standardizing Intelligence: Aligning Generative AI for Regulatory and Operational Compliance
by: Imperial, Joseph Marvin, et al.
Published: (2025)
by: Imperial, Joseph Marvin, et al.
Published: (2025)
Role and Use of Race in AI/ML Models Related to Health
by: Were, Martin C., et al.
Published: (2025)
by: Were, Martin C., et al.
Published: (2025)
Thousands of AI Authors on the Future of AI
by: Grace, Katja, et al.
Published: (2024)
by: Grace, Katja, et al.
Published: (2024)
AI Toolkit: Libraries and Essays for Exploring the Technology and Ethics of AI
by: Ho, Levin, et al.
Published: (2025)
by: Ho, Levin, et al.
Published: (2025)
Lessons for Editors of AI Incidents from the AI Incident Database
by: Paeth, Kevin, et al.
Published: (2024)
by: Paeth, Kevin, et al.
Published: (2024)
Regulating AI Adaptation: An Analysis of AI Medical Device Updates
by: Wu, Kevin, et al.
Published: (2024)
by: Wu, Kevin, et al.
Published: (2024)
Mapping the Potential of Explainable AI for Fairness Along the AI Lifecycle
by: Deck, Luca, et al.
Published: (2024)
by: Deck, Luca, et al.
Published: (2024)
Rigor in AI: Doing Rigorous AI Work Requires a Broader, Responsible AI-Informed Conception of Rigor
by: Olteanu, Alexandra, et al.
Published: (2025)
by: Olteanu, Alexandra, et al.
Published: (2025)
Practical Application and Limitations of AI Certification Catalogues in the Light of the AI Act
by: Autischer, Gregor, et al.
Published: (2025)
by: Autischer, Gregor, et al.
Published: (2025)
Adapting Probabilistic Risk Assessment for AI
by: Wisakanto, Anna Katariina, et al.
Published: (2025)
by: Wisakanto, Anna Katariina, et al.
Published: (2025)
Insuring Uninsurable Risks from AI: Government as Insurer of Last Resort
by: Trout, Cristian
Published: (2024)
by: Trout, Cristian
Published: (2024)
Beware! The AI Act Can Also Apply to Your AI Research Practices
by: Wernick, Alina, et al.
Published: (2025)
by: Wernick, Alina, et al.
Published: (2025)
AI-Cybersecurity Education Through Designing AI-based Cyberharassment Detection Lab
by: Okpala, Ebuka, et al.
Published: (2024)
by: Okpala, Ebuka, et al.
Published: (2024)
Defining AI Models and AI Systems: A Framework to Resolve the Boundary Problem
by: Sun, Yuanyuan, et al.
Published: (2026)
by: Sun, Yuanyuan, et al.
Published: (2026)
Towards AI Transparency and Accountability: A Global Framework for Exchanging Information on AI Systems
by: Buckley, Warren, et al.
Published: (2023)
by: Buckley, Warren, et al.
Published: (2023)
Measuring What AI Systems Might Do: Towards A Measurement Science in AI
by: Voudouris, Konstantinos, et al.
Published: (2026)
by: Voudouris, Konstantinos, et al.
Published: (2026)
Limits of trust in medical AI
by: Hatherley, Joshua
Published: (2025)
by: Hatherley, Joshua
Published: (2025)
International AI Safety Report
by: Bengio, Yoshua, et al.
Published: (2025)
by: Bengio, Yoshua, et al.
Published: (2025)
Towards Environmentally Equitable AI
by: Hajiesmaili, Mohammad, et al.
Published: (2024)
by: Hajiesmaili, Mohammad, et al.
Published: (2024)
AI Alignment at Your Discretion
by: Buyl, Maarten, et al.
Published: (2025)
by: Buyl, Maarten, et al.
Published: (2025)
Exploring the sustainable scaling of AI dilemma: A projective study of corporations' AI environmental impacts
by: Desroches, Clément, et al.
Published: (2025)
by: Desroches, Clément, et al.
Published: (2025)
AI and Generative AI Transforming Disaster Management: A Survey of Damage Assessment and Response Techniques
by: Raj, Aman, et al.
Published: (2025)
by: Raj, Aman, et al.
Published: (2025)
Similar Items
-
Two Types of AI Existential Risk: Decisive and Accumulative
by: Kasirzadeh, Atoosa
Published: (2024) -
Existential Conversations with Large Language Models: Content, Community, and Culture
by: Shanahan, Murray, et al.
Published: (2024) -
AI Consciousness and Existential Risk
by: VanRullen, Rufin
Published: (2025) -
AI-Powered Autonomous Weapons Risk Geopolitical Instability and Threaten AI Research
by: Simmons-Edler, Riley, et al.
Published: (2024) -
How Effective Is Constitutional AI in Small LLMs? A Study on DeepSeek-R1 and Its Peers
by: Menke, Antonio-Gabriel Chacón, et al.
Published: (2025)