Salvato in:
| Autore principale: | Hastings-Woodhouse, Sarah |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2507.21082 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An Approach to Technical AGI Safety and Security
di: Shah, Rohin, et al.
Pubblicazione: (2025)
di: Shah, Rohin, et al.
Pubblicazione: (2025)
Institutional AI: A Governance Framework for Distributional AGI Safety
di: Pierucci, Federico, et al.
Pubblicazione: (2026)
di: Pierucci, Federico, et al.
Pubblicazione: (2026)
AGI, Governments, and Free Societies
di: Bullock, Justin B., et al.
Pubblicazione: (2025)
di: Bullock, Justin B., et al.
Pubblicazione: (2025)
Europe and the Geopolitics of AGI: The Need for a Preparedness Plan
di: Negele, Maximilian, et al.
Pubblicazione: (2026)
di: Negele, Maximilian, et al.
Pubblicazione: (2026)
Several Issues Regarding Data Governance in AGI
di: Hatta, Masayuki
Pubblicazione: (2025)
di: Hatta, Masayuki
Pubblicazione: (2025)
Token Taxes: mitigating AGI's economic risks
di: Irwin, Lucas, et al.
Pubblicazione: (2026)
di: Irwin, Lucas, et al.
Pubblicazione: (2026)
Unsocial Intelligence: an Investigation of the Assumptions of AGI Discourse
di: Blili-Hamelin, Borhane, et al.
Pubblicazione: (2024)
di: Blili-Hamelin, Borhane, et al.
Pubblicazione: (2024)
Misalignment or misuse? The AGI alignment tradeoff
di: Hellrigel-Holderbaum, Max, et al.
Pubblicazione: (2025)
di: Hellrigel-Holderbaum, Max, et al.
Pubblicazione: (2025)
Institutional Management of Information Technology: A Centralised Approach.
di: Dockerill, John
Pubblicazione: (1987)
di: Dockerill, John
Pubblicazione: (1987)
Stop treating `AGI' as the north-star goal of AI research
di: Blili-Hamelin, Borhane, et al.
Pubblicazione: (2025)
di: Blili-Hamelin, Borhane, et al.
Pubblicazione: (2025)
The Trajectory of Romance Scams in the U.S
di: Herrera, LD, et al.
Pubblicazione: (2024)
di: Herrera, LD, et al.
Pubblicazione: (2024)
Policy myopia as a mechanism of gradual disempowerment in Post-AGI governance, Circa 2049
di: Sahoo, Subramanyam
Pubblicazione: (2026)
di: Sahoo, Subramanyam
Pubblicazione: (2026)
High vs. Low AGI: Ontology and Conceptual Taxonomy for Geopolitical Coherence
di: Max, Antonio
Pubblicazione: (2025)
di: Max, Antonio
Pubblicazione: (2025)
Efficiency vs Demand in AI Electricity: Implications for Post-AGI Scaling
di: Kim, Doyi, et al.
Pubblicazione: (2026)
di: Kim, Doyi, et al.
Pubblicazione: (2026)
Against racing to AGI: Cooperation, deterrence, and catastrophic risks
di: Dung, Leonard, et al.
Pubblicazione: (2025)
di: Dung, Leonard, et al.
Pubblicazione: (2025)
The Invisibility Hypothesis: Promises of AGI and the Future of the Global South
di: López, L. Julián Lechuga, et al.
Pubblicazione: (2026)
di: López, L. Julián Lechuga, et al.
Pubblicazione: (2026)
From Checklists to Clusters: A Homeostatic Account of AGI Evaluation
di: Reynolds, Brett
Pubblicazione: (2025)
di: Reynolds, Brett
Pubblicazione: (2025)
How Far Are We From AGI: Are LLMs All We Need?
di: Feng, Tao, et al.
Pubblicazione: (2024)
di: Feng, Tao, et al.
Pubblicazione: (2024)
Towards AI-$45^{\circ}$ Law: A Roadmap to Trustworthy AGI
di: Yang, Chao, et al.
Pubblicazione: (2024)
di: Yang, Chao, et al.
Pubblicazione: (2024)
Keep the Future Human: Why and How We Should Close the Gates to AGI and Superintelligence, and What We Should Build Instead
di: Aguirre, Anthony
Pubblicazione: (2023)
di: Aguirre, Anthony
Pubblicazione: (2023)
Why AI Alignment Failure Is Structural: Learned Human Interaction Structures and AGI as an Endogenous Evolutionary Shock
di: Sornette, Didier, et al.
Pubblicazione: (2026)
di: Sornette, Didier, et al.
Pubblicazione: (2026)
Towards New Benchmark for AI Alignment & Sentiment Analysis in Socially Important Issues: A Comparative Study of Human and LLMs in the Context of AGI
di: Bojic, Ljubisa, et al.
Pubblicazione: (2025)
di: Bojic, Ljubisa, et al.
Pubblicazione: (2025)
Safety Co-Option and Compromised National Security: The Self-Fulfilling Prophecy of Weakened AI Risk Thresholds
di: Khlaaf, Heidy, et al.
Pubblicazione: (2025)
di: Khlaaf, Heidy, et al.
Pubblicazione: (2025)
Dual-Use AI Face Swap Apps Are Mostly Unsafe: A Systematic Safety Audit
di: Daffalla, Alaa, et al.
Pubblicazione: (2026)
di: Daffalla, Alaa, et al.
Pubblicazione: (2026)
Toward Secure and Compliant AI: Organizational Standards and Protocols for NLP Model Lifecycle Management
di: Arora, Sunil, et al.
Pubblicazione: (2025)
di: Arora, Sunil, et al.
Pubblicazione: (2025)
Jolting Technologies: Superexponential Acceleration in AI Capabilities and Implications for AGI
di: Orban, David
Pubblicazione: (2025)
di: Orban, David
Pubblicazione: (2025)
The 2025 AI Agent Index: Documenting Technical and Safety Features of Deployed Agentic AI Systems
di: Staufer, Leon, et al.
Pubblicazione: (2026)
di: Staufer, Leon, et al.
Pubblicazione: (2026)
Safety First: Psychological Safety as the Key to AI Transformation
di: Reich, Aaron, et al.
Pubblicazione: (2026)
di: Reich, Aaron, et al.
Pubblicazione: (2026)
How Should AI Safety Benchmarks Benchmark Safety?
di: Yu, Cheng, et al.
Pubblicazione: (2026)
di: Yu, Cheng, et al.
Pubblicazione: (2026)
Securing Agentic AI Systems -- A Multilayer Security Framework
di: Arora, Sunil, et al.
Pubblicazione: (2025)
di: Arora, Sunil, et al.
Pubblicazione: (2025)
Autonomous Penetration Testing: Solving Capture-the-Flag Challenges with LLMs
di: Bakker, Isabelle, et al.
Pubblicazione: (2025)
di: Bakker, Isabelle, et al.
Pubblicazione: (2025)
The Arrival of AGI? When Expert Personas Exceed Expert Benchmarks
di: Mullens, Drake, et al.
Pubblicazione: (2026)
di: Mullens, Drake, et al.
Pubblicazione: (2026)
Data Augmentation via Diffusion Model to Enhance AI Fairness
di: Blow, Christina Hastings, et al.
Pubblicazione: (2024)
di: Blow, Christina Hastings, et al.
Pubblicazione: (2024)
Some Simple Economics of AGI
di: Catalini, Christian, et al.
Pubblicazione: (2026)
di: Catalini, Christian, et al.
Pubblicazione: (2026)
SafetyAnalyst: Interpretable, Transparent, and Steerable Safety Moderation for AI Behavior
di: Li, Jing-Jing, et al.
Pubblicazione: (2024)
di: Li, Jing-Jing, et al.
Pubblicazione: (2024)
AI Safety for Everyone
di: Gyevnar, Balint, et al.
Pubblicazione: (2025)
di: Gyevnar, Balint, et al.
Pubblicazione: (2025)
The Role of AI Safety Institutes in Contributing to International Standards for Frontier AI Safety
di: Fort, Kristina
Pubblicazione: (2024)
di: Fort, Kristina
Pubblicazione: (2024)
Scalable and Ethical Insider Threat Detection through Data Synthesis and Analysis by LLMs
di: Gelman, Haywood, et al.
Pubblicazione: (2025)
di: Gelman, Haywood, et al.
Pubblicazione: (2025)
Safety cases for frontier AI
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2024)
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2024)
Auditing Agent Harness Safety
di: Liu, Chengzhi, et al.
Pubblicazione: (2026)
di: Liu, Chengzhi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
An Approach to Technical AGI Safety and Security
di: Shah, Rohin, et al.
Pubblicazione: (2025) -
Institutional AI: A Governance Framework for Distributional AGI Safety
di: Pierucci, Federico, et al.
Pubblicazione: (2026) -
AGI, Governments, and Free Societies
di: Bullock, Justin B., et al.
Pubblicazione: (2025) -
Europe and the Geopolitics of AGI: The Need for a Preparedness Plan
di: Negele, Maximilian, et al.
Pubblicazione: (2026) -
Several Issues Regarding Data Governance in AGI
di: Hatta, Masayuki
Pubblicazione: (2025)