Safety cases for frontier AI
Fuente:
arXiv
Saved in:
| Main Authors: | Buhl, Marie Davidsen, Sett, Gaurav, Koessler, Leonie, Schuett, Jonas, Anderljung, Markus |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Risk thresholds for frontier AI
by: Koessler, Leonie, et al.
Published: (2024)
by: Koessler, Leonie, et al.
Published: (2024)
From Principles to Rules: A Regulatory Approach for Frontier AI
by: Schuett, Jonas, et al.
Published: (2024)
by: Schuett, Jonas, et al.
Published: (2024)
Safety case template for frontier AI: A cyber inability argument
by: Goemans, Arthur, et al.
Published: (2024)
by: Goemans, Arthur, et al.
Published: (2024)
A Grading Rubric for AI Safety Frameworks
by: Alaga, Jide, et al.
Published: (2024)
by: Alaga, Jide, et al.
Published: (2024)
On Regulating Downstream AI Developers
by: Williams, Sophie, et al.
Published: (2025)
by: Williams, Sophie, et al.
Published: (2025)
Emerging Practices in Frontier AI Safety Frameworks
by: Buhl, Marie Davidsen, et al.
Published: (2025)
by: Buhl, Marie Davidsen, et al.
Published: (2025)
Safety Cases: A Scalable Approach to Frontier AI Safety
by: Hilton, Benjamin, et al.
Published: (2025)
by: Hilton, Benjamin, et al.
Published: (2025)
Training Compute Thresholds: Features and Functions in AI Regulation
by: Heim, Lennart, et al.
Published: (2024)
by: Heim, Lennart, et al.
Published: (2024)
Defining the scope of AI regulations
by: Schuett, Jonas
Published: (2019)
by: Schuett, Jonas
Published: (2019)
Three lines of defense against risks from AI
by: Schuett, Jonas
Published: (2022)
by: Schuett, Jonas
Published: (2022)
Frontier AI developers need an internal audit function
by: Schuett, Jonas
Published: (2023)
by: Schuett, Jonas
Published: (2023)
Risk management in the Artificial Intelligence Act
by: Schuett, Jonas
Published: (2022)
by: Schuett, Jonas
Published: (2022)
From Turing to Tomorrow: The UK's Approach to AI Regulation
by: Ritchie, Oliver, et al.
Published: (2025)
by: Ritchie, Oliver, et al.
Published: (2025)
Third-party compliance reviews for frontier AI safety frameworks
by: Homewood, Aidan, et al.
Published: (2025)
by: Homewood, Aidan, et al.
Published: (2025)
Dynamic safety cases for frontier AI
by: Cârlan, Carmen, et al.
Published: (2024)
by: Cârlan, Carmen, et al.
Published: (2024)
Measuring AI R&D Automation
by: Chan, Alan, et al.
Published: (2026)
by: Chan, Alan, et al.
Published: (2026)
Assessing confidence in frontier AI safety cases
by: Barrett, Stephen, et al.
Published: (2025)
by: Barrett, Stephen, et al.
Published: (2025)
Societal Adaptation to Advanced AI
by: Bernardi, Jamie, et al.
Published: (2024)
by: Bernardi, Jamie, et al.
Published: (2024)
Towards interactive evaluations for interaction harms in human-AI systems
by: Ibrahim, Lujain, et al.
Published: (2024)
by: Ibrahim, Lujain, et al.
Published: (2024)
Report on the Conference on Ethical and Responsible Design in the National AI Institutes: A Summary of Challenges
by: Conklin, Sherri Lynn, et al.
Published: (2024)
by: Conklin, Sherri Lynn, et al.
Published: (2024)
Responsible Reporting for Frontier AI Development
by: Kolt, Noam, et al.
Published: (2024)
by: Kolt, Noam, et al.
Published: (2024)
The rising costs of training frontier AI models
by: Cottier, Ben, et al.
Published: (2024)
by: Cottier, Ben, et al.
Published: (2024)
Domestic frontier AI regulation, an IAEA for AI, an NPT for AI, and a US-led Allied Public-Private Partnership for AI: Four institutions for governing and developing frontier AI
by: Belfield, Haydn
Published: (2025)
by: Belfield, Haydn
Published: (2025)
Visibility into AI Agents
by: Chan, Alan, et al.
Published: (2024)
by: Chan, Alan, et al.
Published: (2024)
The coordination gap in frontier AI safety policies
by: Mengesha, Isaak
Published: (2026)
by: Mengesha, Isaak
Published: (2026)
Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies
by: Brundage, Miles, et al.
Published: (2026)
by: Brundage, Miles, et al.
Published: (2026)
How frontier AI companies could implement an internal audit function
by: Gomez, Francesca, et al.
Published: (2025)
by: Gomez, Francesca, et al.
Published: (2025)
An alignment safety case sketch based on debate
by: Buhl, Marie Davidsen, et al.
Published: (2025)
by: Buhl, Marie Davidsen, et al.
Published: (2025)
AI Safety for Everyone
by: Gyevnar, Balint, et al.
Published: (2025)
by: Gyevnar, Balint, et al.
Published: (2025)
The Role of AI Safety Institutes in Contributing to International Standards for Frontier AI Safety
by: Fort, Kristina
Published: (2024)
by: Fort, Kristina
Published: (2024)
Safety First: Psychological Safety as the Key to AI Transformation
by: Reich, Aaron, et al.
Published: (2026)
by: Reich, Aaron, et al.
Published: (2026)
How Should AI Safety Benchmarks Benchmark Safety?
by: Yu, Cheng, et al.
Published: (2026)
by: Yu, Cheng, et al.
Published: (2026)
AI Safety, Alignment, and Ethics (AI SAE)
by: Waldner, Dylan
Published: (2025)
by: Waldner, Dylan
Published: (2025)
How malicious AI swarms can threaten democracy: The fusion of agentic AI and LLMs marks a new frontier in information warfare
by: Schroeder, Daniel Thilo, et al.
Published: (2025)
by: Schroeder, Daniel Thilo, et al.
Published: (2025)
Frontier AI developers need an internal audit function
by: Jonas Schuett
Published: (2024)
by: Jonas Schuett
Published: (2024)
The BIG Argument for AI Safety Cases
by: Habli, Ibrahim, et al.
Published: (2025)
by: Habli, Ibrahim, et al.
Published: (2025)
Persuasion and Safety in the Era of Generative AI
by: Kong, Haein
Published: (2025)
by: Kong, Haein
Published: (2025)
International AI Safety Report 2026
by: Bengio, Yoshua, et al.
Published: (2026)
by: Bengio, Yoshua, et al.
Published: (2026)
SafetyAnalyst: Interpretable, Transparent, and Steerable Safety Moderation for AI Behavior
by: Li, Jing-Jing, et al.
Published: (2024)
by: Li, Jing-Jing, et al.
Published: (2024)
AI Safety is Stuck in Technical Terms -- A System Safety Response to the International AI Safety Report
by: Dobbe, Roel
Published: (2025)
by: Dobbe, Roel
Published: (2025)
Similar Items
-
Risk thresholds for frontier AI
by: Koessler, Leonie, et al.
Published: (2024) -
From Principles to Rules: A Regulatory Approach for Frontier AI
by: Schuett, Jonas, et al.
Published: (2024) -
Safety case template for frontier AI: A cyber inability argument
by: Goemans, Arthur, et al.
Published: (2024) -
A Grading Rubric for AI Safety Frameworks
by: Alaga, Jide, et al.
Published: (2024) -
On Regulating Downstream AI Developers
by: Williams, Sophie, et al.
Published: (2025)