Evaluating AI Providers' Frontier Safety Frameworks
Fuente:
arXiv
Saved in:
| Main Authors: | Stelling, Lily, Murray, Malcolm, Galizzi, Bruno, Schaffelder, Max, Campos, Siméon, Papadatos, Henry |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Methodology for Quantitative AI Risk Modeling
by: Murray, Malcolm, et al.
Published: (2025)
by: Murray, Malcolm, et al.
Published: (2025)
The Role of Risk Modeling in Advanced AI Risk Management
by: Touzet, Chloé, et al.
Published: (2025)
by: Touzet, Chloé, et al.
Published: (2025)
Mapping Industry Practices to the EU AI Act's GPAI Code of Practice Safety and Security Measures
by: Stelling, Lily, et al.
Published: (2025)
by: Stelling, Lily, et al.
Published: (2025)
A Frontier AI Risk Management Framework: Bridging the Gap Between Current AI Practices and Established Risk Management
by: Campos, Simeon, et al.
Published: (2025)
by: Campos, Simeon, et al.
Published: (2025)
Lessons from External Review of DeepMind's Scheming Inability Safety Case
by: Barrett, Stephen, et al.
Published: (2026)
by: Barrett, Stephen, et al.
Published: (2026)
Mapping AI Benchmark Data to Quantitative Risk Estimates Through Expert Elicitation
by: Murray, Malcolm, et al.
Published: (2025)
by: Murray, Malcolm, et al.
Published: (2025)
Emerging Practices in Frontier AI Safety Frameworks
by: Buhl, Marie Davidsen, et al.
Published: (2025)
by: Buhl, Marie Davidsen, et al.
Published: (2025)
Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies
by: Brundage, Miles, et al.
Published: (2026)
by: Brundage, Miles, et al.
Published: (2026)
Toward Quantitative Modeling of Cybersecurity Risks Due to AI Misuse
by: Barrett, Steve, et al.
Published: (2025)
by: Barrett, Steve, et al.
Published: (2025)
The Role of AI Safety Institutes in Contributing to International Standards for Frontier AI Safety
by: Fort, Kristina
Published: (2024)
by: Fort, Kristina
Published: (2024)
ForesightSafety Bench: A Frontier Risk Evaluation and Governance Framework towards Safe AI
by: Tong, Haibo, et al.
Published: (2026)
by: Tong, Haibo, et al.
Published: (2026)
Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning
by: Schaffelder, Max, et al.
Published: (2025)
by: Schaffelder, Max, et al.
Published: (2025)
Safety Cases: A Scalable Approach to Frontier AI Safety
by: Hilton, Benjamin, et al.
Published: (2025)
by: Hilton, Benjamin, et al.
Published: (2025)
Enabling Frontier Lab Collaboration to Mitigate AI Safety Risks
by: Felstead, Nicholas
Published: (2025)
by: Felstead, Nicholas
Published: (2025)
Open Problems in Frontier AI Risk Management
by: Ziosi, Marta, et al.
Published: (2026)
by: Ziosi, Marta, et al.
Published: (2026)
Evaluating the Critical Risks of Amazon's Nova Premier under the Frontier Model Safety Framework
by: Krishna, Satyapriya, et al.
Published: (2025)
by: Krishna, Satyapriya, et al.
Published: (2025)
Towards Frontier Safety Policies Plus
by: Pistillo, Matteo
Published: (2025)
by: Pistillo, Matteo
Published: (2025)
FORTRESS: Frontier Risk Evaluation for National Security and Public Safety
by: Knight, Christina Q., et al.
Published: (2025)
by: Knight, Christina Q., et al.
Published: (2025)
Machine Learning for Public Good: Predicting Urban Crime Patterns to Enhance Community Safety
by: Gupta, Sia, et al.
Published: (2024)
by: Gupta, Sia, et al.
Published: (2024)
Expanding External Access To Frontier AI Models For Dangerous Capability Evaluations
by: Charnock, Jacob, et al.
Published: (2026)
by: Charnock, Jacob, et al.
Published: (2026)
Artificially Fluent: Swahili AI Performance Benchmarks Between English-Trained and Natively-Trained Datasets
by: Jaffer, Sophie, et al.
Published: (2025)
by: Jaffer, Sophie, et al.
Published: (2025)
Vernacularizing Taxonomies of Harm is Essential for Operationalizing Holistic AI Safety
by: Kennedy, Wm. Matthew, et al.
Published: (2024)
by: Kennedy, Wm. Matthew, et al.
Published: (2024)
Evaluating the Goal-Directedness of Large Language Models
by: Everitt, Tom, et al.
Published: (2025)
by: Everitt, Tom, et al.
Published: (2025)
A Grading Rubric for AI Safety Frameworks
by: Alaga, Jide, et al.
Published: (2024)
by: Alaga, Jide, et al.
Published: (2024)
The science and practice of proportionality in AI risk evaluations
by: Mougan, Carlos, et al.
Published: (2026)
by: Mougan, Carlos, et al.
Published: (2026)
Clear, Compelling Arguments: Rethinking the Foundations of Frontier AI Safety Cases
by: Feakins, Shaun, et al.
Published: (2026)
by: Feakins, Shaun, et al.
Published: (2026)
The California Report on Frontier AI Policy
by: Bommasani, Rishi, et al.
Published: (2025)
by: Bommasani, Rishi, et al.
Published: (2025)
AI Safety Evaluations Need To Consider Cascading Effects
by: Neumann, Anna, et al.
Published: (2026)
by: Neumann, Anna, et al.
Published: (2026)
Assessing the Case for Africa-Centric AI Safety Evaluations
by: Ireri, Gathoni, et al.
Published: (2026)
by: Ireri, Gathoni, et al.
Published: (2026)
Biothreat Benchmark Generation Framework for Evaluating Frontier AI Models I: The Task-Query Architecture
by: Ackerman, Gary, et al.
Published: (2025)
by: Ackerman, Gary, et al.
Published: (2025)
SafetyAnalyst: Interpretable, Transparent, and Steerable Safety Moderation for AI Behavior
by: Li, Jing-Jing, et al.
Published: (2024)
by: Li, Jing-Jing, et al.
Published: (2024)
Assurance of Frontier AI Built for National Security
by: Pistillo, Matteo, et al.
Published: (2025)
by: Pistillo, Matteo, et al.
Published: (2025)
Evaluating Human-AI Safety: A Framework for Measuring Harmful Capability Uplift
by: Vaccaro, Michelle, et al.
Published: (2026)
by: Vaccaro, Michelle, et al.
Published: (2026)
Institutional AI: A Governance Framework for Distributional AGI Safety
by: Pierucci, Federico, et al.
Published: (2026)
by: Pierucci, Federico, et al.
Published: (2026)
Questionnaire Responses Do not Capture the Safety of AI Agents
by: Hellrigel-Holderbaum, Max, et al.
Published: (2026)
by: Hellrigel-Holderbaum, Max, et al.
Published: (2026)
Frontier Lag: A Bibliometric Audit of Capability Misrepresentation in Academic AI Evaluation
by: Gringras, David, et al.
Published: (2026)
by: Gringras, David, et al.
Published: (2026)
Frontier AI Ethics: Anticipating and Evaluating the Societal Impacts of Language Model Agents
by: Lazar, Seth
Published: (2024)
by: Lazar, Seth
Published: (2024)
Affirmative safety: An approach to risk management for high-risk AI
by: Wasil, Akash R., et al.
Published: (2024)
by: Wasil, Akash R., et al.
Published: (2024)
Frontier AI developers need an internal audit function
by: Schuett, Jonas
Published: (2023)
by: Schuett, Jonas
Published: (2023)
AI Safety Frameworks Should Include Procedures for Model Access Decisions
by: Kembery, Edward, et al.
Published: (2024)
by: Kembery, Edward, et al.
Published: (2024)
Similar Items
-
A Methodology for Quantitative AI Risk Modeling
by: Murray, Malcolm, et al.
Published: (2025) -
The Role of Risk Modeling in Advanced AI Risk Management
by: Touzet, Chloé, et al.
Published: (2025) -
Mapping Industry Practices to the EU AI Act's GPAI Code of Practice Safety and Security Measures
by: Stelling, Lily, et al.
Published: (2025) -
A Frontier AI Risk Management Framework: Bridging the Gap Between Current AI Practices and Established Risk Management
by: Campos, Simeon, et al.
Published: (2025) -
Lessons from External Review of DeepMind's Scheming Inability Safety Case
by: Barrett, Stephen, et al.
Published: (2026)