Third-party compliance reviews for frontier AI safety frameworks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Homewood, Aidan, Williams, Sophie, Dreksler, Noemi, Lidiard, John, Murray, Malcolm, Heim, Lennart, Ziosi, Marta, hÉigeartaigh, Seán Ó, Chen, Michael, Wei, Kevin, Winter, Christoph, Brundage, Miles, Garfinkel, Ben, Schuett, Jonas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Verifying International Agreements on AI: Six Layers of Verification for Rules on Large-Scale AI Development and Deployment
von: Baker, Mauricio, et al.
Veröffentlicht: (2025)
von: Baker, Mauricio, et al.
Veröffentlicht: (2025)
Limits of Safe AI Deployment: Differentiating Oversight and Control
von: Manheim, David, et al.
Veröffentlicht: (2025)
von: Manheim, David, et al.
Veröffentlicht: (2025)
Risk thresholds for frontier AI
von: Koessler, Leonie, et al.
Veröffentlicht: (2024)
von: Koessler, Leonie, et al.
Veröffentlicht: (2024)
The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation
von: Brundage, Miles, et al.
Veröffentlicht: (2018)
von: Brundage, Miles, et al.
Veröffentlicht: (2018)
From Principles to Rules: A Regulatory Approach for Frontier AI
von: Schuett, Jonas, et al.
Veröffentlicht: (2024)
von: Schuett, Jonas, et al.
Veröffentlicht: (2024)
Evidence of What, for Whom? The Socially Contested Role of Algorithmic Bias in a Predictive Policing Tool
von: Ziosi, Marta, et al.
Veröffentlicht: (2024)
von: Ziosi, Marta, et al.
Veröffentlicht: (2024)
Safety cases for frontier AI
von: Buhl, Marie Davidsen, et al.
Veröffentlicht: (2024)
von: Buhl, Marie Davidsen, et al.
Veröffentlicht: (2024)
Local US officials' views on the impacts and governance of AI: Evidence from 2022 and 2023 survey waves
von: Hatz, Sophia, et al.
Veröffentlicht: (2025)
von: Hatz, Sophia, et al.
Veröffentlicht: (2025)
Synthetic Data for Veterinary EHR De-identification: Benefits, Limits, and Safety Trade-offs Under Fixed Compute
von: Brundage, David
Veröffentlicht: (2026)
von: Brundage, David
Veröffentlicht: (2026)
Generating Synthetic Wildlife Health Data from Camera Trap Imagery: A Pipeline for Alopecia and Body Condition Training Data
von: Brundage, David
Veröffentlicht: (2026)
von: Brundage, David
Veröffentlicht: (2026)
The Experience of Academic Library Deans and Directors during the COVID-19 Pandemic: An Interpretive Phenomenological Analysis
von: Kenneth S. Brundage
Veröffentlicht: (2022)
von: Kenneth S. Brundage
Veröffentlicht: (2022)
Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies
von: Brundage, Miles, et al.
Veröffentlicht: (2026)
von: Brundage, Miles, et al.
Veröffentlicht: (2026)
Third party payments in health microinsurance
von: Pascale Le Roy, et al.
Veröffentlicht: (2011)
von: Pascale Le Roy, et al.
Veröffentlicht: (2011)
Editor’s Introduction to Drudgery Divine at 35: Reconsiderations and Applications
von: Nathanael J. Homewood
Veröffentlicht: (2025)
von: Nathanael J. Homewood
Veröffentlicht: (2025)
Editor's Introduction: Approaching Qur'an Commentary through the Academic Study of Religion: A Forum on Tehseen Thaver's Beyond Sectarianism
von: Nathanael J. Homewood
Veröffentlicht: (2025)
von: Nathanael J. Homewood
Veröffentlicht: (2025)
On Regulating Downstream AI Developers
von: Williams, Sophie, et al.
Veröffentlicht: (2025)
von: Williams, Sophie, et al.
Veröffentlicht: (2025)
Harold Garfinkel: Studies of Work in the Sciences
von: Garfinkel, Harold
Veröffentlicht: (2022)
von: Garfinkel, Harold
Veröffentlicht: (2022)
El renacimiento de Buda --
von: Garfinkel, Perry
Veröffentlicht: (2005)
von: Garfinkel, Perry
Veröffentlicht: (2005)
Towards the right standards: The intersection of open science, responsible research and innovation, and standards
von: Michele Garfinkel
Veröffentlicht: (2021)
von: Michele Garfinkel
Veröffentlicht: (2021)
Can Third-parties Read Our Emotions?
von: Li, Jiayi, et al.
Veröffentlicht: (2025)
von: Li, Jiayi, et al.
Veröffentlicht: (2025)
Third party payment mechanism in health microinsurance
von: Pascale Le Roy, et al.
Veröffentlicht: (2011)
von: Pascale Le Roy, et al.
Veröffentlicht: (2011)
Annotated record of the detailed examination of Mn deposits from the Blake Plateau and the U.S.A. East Coast Continental Margin
von: Brundage, W L
Veröffentlicht: (1972)
von: Brundage, W L
Veröffentlicht: (1972)
Application of Stochastic Control Algorithms for the Improvement of the Electron Injection Efficiency of BESSY II
von: Schuett, Alexander
Veröffentlicht: (2024)
von: Schuett, Alexander
Veröffentlicht: (2024)
Risk management in the Artificial Intelligence Act
von: Schuett, Jonas
Veröffentlicht: (2022)
von: Schuett, Jonas
Veröffentlicht: (2022)
Three lines of defense against risks from AI
von: Schuett, Jonas
Veröffentlicht: (2022)
von: Schuett, Jonas
Veröffentlicht: (2022)
Frontier AI developers need an internal audit function
von: Schuett, Jonas
Veröffentlicht: (2023)
von: Schuett, Jonas
Veröffentlicht: (2023)
Defining the scope of AI regulations
von: Schuett, Jonas
Veröffentlicht: (2019)
von: Schuett, Jonas
Veröffentlicht: (2019)
Frontier AI developers need an internal audit function
von: Jonas Schuett
Veröffentlicht: (2024)
von: Jonas Schuett
Veröffentlicht: (2024)
Designing Incident Reporting Systems for Harms from General-Purpose AI
von: Wei, Kevin, et al.
Veröffentlicht: (2025)
von: Wei, Kevin, et al.
Veröffentlicht: (2025)
Training Compute Thresholds: Features and Functions in AI Regulation
von: Heim, Lennart, et al.
Veröffentlicht: (2024)
von: Heim, Lennart, et al.
Veröffentlicht: (2024)
Predictable Artificial Intelligence
von: Zhou, Lexin, et al.
Veröffentlicht: (2023)
von: Zhou, Lexin, et al.
Veröffentlicht: (2023)
Safety case template for frontier AI: A cyber inability argument
von: Goemans, Arthur, et al.
Veröffentlicht: (2024)
von: Goemans, Arthur, et al.
Veröffentlicht: (2024)
Effective Mitigations for Systemic Risks from General-Purpose AI
von: Uuk, Risto, et al.
Veröffentlicht: (2024)
von: Uuk, Risto, et al.
Veröffentlicht: (2024)
Dynamic safety cases for frontier AI
von: Cârlan, Carmen, et al.
Veröffentlicht: (2024)
von: Cârlan, Carmen, et al.
Veröffentlicht: (2024)
Expanding External Access To Frontier AI Models For Dangerous Capability Evaluations
von: Charnock, Jacob, et al.
Veröffentlicht: (2026)
von: Charnock, Jacob, et al.
Veröffentlicht: (2026)
Third-party Provision of Conversion Technologies in Network Markets
von: Arnaud X. Varé
Veröffentlicht: (2009)
von: Arnaud X. Varé
Veröffentlicht: (2009)
The coordination gap in frontier AI safety policies
von: Mengesha, Isaak
Veröffentlicht: (2026)
von: Mengesha, Isaak
Veröffentlicht: (2026)
Assessing confidence in frontier AI safety cases
von: Barrett, Stephen, et al.
Veröffentlicht: (2025)
von: Barrett, Stephen, et al.
Veröffentlicht: (2025)
Third-party management in software development: proposal of a methodology
von: Yeison Núñez-Sánchez
Veröffentlicht: (2020)
von: Yeison Núñez-Sánchez
Veröffentlicht: (2020)
Consider who has responsibility for off‐campus safety
von: David Miles
Veröffentlicht: (2024)
von: David Miles
Veröffentlicht: (2024)
Ähnliche Einträge
-
Verifying International Agreements on AI: Six Layers of Verification for Rules on Large-Scale AI Development and Deployment
von: Baker, Mauricio, et al.
Veröffentlicht: (2025) -
Limits of Safe AI Deployment: Differentiating Oversight and Control
von: Manheim, David, et al.
Veröffentlicht: (2025) -
Risk thresholds for frontier AI
von: Koessler, Leonie, et al.
Veröffentlicht: (2024) -
The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation
von: Brundage, Miles, et al.
Veröffentlicht: (2018) -
From Principles to Rules: A Regulatory Approach for Frontier AI
von: Schuett, Jonas, et al.
Veröffentlicht: (2024)