Responsible Reporting for Frontier AI Development
Fuente:
arXiv
Saved in:
| Main Authors: | Kolt, Noam, Anderljung, Markus, Barnhart, Joslyn, Brass, Asher, Esvelt, Kevin, Hadfield, Gillian K., Heim, Lennart, Rodriguez, Mikel, Sandbrink, Jonas B., Woodside, Thomas |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Consequences of Humiliation
by: Barnhart, Joslyn
Published: (2021)
by: Barnhart, Joslyn
Published: (2021)
Governing AI Agents
by: Kolt, Noam
Published: (2025)
by: Kolt, Noam
Published: (2025)
Visibility into AI Agents
by: Chan, Alan, et al.
Published: (2024)
by: Chan, Alan, et al.
Published: (2024)
IDs for AI Systems
by: Chan, Alan, et al.
Published: (2024)
by: Chan, Alan, et al.
Published: (2024)
Societal Adaptation to Advanced AI
by: Bernardi, Jamie, et al.
Published: (2024)
by: Bernardi, Jamie, et al.
Published: (2024)
On Regulating Downstream AI Developers
by: Williams, Sophie, et al.
Published: (2025)
by: Williams, Sophie, et al.
Published: (2025)
Superintelligence and Law
by: Kolt, Noam
Published: (2026)
by: Kolt, Noam
Published: (2026)
From Principles to Rules: A Regulatory Approach for Frontier AI
by: Schuett, Jonas, et al.
Published: (2024)
by: Schuett, Jonas, et al.
Published: (2024)
Infrastructure for AI Agents
by: Chan, Alan, et al.
Published: (2025)
by: Chan, Alan, et al.
Published: (2025)
Designing Incident Reporting Systems for Harms from General-Purpose AI
by: Wei, Kevin, et al.
Published: (2025)
by: Wei, Kevin, et al.
Published: (2025)
Risk thresholds for frontier AI
by: Koessler, Leonie, et al.
Published: (2024)
by: Koessler, Leonie, et al.
Published: (2024)
A Grading Rubric for AI Safety Frameworks
by: Alaga, Jide, et al.
Published: (2024)
by: Alaga, Jide, et al.
Published: (2024)
Regulating AI Agents
by: Gardhouse, Kathrin, et al.
Published: (2026)
by: Gardhouse, Kathrin, et al.
Published: (2026)
Foundation models may exhibit staged progression in novel CBRN threat disclosure
by: Esvelt, Kevin M
Published: (2025)
by: Esvelt, Kevin M
Published: (2025)
Legal Infrastructure for Transformative AI Governance
by: Hadfield, Gillian K.
Published: (2026)
by: Hadfield, Gillian K.
Published: (2026)
Lessons from complexity theory for AI governance
by: Kolt, Noam, et al.
Published: (2025)
by: Kolt, Noam, et al.
Published: (2025)
An Economy of AI Agents
by: Hadfield, Gillian K., et al.
Published: (2025)
by: Hadfield, Gillian K., et al.
Published: (2025)
Regulatory Markets for AI Safety
by: Clark, Jack, et al.
Published: (2019)
by: Clark, Jack, et al.
Published: (2019)
Regulatory Markets: The Future of AI Governance
by: Hadfield, Gillian K., et al.
Published: (2023)
by: Hadfield, Gillian K., et al.
Published: (2023)
Build Agent Advocates, Not Platform Agents
by: Kapoor, Sayash, et al.
Published: (2025)
by: Kapoor, Sayash, et al.
Published: (2025)
From Turing to Tomorrow: The UK's Approach to AI Regulation
by: Ritchie, Oliver, et al.
Published: (2025)
by: Ritchie, Oliver, et al.
Published: (2025)
Safety cases for frontier AI
by: Buhl, Marie Davidsen, et al.
Published: (2024)
by: Buhl, Marie Davidsen, et al.
Published: (2024)
Training Compute Thresholds: Features and Functions in AI Regulation
by: Heim, Lennart, et al.
Published: (2024)
by: Heim, Lennart, et al.
Published: (2024)
Holistic Safety and Responsibility Evaluations of Advanced AI Models
by: Weidinger, Laura, et al.
Published: (2024)
by: Weidinger, Laura, et al.
Published: (2024)
The 2025 AI Agent Index: Documenting Technical and Safety Features of Deployed Agentic AI Systems
by: Staufer, Leon, et al.
Published: (2026)
by: Staufer, Leon, et al.
Published: (2026)
AI Model Registries: A Foundational Tool for AI Governance
by: McKernon, Elliot, et al.
Published: (2024)
by: McKernon, Elliot, et al.
Published: (2024)
On the formal ribbon extension of a quasitriangular Hopf algebra
by: Kolt, Quinn T.
Published: (2024)
by: Kolt, Quinn T.
Published: (2024)
Rational Silence and False Polarization: How Viewpoint Organizations and Recommender Systems Distort the Expression of Public Opinion
by: Sarkar, Atrisha, et al.
Published: (2024)
by: Sarkar, Atrisha, et al.
Published: (2024)
Measuring AI R&D Automation
by: Chan, Alan, et al.
Published: (2026)
by: Chan, Alan, et al.
Published: (2026)
Increased Compute Efficiency and the Diffusion of AI Capabilities
by: Pilz, Konstantin, et al.
Published: (2023)
by: Pilz, Konstantin, et al.
Published: (2023)
Verifying International Agreements on AI: Six Layers of Verification for Rules on Large-Scale AI Development and Deployment
by: Baker, Mauricio, et al.
Published: (2025)
by: Baker, Mauricio, et al.
Published: (2025)
Legal Alignment for Safe and Ethical AI
by: Kolt, Noam, et al.
Published: (2026)
by: Kolt, Noam, et al.
Published: (2026)
Towards interactive evaluations for interaction harms in human-AI systems
by: Ibrahim, Lujain, et al.
Published: (2024)
by: Ibrahim, Lujain, et al.
Published: (2024)
La música y el diseño sonoro en el cine / Juli n Woodside
by: Woodside, Julián
Published: (2014)
by: Woodside, Julián
Published: (2014)
Cine y memoria cultural: la ilusión del multiculturalismo a partir de dos películas mexicanas de animación
by: Julián Woodside
Published: (2012)
by: Julián Woodside
Published: (2012)
La historicidad del paisaje sonoro y la música popular
by: Julian Woodside
Published: (2008)
by: Julian Woodside
Published: (2008)
Talk Isn't Always Cheap: Understanding Failure Modes in Multi-Agent Debate
by: Wynn, Andrea, et al.
Published: (2025)
by: Wynn, Andrea, et al.
Published: (2025)
The AI Agent Index
by: Casper, Stephen, et al.
Published: (2025)
by: Casper, Stephen, et al.
Published: (2025)
Dynamical Properties of Random Boolean Hypernetworks
by: Stoltz, Kevin M., et al.
Published: (2024)
by: Stoltz, Kevin M., et al.
Published: (2024)
Diverse Preference Learning for Capabilities and Alignment
by: Slocum, Stewart, et al.
Published: (2025)
by: Slocum, Stewart, et al.
Published: (2025)
Similar Items
-
The Consequences of Humiliation
by: Barnhart, Joslyn
Published: (2021) -
Governing AI Agents
by: Kolt, Noam
Published: (2025) -
Visibility into AI Agents
by: Chan, Alan, et al.
Published: (2024) -
IDs for AI Systems
by: Chan, Alan, et al.
Published: (2024) -
Societal Adaptation to Advanced AI
by: Bernardi, Jamie, et al.
Published: (2024)