Designing Incident Reporting Systems for Harms from General-Purpose AI
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wei, Kevin, Heim, Lennart |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Training Compute Thresholds: Features and Functions in AI Regulation
par: Heim, Lennart, et autres
Publié: (2024)
par: Heim, Lennart, et autres
Publié: (2024)
Increased Compute Efficiency and the Diffusion of AI Capabilities
par: Pilz, Konstantin, et autres
Publié: (2023)
par: Pilz, Konstantin, et autres
Publié: (2023)
From Incidents to Insights: Patterns of Responsibility following AI Harms
par: Richards, Isabel, et autres
Publié: (2025)
par: Richards, Isabel, et autres
Publié: (2025)
A Taxonomy of Systemic Risks from General-Purpose AI
par: Uuk, Risto, et autres
Publié: (2024)
par: Uuk, Risto, et autres
Publié: (2024)
Verifying International Agreements on AI: Six Layers of Verification for Rules on Large-Scale AI Development and Deployment
par: Baker, Mauricio, et autres
Publié: (2025)
par: Baker, Mauricio, et autres
Publié: (2025)
Responsible Reporting for Frontier AI Development
par: Kolt, Noam, et autres
Publié: (2024)
par: Kolt, Noam, et autres
Publié: (2024)
Trends in AI Supercomputers
par: Pilz, Konstantin F., et autres
Publié: (2025)
par: Pilz, Konstantin F., et autres
Publié: (2025)
Effective Mitigations for Systemic Risks from General-Purpose AI
par: Uuk, Risto, et autres
Publié: (2024)
par: Uuk, Risto, et autres
Publié: (2024)
Lessons for Editors of AI Incidents from the AI Incident Database
par: Paeth, Kevin, et autres
Publié: (2024)
par: Paeth, Kevin, et autres
Publié: (2024)
Bridging the Artificial Intelligence Governance Gap: The United States' and China's Divergent Approaches to Governing General-Purpose Artificial Intelligence
par: Guest, Oliver, et autres
Publié: (2025)
par: Guest, Oliver, et autres
Publié: (2025)
Societal Adaptation to Advanced AI
par: Bernardi, Jamie, et autres
Publié: (2024)
par: Bernardi, Jamie, et autres
Publié: (2024)
Privacy Risks of General-Purpose AI Systems: A Foundation for Investigating Practitioner Perspectives
par: Meisenbacher, Stephen, et autres
Publié: (2024)
par: Meisenbacher, Stephen, et autres
Publié: (2024)
Disclosure and Evaluation as Fairness Interventions for General-Purpose AI
par: Raman, Vyoma, et autres
Publié: (2025)
par: Raman, Vyoma, et autres
Publié: (2025)
Evaluating General-Purpose AI with Psychometrics
par: Wang, Xiting, et autres
Publié: (2023)
par: Wang, Xiting, et autres
Publié: (2023)
Towards a Harms Taxonomy of AI Likeness Generation
par: Bariach, Ben, et autres
Publié: (2024)
par: Bariach, Ben, et autres
Publié: (2024)
Who is Responsible When AI Fails? Mapping Causes, Entities, and Consequences of AI Privacy and Ethical Incidents
par: Hadan, Hilda, et autres
Publié: (2025)
par: Hadan, Hilda, et autres
Publié: (2025)
Visibility into AI Agents
par: Chan, Alan, et autres
Publié: (2024)
par: Chan, Alan, et autres
Publié: (2024)
Purposeful remixing with generative AI: Constructing designer voice in multimodal composing
par: Tan, Xiao, et autres
Publié: (2024)
par: Tan, Xiao, et autres
Publié: (2024)
Distinguishing Task-Specific and General-Purpose AI in Regulation
par: Wang, Jennifer, et autres
Publié: (2025)
par: Wang, Jennifer, et autres
Publié: (2025)
The Dilemma of Uncertainty Estimation for General Purpose AI in the EU AI Act
par: Valdenegro-Toro, Matias, et autres
Publié: (2024)
par: Valdenegro-Toro, Matias, et autres
Publié: (2024)
Perpetuating Misogyny with Generative AI: How Model Personalization Normalizes Gendered Harm
par: Wagner, Laura, et autres
Publié: (2025)
par: Wagner, Laura, et autres
Publié: (2025)
Governing Through the Cloud: The Intermediary Role of Compute Providers in AI Regulation
par: Heim, Lennart, et autres
Publié: (2024)
par: Heim, Lennart, et autres
Publié: (2024)
Why AI Harms Can't Be Fixed One Identity at a Time: What 5300 Incident Reports Reveal About Intersectionality
par: Bogucka, Edyta, et autres
Publié: (2026)
par: Bogucka, Edyta, et autres
Publié: (2026)
Advancing Trustworthy AI for Sustainable Development: Recommendations for Standardising AI Incident Reporting
par: Agarwal, Avinash, et autres
Publié: (2025)
par: Agarwal, Avinash, et autres
Publié: (2025)
Third-party compliance reviews for frontier AI safety frameworks
par: Homewood, Aidan, et autres
Publié: (2025)
par: Homewood, Aidan, et autres
Publié: (2025)
PluriHarms: Benchmarking the Full Spectrum of Human Judgments on AI Harm
par: Li, Jing-Jing, et autres
Publié: (2026)
par: Li, Jing-Jing, et autres
Publié: (2026)
A Mechanism-Based Approach to Mitigating Harms from Persuasive Generative AI
par: El-Sayed, Seliem, et autres
Publié: (2024)
par: El-Sayed, Seliem, et autres
Publié: (2024)
LLM Harms: A Taxonomy and Discussion
par: Chen, Kevin, et autres
Publié: (2025)
par: Chen, Kevin, et autres
Publié: (2025)
Addressing the Unforeseen Harms of Technology CCC Whitepaper
par: Bliss, Nadya, et autres
Publié: (2024)
par: Bliss, Nadya, et autres
Publié: (2024)
Risk Sources and Risk Management Measures in Support of Standards for General-Purpose AI Systems
par: Gipiškis, Rokas, et autres
Publié: (2024)
par: Gipiškis, Rokas, et autres
Publié: (2024)
AI Generated Child Sexual Abuse Material -- What's the Harm?
par: Ciardha, Caoilte Ó, et autres
Publié: (2025)
par: Ciardha, Caoilte Ó, et autres
Publié: (2025)
Merging AI Incidents Research with Political Misinformation Research: Introducing the Political Deepfakes Incidents Database
par: Walker, Christina P., et autres
Publié: (2024)
par: Walker, Christina P., et autres
Publié: (2024)
Incidental Reverberations: Poetic Similarities in AI Art
par: Grba, Dejan
Publié: (2026)
par: Grba, Dejan
Publié: (2026)
Automating AI Failure Tracking: Semantic Association of Reports in AI Incident Database
par: Russo, Diego, et autres
Publié: (2025)
par: Russo, Diego, et autres
Publié: (2025)
The Case for ESM3 as a General-Purpose AI Model with Systemic Risk Under the EU AI Act
par: Qureshi, Taro, et autres
Publié: (2026)
par: Qureshi, Taro, et autres
Publié: (2026)
Incident Analysis for AI Agents
par: Ezell, Carson, et autres
Publié: (2025)
par: Ezell, Carson, et autres
Publié: (2025)
Self-Regulated Personal Contracts as a Harm Reduction Approach to Generative AI in Undergraduate Programming Education
par: Padiyath, Aadarsh, et autres
Publié: (2026)
par: Padiyath, Aadarsh, et autres
Publié: (2026)
Vernacularizing Taxonomies of Harm is Essential for Operationalizing Holistic AI Safety
par: Kennedy, Wm. Matthew, et autres
Publié: (2024)
par: Kennedy, Wm. Matthew, et autres
Publié: (2024)
Independent Clinical Evaluation of General-Purpose LLM Responses to Signals of Suicide Risk
par: Judd, Nick, et autres
Publié: (2025)
par: Judd, Nick, et autres
Publié: (2025)
Echoes of AI Harms: A Human-LLM Synergistic Framework for Bias-Driven Harm Anticipation
par: Tantalaki, Nicoleta, et autres
Publié: (2025)
par: Tantalaki, Nicoleta, et autres
Publié: (2025)
Documents similaires
-
Training Compute Thresholds: Features and Functions in AI Regulation
par: Heim, Lennart, et autres
Publié: (2024) -
Increased Compute Efficiency and the Diffusion of AI Capabilities
par: Pilz, Konstantin, et autres
Publié: (2023) -
From Incidents to Insights: Patterns of Responsibility following AI Harms
par: Richards, Isabel, et autres
Publié: (2025) -
A Taxonomy of Systemic Risks from General-Purpose AI
par: Uuk, Risto, et autres
Publié: (2024) -
Verifying International Agreements on AI: Six Layers of Verification for Rules on Large-Scale AI Development and Deployment
par: Baker, Mauricio, et autres
Publié: (2025)