Risk Reporting for Developers' Internal AI Model Use
Fuente:
arXiv
Guardado en:
| Autores principales: | Delaney, Oscar, Maheshwari, Sambhav, O'Brien, Joe, Bearman, Theo, Guest, Oliver |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Mapping Technical Safety Research at AI Companies: A literature review and incentives analysis
por: Delaney, Oscar, et al.
Publicado: (2024)
por: Delaney, Oscar, et al.
Publicado: (2024)
Expert Survey: AI Reliability & Security Research Priorities
por: O'Brien, Joe, et al.
Publicado: (2025)
por: O'Brien, Joe, et al.
Publicado: (2025)
Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI
por: O'Brien, Joe, et al.
Publicado: (2024)
por: O'Brien, Joe, et al.
Publicado: (2024)
Responsible Reporting for Frontier AI Development
por: Kolt, Noam, et al.
Publicado: (2024)
por: Kolt, Noam, et al.
Publicado: (2024)
The AI Model Risk Catalog: What Developers and Researchers Miss About Real-World AI Harms
por: Rao, Pooja S. B., et al.
Publicado: (2025)
por: Rao, Pooja S. B., et al.
Publicado: (2025)
When AI Takes the Couch: Psychometric Jailbreaks Reveal Internal Conflict in Frontier Models
por: Khadangi, Afshin, et al.
Publicado: (2025)
por: Khadangi, Afshin, et al.
Publicado: (2025)
Developing Story: Case Studies of Generative AI's Use in Journalism
por: Brigham, Natalie Grace, et al.
Publicado: (2024)
por: Brigham, Natalie Grace, et al.
Publicado: (2024)
AI Cards: Towards an Applied Framework for Machine-Readable AI and Risk Documentation Inspired by the EU AI Act
por: Golpayegani, Delaram, et al.
Publicado: (2024)
por: Golpayegani, Delaram, et al.
Publicado: (2024)
Industrial AI Robustness Card for Time Series Models
por: Windmann, Alexander, et al.
Publicado: (2025)
por: Windmann, Alexander, et al.
Publicado: (2025)
Mapping AI Risk Mitigations: Evidence Scan and Preliminary AI Risk Mitigation Taxonomy
por: Saeri, Alexander K., et al.
Publicado: (2025)
por: Saeri, Alexander K., et al.
Publicado: (2025)
Introducing ELLIPS: An Ethics-Centered Approach to Research on LLM-Based Inference of Psychiatric Conditions
por: Rocca, Roberta, et al.
Publicado: (2024)
por: Rocca, Roberta, et al.
Publicado: (2024)
AI Identity, Empowerment, and Mindfulness in Mitigating Unethical AI Use
por: Shaayesteh, Mayssam Tarighi, et al.
Publicado: (2025)
por: Shaayesteh, Mayssam Tarighi, et al.
Publicado: (2025)
STREAM (ChemBio): A Standard for Transparently Reporting Evaluations in AI Model Reports
por: McCaslin, Tegan, et al.
Publicado: (2025)
por: McCaslin, Tegan, et al.
Publicado: (2025)
AI Consciousness and Existential Risk
por: VanRullen, Rufin
Publicado: (2025)
por: VanRullen, Rufin
Publicado: (2025)
New Tools are Needed for Tracking Adherence to AI Model Behavioral Use Clauses
por: McDuff, Daniel, et al.
Publicado: (2025)
por: McDuff, Daniel, et al.
Publicado: (2025)
Is your AI Model Accurate Enough? The Difficult Choices Behind Rigorous AI Development and the EU AI Act
por: Marin, Lucas G. Uberti-Bona, et al.
Publicado: (2026)
por: Marin, Lucas G. Uberti-Bona, et al.
Publicado: (2026)
Synthetic Data for Robust AI Model Development in Regulated Enterprises
por: Godbole, Aditi
Publicado: (2025)
por: Godbole, Aditi
Publicado: (2025)
Advancing Trustworthy AI for Sustainable Development: Recommendations for Standardising AI Incident Reporting
por: Agarwal, Avinash, et al.
Publicado: (2025)
por: Agarwal, Avinash, et al.
Publicado: (2025)
Prompt Engineering for Responsible Generative AI Use in African Education: A Report from a Three-Day Training Series
por: Quarshie, Benjamin, et al.
Publicado: (2026)
por: Quarshie, Benjamin, et al.
Publicado: (2026)
Measuring AI R&D Automation
por: Chan, Alan, et al.
Publicado: (2026)
por: Chan, Alan, et al.
Publicado: (2026)
Surveys Considered Harmful? Reflecting on the Use of Surveys in AI Research, Development, and Governance
por: Tahaei, Mohammmad, et al.
Publicado: (2024)
por: Tahaei, Mohammmad, et al.
Publicado: (2024)
What are human values, and how do we align AI to them?
por: Klingefjord, Oliver, et al.
Publicado: (2024)
por: Klingefjord, Oliver, et al.
Publicado: (2024)
Evaluation of AI Ethics Tools in Language Models: A Developers' Perspective Case Study
por: Silva, Jhessica, et al.
Publicado: (2025)
por: Silva, Jhessica, et al.
Publicado: (2025)
The Biggest Risk of Embodied AI is Governance Lag
por: Liu, Shaoshan
Publicado: (2026)
por: Liu, Shaoshan
Publicado: (2026)
Brainrot: Deskilling and Addiction are Overlooked AI Risks
por: Chalkidis, Ilias, et al.
Publicado: (2026)
por: Chalkidis, Ilias, et al.
Publicado: (2026)
Algorithmic Administration and the EU AI Act: Legal Principles for Public Sector Use of AI
por: Pavlidis, Georgios, et al.
Publicado: (2026)
por: Pavlidis, Georgios, et al.
Publicado: (2026)
Are Companies Taking AI Risks Seriously? A Systematic Analysis of Companies' AI Risk Disclosures in SEC 10-K forms
por: Marin, Lucas G. Uberti-Bona, et al.
Publicado: (2025)
por: Marin, Lucas G. Uberti-Bona, et al.
Publicado: (2025)
International Scientific Report on the Safety of Advanced AI (Interim Report)
por: Bengio, Yoshua, et al.
Publicado: (2024)
por: Bengio, Yoshua, et al.
Publicado: (2024)
Responsible AI Question Bank: A Comprehensive Tool for AI Risk Assessment
por: Lee, Sung Une, et al.
Publicado: (2024)
por: Lee, Sung Une, et al.
Publicado: (2024)
Anchoring AI Capabilities in Market Valuations: The Capability Realization Rate Model and Valuation Misalignment Risk
por: Fang, Xinmin, et al.
Publicado: (2025)
por: Fang, Xinmin, et al.
Publicado: (2025)
Responsible AI Governance: A Response to UN Interim Report on Governing AI for Humanity
por: Kiden, Sarah, et al.
Publicado: (2024)
por: Kiden, Sarah, et al.
Publicado: (2024)
Adapting Probabilistic Risk Assessment for AI
por: Wisakanto, Anna Katariina, et al.
Publicado: (2025)
por: Wisakanto, Anna Katariina, et al.
Publicado: (2025)
Three Lenses on the AI Revolution: Risk, Transformation, Continuity
por: Makrehchi, Masoud
Publicado: (2025)
por: Makrehchi, Masoud
Publicado: (2025)
Generative Ghosts: Anticipating Benefits and Risks of AI Afterlives
por: Morris, Meredith Ringel, et al.
Publicado: (2024)
por: Morris, Meredith Ringel, et al.
Publicado: (2024)
Threshold Crossings as Tail Events for Catastrophic AI Risk
por: Perrier, Elija
Publicado: (2025)
por: Perrier, Elija
Publicado: (2025)
Now You See Me: Designing Responsible AI Dashboards for Early-Stage Health Innovation
por: Surodina, Svitlana, et al.
Publicado: (2026)
por: Surodina, Svitlana, et al.
Publicado: (2026)
From Turing to Tomorrow: The UK's Approach to AI Regulation
por: Ritchie, Oliver, et al.
Publicado: (2025)
por: Ritchie, Oliver, et al.
Publicado: (2025)
Particip-AI: A Democratic Surveying Framework for Anticipating Future AI Use Cases, Harms and Benefits
por: Mun, Jimin, et al.
Publicado: (2024)
por: Mun, Jimin, et al.
Publicado: (2024)
Report on NSF Workshop on Science of Safe AI
por: Alur, Rajeev, et al.
Publicado: (2025)
por: Alur, Rajeev, et al.
Publicado: (2025)
Developing Strategies to Increase Capacity in AI Education
por: Cowit, Noah Q., et al.
Publicado: (2025)
por: Cowit, Noah Q., et al.
Publicado: (2025)
Ejemplares similares
-
Mapping Technical Safety Research at AI Companies: A literature review and incentives analysis
por: Delaney, Oscar, et al.
Publicado: (2024) -
Expert Survey: AI Reliability & Security Research Priorities
por: O'Brien, Joe, et al.
Publicado: (2025) -
Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI
por: O'Brien, Joe, et al.
Publicado: (2024) -
Responsible Reporting for Frontier AI Development
por: Kolt, Noam, et al.
Publicado: (2024) -
The AI Model Risk Catalog: What Developers and Researchers Miss About Real-World AI Harms
por: Rao, Pooja S. B., et al.
Publicado: (2025)