AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Barrett, Anthony M., Newman, Jessica, Nonnecke, Brandie, Madkour, Nada, Hendrycks, Dan, Murphy, Evan R., Jackson, Krystal, Raman, Deepika |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Toward Risk Thresholds for AI-Enabled Cyber Threats: Enhancing Decision-Making Under Uncertainty with Bayesian Networks
von: Jackson, Krystal, et al.
Veröffentlicht: (2026)
von: Jackson, Krystal, et al.
Veröffentlicht: (2026)
Benchmark Early and Red Team Often: A Framework for Assessing and Managing Dual-Use Hazards of AI Foundation Models
von: Barrett, Anthony M., et al.
Veröffentlicht: (2024)
von: Barrett, Anthony M., et al.
Veröffentlicht: (2024)
Intolerable Risk Threshold Recommendations for Artificial Intelligence
von: Raman, Deepika, et al.
Veröffentlicht: (2025)
von: Raman, Deepika, et al.
Veröffentlicht: (2025)
Responsible Generative AI Use by Product Managers: Recoupling Ethical Principles and Practices
von: Smith, Genevieve, et al.
Veröffentlicht: (2025)
von: Smith, Genevieve, et al.
Veröffentlicht: (2025)
Watermarking Without Standards Is Not AI Governance
von: Nemecek, Alexander, et al.
Veröffentlicht: (2025)
von: Nemecek, Alexander, et al.
Veröffentlicht: (2025)
SL5 Standard for AI Security
von: Thiergart, Lisa, et al.
Veröffentlicht: (2026)
von: Thiergart, Lisa, et al.
Veröffentlicht: (2026)
Fundamental Risks in the Current Deployment of General-Purpose AI Models: What Have We (Not) Learnt From Cybersecurity?
von: Fritz, Mario
Veröffentlicht: (2024)
von: Fritz, Mario
Veröffentlicht: (2024)
Introduction to AI Safety, Ethics, and Society
von: Hendrycks, Dan
Veröffentlicht: (2024)
von: Hendrycks, Dan
Veröffentlicht: (2024)
Introduction to AI Safety, Ethics, and Society
von: Hendrycks, Dan
Veröffentlicht: (2024)
von: Hendrycks, Dan
Veröffentlicht: (2024)
HonestCyberEval: An AI Cyber Risk Benchmark for Automated Software Exploitation
von: Ristea, Dan, et al.
Veröffentlicht: (2024)
von: Ristea, Dan, et al.
Veröffentlicht: (2024)
AI Risk Management Should Incorporate Both Safety and Security
von: Qi, Xiangyu, et al.
Veröffentlicht: (2024)
von: Qi, Xiangyu, et al.
Veröffentlicht: (2024)
AI Identity: Standards, Gaps, and Research Directions for AI Agents
von: Otsuka, Takumi, et al.
Veröffentlicht: (2026)
von: Otsuka, Takumi, et al.
Veröffentlicht: (2026)
GPAI Evaluations Standards Taskforce: Towards Effective AI Governance
von: Paskov, Patricia, et al.
Veröffentlicht: (2024)
von: Paskov, Patricia, et al.
Veröffentlicht: (2024)
Aggressive Compression Enables LLM Weight Theft
von: Brown, Davis, et al.
Veröffentlicht: (2026)
von: Brown, Davis, et al.
Veröffentlicht: (2026)
Extending the Formalism and Theoretical Foundations of Cryptography to AI
von: Villa, Federico, et al.
Veröffentlicht: (2026)
von: Villa, Federico, et al.
Veröffentlicht: (2026)
Position: Mind the Gap-AI Security and the Limits of Current Reporting Standards
von: Bieringer, Lukas, et al.
Veröffentlicht: (2024)
von: Bieringer, Lukas, et al.
Veröffentlicht: (2024)
Identifying the Supply Chain of AI for Trustworthiness and Risk Management in Critical Applications
von: Sheh, Raymond K., et al.
Veröffentlicht: (2025)
von: Sheh, Raymond K., et al.
Veröffentlicht: (2025)
Best Practices for Biorisk Evaluations on Open-Weight Bio-Foundation Models
von: Wei, Boyi, et al.
Veröffentlicht: (2025)
von: Wei, Boyi, et al.
Veröffentlicht: (2025)
Preventing Adversarial AI Attacks Against Autonomous Situational Awareness: A Maritime Case Study
von: Walter, Mathew J., et al.
Veröffentlicht: (2025)
von: Walter, Mathew J., et al.
Veröffentlicht: (2025)
The Agent Economy: A Blockchain-Based Foundation for Autonomous AI Agents
von: Xu, Minghui
Veröffentlicht: (2026)
von: Xu, Minghui
Veröffentlicht: (2026)
Timing and Memory Telemetry on GPUs for AI Governance
von: Monfared, Saleh K., et al.
Veröffentlicht: (2026)
von: Monfared, Saleh K., et al.
Veröffentlicht: (2026)
Security-First AI: Foundations for Robust and Trustworthy Systems
von: Tallam, Krti
Veröffentlicht: (2025)
von: Tallam, Krti
Veröffentlicht: (2025)
AgenticVM: Agentic AI for Adaptive Software Vulnerability Management
von: Arifin, Asrul, et al.
Veröffentlicht: (2026)
von: Arifin, Asrul, et al.
Veröffentlicht: (2026)
Establishing Minimum Elements for Effective Vulnerability Management in AI Software
von: Fazelnia, Mohamad, et al.
Veröffentlicht: (2024)
von: Fazelnia, Mohamad, et al.
Veröffentlicht: (2024)
Referential Security as a New Paradigm for AI Evaluations
von: Ristea, Dan, et al.
Veröffentlicht: (2026)
von: Ristea, Dan, et al.
Veröffentlicht: (2026)
Security Risks Concerns of Generative AI in the IoT
von: Xu, Honghui, et al.
Veröffentlicht: (2024)
von: Xu, Honghui, et al.
Veröffentlicht: (2024)
Edge AI-based Radio Frequency Fingerprinting for IoT Networks
von: Hussain, Ahmed Mohamed, et al.
Veröffentlicht: (2024)
von: Hussain, Ahmed Mohamed, et al.
Veröffentlicht: (2024)
When AI Meets the Web: Prompt Injection Risks in Third-Party AI Chatbot Plugins
von: Kaya, Yigitcan, et al.
Veröffentlicht: (2025)
von: Kaya, Yigitcan, et al.
Veröffentlicht: (2025)
Securing Agentic AI: Threat Modeling and Risk Analysis for Network Monitoring Agentic AI System
von: Zambare, Pallavi, et al.
Veröffentlicht: (2025)
von: Zambare, Pallavi, et al.
Veröffentlicht: (2025)
Trinity: A General Purpose FHE Accelerator
von: Deng, Xianglong, et al.
Veröffentlicht: (2024)
von: Deng, Xianglong, et al.
Veröffentlicht: (2024)
Securing the Future of GenAI: Policy and Technology
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2024)
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2024)
Prevalence of Security and Privacy Risk-Inducing Usage of AI-based Conversational Agents
von: Grosse, Kathrin, et al.
Veröffentlicht: (2025)
von: Grosse, Kathrin, et al.
Veröffentlicht: (2025)
Foundations for Agentic AI Investigations from the Forensic Analysis of OpenClaw
von: Gruber, Jan, et al.
Veröffentlicht: (2026)
von: Gruber, Jan, et al.
Veröffentlicht: (2026)
Emerging Cyber Attack Risks of Medical AI Agents
von: Qiu, Jianing, et al.
Veröffentlicht: (2025)
von: Qiu, Jianing, et al.
Veröffentlicht: (2025)
CIA+TA Risk Assessment for AI Reasoning Vulnerabilities
von: Aydin, Yuksel
Veröffentlicht: (2025)
von: Aydin, Yuksel
Veröffentlicht: (2025)
AgenTRIM: Tool Risk Mitigation for Agentic AI
von: Betser, Roy, et al.
Veröffentlicht: (2026)
von: Betser, Roy, et al.
Veröffentlicht: (2026)
Monero Traceability Heuristics: Wallet Application Bugs and the Mordinal-P2Pool Perspective
von: Hammad, Nada, et al.
Veröffentlicht: (2024)
von: Hammad, Nada, et al.
Veröffentlicht: (2024)
An Intelligent Quantum Cyber-Security Framework for Healthcare Data Management
von: Gupta, Kishu, et al.
Veröffentlicht: (2024)
von: Gupta, Kishu, et al.
Veröffentlicht: (2024)
FIDEM: A Standard-Compliant Framework for Secure Binding of MUD Profiles to IoT Devices
von: Lotto, Alessandro, et al.
Veröffentlicht: (2026)
von: Lotto, Alessandro, et al.
Veröffentlicht: (2026)
TAIBOM: Bringing Trustworthiness to AI-Enabled Systems
von: Safronov, Vadim, et al.
Veröffentlicht: (2025)
von: Safronov, Vadim, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Toward Risk Thresholds for AI-Enabled Cyber Threats: Enhancing Decision-Making Under Uncertainty with Bayesian Networks
von: Jackson, Krystal, et al.
Veröffentlicht: (2026) -
Benchmark Early and Red Team Often: A Framework for Assessing and Managing Dual-Use Hazards of AI Foundation Models
von: Barrett, Anthony M., et al.
Veröffentlicht: (2024) -
Intolerable Risk Threshold Recommendations for Artificial Intelligence
von: Raman, Deepika, et al.
Veröffentlicht: (2025) -
Responsible Generative AI Use by Product Managers: Recoupling Ethical Principles and Practices
von: Smith, Genevieve, et al.
Veröffentlicht: (2025) -
Watermarking Without Standards Is Not AI Governance
von: Nemecek, Alexander, et al.
Veröffentlicht: (2025)