AI Risk Categorization Decoded (AIR 2024): From Government Regulations to Corporate Policies
Fuente:
arXiv
Saved in:
| Main Authors: | Zeng, Yi, Klyman, Kevin, Zhou, Andy, Yang, Yu, Pan, Minzhou, Jia, Ruoxi, Song, Dawn, Liang, Percy, Li, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
by: Zeng, Yi, et al.
Published: (2024)
by: Zeng, Yi, et al.
Published: (2024)
From Symptoms to Systems: An Expert-Guided Approach to Understanding Risks of Generative AI for Eating Disorders
by: Winecoff, Amy, et al.
Published: (2025)
by: Winecoff, Amy, et al.
Published: (2025)
SpecEval: Evaluating Model Adherence to Behavior Specifications
by: Ahmed, Ahmed, et al.
Published: (2025)
by: Ahmed, Ahmed, et al.
Published: (2025)
Acceptable Use Policies for Foundation Models
by: Klyman, Kevin
Published: (2024)
by: Klyman, Kevin
Published: (2024)
JIGMARK: A Black-Box Approach for Enhancing Image Watermarks against Diffusion Model Edits
by: Pan, Minzhou, et al.
Published: (2024)
by: Pan, Minzhou, et al.
Published: (2024)
The 2024 Foundation Model Transparency Index
by: Bommasani, Rishi, et al.
Published: (2024)
by: Bommasani, Rishi, et al.
Published: (2024)
SafeWatch: An Efficient Safety-Policy Following Video Guardrail Model with Transparent Explanations
by: Chen, Zhaorun, et al.
Published: (2024)
by: Chen, Zhaorun, et al.
Published: (2024)
Unsafer in Many Turns: Benchmarking and Defending Multi-Turn Safety Risks in Tool-Using Agents
by: Li, Xu, et al.
Published: (2026)
by: Li, Xu, et al.
Published: (2026)
Language model developers should report train-test overlap
by: Zhang, Andy K, et al.
Published: (2024)
by: Zhang, Andy K, et al.
Published: (2024)
BEEAR: Embedding-based Adversarial Removal of Safety Backdoors in Instruction-tuned Language Models
by: Zeng, Yi, et al.
Published: (2024)
by: Zeng, Yi, et al.
Published: (2024)
Comparing Apples to Oranges: A Taxonomy for Navigating the Global Landscape of AI Regulation
by: Alanoca, Sacha, et al.
Published: (2025)
by: Alanoca, Sacha, et al.
Published: (2025)
RigorLLM: Resilient Guardrails for Large Language Models against Undesired Content
by: Yuan, Zhuowen, et al.
Published: (2024)
by: Yuan, Zhuowen, et al.
Published: (2024)
SafeVision: Efficient Image Guardrail with Robust Policy Adherence and Explainability
by: Xu, Peiyang, et al.
Published: (2025)
by: Xu, Peiyang, et al.
Published: (2025)
Is Decentralized AI Governable? From Regulative Policy to Constitutive Protocol
by: Hu, Botao Amber, et al.
Published: (2026)
by: Hu, Botao Amber, et al.
Published: (2026)
Benchmarking Zero-Shot Robustness of Multimodal Foundation Models: A Pilot Study
by: Wang, Chenguang, et al.
Published: (2024)
by: Wang, Chenguang, et al.
Published: (2024)
Data Shapley in One Training Run
by: Wang, Jiachen T., et al.
Published: (2024)
by: Wang, Jiachen T., et al.
Published: (2024)
Do AI Companies Make Good on Voluntary Commitments to the White House?
by: Wang, Jennifer, et al.
Published: (2025)
by: Wang, Jennifer, et al.
Published: (2025)
The Limits of AI Data Transparency Policy: Three Disclosure Fallacies
by: Shen, Judy Hanwen, et al.
Published: (2026)
by: Shen, Judy Hanwen, et al.
Published: (2026)
Evaluating and Mitigating IP Infringement in Visual Generative AI
by: Wang, Zhenting, et al.
Published: (2024)
by: Wang, Zhenting, et al.
Published: (2024)
Chapter 6 The Role of Corporate Governance in Macro-Prudential Regulation of Systemic Risk
by: Dill, Alexander
Published: (2020)
by: Dill, Alexander
Published: (2020)
A Sustainable AI Economy Needs Data Deals That Work for Generators
by: Jia, Ruoxi, et al.
Published: (2026)
by: Jia, Ruoxi, et al.
Published: (2026)
User Privacy and Large Language Models: An Analysis of Frontier Developers' Privacy Policies
by: King, Jennifer, et al.
Published: (2025)
by: King, Jennifer, et al.
Published: (2025)
Foundation Model Transparency Reports
by: Bommasani, Rishi, et al.
Published: (2024)
by: Bommasani, Rishi, et al.
Published: (2024)
The 2025 Foundation Model Transparency Index
by: Wan, Alexander, et al.
Published: (2025)
by: Wan, Alexander, et al.
Published: (2025)
Capturing the Temporal Dependence of Training Data Influence
by: Wang, Jiachen T., et al.
Published: (2024)
by: Wang, Jiachen T., et al.
Published: (2024)
New Tools are Needed for Tracking Adherence to AI Model Behavioral Use Clauses
by: McDuff, Daniel, et al.
Published: (2025)
by: McDuff, Daniel, et al.
Published: (2025)
Recourse, Repair, Reparation, & Prevention: A Stakeholder Analysis of AI Supply Chains
by: Hopkins, Aspen K., et al.
Published: (2025)
by: Hopkins, Aspen K., et al.
Published: (2025)
ICTs and Rural E‐Governance: From Digital Design to Public Policy
by: Ning Su, et al.
Published: (2025)
by: Ning Su, et al.
Published: (2025)
Government‐Level Political Risk Perception and Corporate Violations: Evidence From China
by: Lin Zhang, et al.
Published: (2025)
by: Lin Zhang, et al.
Published: (2025)
The Governance of Insurance Undertakings Corporate Law and Insurance Regulation
by: Pierpaolo Marano
by: Pierpaolo Marano
Poly-Guard: Massive Multi-Domain Safety Policy-Grounded Guardrail Dataset
by: Kang, Mintong, et al.
Published: (2025)
by: Kang, Mintong, et al.
Published: (2025)
A Safe Harbor for AI Evaluation and Red Teaming
by: Longpre, Shayne, et al.
Published: (2024)
by: Longpre, Shayne, et al.
Published: (2024)
Corporate Governance and Dividend Policy: Evidence from Colombia
by: Samuel Mongrut Montalvan
Published: (2023)
by: Samuel Mongrut Montalvan
Published: (2023)
Corporate Governance and Dividend Policy in Peru: Is there any link?
by: Samuel Mongrut Montalvan
Published: (2017)
by: Samuel Mongrut Montalvan
Published: (2017)
From Industrial Electrification to Artificial Intelligence: Institutional Lessons from Construction Governance for AI Risk Regulation
by: Kc, Subodh
Published: (2026)
by: Kc, Subodh
Published: (2026)
DecodingTrust-Agent Platform (DTap): A Controllable and Interactive Red-Teaming Platform for AI Agents
by: Chen, Zhaorun, et al.
Published: (2026)
by: Chen, Zhaorun, et al.
Published: (2026)
How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs
by: Zeng, Yi, et al.
Published: (2024)
by: Zeng, Yi, et al.
Published: (2024)
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
by: Guo, Chengquan, et al.
Published: (2024)
by: Guo, Chengquan, et al.
Published: (2024)
AI Governance InternationaL Evaluation Index (AGILE Index) 2024
by: Zeng, Yi, et al.
Published: (2025)
by: Zeng, Yi, et al.
Published: (2025)
Mandatory Corporate Responsibility Regulation, Geopolitical Risk, and Firm Performance: Evidence From India
by: Niharika Kasaudhan, et al.
Published: (2025)
by: Niharika Kasaudhan, et al.
Published: (2025)
Similar Items
-
AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
by: Zeng, Yi, et al.
Published: (2024) -
From Symptoms to Systems: An Expert-Guided Approach to Understanding Risks of Generative AI for Eating Disorders
by: Winecoff, Amy, et al.
Published: (2025) -
SpecEval: Evaluating Model Adherence to Behavior Specifications
by: Ahmed, Ahmed, et al.
Published: (2025) -
Acceptable Use Policies for Foundation Models
by: Klyman, Kevin
Published: (2024) -
JIGMARK: A Black-Box Approach for Enhancing Image Watermarks against Diffusion Model Edits
by: Pan, Minzhou, et al.
Published: (2024)