Vernacularizing Taxonomies of Harm is Essential for Operationalizing Holistic AI Safety
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kennedy, Wm. Matthew, Campos, Daniel Vargas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards a Harms Taxonomy of AI Likeness Generation
von: Bariach, Ben, et al.
Veröffentlicht: (2024)
von: Bariach, Ben, et al.
Veröffentlicht: (2024)
Towards an Evaluation Methodology for AI in Second Language Education: Lessons Learned from Developing L2-Bench
von: Edgell, James, et al.
Veröffentlicht: (2026)
von: Edgell, James, et al.
Veröffentlicht: (2026)
LLM Harms: A Taxonomy and Discussion
von: Chen, Kevin, et al.
Veröffentlicht: (2025)
von: Chen, Kevin, et al.
Veröffentlicht: (2025)
Not My Voice! A Taxonomy of Ethical and Safety Harms of Speech Generators
von: Hutiri, Wiebke, et al.
Veröffentlicht: (2024)
von: Hutiri, Wiebke, et al.
Veröffentlicht: (2024)
The Dark Side of AI Companionship: A Taxonomy of Harmful Algorithmic Behaviors in Human-AI Relationships
von: Zhang, Renwen, et al.
Veröffentlicht: (2024)
von: Zhang, Renwen, et al.
Veröffentlicht: (2024)
AI, Climate, and Transparency: Operationalizing and Improving the AI Act
von: Alder, Nicolas, et al.
Veröffentlicht: (2024)
von: Alder, Nicolas, et al.
Veröffentlicht: (2024)
Reinforcing Stereotypes of Anger: Emotion AI on African American Vernacular English
von: Dorn, Rebecca, et al.
Veröffentlicht: (2025)
von: Dorn, Rebecca, et al.
Veröffentlicht: (2025)
Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents
von: Li, Miles Q., et al.
Veröffentlicht: (2026)
von: Li, Miles Q., et al.
Veröffentlicht: (2026)
In Quest of an Extensible Multi-Level Harm Taxonomy for Adversarial AI: Heart of Security, Ethical Risk Scoring and Resilience Analytics
von: Khan, Javed I., et al.
Veröffentlicht: (2026)
von: Khan, Javed I., et al.
Veröffentlicht: (2026)
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations
von: Davani, Aida, et al.
Veröffentlicht: (2025)
von: Davani, Aida, et al.
Veröffentlicht: (2025)
Harmful YouTube Video Detection: A Taxonomy of Online Harm and MLLMs as Alternative Annotators
von: Jo, Claire Wonjeong, et al.
Veröffentlicht: (2024)
von: Jo, Claire Wonjeong, et al.
Veröffentlicht: (2024)
The Bidirectional Relationship Between XAI and Regulation: Operationalizing XAI for the AI Act
von: Hummel, Anton, et al.
Veröffentlicht: (2025)
von: Hummel, Anton, et al.
Veröffentlicht: (2025)
AI TIPS 2.0: A Comprehensive Framework for Operationalizing AI Governance
von: Gupta, Pamela
Veröffentlicht: (2025)
von: Gupta, Pamela
Veröffentlicht: (2025)
Making Power Explicable in AI: Analyzing, Understanding, and Redirecting Power to Operationalize Ethics in AI Technical Practice
von: Jin, Weina, et al.
Veröffentlicht: (2025)
von: Jin, Weina, et al.
Veröffentlicht: (2025)
Cascade! Human in the loop shortcomings can increase the risk of failures in recommender systems
von: Kennedy, Wm. Matthew, et al.
Veröffentlicht: (2025)
von: Kennedy, Wm. Matthew, et al.
Veröffentlicht: (2025)
Disentangling AI Alignment: A Structured Taxonomy Beyond Safety and Ethics
von: Baum, Kevin
Veröffentlicht: (2025)
von: Baum, Kevin
Veröffentlicht: (2025)
Evaluating Human-AI Safety: A Framework for Measuring Harmful Capability Uplift
von: Vaccaro, Michelle, et al.
Veröffentlicht: (2026)
von: Vaccaro, Michelle, et al.
Veröffentlicht: (2026)
Operationalizing Justice: Towards the Development of a Principle Based Design Framework for Human Services AI
von: Rodriguez, Maria Y., et al.
Veröffentlicht: (2025)
von: Rodriguez, Maria Y., et al.
Veröffentlicht: (2025)
A Taxonomy of Systemic Risks from General-Purpose AI
von: Uuk, Risto, et al.
Veröffentlicht: (2024)
von: Uuk, Risto, et al.
Veröffentlicht: (2024)
PluriHarms: Benchmarking the Full Spectrum of Human Judgments on AI Harm
von: Li, Jing-Jing, et al.
Veröffentlicht: (2026)
von: Li, Jing-Jing, et al.
Veröffentlicht: (2026)
IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures
von: Gringras, David
Veröffentlicht: (2026)
von: Gringras, David
Veröffentlicht: (2026)
From Representational Harms to Quality-of-Service Harms: A Case Study on Llama 2 Safety Safeguards
von: Chehbouni, Khaoula, et al.
Veröffentlicht: (2024)
von: Chehbouni, Khaoula, et al.
Veröffentlicht: (2024)
Evaluating AI Providers' Frontier Safety Frameworks
von: Stelling, Lily, et al.
Veröffentlicht: (2025)
von: Stelling, Lily, et al.
Veröffentlicht: (2025)
Operationalizing Ethics for AI Agents: How Developers Encode Values into Repository Context Files
von: Treude, Christoph, et al.
Veröffentlicht: (2026)
von: Treude, Christoph, et al.
Veröffentlicht: (2026)
Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
Addressing the Unforeseen Harms of Technology CCC Whitepaper
von: Bliss, Nadya, et al.
Veröffentlicht: (2024)
von: Bliss, Nadya, et al.
Veröffentlicht: (2024)
Gendered Inequalities in Online Harms: Fear, Safety Work, and Online Participation
von: Enock, Florence E., et al.
Veröffentlicht: (2024)
von: Enock, Florence E., et al.
Veröffentlicht: (2024)
AI Safety Training Can be Clinically Harmful
von: BN, Suhas, et al.
Veröffentlicht: (2026)
von: BN, Suhas, et al.
Veröffentlicht: (2026)
An AI System Evaluation Framework for Advancing AI Safety: Terminology, Taxonomy, Lifecycle Mapping
von: Xia, Boming, et al.
Veröffentlicht: (2024)
von: Xia, Boming, et al.
Veröffentlicht: (2024)
From Incidents to Insights: Patterns of Responsibility following AI Harms
von: Richards, Isabel, et al.
Veröffentlicht: (2025)
von: Richards, Isabel, et al.
Veröffentlicht: (2025)
A Collaborative, Human-Centred Taxonomy of AI, Algorithmic, and Automation Harms
von: Abercrombie, Gavin, et al.
Veröffentlicht: (2024)
von: Abercrombie, Gavin, et al.
Veröffentlicht: (2024)
Specification, Application, and Operationalization of a Metamodel of Fairness
von: Mendez, Julian Alfredo, et al.
Veröffentlicht: (2025)
von: Mendez, Julian Alfredo, et al.
Veröffentlicht: (2025)
Operationalizing the Blueprint for an AI Bill of Rights: Recommendations for Practitioners, Researchers, and Policy Makers
von: Oesterling, Alex, et al.
Veröffentlicht: (2024)
von: Oesterling, Alex, et al.
Veröffentlicht: (2024)
GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory
von: Cobben, Pepijn, et al.
Veröffentlicht: (2026)
von: Cobben, Pepijn, et al.
Veröffentlicht: (2026)
Echoes of AI Harms: A Human-LLM Synergistic Framework for Bias-Driven Harm Anticipation
von: Tantalaki, Nicoleta, et al.
Veröffentlicht: (2025)
von: Tantalaki, Nicoleta, et al.
Veröffentlicht: (2025)
A Mechanism-Based Approach to Mitigating Harms from Persuasive Generative AI
von: El-Sayed, Seliem, et al.
Veröffentlicht: (2024)
von: El-Sayed, Seliem, et al.
Veröffentlicht: (2024)
Designing Incident Reporting Systems for Harms from General-Purpose AI
von: Wei, Kevin, et al.
Veröffentlicht: (2025)
von: Wei, Kevin, et al.
Veröffentlicht: (2025)
Operationalizing Perceptions of Agent Gender: Foundations and Guidelines
von: Seaborn, Katie, et al.
Veröffentlicht: (2026)
von: Seaborn, Katie, et al.
Veröffentlicht: (2026)
AI Risk Atlas: Taxonomy and Tooling for Navigating AI Risks and Resources
von: Bagehorn, Frank, et al.
Veröffentlicht: (2025)
von: Bagehorn, Frank, et al.
Veröffentlicht: (2025)
Expected Harm: Rethinking Safety Evaluation of (Mis)Aligned LLMs
von: Chen, Yen-Shan, et al.
Veröffentlicht: (2026)
von: Chen, Yen-Shan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards a Harms Taxonomy of AI Likeness Generation
von: Bariach, Ben, et al.
Veröffentlicht: (2024) -
Towards an Evaluation Methodology for AI in Second Language Education: Lessons Learned from Developing L2-Bench
von: Edgell, James, et al.
Veröffentlicht: (2026) -
LLM Harms: A Taxonomy and Discussion
von: Chen, Kevin, et al.
Veröffentlicht: (2025) -
Not My Voice! A Taxonomy of Ethical and Safety Harms of Speech Generators
von: Hutiri, Wiebke, et al.
Veröffentlicht: (2024) -
The Dark Side of AI Companionship: A Taxonomy of Harmful Algorithmic Behaviors in Human-AI Relationships
von: Zhang, Renwen, et al.
Veröffentlicht: (2024)