Measuring What Matters: Connecting AI Ethics Evaluations to System Attributes, Hazards, and Harms
Fuente:
arXiv
Saved in:
| Main Authors: | Rismani, Shalaleh, Shelby, Renee, Davis, Leah, Rostamzadeh, Negar, Moon, AJung |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Silos to Systems: Process-Oriented Hazard Analysis for AI Systems
by: Rismani, Shalaleh, et al.
Published: (2024)
by: Rismani, Shalaleh, et al.
Published: (2024)
From Use to Oversight: How Mental Models Influence User Behavior and Output in AI Writing Assistants
by: Rismani, Shalaleh, et al.
Published: (2026)
by: Rismani, Shalaleh, et al.
Published: (2026)
Towards A Framework for Levels of Anthropomorphic Deception in Robots and AI
by: Babel, Franziska, et al.
Published: (2026)
by: Babel, Franziska, et al.
Published: (2026)
Evaluating Human-AI Safety: A Framework for Measuring Harmful Capability Uplift
by: Vaccaro, Michelle, et al.
Published: (2026)
by: Vaccaro, Michelle, et al.
Published: (2026)
How Tech Workers Contend with Hazards of Humanlikeness in Generative AI
by: Díaz, Mark, et al.
Published: (2025)
by: Díaz, Mark, et al.
Published: (2025)
Harmful Traits of AI Companions
by: Knox, W. Bradley, et al.
Published: (2025)
by: Knox, W. Bradley, et al.
Published: (2025)
AI Generated Child Sexual Abuse Material -- What's the Harm?
by: Ciardha, Caoilte Ó, et al.
Published: (2025)
by: Ciardha, Caoilte Ó, et al.
Published: (2025)
AI Mismatches: Identifying Potential Algorithmic Harms Before AI Development
by: Saxena, Devansh, et al.
Published: (2025)
by: Saxena, Devansh, et al.
Published: (2025)
Interaction-Centered Intelligence: Toward Interaction as the Primary Unit of Analysis in Co-Creative AI and Human-AI Systems
by: Davis, Nicholas
Published: (2026)
by: Davis, Nicholas
Published: (2026)
The Widening Gap: The Benefits and Harms of Generative AI for Novice Programmers
by: Prather, James, et al.
Published: (2024)
by: Prather, James, et al.
Published: (2024)
AI's Regimes of Representation: A Community-centered Study of Text-to-Image Models in South Asia
by: Qadri, Rida, et al.
Published: (2023)
by: Qadri, Rida, et al.
Published: (2023)
Attribution Gradients: Incrementally Unfolding Citations for Critical Examination of Attributed AI Answers
by: Kambhamettu, Hita, et al.
Published: (2025)
by: Kambhamettu, Hita, et al.
Published: (2025)
Biased Error Attribution in Multi-Agent Human-AI Systems Under Delayed Feedback
by: Parakh, Teerthaa, et al.
Published: (2026)
by: Parakh, Teerthaa, et al.
Published: (2026)
AI Chatbots for Mental Health: Values and Harms from Lived Experiences of Depression
by: Yoo, Dong Whi, et al.
Published: (2025)
by: Yoo, Dong Whi, et al.
Published: (2025)
DisaBench: A Participatory Evaluation Framework for Disability Harms in Language Models
by: Kim, Eugenia, et al.
Published: (2026)
by: Kim, Eugenia, et al.
Published: (2026)
Positioning AI Tools to Support Online Harm Reduction Practice: Applications and Design Directions
by: Wang, Kaixuan, et al.
Published: (2025)
by: Wang, Kaixuan, et al.
Published: (2025)
Secondary Stakeholders in AI: Fighting for, Brokering, and Navigating Agency
by: Ajmani, Leah Hope, et al.
Published: (2025)
by: Ajmani, Leah Hope, et al.
Published: (2025)
What Is Required for Empathic AI? It Depends, and Why That Matters for AI Developers and Users
by: Borg, Jana Schaich, et al.
Published: (2024)
by: Borg, Jana Schaich, et al.
Published: (2024)
SPHERE: An Evaluation Card for Human-AI Systems
by: Ma, Qianou, et al.
Published: (2025)
by: Ma, Qianou, et al.
Published: (2025)
Detecting and Preventing Harmful Behaviors in AI Companions: Development and Evaluation of the SHIELD Supervisory System
by: Ben-Zion, Ziv, et al.
Published: (2025)
by: Ben-Zion, Ziv, et al.
Published: (2025)
AI Ethics and Governance in Practice: An Introduction
by: Leslie, David, et al.
Published: (2024)
by: Leslie, David, et al.
Published: (2024)
Underspecified Human Decision Experiments Considered Harmful
by: Hullman, Jessica, et al.
Published: (2024)
by: Hullman, Jessica, et al.
Published: (2024)
Causal Responsibility Attribution for Human-AI Collaboration
by: Qi, Yahang, et al.
Published: (2024)
by: Qi, Yahang, et al.
Published: (2024)
What Makes for a Good Saliency Map? Comparing Strategies for Evaluating Saliency Maps in Explainable AI (XAI)
by: Kares, Felix, et al.
Published: (2025)
by: Kares, Felix, et al.
Published: (2025)
Dubito Ergo Sum: Exploring AI Ethics
by: Dorfler, Viktor, et al.
Published: (2025)
by: Dorfler, Viktor, et al.
Published: (2025)
EthicAlly: a Prototype for AI-Powered Research Ethics Support for the Social Sciences and Humanities
by: Grohmann, Steph
Published: (2025)
by: Grohmann, Steph
Published: (2025)
The Generative AI Paradox on Evaluation: What It Can Solve, It May Not Evaluate
by: Oh, Juhyun, et al.
Published: (2024)
by: Oh, Juhyun, et al.
Published: (2024)
AI Cat Narrator: Designing an AI Tool for Exploring the Shared World and Social Connection with a Cat
by: Lai, Zhenchi, et al.
Published: (2024)
by: Lai, Zhenchi, et al.
Published: (2024)
Interactive Counterfactual Exploration of Algorithmic Harms in Recommender Systems
by: Ahn, Yongsu, et al.
Published: (2024)
by: Ahn, Yongsu, et al.
Published: (2024)
Feeling Machines: Ethics, Culture, and the Rise of Emotional AI
by: Chavan, Vivek, et al.
Published: (2025)
by: Chavan, Vivek, et al.
Published: (2025)
AI Ethics and Social Norms: Exploring ChatGPT's Capabilities From What to How
by: Veisi, Omid, et al.
Published: (2025)
by: Veisi, Omid, et al.
Published: (2025)
Assessing the Quality of Mental Health Support in LLM Responses through Multi-Attribute Human Evaluation
by: Badawi, Abeer, et al.
Published: (2026)
by: Badawi, Abeer, et al.
Published: (2026)
How Culture Shapes What People Want From AI
by: Ge, Xiao, et al.
Published: (2024)
by: Ge, Xiao, et al.
Published: (2024)
What happens when reviewers receive AI feedback in their reviews?
by: Chen, Shiping, et al.
Published: (2026)
by: Chen, Shiping, et al.
Published: (2026)
From Future of Work to Future of Workers: Addressing Asymptomatic AI Harms for Dignified Human-AI Interaction
by: Ehsan, Upol, et al.
Published: (2026)
by: Ehsan, Upol, et al.
Published: (2026)
Surveys Considered Harmful? Reflecting on the Use of Surveys in AI Research, Development, and Governance
by: Tahaei, Mohammmad, et al.
Published: (2024)
by: Tahaei, Mohammmad, et al.
Published: (2024)
A Decision Theoretic Framework for Measuring AI Reliance
by: Guo, Ziyang, et al.
Published: (2024)
by: Guo, Ziyang, et al.
Published: (2024)
Human and AI Trust: Trust Attitude Measurement Instrument
by: Larasati, Retno
Published: (2025)
by: Larasati, Retno
Published: (2025)
Cross-subject Brain Functional Connectivity Analysis for Multi-task Cognitive State Evaluation
by: Chen, Jun, et al.
Published: (2024)
by: Chen, Jun, et al.
Published: (2024)
How do Visual Attributes Influence Web Agents? A Comprehensive Evaluation of User Interface Design Factors
by: Yu, Kuai, et al.
Published: (2026)
by: Yu, Kuai, et al.
Published: (2026)
Similar Items
-
From Silos to Systems: Process-Oriented Hazard Analysis for AI Systems
by: Rismani, Shalaleh, et al.
Published: (2024) -
From Use to Oversight: How Mental Models Influence User Behavior and Output in AI Writing Assistants
by: Rismani, Shalaleh, et al.
Published: (2026) -
Towards A Framework for Levels of Anthropomorphic Deception in Robots and AI
by: Babel, Franziska, et al.
Published: (2026) -
Evaluating Human-AI Safety: A Framework for Measuring Harmful Capability Uplift
by: Vaccaro, Michelle, et al.
Published: (2026) -
How Tech Workers Contend with Hazards of Humanlikeness in Generative AI
by: Díaz, Mark, et al.
Published: (2025)