AI red-teaming is a sociotechnical problem: on values, labor, and harms
Fuente:
arXiv
Salvato in:
| Autori principali: | Gillespie, Tarleton, Shaw, Ryland, Gray, Mary L., Suh, Jina |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Human Factor in AI Red Teaming: Perspectives from Social and Collaborative Computing
di: Zhang, Alice Qian, et al.
Pubblicazione: (2024)
di: Zhang, Alice Qian, et al.
Pubblicazione: (2024)
Locating Risk: Task Designers and the Challenge of Risk Disclosure in RAI Content Work
di: Qian, Alice, et al.
Pubblicazione: (2025)
di: Qian, Alice, et al.
Pubblicazione: (2025)
Worker Discretion Advised: Co-designing Risk Disclosure in Crowdsourced Responsible AI (RAI) Content Work
di: Qian, Alice, et al.
Pubblicazione: (2025)
di: Qian, Alice, et al.
Pubblicazione: (2025)
Effective Automation to Support the Human Infrastructure in AI Red Teaming
di: Zhang, Alice Qian, et al.
Pubblicazione: (2025)
di: Zhang, Alice Qian, et al.
Pubblicazione: (2025)
AURA: Amplifying Understanding, Resilience, and Awareness for Responsible AI Content Work
di: Zhang, Alice Qian, et al.
Pubblicazione: (2024)
di: Zhang, Alice Qian, et al.
Pubblicazione: (2024)
Human-centered explanation does not fit all: The interplay of sociotechnical, cognitive, and individual factors in the effect AI explanations in algorithmic decision-making
di: Ahn, Yongsu, et al.
Pubblicazione: (2025)
di: Ahn, Yongsu, et al.
Pubblicazione: (2025)
You Can't Get There From Here: Redefining Information Science to address our sociotechnical futures
di: Humr, Scott, et al.
Pubblicazione: (2025)
di: Humr, Scott, et al.
Pubblicazione: (2025)
When Testing AI Tests Us: Safeguarding Mental Health on the Digital Frontlines
di: Pendse, Sachin R., et al.
Pubblicazione: (2025)
di: Pendse, Sachin R., et al.
Pubblicazione: (2025)
Longitudinal Study on Social and Emotional Use of AI Conversational Agent
di: Chandra, Mohit, et al.
Pubblicazione: (2025)
di: Chandra, Mohit, et al.
Pubblicazione: (2025)
Towards interactive evaluations for interaction harms in human-AI systems
di: Ibrahim, Lujain, et al.
Pubblicazione: (2024)
di: Ibrahim, Lujain, et al.
Pubblicazione: (2024)
Seeking Late Night Life Lines: Experiences of Conversational AI Use in Mental Health Crisis
di: Ajmani, Leah Hope, et al.
Pubblicazione: (2025)
di: Ajmani, Leah Hope, et al.
Pubblicazione: (2025)
Characterizing and modeling harms from interactions with design patterns in AI interfaces
di: Ibrahim, Lujain, et al.
Pubblicazione: (2024)
di: Ibrahim, Lujain, et al.
Pubblicazione: (2024)
From Lived Experience to Insight: Unpacking the Psychological Risks of Using AI Conversational Agents
di: Chandra, Mohit, et al.
Pubblicazione: (2024)
di: Chandra, Mohit, et al.
Pubblicazione: (2024)
From Measurement to Expertise: Empathetic Expert Adapters for Context-Based Empathy in Conversational AI Agents
di: Shayegani, Erfan, et al.
Pubblicazione: (2025)
di: Shayegani, Erfan, et al.
Pubblicazione: (2025)
"It's not a representation of me": Examining Accent Bias and Digital Exclusion in Synthetic AI Voice Services
di: Michel, Shira, et al.
Pubblicazione: (2025)
di: Michel, Shira, et al.
Pubblicazione: (2025)
AI and Identity
di: Tadimalla, Sri Yash, et al.
Pubblicazione: (2024)
di: Tadimalla, Sri Yash, et al.
Pubblicazione: (2024)
Cultural Bias in Explainable AI Research: A Systematic Analysis
di: Peters, Uwe, et al.
Pubblicazione: (2024)
di: Peters, Uwe, et al.
Pubblicazione: (2024)
SENSE-7: Taxonomy and Dataset for Measuring User Perceptions of Empathy in Sustained Human-AI Conversations
di: Suh, Jina, et al.
Pubblicazione: (2025)
di: Suh, Jina, et al.
Pubblicazione: (2025)
From Stem to Stern: Contestability Along AI Value Chains
di: Balayn, Agathe, et al.
Pubblicazione: (2024)
di: Balayn, Agathe, et al.
Pubblicazione: (2024)
Generative AI in K-12 Education: The CyberScholar Initiative
di: Castro, Vania, et al.
Pubblicazione: (2025)
di: Castro, Vania, et al.
Pubblicazione: (2025)
Participation in the age of foundation models
di: Suresh, Harini, et al.
Pubblicazione: (2024)
di: Suresh, Harini, et al.
Pubblicazione: (2024)
A Rational Analysis of the Effects of Sycophantic AI
di: Batista, Rafael M., et al.
Pubblicazione: (2026)
di: Batista, Rafael M., et al.
Pubblicazione: (2026)
Exploring Human-AI Collaboration Using Mental Models of Early Adopters of Multi-Agent Generative AI Tools
di: Naik, Suchismita, et al.
Pubblicazione: (2025)
di: Naik, Suchismita, et al.
Pubblicazione: (2025)
"This is not a data problem": Algorithms and Power in Public Higher Education in Canada
di: McConvey, Kelly, et al.
Pubblicazione: (2024)
di: McConvey, Kelly, et al.
Pubblicazione: (2024)
AI-rays: Exploring Bias in the Gaze of AI Through a Multimodal Interactive Installation
di: Gao, Ziyao, et al.
Pubblicazione: (2024)
di: Gao, Ziyao, et al.
Pubblicazione: (2024)
The Right to AI
di: Mushkani, Rashid, et al.
Pubblicazione: (2025)
di: Mushkani, Rashid, et al.
Pubblicazione: (2025)
Digital Companionship: Overlapping Uses of AI Companions and AI Assistants
di: Manoli, Aikaterina, et al.
Pubblicazione: (2025)
di: Manoli, Aikaterina, et al.
Pubblicazione: (2025)
AI Sustainability in Practice Part One: Foundations for Sustainable AI Projects
di: Leslie, David, et al.
Pubblicazione: (2024)
di: Leslie, David, et al.
Pubblicazione: (2024)
AI Sustainability in Practice Part Two: Sustainability Throughout the AI Workflow
di: Leslie, David, et al.
Pubblicazione: (2024)
di: Leslie, David, et al.
Pubblicazione: (2024)
The Who in XAI: How AI Background Shapes Perceptions of AI Explanations
di: Ehsan, Upol, et al.
Pubblicazione: (2021)
di: Ehsan, Upol, et al.
Pubblicazione: (2021)
Generative AI User Experience: Developing Human--AI Epistemic Partnership
di: Zhai, Xiaoming
Pubblicazione: (2026)
di: Zhai, Xiaoming
Pubblicazione: (2026)
AI Meets Mathematics Education: A Case Study on Supporting an Instructor in a Large Mathematics Class with Context-Aware AI
di: Barghorn, Jérémy, et al.
Pubblicazione: (2026)
di: Barghorn, Jérémy, et al.
Pubblicazione: (2026)
Can AI be a moral victim? The role of moral patiency and ownership perceptions in ethical judgments of using AI-generated content
di: Choung, Hyesun, et al.
Pubblicazione: (2026)
di: Choung, Hyesun, et al.
Pubblicazione: (2026)
AI Fairness in Practice
di: Leslie, David, et al.
Pubblicazione: (2024)
di: Leslie, David, et al.
Pubblicazione: (2024)
Can AI be Auditable?
di: Verma, Himanshu, et al.
Pubblicazione: (2025)
di: Verma, Himanshu, et al.
Pubblicazione: (2025)
The AI Criminal Mastermind
di: Krook, Joshua
Pubblicazione: (2026)
di: Krook, Joshua
Pubblicazione: (2026)
What Is Required for Empathic AI? It Depends, and Why That Matters for AI Developers and Users
di: Borg, Jana Schaich, et al.
Pubblicazione: (2024)
di: Borg, Jana Schaich, et al.
Pubblicazione: (2024)
Advancing Trustworthy AI for Sustainable Development: Recommendations for Standardising AI Incident Reporting
di: Agarwal, Avinash, et al.
Pubblicazione: (2025)
di: Agarwal, Avinash, et al.
Pubblicazione: (2025)
Keeping an Eye on AI: A Framework for Effective Human Oversight of AI Systems
di: Gaube, Susanne, et al.
Pubblicazione: (2026)
di: Gaube, Susanne, et al.
Pubblicazione: (2026)
The Missing Knowledge Layer in AI: A Framework for Stable Human-AI Reasoning
di: Rosenbacke, Rikard, et al.
Pubblicazione: (2026)
di: Rosenbacke, Rikard, et al.
Pubblicazione: (2026)
Documenti analoghi
-
The Human Factor in AI Red Teaming: Perspectives from Social and Collaborative Computing
di: Zhang, Alice Qian, et al.
Pubblicazione: (2024) -
Locating Risk: Task Designers and the Challenge of Risk Disclosure in RAI Content Work
di: Qian, Alice, et al.
Pubblicazione: (2025) -
Worker Discretion Advised: Co-designing Risk Disclosure in Crowdsourced Responsible AI (RAI) Content Work
di: Qian, Alice, et al.
Pubblicazione: (2025) -
Effective Automation to Support the Human Infrastructure in AI Red Teaming
di: Zhang, Alice Qian, et al.
Pubblicazione: (2025) -
AURA: Amplifying Understanding, Resilience, and Awareness for Responsible AI Content Work
di: Zhang, Alice Qian, et al.
Pubblicazione: (2024)