Effective Automation to Support the Human Infrastructure in AI Red Teaming
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Alice Qian, Suh, Jina, Gray, Mary L., Shen, Hong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AURA: Amplifying Understanding, Resilience, and Awareness for Responsible AI Content Work
by: Zhang, Alice Qian, et al.
Published: (2024)
by: Zhang, Alice Qian, et al.
Published: (2024)
The Human Factor in AI Red Teaming: Perspectives from Social and Collaborative Computing
by: Zhang, Alice Qian, et al.
Published: (2024)
by: Zhang, Alice Qian, et al.
Published: (2024)
Worker Discretion Advised: Co-designing Risk Disclosure in Crowdsourced Responsible AI (RAI) Content Work
by: Qian, Alice, et al.
Published: (2025)
by: Qian, Alice, et al.
Published: (2025)
Locating Risk: Task Designers and the Challenge of Risk Disclosure in RAI Content Work
by: Qian, Alice, et al.
Published: (2025)
by: Qian, Alice, et al.
Published: (2025)
AI red-teaming is a sociotechnical problem: on values, labor, and harms
by: Gillespie, Tarleton, et al.
Published: (2024)
by: Gillespie, Tarleton, et al.
Published: (2024)
PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI
by: Deng, Wesley Hanwen, et al.
Published: (2026)
by: Deng, Wesley Hanwen, et al.
Published: (2026)
From Risk Avoidance to User Empowerment in AI Mental Health Crisis Support
by: Kaveladze, Benjamin, et al.
Published: (2026)
by: Kaveladze, Benjamin, et al.
Published: (2026)
Beyond the Single Turn: Reframing Refusals as Dynamic Experiences Embedded in the Context of Mental Health Support Interactions with LLMs
by: Tang, Ningjing, et al.
Published: (2026)
by: Tang, Ningjing, et al.
Published: (2026)
Red-Teaming for Generative AI: Silver Bullet or Security Theater?
by: Feffer, Michael, et al.
Published: (2024)
by: Feffer, Michael, et al.
Published: (2024)
Assessing Risks of Large Language Models in Mental Health Support: A Framework for Automated Clinical AI Red Teaming
by: Steenstra, Ian, et al.
Published: (2026)
by: Steenstra, Ian, et al.
Published: (2026)
From Measurement to Expertise: Empathetic Expert Adapters for Context-Based Empathy in Conversational AI Agents
by: Shayegani, Erfan, et al.
Published: (2025)
by: Shayegani, Erfan, et al.
Published: (2025)
When Testing AI Tests Us: Safeguarding Mental Health on the Digital Frontlines
by: Pendse, Sachin R., et al.
Published: (2025)
by: Pendse, Sachin R., et al.
Published: (2025)
Longitudinal Study on Social and Emotional Use of AI Conversational Agent
by: Chandra, Mohit, et al.
Published: (2025)
by: Chandra, Mohit, et al.
Published: (2025)
OpenAI's Approach to External Red Teaming for AI Models and Systems
by: Ahmad, Lama, et al.
Published: (2025)
by: Ahmad, Lama, et al.
Published: (2025)
Sovereign Agents: Towards Infrastructural Sovereignty and Diffused Accountability in Decentralized AI
by: Hu, Botao Amber, et al.
Published: (2026)
by: Hu, Botao Amber, et al.
Published: (2026)
Constructing Algorithmic Authority: How Multi-Channel Networks (MCNs) Govern Live-Streaming Labor in China
by: Xiao, Qing, et al.
Published: (2025)
by: Xiao, Qing, et al.
Published: (2025)
Seeking Late Night Life Lines: Experiences of Conversational AI Use in Mental Health Crisis
by: Ajmani, Leah Hope, et al.
Published: (2025)
by: Ajmani, Leah Hope, et al.
Published: (2025)
From Lived Experience to Insight: Unpacking the Psychological Risks of Using AI Conversational Agents
by: Chandra, Mohit, et al.
Published: (2024)
by: Chandra, Mohit, et al.
Published: (2024)
Talking to a Human as an Attitudinal Barrier: A Mixed Methods Evaluation of Stigma, Access, and the Appeal of AI Mental Health Support
by: Stamatis, Caitlin A., et al.
Published: (2026)
by: Stamatis, Caitlin A., et al.
Published: (2026)
Keeping an Eye on AI: A Framework for Effective Human Oversight of AI Systems
by: Gaube, Susanne, et al.
Published: (2026)
by: Gaube, Susanne, et al.
Published: (2026)
What's So Human about Human-AI Collaboration, Anyway? Generative AI and Human-Computer Interaction
by: Watkins, Elizabeth Anne, et al.
Published: (2025)
by: Watkins, Elizabeth Anne, et al.
Published: (2025)
Can GenAI Move from Individual Use to Collaborative Work? Experiences, Challenges, and Opportunities of Coordinating GenAI into Collaborative Newswork
by: Xiao, Qing, et al.
Published: (2025)
by: Xiao, Qing, et al.
Published: (2025)
AI-Driven Human-Autonomy Teaming in Tactical Operations: Proposed Framework, Challenges, and Future Directions
by: Hagos, Desta Haileselassie, et al.
Published: (2024)
by: Hagos, Desta Haileselassie, et al.
Published: (2024)
When Your Boss Is an AI Bot: Exploring Opportunities and Risks of Manager Clone Agents in the Future Workplace
by: Hu, Qing, et al.
Published: (2025)
by: Hu, Qing, et al.
Published: (2025)
Designing Human-AI Collaboration to Support Learning in Counterspeech Writing
by: Ding, Xiaohan, et al.
Published: (2024)
by: Ding, Xiaohan, et al.
Published: (2024)
From Regulation to Support: Centering Humans in Technology-Mediated Emotion Intervention in Care Contexts
by: Liu, Jiaying "Lizzy", et al.
Published: (2025)
by: Liu, Jiaying "Lizzy", et al.
Published: (2025)
Confident-Knowledge Diversity Drives Human-Human and Human-AI Free Discussion Synergy and Reveals Pure-AI Discussion Shortfalls
by: Sheffer, Tom, et al.
Published: (2025)
by: Sheffer, Tom, et al.
Published: (2025)
DISCERN: Designing Decision Support Interfaces to Investigate the Complexities of Workplace Social Decision-Making With Line Managers
by: Khadpe, Pranav, et al.
Published: (2024)
by: Khadpe, Pranav, et al.
Published: (2024)
Towards Apples to Apples for AI Evaluations: From Real-World Use Cases to Evaluation Scenarios
by: Choong, Yee-Yin, et al.
Published: (2026)
by: Choong, Yee-Yin, et al.
Published: (2026)
The Dark Side of AI Companionship: A Taxonomy of Harmful Algorithmic Behaviors in Human-AI Relationships
by: Zhang, Renwen, et al.
Published: (2024)
by: Zhang, Renwen, et al.
Published: (2024)
Interactive AI and Human Behavior: Challenges and Pathways for AI Governance
by: Pi, Yulu, et al.
Published: (2025)
by: Pi, Yulu, et al.
Published: (2025)
The Agency Gap: How Generative AI Literacy Shapes Independent Writing after AI Support
by: Jin, Yueqiao, et al.
Published: (2025)
by: Jin, Yueqiao, et al.
Published: (2025)
Human-Centered Human-AI Collaboration (HCHAC)
by: Gao, Qi, et al.
Published: (2025)
by: Gao, Qi, et al.
Published: (2025)
Privacy Perspectives and Practices of Chinese Smart Home Product Teams
by: He, Shijing, et al.
Published: (2025)
by: He, Shijing, et al.
Published: (2025)
Protecting Human Cognition in the Age of AI
by: Singh, Anjali, et al.
Published: (2025)
by: Singh, Anjali, et al.
Published: (2025)
Bidirectional Human-AI Alignment in Education for Trustworthy Learning Environments
by: Shen, Hua
Published: (2025)
by: Shen, Hua
Published: (2025)
Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South
by: Rastogi, Charvi, et al.
Published: (2026)
by: Rastogi, Charvi, et al.
Published: (2026)
Understanding and Supporting Peer Review Using AI-reframed Positive Summary
by: Yang, Chi-Lan, et al.
Published: (2025)
by: Yang, Chi-Lan, et al.
Published: (2025)
Navigating Privacy and Trust: AI Assistants as Social Support for Older Adults
by: LaRubbio, Karina, et al.
Published: (2025)
by: LaRubbio, Karina, et al.
Published: (2025)
IMBUE: Improving Interpersonal Effectiveness through Simulation and Just-in-time Feedback with Human-Language Model Interaction
by: Lin, Inna Wanyin, et al.
Published: (2024)
by: Lin, Inna Wanyin, et al.
Published: (2024)
Similar Items
-
AURA: Amplifying Understanding, Resilience, and Awareness for Responsible AI Content Work
by: Zhang, Alice Qian, et al.
Published: (2024) -
The Human Factor in AI Red Teaming: Perspectives from Social and Collaborative Computing
by: Zhang, Alice Qian, et al.
Published: (2024) -
Worker Discretion Advised: Co-designing Risk Disclosure in Crowdsourced Responsible AI (RAI) Content Work
by: Qian, Alice, et al.
Published: (2025) -
Locating Risk: Task Designers and the Challenge of Risk Disclosure in RAI Content Work
by: Qian, Alice, et al.
Published: (2025) -
AI red-teaming is a sociotechnical problem: on values, labor, and harms
by: Gillespie, Tarleton, et al.
Published: (2024)