The Human Factor in AI Red Teaming: Perspectives from Social and Collaborative Computing
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Alice Qian, Shaw, Ryland, Anthis, Jacy Reese, Milton, Ashlee, Tseng, Emily, Suh, Jina, Ahmad, Lama, Kumar, Ram Shankar Siva, Posada, Julian, Shestakofsky, Benjamin, Roberts, Sarah T., Gray, Mary L. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AI red-teaming is a sociotechnical problem: on values, labor, and harms
por: Gillespie, Tarleton, et al.
Publicado: (2024)
por: Gillespie, Tarleton, et al.
Publicado: (2024)
Effective Automation to Support the Human Infrastructure in AI Red Teaming
por: Zhang, Alice Qian, et al.
Publicado: (2025)
por: Zhang, Alice Qian, et al.
Publicado: (2025)
Locating Risk: Task Designers and the Challenge of Risk Disclosure in RAI Content Work
por: Qian, Alice, et al.
Publicado: (2025)
por: Qian, Alice, et al.
Publicado: (2025)
Which Artificial Intelligences Do People Care About Most? A Conjoint Experiment on Moral Consideration
por: Ladak, Ali, et al.
Publicado: (2024)
por: Ladak, Ali, et al.
Publicado: (2024)
Worker Discretion Advised: Co-designing Risk Disclosure in Crowdsourced Responsible AI (RAI) Content Work
por: Qian, Alice, et al.
Publicado: (2025)
por: Qian, Alice, et al.
Publicado: (2025)
The AI Double Standard: Humans Judge All AIs for the Actions of One
por: Manoli, Aikaterina, et al.
Publicado: (2024)
por: Manoli, Aikaterina, et al.
Publicado: (2024)
HumanAgencyBench: Scalable Evaluation of Human Agency Support in AI Assistants
por: Sturgeon, Benjamin, et al.
Publicado: (2025)
por: Sturgeon, Benjamin, et al.
Publicado: (2025)
Robots, Chatbots, Self-Driving Cars: Perceptions of Mind and Morality Across Artificial Intelligences
por: Ladak, Ali, et al.
Publicado: (2025)
por: Ladak, Ali, et al.
Publicado: (2025)
Perceptions of Sentient AI and Other Digital Minds: Evidence from the AI, Morality, and Sentience (AIMS) Survey
por: Anthis, Jacy Reese, et al.
Publicado: (2024)
por: Anthis, Jacy Reese, et al.
Publicado: (2024)
Mental Models of Autonomy and Sentience Shape Reactions to AI
por: Pauketat, Janet V. T., et al.
Publicado: (2025)
por: Pauketat, Janet V. T., et al.
Publicado: (2025)
Effects of Generative AI Errors on User Reliance Across Task Difficulty
por: Anthis, Jacy Reese, et al.
Publicado: (2026)
por: Anthis, Jacy Reese, et al.
Publicado: (2026)
AURA: Amplifying Understanding, Resilience, and Awareness for Responsible AI Content Work
por: Zhang, Alice Qian, et al.
Publicado: (2024)
por: Zhang, Alice Qian, et al.
Publicado: (2024)
How AI Companionship Develops: Evidence from a Longitudinal Study
por: Hwang, Angel Hsing-Chi, et al.
Publicado: (2025)
por: Hwang, Angel Hsing-Chi, et al.
Publicado: (2025)
Bias in Language Models: Beyond Trick Tests and Toward RUTEd Evaluation
por: Lum, Kristian, et al.
Publicado: (2024)
por: Lum, Kristian, et al.
Publicado: (2024)
Public Opinion and The Rise of Digital Minds: Perceived Risk, Trust, and Regulation Support
por: Bullock, Justin B., et al.
Publicado: (2025)
por: Bullock, Justin B., et al.
Publicado: (2025)
Digital Companionship: Overlapping Uses of AI Companions and AI Assistants
por: Manoli, Aikaterina, et al.
Publicado: (2025)
por: Manoli, Aikaterina, et al.
Publicado: (2025)
The Impossibility of Fair LLMs
por: Anthis, Jacy, et al.
Publicado: (2024)
por: Anthis, Jacy, et al.
Publicado: (2024)
LLM Social Simulations Are a Promising Research Method
por: Anthis, Jacy Reese, et al.
Publicado: (2025)
por: Anthis, Jacy Reese, et al.
Publicado: (2025)
The Dynamics of Delusion: Modeling Bidirectional False Belief Amplification in Human-Chatbot Dialogue
por: Mehta, Ashish, et al.
Publicado: (2026)
por: Mehta, Ashish, et al.
Publicado: (2026)
What do Large Language Models Say About Animals? Investigating Risks of Animal Harm in Generated Text
por: Kanepajs, Arturs, et al.
Publicado: (2025)
por: Kanepajs, Arturs, et al.
Publicado: (2025)
OpenAI's Approach to External Red Teaming for AI Models and Systems
por: Ahmad, Lama, et al.
Publicado: (2025)
por: Ahmad, Lama, et al.
Publicado: (2025)
Reckoning with the Political Economy of AI: Avoiding Decoys in Pursuit of Accountability
por: Vertesi, Janet, et al.
Publicado: (2026)
por: Vertesi, Janet, et al.
Publicado: (2026)
Participation in the age of foundation models
por: Suresh, Harini, et al.
Publicado: (2024)
por: Suresh, Harini, et al.
Publicado: (2024)
User Perceptions of an LLM-Based Chatbot for Cognitive Reappraisal of Stress: Feasibility Study
por: Bhattacharjee, Ananya, et al.
Publicado: (2026)
por: Bhattacharjee, Ananya, et al.
Publicado: (2026)
Red Teaming AI Red Teaming
por: Majumdar, Subhabrata, et al.
Publicado: (2025)
por: Majumdar, Subhabrata, et al.
Publicado: (2025)
From Measurement to Expertise: Empathetic Expert Adapters for Context-Based Empathy in Conversational AI Agents
por: Shayegani, Erfan, et al.
Publicado: (2025)
por: Shayegani, Erfan, et al.
Publicado: (2025)
Prompt Optimization and Evaluation for LLM Automated Red Teaming
por: Freenor, Michael, et al.
Publicado: (2025)
por: Freenor, Michael, et al.
Publicado: (2025)
PyRIT: A Framework for Security Risk Identification and Red Teaming in Generative AI System
por: Munoz, Gary D. Lopez, et al.
Publicado: (2024)
por: Munoz, Gary D. Lopez, et al.
Publicado: (2024)
Unraveling Entangled Feeds: Rethinking Social Media Design to Enhance User Well-being
por: Milton, Ashlee, et al.
Publicado: (2026)
por: Milton, Ashlee, et al.
Publicado: (2026)
PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI
por: Deng, Wesley Hanwen, et al.
Publicado: (2026)
por: Deng, Wesley Hanwen, et al.
Publicado: (2026)
DISCERN: Designing Decision Support Interfaces to Investigate the Complexities of Workplace Social Decision-Making With Line Managers
por: Khadpe, Pranav, et al.
Publicado: (2024)
por: Khadpe, Pranav, et al.
Publicado: (2024)
Adaptive Instruction Composition for Automated LLM Red-Teaming
por: Zymet, Jesse, et al.
Publicado: (2026)
por: Zymet, Jesse, et al.
Publicado: (2026)
Red Teaming Visual Language Models
por: Li, Mukai, et al.
Publicado: (2024)
por: Li, Mukai, et al.
Publicado: (2024)
Characterizing Delusional Spirals through Human-LLM Chat Logs
por: Moore, Jared, et al.
Publicado: (2026)
por: Moore, Jared, et al.
Publicado: (2026)
"I'm Petting the Laptop, Which Has You Inside It": Reflecting on Lived Experiences of Online Friendship
por: Yong, Seraphina, et al.
Publicado: (2025)
por: Yong, Seraphina, et al.
Publicado: (2025)
Towards Red Teaming in Multimodal and Multilingual Translation
por: Ropers, Christophe, et al.
Publicado: (2024)
por: Ropers, Christophe, et al.
Publicado: (2024)
Automated Progressive Red Teaming
por: Jiang, Bojian, et al.
Publicado: (2024)
por: Jiang, Bojian, et al.
Publicado: (2024)
Anecdoctoring: Automated Red-Teaming Across Language and Place
por: Cuevas, Alejandro, et al.
Publicado: (2025)
por: Cuevas, Alejandro, et al.
Publicado: (2025)
IMBUE: Improving Interpersonal Effectiveness through Simulation and Just-in-time Feedback with Human-Language Model Interaction
por: Lin, Inna Wanyin, et al.
Publicado: (2024)
por: Lin, Inna Wanyin, et al.
Publicado: (2024)
Supporting the Digital Autonomy of Elders Through LLM Assistance
por: Roberts, Jesse, et al.
Publicado: (2024)
por: Roberts, Jesse, et al.
Publicado: (2024)
Ejemplares similares
-
AI red-teaming is a sociotechnical problem: on values, labor, and harms
por: Gillespie, Tarleton, et al.
Publicado: (2024) -
Effective Automation to Support the Human Infrastructure in AI Red Teaming
por: Zhang, Alice Qian, et al.
Publicado: (2025) -
Locating Risk: Task Designers and the Challenge of Risk Disclosure in RAI Content Work
por: Qian, Alice, et al.
Publicado: (2025) -
Which Artificial Intelligences Do People Care About Most? A Conjoint Experiment on Moral Consideration
por: Ladak, Ali, et al.
Publicado: (2024) -
Worker Discretion Advised: Co-designing Risk Disclosure in Crowdsourced Responsible AI (RAI) Content Work
por: Qian, Alice, et al.
Publicado: (2025)