PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Deng, Wesley Hanwen, Yan, Mingxi, Kim, Sunnie S. Y., Jha, Akshita, Wilcox, Lauren, Holstein, Kenneth, Eslami, Motahhare, Gatys, Leon A. |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
PersonaTeaming: Exploring How Introducing Personas Can Improve Automated AI Red-Teaming
par: Deng, Wesley Hanwen, et autres
Publié: (2025)
par: Deng, Wesley Hanwen, et autres
Publié: (2025)
MIRAGE: Multi-model Interface for Reviewing and Auditing Generative Text-to-Image AI
par: Maldaner, Matheus Kunzler, et autres
Publié: (2025)
par: Maldaner, Matheus Kunzler, et autres
Publié: (2025)
WeAudit: Scaffolding User Auditors and AI Practitioners in Auditing Generative AI
par: Deng, Wesley Hanwen, et autres
Publié: (2025)
par: Deng, Wesley Hanwen, et autres
Publié: (2025)
Seeing Twice: How Side-by-Side T2I Comparison Changes Auditing Strategies
par: Maldaner, Matheus Kunzler, et autres
Publié: (2025)
par: Maldaner, Matheus Kunzler, et autres
Publié: (2025)
"I Don't Think RAI Applies to My Model'' -- Engaging Non-champions with Sticky Stories for Responsible AI Work
par: Nahar, Nadia, et autres
Publié: (2025)
par: Nahar, Nadia, et autres
Publié: (2025)
Red-Teaming for Generative AI: Silver Bullet or Security Theater?
par: Feffer, Michael, et autres
Publié: (2024)
par: Feffer, Michael, et autres
Publié: (2024)
Critical or Compliant? The Double-Edged Sword of Reasoning in Chain-of-Thought Explanations
par: Park, Eunkyu, et autres
Publié: (2025)
par: Park, Eunkyu, et autres
Publié: (2025)
Vipera: Blending Visual and LLM-Driven Guidance for Systematic Auditing of Text-to-Image Generative AI
par: Huang, Yanwei, et autres
Publié: (2025)
par: Huang, Yanwei, et autres
Publié: (2025)
Investigating Youth AI Auditing
par: Solyst, Jaemarie, et autres
Publié: (2025)
par: Solyst, Jaemarie, et autres
Publié: (2025)
Vipera: Towards systematic auditing of generative text-to-image models at scale
par: Huang, Yanwei, et autres
Publié: (2025)
par: Huang, Yanwei, et autres
Publié: (2025)
The Social Blindspot in Human-AI Collaboration: How Undetected AI Personas Reshape Team Dynamics
par: Yan, Lixiang, et autres
Publié: (2025)
par: Yan, Lixiang, et autres
Publié: (2025)
Effective Automation to Support the Human Infrastructure in AI Red Teaming
par: Zhang, Alice Qian, et autres
Publié: (2025)
par: Zhang, Alice Qian, et autres
Publié: (2025)
MM-SCALE: Grounded Multimodal Moral Reasoning via Scalar Judgment and Listwise Alignment
par: Park, Eunkyu, et autres
Publié: (2026)
par: Park, Eunkyu, et autres
Publié: (2026)
Strategies for Designing Responsibly within a Capitalist Enterprise
par: Xie, Shixian, et autres
Publié: (2026)
par: Xie, Shixian, et autres
Publié: (2026)
CulturalTeaming: AI-Assisted Interactive Red-Teaming for Challenging LLMs' (Lack of) Multicultural Knowledge
par: Chiu, Yu Ying, et autres
Publié: (2024)
par: Chiu, Yu Ying, et autres
Publié: (2024)
TeamVision: An AI-powered Learning Analytics System for Supporting Reflection in Team-based Healthcare Simulation
par: Echeverria, Vanessa, et autres
Publié: (2025)
par: Echeverria, Vanessa, et autres
Publié: (2025)
Supporting Industry Computing Researchers in Assessing, Articulating, and Addressing the Potential Negative Societal Impact of Their Work
par: Deng, Wesley Hanwen, et autres
Publié: (2024)
par: Deng, Wesley Hanwen, et autres
Publié: (2024)
Children's Overtrust and Shifting Perspectives of Generative AI
par: Solyst, Jaemarie, et autres
Publié: (2024)
par: Solyst, Jaemarie, et autres
Publié: (2024)
Breaking Political Filter Bubbles via Social Comparison
par: Soliman, Nouran, et autres
Publié: (2024)
par: Soliman, Nouran, et autres
Publié: (2024)
Summon a Demon and Bind it: A Grounded Theory of LLM Red Teaming
par: Inie, Nanna, et autres
Publié: (2023)
par: Inie, Nanna, et autres
Publié: (2023)
Personas with Attitudes: Controlling LLMs for Diverse Data Annotation
par: Fröhling, Leon, et autres
Publié: (2024)
par: Fröhling, Leon, et autres
Publié: (2024)
Supporting AI-Augmented Meta-Decision Making with InDecision
par: Castañeda, Chance, et autres
Publié: (2025)
par: Castañeda, Chance, et autres
Publié: (2025)
Troubling Taxonomies in GenAI Evaluation
par: Berman, Glen, et autres
Publié: (2024)
par: Berman, Glen, et autres
Publié: (2024)
OpenAI's Approach to External Red Teaming for AI Models and Systems
par: Ahmad, Lama, et autres
Publié: (2025)
par: Ahmad, Lama, et autres
Publié: (2025)
Legacy Procurement Practices Shape How U.S. Cities Govern AI: Understanding Government Employees' Practices, Challenges, and Needs
par: Johnson, Nari, et autres
Publié: (2024)
par: Johnson, Nari, et autres
Publié: (2024)
Improving the State of the Art for Training Human-AI Teams: Technical Report #5 -- Individual Differences and Team Qualities to Measure in a Human-AI Teaming Testbed
par: Asiala, Lillian, et autres
Publié: (2025)
par: Asiala, Lillian, et autres
Publié: (2025)
The Human Factor in AI Red Teaming: Perspectives from Social and Collaborative Computing
par: Zhang, Alice Qian, et autres
Publié: (2024)
par: Zhang, Alice Qian, et autres
Publié: (2024)
The Situate AI Guidebook: Co-Designing a Toolkit to Support Multi-Stakeholder Early-stage Deliberations Around Public Sector AI Proposals
par: Kawakami, Anna, et autres
Publié: (2024)
par: Kawakami, Anna, et autres
Publié: (2024)
Organization Matters: A Qualitative Study of Organizational Dynamics in Red Teaming Practices for Generative AI
par: Ren, Bixuan, et autres
Publié: (2025)
par: Ren, Bixuan, et autres
Publié: (2025)
Proxona: Supporting Creators' Sensemaking and Ideation with LLM-Powered Audience Personas
par: Choi, Yoonseo, et autres
Publié: (2024)
par: Choi, Yoonseo, et autres
Publié: (2024)
Human-AI Teaming Under Deception: An Implicit BCI Safeguards Drone Team Performance in Virtual Reality
par: Baker, Christopher, et autres
Publié: (2025)
par: Baker, Christopher, et autres
Publié: (2025)
Prototyping Multimodal GenAI Real-Time Agents with Counterfactual Replays and Hybrid Wizard-of-Oz
par: Gmeiner, Frederic, et autres
Publié: (2025)
par: Gmeiner, Frederic, et autres
Publié: (2025)
Putting Privacy to the Test: Introducing Red Teaming for Research Data Anonymization
par: Jansen, Luisa, et autres
Publié: (2026)
par: Jansen, Luisa, et autres
Publié: (2026)
Assessing Risks of Large Language Models in Mental Health Support: A Framework for Automated Clinical AI Red Teaming
par: Steenstra, Ian, et autres
Publié: (2026)
par: Steenstra, Ian, et autres
Publié: (2026)
ConSearcher: Supporting Conversational Information Seeking in Online Communities with Member Personas
par: Wu, Shiwei, et autres
Publié: (2026)
par: Wu, Shiwei, et autres
Publié: (2026)
PersonaAI: Leveraging Retrieval-Augmented Generation and Personalized Context for AI-Driven Digital Avatars
par: Kimara, Elvis, et autres
Publié: (2025)
par: Kimara, Elvis, et autres
Publié: (2025)
Advising Agent for Supporting Human-Multi-Drone Team Collaboration
par: Barr, Hodaya, et autres
Publié: (2025)
par: Barr, Hodaya, et autres
Publié: (2025)
Exploring the Potential of Metacognitive Support Agents for Human-AI Co-Creation
par: Gmeiner, Frederic, et autres
Publié: (2025)
par: Gmeiner, Frederic, et autres
Publié: (2025)
Runtime Advocates: A Persona-Driven Framework for Requirements@Runtime Decision Support
par: Hernandez, Demetrius, et autres
Publié: (2025)
par: Hernandez, Demetrius, et autres
Publié: (2025)
What Comes After Harm? Mapping Reparative Actions in AI through Justice Frameworks
par: Xiao, Sijia, et autres
Publié: (2025)
par: Xiao, Sijia, et autres
Publié: (2025)
Documents similaires
-
PersonaTeaming: Exploring How Introducing Personas Can Improve Automated AI Red-Teaming
par: Deng, Wesley Hanwen, et autres
Publié: (2025) -
MIRAGE: Multi-model Interface for Reviewing and Auditing Generative Text-to-Image AI
par: Maldaner, Matheus Kunzler, et autres
Publié: (2025) -
WeAudit: Scaffolding User Auditors and AI Practitioners in Auditing Generative AI
par: Deng, Wesley Hanwen, et autres
Publié: (2025) -
Seeing Twice: How Side-by-Side T2I Comparison Changes Auditing Strategies
par: Maldaner, Matheus Kunzler, et autres
Publié: (2025) -
"I Don't Think RAI Applies to My Model'' -- Engaging Non-champions with Sticky Stories for Responsible AI Work
par: Nahar, Nadia, et autres
Publié: (2025)