Red Teaming LLMs as Socio-Technical Practice: From Exploration and Data Creation to Evaluation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Garcia, Adriana Alvarado, Wan, Ruyuan, Oguine, Ozioma C., Badillo-Urquiola, Karla |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
'Teens Need to Be Educated on the Danger': Digital Access, Online Risks, and Safety Practices Among Nigerian Adolescents
par: Oguine, Munachimso B., et autres
Publié: (2025)
par: Oguine, Munachimso B., et autres
Publié: (2025)
Online Safety for All: Sociocultural Insights from a Systematic Review of Youth Online Safety in the Global South
par: Oguine, Ozioma C., et autres
Publié: (2025)
par: Oguine, Ozioma C., et autres
Publié: (2025)
How the Internet Facilitates Adverse Childhood Experiences for Youth Who Self-Identify as in Need of Services
par: Oguine, Ozioma C., et autres
Publié: (2024)
par: Oguine, Ozioma C., et autres
Publié: (2024)
A System of Care, Not Control: Co-Designing Online Safety and Wellbeing Solutions with Guardians ad Litem for Youth in Child Welfare
par: Olesk, Johanna, et autres
Publié: (2026)
par: Olesk, Johanna, et autres
Publié: (2026)
Participation and Power: A Case Study of Using Ecological Momentary Assessment to Engage Adolescents in Academic Research
par: Oguine, Ozioma C., et autres
Publié: (2026)
par: Oguine, Ozioma C., et autres
Publié: (2026)
CoCo Matrix: Taxonomy of Cognitive Contributions in Co-writing with Intelligent Agents
par: Wan, Ruyuan, et autres
Publié: (2024)
par: Wan, Ruyuan, et autres
Publié: (2024)
Design Patterns for the Common Good: Building Better Technologies Using the Wisdom of Virtue Ethics
par: Conwill, Louisa, et autres
Publié: (2025)
par: Conwill, Louisa, et autres
Publié: (2025)
STAR: SocioTechnical Approach to Red Teaming Language Models
par: Weidinger, Laura, et autres
Publié: (2024)
par: Weidinger, Laura, et autres
Publié: (2024)
The Missing Variable: Socio-Technical Alignment in Risk Evaluation
par: Flehmig, Niclas, et autres
Publié: (2025)
par: Flehmig, Niclas, et autres
Publié: (2025)
Human Perception of LLM-generated Text Content in Social Media Environments
par: Radivojevic, Kristina, et autres
Publié: (2024)
par: Radivojevic, Kristina, et autres
Publié: (2024)
Effective Automation to Support the Human Infrastructure in AI Red Teaming
par: Zhang, Alice Qian, et autres
Publié: (2025)
par: Zhang, Alice Qian, et autres
Publié: (2025)
Evaluation of LLMs Biases Towards Elite Universities: A Persona-Based Exploration
par: Gupta, Shailja, et autres
Publié: (2024)
par: Gupta, Shailja, et autres
Publié: (2024)
Privacy Perspectives and Practices of Chinese Smart Home Product Teams
par: He, Shijing, et autres
Publié: (2025)
par: He, Shijing, et autres
Publié: (2025)
The Challenges and Benefits of Bringing Religious Values Into Design
par: Conwill, Louisa, et autres
Publié: (2025)
par: Conwill, Louisa, et autres
Publié: (2025)
PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI
par: Deng, Wesley Hanwen, et autres
Publié: (2026)
par: Deng, Wesley Hanwen, et autres
Publié: (2026)
Red-Teaming for Generative AI: Silver Bullet or Security Theater?
par: Feffer, Michael, et autres
Publié: (2024)
par: Feffer, Michael, et autres
Publié: (2024)
The Human Factor in AI Red Teaming: Perspectives from Social and Collaborative Computing
par: Zhang, Alice Qian, et autres
Publié: (2024)
par: Zhang, Alice Qian, et autres
Publié: (2024)
Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South
par: Rastogi, Charvi, et autres
Publié: (2026)
par: Rastogi, Charvi, et autres
Publié: (2026)
Managing Project Teams in an Online Class of 1000+ Students
par: Anaraki, Nazanin Tabatabaei, et autres
Publié: (2024)
par: Anaraki, Nazanin Tabatabaei, et autres
Publié: (2024)
A Comparative Study of Technical Writing Feedback Quality: Evaluating LLMs, SLMs, and Humans in Computer Science Topics
par: Liu, Suqing, et autres
Publié: (2025)
par: Liu, Suqing, et autres
Publié: (2025)
Convivial Fabrication: Towards Relational Computational Tools For and From Craft Practices
par: Batra, Ritik, et autres
Publié: (2026)
par: Batra, Ritik, et autres
Publié: (2026)
From Script to Stage: Automating Experimental Design for Social Simulations with LLMs
par: Guo, Yuwei, et autres
Publié: (2025)
par: Guo, Yuwei, et autres
Publié: (2025)
Exploring Socio-Cultural Challenges and Opportunities in Designing Mental Health Chatbots for Adolescents in India
par: Sehgal, Neil K. R., et autres
Publié: (2025)
par: Sehgal, Neil K. R., et autres
Publié: (2025)
Beyond Western Politics: Cross-Cultural Benchmarks for Evaluating Partisan Associations in LLMs
par: Kumar, Divyanshu, et autres
Publié: (2025)
par: Kumar, Divyanshu, et autres
Publié: (2025)
Ethics and Technical Aspects of Generative AI Models in Digital Content Creation
par: Karagoz, Atahan
Publié: (2024)
par: Karagoz, Atahan
Publié: (2024)
From Framework to Practice: Designing a Real-World Telehealth Application for Palliative Care
par: Zhou, Wei, et autres
Publié: (2025)
par: Zhou, Wei, et autres
Publié: (2025)
Organizational Practices and Socio-Technical Design of Human-Centered AI
par: Herrmann, Thomas
Publié: (2026)
par: Herrmann, Thomas
Publié: (2026)
Comparing Human and AI Performance in Visual Storytelling through Creation of Comic Strips: A Case Study
par: Önal, Uğur, et autres
Publié: (2025)
par: Önal, Uğur, et autres
Publié: (2025)
Who Decides How Knowing Becomes Doing? Redistributing Authority in Human-AI Music Co-Creation
par: Hu, Zhejing, et autres
Publié: (2025)
par: Hu, Zhejing, et autres
Publié: (2025)
The Jade Gateway to Trust: Exploring How Socio-Cultural Perspectives Shape Trust Within Chinese NFT Communities
par: Cao, Yi-Fan, et autres
Publié: (2025)
par: Cao, Yi-Fan, et autres
Publié: (2025)
SaludConectaMX: Lessons Learned from Deploying a Cooperative Mobile Health System for Pediatric Cancer Care in Mexico
par: Schnur, Jennifer J., et autres
Publié: (2024)
par: Schnur, Jennifer J., et autres
Publié: (2024)
From Clicks to Consensus: Collective Consent Assemblies for Data Governance
par: Kyi, Lin, et autres
Publié: (2026)
par: Kyi, Lin, et autres
Publié: (2026)
"Newspaper Eat" Means "Not Tasty": A Taxonomy and Benchmark for Coded Language in Real-World Chinese Online Reviews
par: Wan, Ruyuan, et autres
Publié: (2026)
par: Wan, Ruyuan, et autres
Publié: (2026)
OpenAI's Approach to External Red Teaming for AI Models and Systems
par: Ahmad, Lama, et autres
Publié: (2025)
par: Ahmad, Lama, et autres
Publié: (2025)
Data Ethics Emergency Drill: A Toolbox for Discussing Responsible AI for Industry Teams
par: Hanschke, Vanessa Aisyahsari, et autres
Publié: (2024)
par: Hanschke, Vanessa Aisyahsari, et autres
Publié: (2024)
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
par: Arita, Takaya, et autres
Publié: (2025)
par: Arita, Takaya, et autres
Publié: (2025)
From Informal Addresses to Reliable Places: Participatory Data Governance of Civic Addressing in Puerto Rico
par: Padilla, Juan A.
Publié: (2026)
par: Padilla, Juan A.
Publié: (2026)
When LLMs Can't Help: Real-World Evaluation of LLMs in Nutrition
par: Li, Karen Jia-Hui, et autres
Publié: (2025)
par: Li, Karen Jia-Hui, et autres
Publié: (2025)
FATE in MMLA: A Student-Centred Exploration of Fairness, Accountability, Transparency, and Ethics in Multimodal Learning Analytics
par: Jin, Yueqiao, et autres
Publié: (2024)
par: Jin, Yueqiao, et autres
Publié: (2024)
Characterizing Scam-Driven Human Trafficking Across Chinese Borders and Online Community Responses on RedNote
par: Zheng, Jiamin, et autres
Publié: (2026)
par: Zheng, Jiamin, et autres
Publié: (2026)
Documents similaires
-
'Teens Need to Be Educated on the Danger': Digital Access, Online Risks, and Safety Practices Among Nigerian Adolescents
par: Oguine, Munachimso B., et autres
Publié: (2025) -
Online Safety for All: Sociocultural Insights from a Systematic Review of Youth Online Safety in the Global South
par: Oguine, Ozioma C., et autres
Publié: (2025) -
How the Internet Facilitates Adverse Childhood Experiences for Youth Who Self-Identify as in Need of Services
par: Oguine, Ozioma C., et autres
Publié: (2024) -
A System of Care, Not Control: Co-Designing Online Safety and Wellbeing Solutions with Guardians ad Litem for Youth in Child Welfare
par: Olesk, Johanna, et autres
Publié: (2026) -
Participation and Power: A Case Study of Using Ecological Momentary Assessment to Engage Adolescents in Academic Research
par: Oguine, Ozioma C., et autres
Publié: (2026)