The Elephant in the Room -- Why AI Safety Demands Diverse Teams
Fuente:
arXiv
Salvato in:
| Autori principali: | Rostcheck, David, Scheibling, Lara |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Illusion of Friendship: Why Generative AI Demands Unprecedented Ethical Vigilance
di: Islam, Md Zahidul
Pubblicazione: (2026)
di: Islam, Md Zahidul
Pubblicazione: (2026)
Catalyzing Equity in STEM Teams: Harnessing Generative AI for Inclusion and Diversity
di: Nixon, Nia, et al.
Pubblicazione: (2024)
di: Nixon, Nia, et al.
Pubblicazione: (2024)
Towards responsible AI for education: Hybrid human-AI to confront the Elephant in the room
di: Hooshyar, Danial, et al.
Pubblicazione: (2025)
di: Hooshyar, Danial, et al.
Pubblicazione: (2025)
The Homogenization Problem in LLMs: Towards Meaningful Diversity in AI Safety
di: Rios-Sialer, Ian
Pubblicazione: (2026)
di: Rios-Sialer, Ian
Pubblicazione: (2026)
Enhancing Team Diversity with Generative AI: A Novel Project Management Framework
di: Chan, Johnny, et al.
Pubblicazione: (2025)
di: Chan, Johnny, et al.
Pubblicazione: (2025)
Safety Cases: How to Justify the Safety of Advanced AI Systems
di: Clymer, Joshua, et al.
Pubblicazione: (2024)
di: Clymer, Joshua, et al.
Pubblicazione: (2024)
Red Teaming AI Red Teaming
di: Majumdar, Subhabrata, et al.
Pubblicazione: (2025)
di: Majumdar, Subhabrata, et al.
Pubblicazione: (2025)
Red Teaming AI Policy: A Taxonomy of Avoision and the EU AI Act
di: Yew, Rui-Jie, et al.
Pubblicazione: (2025)
di: Yew, Rui-Jie, et al.
Pubblicazione: (2025)
AI Safety is Stuck in Technical Terms -- A System Safety Response to the International AI Safety Report
di: Dobbe, Roel
Pubblicazione: (2025)
di: Dobbe, Roel
Pubblicazione: (2025)
PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI
di: Deng, Wesley Hanwen, et al.
Pubblicazione: (2026)
di: Deng, Wesley Hanwen, et al.
Pubblicazione: (2026)
International Agreements on AI Safety: Review and Recommendations for a Conditional AI Safety Treaty
di: Scholefield, Rebecca, et al.
Pubblicazione: (2025)
di: Scholefield, Rebecca, et al.
Pubblicazione: (2025)
Baseline Performance of AI Tools in Classifying Cognitive Demand of Mathematical Tasks
di: Fox, Danielle S., et al.
Pubblicazione: (2026)
di: Fox, Danielle S., et al.
Pubblicazione: (2026)
Safety Cases: A Scalable Approach to Frontier AI Safety
di: Hilton, Benjamin, et al.
Pubblicazione: (2025)
di: Hilton, Benjamin, et al.
Pubblicazione: (2025)
Interoperability in AI Safety Governance: Ethics, Regulations, and Standards
di: Chin, Yik Chan, et al.
Pubblicazione: (2026)
di: Chin, Yik Chan, et al.
Pubblicazione: (2026)
How English Print Media Frames Human-Elephant Conflicts in India
di: Punith, Bonala Sai, et al.
Pubblicazione: (2026)
di: Punith, Bonala Sai, et al.
Pubblicazione: (2026)
Toward an African Agenda for AI Safety
di: Segun, Samuel T., et al.
Pubblicazione: (2025)
di: Segun, Samuel T., et al.
Pubblicazione: (2025)
Concrete Problems in AI Safety, Revisited
di: Raji, Inioluwa Deborah, et al.
Pubblicazione: (2023)
di: Raji, Inioluwa Deborah, et al.
Pubblicazione: (2023)
Smaller, Younger, and More Impactful: How AI-Assisted Writing Transforms Research Teams
di: Wang, Haoyang, et al.
Pubblicazione: (2026)
di: Wang, Haoyang, et al.
Pubblicazione: (2026)
Position: The AI Conference Peer Review Crisis Demands Author Feedback and Reviewer Rewards
di: Kim, Jaeho, et al.
Pubblicazione: (2025)
di: Kim, Jaeho, et al.
Pubblicazione: (2025)
AI Safety: Necessary, but insufficient and possibly problematic
di: P, Deepak
Pubblicazione: (2024)
di: P, Deepak
Pubblicazione: (2024)
Emerging Practices in Frontier AI Safety Frameworks
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2025)
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2025)
Why Agents Compromise Safety Under Pressure
di: Jiang, Hengle, et al.
Pubblicazione: (2026)
di: Jiang, Hengle, et al.
Pubblicazione: (2026)
Why can't Epidemiology be automated (yet)?
di: Bann, David, et al.
Pubblicazione: (2025)
di: Bann, David, et al.
Pubblicazione: (2025)
Data Ethics Emergency Drill: A Toolbox for Discussing Responsible AI for Industry Teams
di: Hanschke, Vanessa Aisyahsari, et al.
Pubblicazione: (2024)
di: Hanschke, Vanessa Aisyahsari, et al.
Pubblicazione: (2024)
Racial/Ethnic Categories in AI and Algorithmic Fairness: Why They Matter and What They Represent
di: Mickel, Jennifer
Pubblicazione: (2024)
di: Mickel, Jennifer
Pubblicazione: (2024)
AI for Just Work: Constructing Diverse Imaginations of AI beyond "Replacing Humans"
di: Jin, Weina, et al.
Pubblicazione: (2025)
di: Jin, Weina, et al.
Pubblicazione: (2025)
Why Trust in AI May Be Inevitable
di: Truong, Nghi, et al.
Pubblicazione: (2025)
di: Truong, Nghi, et al.
Pubblicazione: (2025)
Upstream and Downstream AI Safety: Both on the Same River?
di: McDermid, John, et al.
Pubblicazione: (2024)
di: McDermid, John, et al.
Pubblicazione: (2024)
Probabilistic Analysis of Copyright Disputes and Generative AI Safety
di: Chiba-Okabe, Hiroaki
Pubblicazione: (2024)
di: Chiba-Okabe, Hiroaki
Pubblicazione: (2024)
The Ghost in the Grammar: Methodological Anthropomorphism in AI Safety Evaluations
di: Costa, Mariana Lins
Pubblicazione: (2026)
di: Costa, Mariana Lins
Pubblicazione: (2026)
Combining Cost-Constrained Runtime Monitors for AI Safety
di: Hua, Tim Tian, et al.
Pubblicazione: (2025)
di: Hua, Tim Tian, et al.
Pubblicazione: (2025)
Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents
di: Li, Miles Q., et al.
Pubblicazione: (2026)
di: Li, Miles Q., et al.
Pubblicazione: (2026)
Agentic Microphysics: A Manifesto for Generative AI Safety
di: Pierucci, Federico, et al.
Pubblicazione: (2026)
di: Pierucci, Federico, et al.
Pubblicazione: (2026)
Building Effective Safety Guardrails in AI Education Tools
di: Clark, Hannah-Beth, et al.
Pubblicazione: (2025)
di: Clark, Hannah-Beth, et al.
Pubblicazione: (2025)
What Is AI Safety? What Do We Want It to Be?
di: Harding, Jacqueline, et al.
Pubblicazione: (2025)
di: Harding, Jacqueline, et al.
Pubblicazione: (2025)
The Singapore Consensus on Global AI Safety Research Priorities
di: Bengio, Yoshua, et al.
Pubblicazione: (2025)
di: Bengio, Yoshua, et al.
Pubblicazione: (2025)
International Scientific Report on the Safety of Advanced AI (Interim Report)
di: Bengio, Yoshua, et al.
Pubblicazione: (2024)
di: Bengio, Yoshua, et al.
Pubblicazione: (2024)
Red Teaming for Generative AI, Report on a Copyright-Focused Exercise Completed in an Academic Medical Center
di: Wen, James, et al.
Pubblicazione: (2025)
di: Wen, James, et al.
Pubblicazione: (2025)
Prompting Diverse Ideas: Increasing AI Idea Variance
di: Meincke, Lennart, et al.
Pubblicazione: (2024)
di: Meincke, Lennart, et al.
Pubblicazione: (2024)
Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South
di: Rastogi, Charvi, et al.
Pubblicazione: (2026)
di: Rastogi, Charvi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
The Illusion of Friendship: Why Generative AI Demands Unprecedented Ethical Vigilance
di: Islam, Md Zahidul
Pubblicazione: (2026) -
Catalyzing Equity in STEM Teams: Harnessing Generative AI for Inclusion and Diversity
di: Nixon, Nia, et al.
Pubblicazione: (2024) -
Towards responsible AI for education: Hybrid human-AI to confront the Elephant in the room
di: Hooshyar, Danial, et al.
Pubblicazione: (2025) -
The Homogenization Problem in LLMs: Towards Meaningful Diversity in AI Safety
di: Rios-Sialer, Ian
Pubblicazione: (2026) -
Enhancing Team Diversity with Generative AI: A Novel Project Management Framework
di: Chan, Johnny, et al.
Pubblicazione: (2025)