A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations
Fuente:
arXiv
Salvato in:
| Autori principali: | Davani, Aida, Dev, Sunipa, Pérez-Urbina, Héctor, Prabhakaran, Vinodkumar |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GeniL: A Multilingual Dataset on Generalizing Language
di: Davani, Aida Mostafazadeh, et al.
Pubblicazione: (2024)
di: Davani, Aida Mostafazadeh, et al.
Pubblicazione: (2024)
Adaptive Data Collection for Latin-American Community-sourced Evaluation of Stereotypes (LACES)
di: Ivetta, Guido, et al.
Pubblicazione: (2025)
di: Ivetta, Guido, et al.
Pubblicazione: (2025)
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
di: Jha, Akshita, et al.
Pubblicazione: (2024)
di: Jha, Akshita, et al.
Pubblicazione: (2024)
SeeGULL Multilingual: a Dataset of Geo-Culturally Situated Stereotypes
di: Bhutani, Mukul, et al.
Pubblicazione: (2024)
di: Bhutani, Mukul, et al.
Pubblicazione: (2024)
Cultural Compass: A Framework for Organizing Societal Norms to Detect Violations in Human-AI Conversations
di: Cheng, Myra, et al.
Pubblicazione: (2026)
di: Cheng, Myra, et al.
Pubblicazione: (2026)
A Unified Framework to Quantify Cultural Intelligence of AI
di: Dev, Sunipa, et al.
Pubblicazione: (2026)
di: Dev, Sunipa, et al.
Pubblicazione: (2026)
SAFARI: A Community-Engaged Approach and Dataset of Stereotype Resources in the Sub-Saharan African Context
di: Verma, Aishwarya, et al.
Pubblicazione: (2026)
di: Verma, Aishwarya, et al.
Pubblicazione: (2026)
D3CODE: Disentangling Disagreements in Data across Cultures on Offensiveness Detection and Evaluation
di: Davani, Aida Mostafazadeh, et al.
Pubblicazione: (2024)
di: Davani, Aida Mostafazadeh, et al.
Pubblicazione: (2024)
Scaling Cultural Resources for Improving Generative Models
di: Stepanyan, Hayk, et al.
Pubblicazione: (2025)
di: Stepanyan, Hayk, et al.
Pubblicazione: (2025)
Risks of Cultural Erasure in Large Language Models
di: Qadri, Rida, et al.
Pubblicazione: (2025)
di: Qadri, Rida, et al.
Pubblicazione: (2025)
Humanlike AI Design Increases Anthropomorphism but Yields Divergent Outcomes on Engagement and Trust Globally
di: Schimmelpfennig, Robin, et al.
Pubblicazione: (2025)
di: Schimmelpfennig, Robin, et al.
Pubblicazione: (2025)
Towards Geo-Culturally Grounded LLM Generations
di: Lertvittayakumjorn, Piyawat, et al.
Pubblicazione: (2025)
di: Lertvittayakumjorn, Piyawat, et al.
Pubblicazione: (2025)
MiTTenS: A Dataset for Evaluating Gender Mistranslation
di: Robinson, Kevin, et al.
Pubblicazione: (2024)
di: Robinson, Kevin, et al.
Pubblicazione: (2024)
Tracing the Techno-Supremacy Doctrine: A Critical Discourse Analysis of the AI Executive Elite
di: Pérez-Urbina, Héctor
Pubblicazione: (2025)
di: Pérez-Urbina, Héctor
Pubblicazione: (2025)
AI TIPS 2.0: A Comprehensive Framework for Operationalizing AI Governance
di: Gupta, Pamela
Pubblicazione: (2025)
di: Gupta, Pamela
Pubblicazione: (2025)
Cultural Authenticity: Comparing LLM Cultural Representations to Native Human Expectations
di: van Liemt, Erin MacMurray, et al.
Pubblicazione: (2026)
di: van Liemt, Erin MacMurray, et al.
Pubblicazione: (2026)
JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors
di: Jin, Jiho, et al.
Pubblicazione: (2026)
di: Jin, Jiho, et al.
Pubblicazione: (2026)
Framework, Standards, Applications and Best practices of Responsible AI : A Comprehensive Survey
di: Gadekallu, Thippa Reddy, et al.
Pubblicazione: (2025)
di: Gadekallu, Thippa Reddy, et al.
Pubblicazione: (2025)
Operationalizing Justice: Towards the Development of a Principle Based Design Framework for Human Services AI
di: Rodriguez, Maria Y., et al.
Pubblicazione: (2025)
di: Rodriguez, Maria Y., et al.
Pubblicazione: (2025)
AI, Climate, and Transparency: Operationalizing and Improving the AI Act
di: Alder, Nicolas, et al.
Pubblicazione: (2024)
di: Alder, Nicolas, et al.
Pubblicazione: (2024)
Vernacularizing Taxonomies of Harm is Essential for Operationalizing Holistic AI Safety
di: Kennedy, Wm. Matthew, et al.
Pubblicazione: (2024)
di: Kennedy, Wm. Matthew, et al.
Pubblicazione: (2024)
GRASP: A Disagreement Analysis Framework to Assess Group Associations in Perspectives
di: Prabhakaran, Vinodkumar, et al.
Pubblicazione: (2023)
di: Prabhakaran, Vinodkumar, et al.
Pubblicazione: (2023)
Generative AI Literacy: A Comprehensive Framework for Literacy and Responsible Use
di: Zhang, Chengzhi, et al.
Pubblicazione: (2025)
di: Zhang, Chengzhi, et al.
Pubblicazione: (2025)
Decoding Safety Feedback from Diverse Raters: A Data-driven Lens on Responsiveness to Severity
di: Mishra, Pushkar, et al.
Pubblicazione: (2025)
di: Mishra, Pushkar, et al.
Pubblicazione: (2025)
Comprehensive Framework for Evaluating Conversational AI Chatbots
di: Gupta, Shailja, et al.
Pubblicazione: (2025)
di: Gupta, Shailja, et al.
Pubblicazione: (2025)
SESGO: Spanish Evaluation of Stereotypical Generative Outputs
di: Robles, Melissa, et al.
Pubblicazione: (2025)
di: Robles, Melissa, et al.
Pubblicazione: (2025)
Bias and Volatility: A Statistical Framework for Evaluating Large Language Model's Stereotypes and the Associated Generation Inconsistency
di: Liu, Yiran, et al.
Pubblicazione: (2024)
di: Liu, Yiran, et al.
Pubblicazione: (2024)
The Bidirectional Relationship Between XAI and Regulation: Operationalizing XAI for the AI Act
di: Hummel, Anton, et al.
Pubblicazione: (2025)
di: Hummel, Anton, et al.
Pubblicazione: (2025)
Making Power Explicable in AI: Analyzing, Understanding, and Redirecting Power to Operationalize Ethics in AI Technical Practice
di: Jin, Weina, et al.
Pubblicazione: (2025)
di: Jin, Weina, et al.
Pubblicazione: (2025)
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes
di: Nejadgholi, Isar, et al.
Pubblicazione: (2024)
di: Nejadgholi, Isar, et al.
Pubblicazione: (2024)
(Unfair) Norms in Fairness Research: A Meta-Analysis
di: Chien, Jennifer, et al.
Pubblicazione: (2024)
di: Chien, Jennifer, et al.
Pubblicazione: (2024)
Surfacing Subtle Stereotypes: A Multilingual, Debate-Oriented Evaluation of Modern LLMs
di: Saeed, Muhammed, et al.
Pubblicazione: (2025)
di: Saeed, Muhammed, et al.
Pubblicazione: (2025)
SAIF: A Comprehensive Framework for Evaluating the Risks of Generative AI in the Public Sector
di: Lee, Kyeongryul, et al.
Pubblicazione: (2025)
di: Lee, Kyeongryul, et al.
Pubblicazione: (2025)
Evaluating Chinese Large Language Models: The Influence of Persona Assignment on Stereotypes and Safeguards
di: Liu, Geng, et al.
Pubblicazione: (2025)
di: Liu, Geng, et al.
Pubblicazione: (2025)
Operationalizing Ethics for AI Agents: How Developers Encode Values into Repository Context Files
di: Treude, Christoph, et al.
Pubblicazione: (2026)
di: Treude, Christoph, et al.
Pubblicazione: (2026)
From Perceived Effectiveness to Measured Impact: Identity-Aware Evaluation of Automated Counter-Stereotypes
di: Kiritchenko, Svetlana, et al.
Pubblicazione: (2025)
di: Kiritchenko, Svetlana, et al.
Pubblicazione: (2025)
MisgenderMender: A Community-Informed Approach to Interventions for Misgendering
di: Hossain, Tamanna, et al.
Pubblicazione: (2024)
di: Hossain, Tamanna, et al.
Pubblicazione: (2024)
StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs
di: Jeune, Pierre Le, et al.
Pubblicazione: (2026)
di: Jeune, Pierre Le, et al.
Pubblicazione: (2026)
Responsible AI Question Bank: A Comprehensive Tool for AI Risk Assessment
di: Lee, Sung Une, et al.
Pubblicazione: (2024)
di: Lee, Sung Une, et al.
Pubblicazione: (2024)
Specification, Application, and Operationalization of a Metamodel of Fairness
di: Mendez, Julian Alfredo, et al.
Pubblicazione: (2025)
di: Mendez, Julian Alfredo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
GeniL: A Multilingual Dataset on Generalizing Language
di: Davani, Aida Mostafazadeh, et al.
Pubblicazione: (2024) -
Adaptive Data Collection for Latin-American Community-sourced Evaluation of Stereotypes (LACES)
di: Ivetta, Guido, et al.
Pubblicazione: (2025) -
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
di: Jha, Akshita, et al.
Pubblicazione: (2024) -
SeeGULL Multilingual: a Dataset of Geo-Culturally Situated Stereotypes
di: Bhutani, Mukul, et al.
Pubblicazione: (2024) -
Cultural Compass: A Framework for Organizing Societal Norms to Detect Violations in Human-AI Conversations
di: Cheng, Myra, et al.
Pubblicazione: (2026)