GeniL: A Multilingual Dataset on Generalizing Language
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Davani, Aida Mostafazadeh, Gubbi, Sagar, Dev, Sunipa, Dave, Shachi, Prabhakaran, Vinodkumar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SeeGULL Multilingual: a Dataset of Geo-Culturally Situated Stereotypes
von: Bhutani, Mukul, et al.
Veröffentlicht: (2024)
von: Bhutani, Mukul, et al.
Veröffentlicht: (2024)
D3CODE: Disentangling Disagreements in Data across Cultures on Offensiveness Detection and Evaluation
von: Davani, Aida Mostafazadeh, et al.
Veröffentlicht: (2024)
von: Davani, Aida Mostafazadeh, et al.
Veröffentlicht: (2024)
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations
von: Davani, Aida, et al.
Veröffentlicht: (2025)
von: Davani, Aida, et al.
Veröffentlicht: (2025)
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
von: Jha, Akshita, et al.
Veröffentlicht: (2024)
von: Jha, Akshita, et al.
Veröffentlicht: (2024)
Risks of Cultural Erasure in Large Language Models
von: Qadri, Rida, et al.
Veröffentlicht: (2025)
von: Qadri, Rida, et al.
Veröffentlicht: (2025)
Towards Geo-Culturally Grounded LLM Generations
von: Lertvittayakumjorn, Piyawat, et al.
Veröffentlicht: (2025)
von: Lertvittayakumjorn, Piyawat, et al.
Veröffentlicht: (2025)
SAFARI: A Community-Engaged Approach and Dataset of Stereotype Resources in the Sub-Saharan African Context
von: Verma, Aishwarya, et al.
Veröffentlicht: (2026)
von: Verma, Aishwarya, et al.
Veröffentlicht: (2026)
Cultural Authenticity: Comparing LLM Cultural Representations to Native Human Expectations
von: van Liemt, Erin MacMurray, et al.
Veröffentlicht: (2026)
von: van Liemt, Erin MacMurray, et al.
Veröffentlicht: (2026)
GRASP: A Disagreement Analysis Framework to Assess Group Associations in Perspectives
von: Prabhakaran, Vinodkumar, et al.
Veröffentlicht: (2023)
von: Prabhakaran, Vinodkumar, et al.
Veröffentlicht: (2023)
JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors
von: Jin, Jiho, et al.
Veröffentlicht: (2026)
von: Jin, Jiho, et al.
Veröffentlicht: (2026)
Cultural Compass: A Framework for Organizing Societal Norms to Detect Violations in Human-AI Conversations
von: Cheng, Myra, et al.
Veröffentlicht: (2026)
von: Cheng, Myra, et al.
Veröffentlicht: (2026)
MisgenderMender: A Community-Informed Approach to Interventions for Misgendering
von: Hossain, Tamanna, et al.
Veröffentlicht: (2024)
von: Hossain, Tamanna, et al.
Veröffentlicht: (2024)
MiTTenS: A Dataset for Evaluating Gender Mistranslation
von: Robinson, Kevin, et al.
Veröffentlicht: (2024)
von: Robinson, Kevin, et al.
Veröffentlicht: (2024)
Humanlike AI Design Increases Anthropomorphism but Yields Divergent Outcomes on Engagement and Trust Globally
von: Schimmelpfennig, Robin, et al.
Veröffentlicht: (2025)
von: Schimmelpfennig, Robin, et al.
Veröffentlicht: (2025)
Taxonomy of User Needs and Actions
von: Shelby, Renee, et al.
Veröffentlicht: (2025)
von: Shelby, Renee, et al.
Veröffentlicht: (2025)
A Unified Framework to Quantify Cultural Intelligence of AI
von: Dev, Sunipa, et al.
Veröffentlicht: (2026)
von: Dev, Sunipa, et al.
Veröffentlicht: (2026)
Whose View of Safety? A Deep DIVE Dataset for Pluralistic Alignment of Text-to-Image Models
von: Rastogi, Charvi, et al.
Veröffentlicht: (2025)
von: Rastogi, Charvi, et al.
Veröffentlicht: (2025)
Genie: Achieving Human Parity in Content-Grounded Datasets Generation
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2024)
Scaling Cultural Resources for Improving Generative Models
von: Stepanyan, Hayk, et al.
Veröffentlicht: (2025)
von: Stepanyan, Hayk, et al.
Veröffentlicht: (2025)
Adaptive Data Collection for Latin-American Community-sourced Evaluation of Stereotypes (LACES)
von: Ivetta, Guido, et al.
Veröffentlicht: (2025)
von: Ivetta, Guido, et al.
Veröffentlicht: (2025)
Beyond Aesthetics: Cultural Competence in Text-to-Image Models
von: Kannen, Nithish, et al.
Veröffentlicht: (2024)
von: Kannen, Nithish, et al.
Veröffentlicht: (2024)
ELSA: A Style Aligned Dataset for Emotionally Intelligent Language Generation
von: Gandhi, Vishal, et al.
Veröffentlicht: (2025)
von: Gandhi, Vishal, et al.
Veröffentlicht: (2025)
Arabizi vs LLMs: Can the Genie Understand the Language of Aladdin?
von: Almaoui, Perla Al, et al.
Veröffentlicht: (2025)
von: Almaoui, Perla Al, et al.
Veröffentlicht: (2025)
QuesGenie: Intelligent Multimodal Question Generation
von: Mubarak, Ahmed, et al.
Veröffentlicht: (2025)
von: Mubarak, Ahmed, et al.
Veröffentlicht: (2025)
TALES: A Taxonomy and Analysis of Cultural Representations in LLM-generated Stories
von: Bhagat, Kirti, et al.
Veröffentlicht: (2025)
von: Bhagat, Kirti, et al.
Veröffentlicht: (2025)
Efficiently Identifying Low-Quality Language Subsets in Multilingual Datasets: A Case Study on a Large-Scale Multilingual Audio Dataset
von: Samir, Farhan, et al.
Veröffentlicht: (2024)
von: Samir, Farhan, et al.
Veröffentlicht: (2024)
ReactGenie: A Development Framework for Complex Multimodal Interactions Using Large Language Models
von: Yang, Jackie Junrui, et al.
Veröffentlicht: (2023)
von: Yang, Jackie Junrui, et al.
Veröffentlicht: (2023)
A New Massive Multilingual Dataset for High-Performance Language Technologies
von: de Gibert, Ona, et al.
Veröffentlicht: (2024)
von: de Gibert, Ona, et al.
Veröffentlicht: (2024)
Taxi1500: A Multilingual Dataset for Text Classification in 1500 Languages
von: Ma, Chunlan, et al.
Veröffentlicht: (2023)
von: Ma, Chunlan, et al.
Veröffentlicht: (2023)
AfriVoices-KE: A Multilingual Speech Dataset for Kenyan Languages
von: Wanzare, Lilian, et al.
Veröffentlicht: (2026)
von: Wanzare, Lilian, et al.
Veröffentlicht: (2026)
LexGenie: Automated Generation of Structured Reports for European Court of Human Rights Case Law
von: Santosh, T. Y. S. S, et al.
Veröffentlicht: (2025)
von: Santosh, T. Y. S. S, et al.
Veröffentlicht: (2025)
Multilingual Text Style Transfer: Datasets & Models for Indian Languages
von: Mukherjee, Sourabrata, et al.
Veröffentlicht: (2024)
von: Mukherjee, Sourabrata, et al.
Veröffentlicht: (2024)
Multilingual Synopses of Movie Narratives: A Dataset for Vision-Language Story Understanding
von: Sun, Yidan, et al.
Veröffentlicht: (2024)
von: Sun, Yidan, et al.
Veröffentlicht: (2024)
Abstractive Summarization of Low resourced Nepali language using Multilingual Transformers
von: Dhakal, Prakash, et al.
Veröffentlicht: (2024)
von: Dhakal, Prakash, et al.
Veröffentlicht: (2024)
AfriHate: A Multilingual Collection of Hate Speech and Abusive Language Datasets for African Languages
von: Muhammad, Shamsuddeen Hassan, et al.
Veröffentlicht: (2025)
von: Muhammad, Shamsuddeen Hassan, et al.
Veröffentlicht: (2025)
An Expanded Massive Multilingual Dataset for High-Performance Language Technologies (HPLT)
von: Burchell, Laurie, et al.
Veröffentlicht: (2025)
von: Burchell, Laurie, et al.
Veröffentlicht: (2025)
GenieBlue: Integrating both Linguistic and Multimodal Capabilities for Large Language Models on Mobile Devices
von: Lu, Xudong, et al.
Veröffentlicht: (2025)
von: Lu, Xudong, et al.
Veröffentlicht: (2025)
VLR-Bench: Multilingual Benchmark Dataset for Vision-Language Retrieval Augmented Generation
von: Lim, Hyeonseok, et al.
Veröffentlicht: (2024)
von: Lim, Hyeonseok, et al.
Veröffentlicht: (2024)
Constructing Multilingual Visual-Text Datasets Revealing Visual Multilingual Ability of Vision Language Models
von: Atuhurra, Jesse, et al.
Veröffentlicht: (2024)
von: Atuhurra, Jesse, et al.
Veröffentlicht: (2024)
MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SeeGULL Multilingual: a Dataset of Geo-Culturally Situated Stereotypes
von: Bhutani, Mukul, et al.
Veröffentlicht: (2024) -
D3CODE: Disentangling Disagreements in Data across Cultures on Offensiveness Detection and Evaluation
von: Davani, Aida Mostafazadeh, et al.
Veröffentlicht: (2024) -
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations
von: Davani, Aida, et al.
Veröffentlicht: (2025) -
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
von: Jha, Akshita, et al.
Veröffentlicht: (2024) -
Risks of Cultural Erasure in Large Language Models
von: Qadri, Rida, et al.
Veröffentlicht: (2025)