Common to Whom? Regional Cultural Commonsense and LLM Bias in India
Fuente:
arXiv
Guardado en:
| Autores principales: | Madhusudan, Sangmitra, More, Trush Shashank, Buongiorno, Steph, Dividino, Renata, Kabbara, Jad, Emami, Ali |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Dog the Cat Chased Stumped the Model: Measuring When Language Models Abandon Structure for Shortcuts
por: Madhusudan, Sangmitra, et al.
Publicado: (2025)
por: Madhusudan, Sangmitra, et al.
Publicado: (2025)
Fine-Tuned LLMs are "Time Capsules" for Tracking Societal Bias Through Books
por: Madhusudan, Sangmitra, et al.
Publicado: (2025)
por: Madhusudan, Sangmitra, et al.
Publicado: (2025)
Which Words Matter Most in Zero-Shot Prompts?
por: Sadr, Nikta Gohari, et al.
Publicado: (2025)
por: Sadr, Nikta Gohari, et al.
Publicado: (2025)
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
por: Morabito, Robert, et al.
Publicado: (2024)
por: Morabito, Robert, et al.
Publicado: (2024)
DART: Mitigating Harm Drift in Difference-Aware LLMs via Distill-Audit-Repair Training
por: Pan, Ziwen, et al.
Publicado: (2026)
por: Pan, Ziwen, et al.
Publicado: (2026)
Confidence Under the Hood: An Investigation into the Confidence-Probability Alignment in Large Language Models
por: Kumar, Abhishek, et al.
Publicado: (2024)
por: Kumar, Abhishek, et al.
Publicado: (2024)
LLM Targeted Underperformance Disproportionately Impacts Vulnerable Users
por: Poole-Dayan, Elinor, et al.
Publicado: (2024)
por: Poole-Dayan, Elinor, et al.
Publicado: (2024)
An AI-Powered Framework for Analyzing Collective Idea Evolution in Deliberative Assemblies
por: Poole-Dayan, Elinor, et al.
Publicado: (2025)
por: Poole-Dayan, Elinor, et al.
Publicado: (2025)
On the Relationship between Truth and Political Bias in Language Models
por: Fulay, Suyash, et al.
Publicado: (2024)
por: Fulay, Suyash, et al.
Publicado: (2024)
Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset
por: Yamamoto, Taisei, et al.
Publicado: (2025)
por: Yamamoto, Taisei, et al.
Publicado: (2025)
PersonaLLM: Investigating the Ability of Large Language Models to Express Personality Traits
por: Jiang, Hang, et al.
Publicado: (2023)
por: Jiang, Hang, et al.
Publicado: (2023)
Translate With Care: Addressing Gender Bias, Neutrality, and Reasoning in Large Language Model Translations
por: Zahraei, Pardis Sadat, et al.
Publicado: (2025)
por: Zahraei, Pardis Sadat, et al.
Publicado: (2025)
I Am Aligned, But With Whom? MENA Values Benchmark for Evaluating Cultural Alignment and Multilingual Bias in LLMs
por: Zahraei, Pardis Sadat, et al.
Publicado: (2025)
por: Zahraei, Pardis Sadat, et al.
Publicado: (2025)
Computational Analysis of Conversation Dynamics through Participant Responsivity
por: Hughes, Margaret, et al.
Publicado: (2025)
por: Hughes, Margaret, et al.
Publicado: (2025)
Commonsense Reasoning in Arab Culture
por: Sadallah, Abdelrahman, et al.
Publicado: (2025)
por: Sadallah, Abdelrahman, et al.
Publicado: (2025)
PANGeA: Procedural Artificial Narrative using Generative AI for Turn-Based Video Games
por: Buongiorno, Steph, et al.
Publicado: (2024)
por: Buongiorno, Steph, et al.
Publicado: (2024)
Cultural Commonsense Knowledge for Intercultural Dialogues
por: Nguyen, Tuan-Phong, et al.
Publicado: (2024)
por: Nguyen, Tuan-Phong, et al.
Publicado: (2024)
A Framework for Leveraging Human Computation Gaming to Enhance Knowledge Graphs for Accuracy Critical Generative AI Applications
por: Buongiorno, Steph, et al.
Publicado: (2024)
por: Buongiorno, Steph, et al.
Publicado: (2024)
Just Put a Human in the Loop? Investigating LLM-Assisted Annotation for Subjective Tasks
por: Schroeder, Hope, et al.
Publicado: (2025)
por: Schroeder, Hope, et al.
Publicado: (2025)
Bridging Context Gaps: Enhancing Comprehension in Long-Form Social Conversations Through Contextualized Excerpts
por: Mohanty, Shrestha, et al.
Publicado: (2024)
por: Mohanty, Shrestha, et al.
Publicado: (2024)
LLMs as Cultural Archives: Cultural Commonsense Knowledge Graph Extraction
por: Tonga, Junior Cedric, et al.
Publicado: (2026)
por: Tonga, Junior Cedric, et al.
Publicado: (2026)
Modelling Commonsense Commonalities with Multi-Facet Concept Embeddings
por: Kteich, Hanane, et al.
Publicado: (2024)
por: Kteich, Hanane, et al.
Publicado: (2024)
ConGraT: Self-Supervised Contrastive Pretraining for Joint Graph and Text Embeddings
por: Brannon, William, et al.
Publicado: (2023)
por: Brannon, William, et al.
Publicado: (2023)
Can LLM Generate Culturally Relevant Commonsense QA Data? Case Study in Indonesian and Sundanese
por: Putri, Rifki Afina, et al.
Publicado: (2024)
por: Putri, Rifki Afina, et al.
Publicado: (2024)
Subtle Biases Need Subtler Measures: Dual Metrics for Evaluating Representative and Affinity Bias in Large Language Models
por: Kumar, Abhishek, et al.
Publicado: (2024)
por: Kumar, Abhishek, et al.
Publicado: (2024)
Understanding the Capabilities and Limitations of Large Language Models for Cultural Commonsense
por: Shen, Siqi, et al.
Publicado: (2024)
por: Shen, Siqi, et al.
Publicado: (2024)
Memory Dial: A Training Framework for Controllable Memorization in Language Models
por: Zhang, Xiangbo, et al.
Publicado: (2026)
por: Zhang, Xiangbo, et al.
Publicado: (2026)
Global PIQA: Evaluating Commonsense Reasoning Across 100+ Languages and Cultures
por: Chang, Tyler A., et al.
Publicado: (2025)
por: Chang, Tyler A., et al.
Publicado: (2025)
Metamorphic Testing for Fairness Evaluation in Large Language Models: Identifying Intersectional Bias in LLaMA and GPT
por: Reddy, Harishwar, et al.
Publicado: (2025)
por: Reddy, Harishwar, et al.
Publicado: (2025)
IndoCulture: Exploring Geographically-Influenced Cultural Commonsense Reasoning Across Eleven Indonesian Provinces
por: Koto, Fajri, et al.
Publicado: (2024)
por: Koto, Fajri, et al.
Publicado: (2024)
PizzaCommonSense: Learning to Model Commonsense Reasoning about Intermediate Steps in Cooking Recipes
por: Diallo, Aissatou, et al.
Publicado: (2024)
por: Diallo, Aissatou, et al.
Publicado: (2024)
EvoGrad: A Dynamic Take on the Winograd Schema Challenge with Human Adversaries
por: Sun, Jing Han, et al.
Publicado: (2024)
por: Sun, Jing Han, et al.
Publicado: (2024)
Ko-PIQA: A Korean Physical Commonsense Reasoning Dataset with Cultural Context
por: Choi, Dasol, et al.
Publicado: (2025)
por: Choi, Dasol, et al.
Publicado: (2025)
CommonWhy: A Dataset for Evaluating Entity-Based Causal Commonsense Reasoning in Large Language Models
por: Toroghi, Armin, et al.
Publicado: (2026)
por: Toroghi, Armin, et al.
Publicado: (2026)
Personality Matters: User Traits Predict LLM Preferences in Multi-Turn Collaborative Tasks
por: Yunusov, Sarfaroz, et al.
Publicado: (2025)
por: Yunusov, Sarfaroz, et al.
Publicado: (2025)
Susu Box or Piggy Bank: Assessing Cultural Commonsense Knowledge between Ghana and the U.S
por: Acquaye, Christabel, et al.
Publicado: (2024)
por: Acquaye, Christabel, et al.
Publicado: (2024)
Which Feedback Works for Whom? Differential Effects of LLM-Generated Feedback Elements Across Learner Profiles
por: Furuhashi, Momoka, et al.
Publicado: (2026)
por: Furuhashi, Momoka, et al.
Publicado: (2026)
Driving Generative Agents With Their Personality
por: Klinkert, Lawrence J., et al.
Publicado: (2024)
por: Klinkert, Lawrence J., et al.
Publicado: (2024)
Difficult for Whom? A Study of Japanese Lexical Complexity
por: Nohejl, Adam, et al.
Publicado: (2024)
por: Nohejl, Adam, et al.
Publicado: (2024)
We Politely Insist: Your LLM Must Learn the Persian Art of Taarof
por: Sadr, Nikta Gohari, et al.
Publicado: (2025)
por: Sadr, Nikta Gohari, et al.
Publicado: (2025)
Ejemplares similares
-
The Dog the Cat Chased Stumped the Model: Measuring When Language Models Abandon Structure for Shortcuts
por: Madhusudan, Sangmitra, et al.
Publicado: (2025) -
Fine-Tuned LLMs are "Time Capsules" for Tracking Societal Bias Through Books
por: Madhusudan, Sangmitra, et al.
Publicado: (2025) -
Which Words Matter Most in Zero-Shot Prompts?
por: Sadr, Nikta Gohari, et al.
Publicado: (2025) -
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
por: Morabito, Robert, et al.
Publicado: (2024) -
DART: Mitigating Harm Drift in Difference-Aware LLMs via Distill-Audit-Repair Training
por: Pan, Ziwen, et al.
Publicado: (2026)