Beyond Western Politics: Cross-Cultural Benchmarks for Evaluating Partisan Associations in LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Divyanshu, Gupta, Ishita, Birur, Nitin Aravind, Baswa, Tanay, Agarwal, Sahil, Harshangi, Prashanth |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SocioEval: A Template-Based Framework for Evaluating Socioeconomic Status Bias in Foundation Models
by: Kumar, Divyanshu, et al.
Published: (2026)
by: Kumar, Divyanshu, et al.
Published: (2026)
Redirected, Not Removed: Task-Dependent Stereotyping Reveals the Limits of LLM Alignments
by: Kumar, Divyanshu, et al.
Published: (2026)
by: Kumar, Divyanshu, et al.
Published: (2026)
VERA: Validation and Enhancement for Retrieval Augmented systems
by: Birur, Nitin Aravind, et al.
Published: (2024)
by: Birur, Nitin Aravind, et al.
Published: (2024)
No Free Lunch with Guardrails
by: Kumar, Divyanshu, et al.
Published: (2025)
by: Kumar, Divyanshu, et al.
Published: (2025)
Quantifying CBRN Risk in Frontier Models
by: Kumar, Divyanshu, et al.
Published: (2025)
by: Kumar, Divyanshu, et al.
Published: (2025)
SAGE-RT: Synthetic Alignment data Generation for Safety Evaluation and Red Teaming
by: Kumar, Anurakt, et al.
Published: (2024)
by: Kumar, Anurakt, et al.
Published: (2024)
Beyond Text: Multimodal Jailbreaking of Vision-Language and Audio Models through Perceptually Simple Transformations
by: Kumar, Divyanshu, et al.
Published: (2025)
by: Kumar, Divyanshu, et al.
Published: (2025)
Investigating Implicit Bias in Large Language Models: A Large-Scale Study of Over 50 LLMs
by: Kumar, Divyanshu, et al.
Published: (2024)
by: Kumar, Divyanshu, et al.
Published: (2024)
Fine-Tuning, Quantization, and LLMs: Navigating Unintended Outcomes
by: Kumar, Divyanshu, et al.
Published: (2024)
by: Kumar, Divyanshu, et al.
Published: (2024)
The Role of Partisan Culture in Mental Health Language Online
by: Pendse, Sachin R., et al.
Published: (2025)
by: Pendse, Sachin R., et al.
Published: (2025)
Provocation on Expertise in Social Impact Evaluations of Generative AI (and Beyond)
by: Kahn, Zoe, et al.
Published: (2024)
by: Kahn, Zoe, et al.
Published: (2024)
'Since Lawyers are Males..': Examining Implicit Gender Bias in Hindi Language Generation by LLMs
by: Joshi, Ishika, et al.
Published: (2024)
by: Joshi, Ishika, et al.
Published: (2024)
Partisan Fact-Checkers' Warnings Can Effectively Correct Individuals' Misbeliefs About Political Misinformation
by: Lee, Sian, et al.
Published: (2025)
by: Lee, Sian, et al.
Published: (2025)
AI Suggestions Homogenize Writing Toward Western Styles and Diminish Cultural Nuances
by: Agarwal, Dhruv, et al.
Published: (2024)
by: Agarwal, Dhruv, et al.
Published: (2024)
Bridging Dictionary: AI-Generated Dictionary of Partisan Language Use
by: Jiang, Hang, et al.
Published: (2024)
by: Jiang, Hang, et al.
Published: (2024)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
by: Vo, Truong, et al.
Published: (2025)
by: Vo, Truong, et al.
Published: (2025)
Beyond Categories of Caste: Examining Caste Bias and Morality in Text-to-Image AI Models
by: Singh, Divyanshu Kumar, et al.
Published: (2026)
by: Singh, Divyanshu Kumar, et al.
Published: (2026)
"Which LLM should I use?": Evaluating LLMs for tasks performed by Undergraduate Computer Science Students
by: Agarwal, Vibhor, et al.
Published: (2024)
by: Agarwal, Vibhor, et al.
Published: (2024)
Disability Across Cultures: A Human-Centered Audit of Ableism in Western and Indic LLMs
by: Phutane, Mahika, et al.
Published: (2025)
by: Phutane, Mahika, et al.
Published: (2025)
Evaluation of LLMs Biases Towards Elite Universities: A Persona-Based Exploration
by: Gupta, Shailja, et al.
Published: (2024)
by: Gupta, Shailja, et al.
Published: (2024)
Beyond Benchmarks: How Users Evaluate AI Chat Assistants
by: Awan, Moiz Sadiq, et al.
Published: (2026)
by: Awan, Moiz Sadiq, et al.
Published: (2026)
"I Would Never Trust Anything Western": Kumu (Educator) Perspectives on Use of LLMs for Culturally Revitalizing CS Education in Hawaiian Schools
by: Mhasakar, Manas, et al.
Published: (2025)
by: Mhasakar, Manas, et al.
Published: (2025)
When Benchmarks Talk: Re-Evaluating Code LLMs with Interactive Feedback
by: Pan, Jane, et al.
Published: (2025)
by: Pan, Jane, et al.
Published: (2025)
Questionnaires for Everyone: Streamlining Cross-Cultural Questionnaire Adaptation with GPT-Based Translation Quality Evaluation
by: Haavisto, Otso, et al.
Published: (2024)
by: Haavisto, Otso, et al.
Published: (2024)
User-Centered Design with AI in the Loop: A Case Study of Rapid User Interface Prototyping with "Vibe Coding"
by: Li, Tianyi, et al.
Published: (2025)
by: Li, Tianyi, et al.
Published: (2025)
Designing Culturally Aligned AI Systems For Social Good in Non-Western Contexts
by: Dennison, Deepak Varuvel, et al.
Published: (2025)
by: Dennison, Deepak Varuvel, et al.
Published: (2025)
Cross-Cultural Communication in the Digital Age: An Analysis of Cultural Representation and Inclusivity in Emojis
by: Li, Lingfeng, et al.
Published: (2024)
by: Li, Lingfeng, et al.
Published: (2024)
SparseEMG: Computational Design of Sparse EMG Layouts for Sensing Gestures
by: Kumar, Anand, et al.
Published: (2025)
by: Kumar, Anand, et al.
Published: (2025)
Accessibility evaluation of major assistive mobile applications available for the visually impaired
by: Bhagat, Saidarshan, et al.
Published: (2024)
by: Bhagat, Saidarshan, et al.
Published: (2024)
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
by: Badawi, Abeer, et al.
Published: (2025)
by: Badawi, Abeer, et al.
Published: (2025)
CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs
by: Liu, Hongtao, et al.
Published: (2025)
by: Liu, Hongtao, et al.
Published: (2025)
SportsBuddy: Designing and Evaluating an AI-Powered Sports Video Storytelling Tool Through Real-World Deployment
by: Lin, Tica, et al.
Published: (2025)
by: Lin, Tica, et al.
Published: (2025)
Beyond Anthropomorphism: a Spectrum of Interface Metaphors for LLMs
by: So, Jianna, et al.
Published: (2026)
by: So, Jianna, et al.
Published: (2026)
Cross-Cultural Validation of Partner Models for Voice User Interfaces
by: Seaborn, Katie, et al.
Published: (2024)
by: Seaborn, Katie, et al.
Published: (2024)
AI Eyes on the Road: Cross-Cultural Perspectives on Traffic Surveillance
by: Wang, Ziming, et al.
Published: (2025)
by: Wang, Ziming, et al.
Published: (2025)
Play Across Boundaries: Exploring Cross-Cultural Maldaimonic Game Experiences
by: Seaborn, Katie, et al.
Published: (2024)
by: Seaborn, Katie, et al.
Published: (2024)
Sketch2Colab: Sketch-Conditioned Multi-Human Animation via Controllable Flow Distillation
by: Daiya, Divyanshu, et al.
Published: (2026)
by: Daiya, Divyanshu, et al.
Published: (2026)
Benchmarking PDF Accessibility Evaluation A Dataset and Framework for Assessing Automated and LLM-Based Approaches for Accessibility Testing
by: Kumar, Anukriti, et al.
Published: (2025)
by: Kumar, Anukriti, et al.
Published: (2025)
Beyond Final Answers: Evaluating Large Language Models for Math Tutoring
by: Gupta, Adit, et al.
Published: (2025)
by: Gupta, Adit, et al.
Published: (2025)
Toward Cultural Interpretability: A Linguistic Anthropological Framework for Describing and Evaluating Large Language Models (LLMs)
by: Jones, Graham M., et al.
Published: (2024)
by: Jones, Graham M., et al.
Published: (2024)
Similar Items
-
SocioEval: A Template-Based Framework for Evaluating Socioeconomic Status Bias in Foundation Models
by: Kumar, Divyanshu, et al.
Published: (2026) -
Redirected, Not Removed: Task-Dependent Stereotyping Reveals the Limits of LLM Alignments
by: Kumar, Divyanshu, et al.
Published: (2026) -
VERA: Validation and Enhancement for Retrieval Augmented systems
by: Birur, Nitin Aravind, et al.
Published: (2024) -
No Free Lunch with Guardrails
by: Kumar, Divyanshu, et al.
Published: (2025) -
Quantifying CBRN Risk in Frontier Models
by: Kumar, Divyanshu, et al.
Published: (2025)