Culturally-Aware Conversations: A Framework & Benchmark for LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Havaldar, Shreya, Rai, Sunny, Cho, Young-Min, Ungar, Lyle |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Building Knowledge-Guided Lexica to Model Cultural Variation
by: Havaldar, Shreya, et al.
Published: (2024)
by: Havaldar, Shreya, et al.
Published: (2024)
Comparing Styles across Languages: A Cross-Cultural Exploration of Politeness
by: Havaldar, Shreya, et al.
Published: (2023)
by: Havaldar, Shreya, et al.
Published: (2023)
Towards Style Alignment in Cross-Cultural Translation
by: Havaldar, Shreya, et al.
Published: (2025)
by: Havaldar, Shreya, et al.
Published: (2025)
Social Norms in Cinema: A Cross-Cultural Analysis of Shame, Pride and Prejudice
by: Rai, Sunny, et al.
Published: (2024)
by: Rai, Sunny, et al.
Published: (2024)
A Concise Agent is Less Expert: Revealing Side Effects of Using Style Features on Conversational Agents
by: Cho, Young-Min, et al.
Published: (2026)
by: Cho, Young-Min, et al.
Published: (2026)
Know Me, Respond to Me: Benchmarking LLMs for Dynamic User Profiling and Personalized Responses at Scale
by: Jiang, Bowen, et al.
Published: (2025)
by: Jiang, Bowen, et al.
Published: (2025)
Conceptors for Semantic Steering
by: Triantafyllopoulos, Ilias, et al.
Published: (2026)
by: Triantafyllopoulos, Ilias, et al.
Published: (2026)
Language-based Valence and Arousal Expressions between the United States and China: a Cross-Cultural Examination
by: Cho, Young-Min, et al.
Published: (2024)
by: Cho, Young-Min, et al.
Published: (2024)
Multilingual LLMs Are Not Multilingual Thinkers: Evidence from Hindi Analogy Evaluation
by: Gupta, Ashray, et al.
Published: (2025)
by: Gupta, Ashray, et al.
Published: (2025)
Cross-Cultural Differences in Mental Health Expressions on Social Media
by: Rai, Sunny, et al.
Published: (2024)
by: Rai, Sunny, et al.
Published: (2024)
Closing the Confidence-Faithfulness Gap in Large Language Models
by: Miao, Miranda Muqing, et al.
Published: (2026)
by: Miao, Miranda Muqing, et al.
Published: (2026)
Entailed Between the Lines: Incorporating Implication into NLI
by: Havaldar, Shreya, et al.
Published: (2025)
by: Havaldar, Shreya, et al.
Published: (2025)
Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation
by: Mukherjee, Arka, et al.
Published: (2025)
by: Mukherjee, Arka, et al.
Published: (2025)
T-FIX: Text-Based Explanations with Features Interpretable to eXperts
by: Havaldar, Shreya, et al.
Published: (2025)
by: Havaldar, Shreya, et al.
Published: (2025)
Adaptively profiling models with task elicitation
by: Brown, Davis, et al.
Published: (2025)
by: Brown, Davis, et al.
Published: (2025)
DiverseDialogue: A Methodology for Designing Chatbots with Human-Like Diversity
by: Lin, Xiaoyu, et al.
Published: (2024)
by: Lin, Xiaoyu, et al.
Published: (2024)
Benchmarking Machine Translation with Cultural Awareness
by: Yao, Binwei, et al.
Published: (2023)
by: Yao, Binwei, et al.
Published: (2023)
MAKIEval: A Multilingual Automatic WiKidata-based Framework for Cultural Awareness Evaluation for LLMs
by: Zhao, Raoyuan, et al.
Published: (2025)
by: Zhao, Raoyuan, et al.
Published: (2025)
Probabilistic Soundness Guarantees in LLM Reasoning Chains
by: You, Weiqiu, et al.
Published: (2025)
by: You, Weiqiu, et al.
Published: (2025)
Can LLMs Grasp Implicit Cultural Values? Benchmarking LLMs' Cultural Intelligence with CQ-Bench
by: Liu, Ziyi, et al.
Published: (2025)
by: Liu, Ziyi, et al.
Published: (2025)
Correctness-Optimized Residual Activation Lens (CORAL): Transferrable and Calibration-Aware Inference-Time Steering
by: Miao, Miranda Muqing, et al.
Published: (2026)
by: Miao, Miranda Muqing, et al.
Published: (2026)
Modeling Human Subjectivity in LLMs Using Explicit and Implicit Human Factors in Personas
by: Giorgi, Salvatore, et al.
Published: (2024)
by: Giorgi, Salvatore, et al.
Published: (2024)
RENOVI: A Benchmark Towards Remediating Norm Violations in Socio-Cultural Conversations
by: Zhan, Haolan, et al.
Published: (2024)
by: Zhan, Haolan, et al.
Published: (2024)
Probing Cultural Awareness in LLMs: A Case Study of Cross-Culture Aesthetic Stylistics
by: Wang, Jiashuo, et al.
Published: (2026)
by: Wang, Jiashuo, et al.
Published: (2026)
Camellia: Benchmarking Cultural Biases in LLMs for Asian Languages
by: Naous, Tarek, et al.
Published: (2025)
by: Naous, Tarek, et al.
Published: (2025)
CULEMO: Cultural Lenses on Emotion -- Benchmarking LLMs for Cross-Cultural Emotion Understanding
by: Belay, Tadesse Destaw, et al.
Published: (2025)
by: Belay, Tadesse Destaw, et al.
Published: (2025)
LingBench++: A Linguistically-Informed Benchmark and Reasoning Framework for Multi-Step and Cross-Cultural Inference with LLMs
by: Lian, Da-Chen, et al.
Published: (2025)
by: Lian, Da-Chen, et al.
Published: (2025)
mmJEE-Eval: A Bilingual Multimodal Benchmark for Evaluating Scientific Reasoning in Vision-Language Models
by: Mukherjee, Arka, et al.
Published: (2025)
by: Mukherjee, Arka, et al.
Published: (2025)
Do LLMs Understand Wine Descriptors Across Cultures? A Benchmark for Cultural Adaptations of Wine Reviews
by: Zou, Chenye, et al.
Published: (2025)
by: Zou, Chenye, et al.
Published: (2025)
Intent Detection in the Age of LLMs
by: Arora, Gaurav, et al.
Published: (2024)
by: Arora, Gaurav, et al.
Published: (2024)
Interactive Concept Learning for Uncovering Latent Themes in Large Text Collections
by: Pacheco, Maria Leonor, et al.
Published: (2023)
by: Pacheco, Maria Leonor, et al.
Published: (2023)
Vernacular? I Barely Know Her: Challenges with Style Control and Stereotyping
by: Aich, Ankit, et al.
Published: (2024)
by: Aich, Ankit, et al.
Published: (2024)
BLEnD: A Benchmark for LLMs on Everyday Knowledge in Diverse Cultures and Languages
by: Myung, Junho, et al.
Published: (2024)
by: Myung, Junho, et al.
Published: (2024)
CaMMT: Benchmarking Culturally Aware Multimodal Machine Translation
by: Villa-Cueva, Emilio, et al.
Published: (2025)
by: Villa-Cueva, Emilio, et al.
Published: (2025)
MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks
by: Ruan, Junhao, et al.
Published: (2026)
by: Ruan, Junhao, et al.
Published: (2026)
"Be My Cheese?": Cultural Nuance Benchmarking for Machine Translation in Multilingual LLMs
by: Van Doren, Madison, et al.
Published: (2026)
by: Van Doren, Madison, et al.
Published: (2026)
JUBAKU: An Adversarial Benchmark for Exposing Culturally Grounded Stereotypes in Japanese LLMs
by: Shiotani, Taihei, et al.
Published: (2026)
by: Shiotani, Taihei, et al.
Published: (2026)
FairMT-Bench: Benchmarking Fairness for Multi-turn Dialogue in Conversational LLMs
by: Fan, Zhiting, et al.
Published: (2024)
by: Fan, Zhiting, et al.
Published: (2024)
HKCanto-Eval: A Benchmark for Evaluating Cantonese Language Understanding and Cultural Comprehension in LLMs
by: Cheng, Tsz Chung, et al.
Published: (2025)
by: Cheng, Tsz Chung, et al.
Published: (2025)
Evaluating Cultural Awareness of LLMs for Yoruba, Malayalam, and English
by: Dawson, Fiifi, et al.
Published: (2024)
by: Dawson, Fiifi, et al.
Published: (2024)
Similar Items
-
Building Knowledge-Guided Lexica to Model Cultural Variation
by: Havaldar, Shreya, et al.
Published: (2024) -
Comparing Styles across Languages: A Cross-Cultural Exploration of Politeness
by: Havaldar, Shreya, et al.
Published: (2023) -
Towards Style Alignment in Cross-Cultural Translation
by: Havaldar, Shreya, et al.
Published: (2025) -
Social Norms in Cinema: A Cross-Cultural Analysis of Shame, Pride and Prejudice
by: Rai, Sunny, et al.
Published: (2024) -
A Concise Agent is Less Expert: Revealing Side Effects of Using Style Features on Conversational Agents
by: Cho, Young-Min, et al.
Published: (2026)