When Tom Eats Kimchi: Evaluating Cultural Bias of Multimodal Large Language Models in Cultural Mixture Contexts
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Jun Seong, Thu, Kyaw Ye, Ismayilzada, Javad, Park, Junyeong, Kim, Eunsu, Ahmad, Huzama, An, Na Min, Thorne, James, Oh, Alice |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Diffusion Models Through a Global Lens: Are They Culturally Inclusive?
di: Bayramli, Zahra, et al.
Pubblicazione: (2025)
di: Bayramli, Zahra, et al.
Pubblicazione: (2025)
CLIcK: A Benchmark Dataset of Cultural and Linguistic Intelligence in Korean
di: Kim, Eunsu, et al.
Pubblicazione: (2024)
di: Kim, Eunsu, et al.
Pubblicazione: (2024)
World in a Frame: Understanding Culture Mixing as a New Challenge for Vision-Language Models
di: Kim, Eunsu, et al.
Pubblicazione: (2025)
di: Kim, Eunsu, et al.
Pubblicazione: (2025)
Are they lovers or friends? Evaluating LLMs' Social Reasoning in English and Korean Dialogues
di: Kim, Eunsu, et al.
Pubblicazione: (2025)
di: Kim, Eunsu, et al.
Pubblicazione: (2025)
Flex-TravelPlanner: A Benchmark for Flexible Planning with Language Agents
di: Oh, Juhyun, et al.
Pubblicazione: (2025)
di: Oh, Juhyun, et al.
Pubblicazione: (2025)
LLM-C3MOD: A Human-LLM Collaborative System for Cross-Cultural Hate Speech Moderation
di: Park, Junyeong, et al.
Pubblicazione: (2025)
di: Park, Junyeong, et al.
Pubblicazione: (2025)
Multi-FAct: Assessing Factuality of Multilingual LLMs using FActScore
di: Shafayat, Sheikh, et al.
Pubblicazione: (2024)
di: Shafayat, Sheikh, et al.
Pubblicazione: (2024)
The Generative AI Paradox on Evaluation: What It Can Solve, It May Not Evaluate
di: Oh, Juhyun, et al.
Pubblicazione: (2024)
di: Oh, Juhyun, et al.
Pubblicazione: (2024)
Designing “Korean” Kimchi: Speculative Configuration of Distance and Commodity Value in the Chinese Kimchi Industry
di: Heangjin Park
Pubblicazione: (2025)
di: Heangjin Park
Pubblicazione: (2025)
BLUCK: A Benchmark Dataset for Bengali Linguistic Understanding and Cultural Knowledge
di: Kabir, Daeen, et al.
Pubblicazione: (2025)
di: Kabir, Daeen, et al.
Pubblicazione: (2025)
Physicochemical Property Analyses of Deep‐Frozen Kimchi Cabbage during Long‐Term Storage
di: Dong Hyeon Park, et al.
Pubblicazione: (2024)
di: Dong Hyeon Park, et al.
Pubblicazione: (2024)
JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors
di: Jin, Jiho, et al.
Pubblicazione: (2026)
di: Jin, Jiho, et al.
Pubblicazione: (2026)
Spotting Out-of-Character Behavior: Atomic-Level Evaluation of Persona Fidelity in Open-Ended Generation
di: Shin, Jisu, et al.
Pubblicazione: (2025)
di: Shin, Jisu, et al.
Pubblicazione: (2025)
LLM-as-an-Interviewer: Beyond Static Testing Through Dynamic LLM Evaluation
di: Kim, Eunsu, et al.
Pubblicazione: (2024)
di: Kim, Eunsu, et al.
Pubblicazione: (2024)
QEVA: A Reference-Free Evaluation Metric for Narrative Video Summarization with Multimodal Question Answering
di: Jung, Woojun, et al.
Pubblicazione: (2026)
di: Jung, Woojun, et al.
Pubblicazione: (2026)
Survey of Cultural Awareness in Language Models: Text and Beyond
di: Pawar, Siddhesh, et al.
Pubblicazione: (2024)
di: Pawar, Siddhesh, et al.
Pubblicazione: (2024)
LoCar: Localization-Aware Evaluation of In-Vehicle Assistants through Fine-Grained Sociolinguistic Control
di: Jeong, Seogyeong, et al.
Pubblicazione: (2026)
di: Jeong, Seogyeong, et al.
Pubblicazione: (2026)
Uncovering Factor Level Preferences to Improve Human-Model Alignment
di: Oh, Juhyun, et al.
Pubblicazione: (2024)
di: Oh, Juhyun, et al.
Pubblicazione: (2024)
A Dual-Layered Evaluation of Geopolitical and Cultural Bias in LLMs
di: Kim, Sean, et al.
Pubblicazione: (2025)
di: Kim, Sean, et al.
Pubblicazione: (2025)
Spicy or Not? Exploring Kimchi's Spiciness Perception Across Spicy Food Tolerant and Sensitive Groups
di: Seo‐yeong Chon, et al.
Pubblicazione: (2025)
di: Seo‐yeong Chon, et al.
Pubblicazione: (2025)
Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models
di: Seo, Huichan, et al.
Pubblicazione: (2025)
di: Seo, Huichan, et al.
Pubblicazione: (2025)
Culture is Everywhere: A Call for Intentionally Cultural Evaluation
di: Oh, Juhyun, et al.
Pubblicazione: (2025)
di: Oh, Juhyun, et al.
Pubblicazione: (2025)
MUG-Eval: A Proxy Evaluation Framework for Multilingual Generation Capabilities in Any Language
di: Song, Seyoung, et al.
Pubblicazione: (2025)
di: Song, Seyoung, et al.
Pubblicazione: (2025)
RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity
di: Shin, Jisu, et al.
Pubblicazione: (2025)
di: Shin, Jisu, et al.
Pubblicazione: (2025)
BenchHub: A Unified Benchmark Suite for Holistic and Customizable LLM Evaluation
di: Kim, Eunsu, et al.
Pubblicazione: (2025)
di: Kim, Eunsu, et al.
Pubblicazione: (2025)
Understanding EFL Learners' Code-Switching and Teachers' Pedagogical Approaches in LLM-Supported Speaking Practice
di: Park, Junyeong, et al.
Pubblicazione: (2025)
di: Park, Junyeong, et al.
Pubblicazione: (2025)
AI Should Sense Better, Not Just Scale Bigger: Adaptive Sensing as a Paradigm Shift
di: Baek, Eunsu, et al.
Pubblicazione: (2025)
di: Baek, Eunsu, et al.
Pubblicazione: (2025)
Unexplored Faces of Robustness and Out-of-Distribution: Covariate Shifts in Environment and Sensor Domains
di: Baek, Eunsu, et al.
Pubblicazione: (2024)
di: Baek, Eunsu, et al.
Pubblicazione: (2024)
MrSteve: Instruction-Following Agents in Minecraft with What-Where-When Memory
di: Park, Junyeong, et al.
Pubblicazione: (2024)
di: Park, Junyeong, et al.
Pubblicazione: (2024)
What, When, and Where America Eats
di: Sloan, E
Pubblicazione: (2010)
di: Sloan, E
Pubblicazione: (2010)
What, When, and Where America Eats
di: Sloan, E. A
Pubblicazione: (2012)
di: Sloan, E. A
Pubblicazione: (2012)
Context Filtering with Reward Modeling in Question Answering
di: Kim, Sangryul, et al.
Pubblicazione: (2024)
di: Kim, Sangryul, et al.
Pubblicazione: (2024)
To Eat or Not to Eat
di: Altmann, Peter, et al.
Pubblicazione: (2024)
di: Altmann, Peter, et al.
Pubblicazione: (2024)
On properness of moduli stacks of $D^{\times}$-shtukas over ramified legs
di: Choi, Yong-Gyu, et al.
Pubblicazione: (2025)
di: Choi, Yong-Gyu, et al.
Pubblicazione: (2025)
68‐2: A New PWM Micro‐LED Pixel Circuit Using LTPO TFTs with Threshold Voltage and IR‐Drop Compensations
di: Junyeong Kim, et al.
Pubblicazione: (2024)
di: Junyeong Kim, et al.
Pubblicazione: (2024)
When Will It Fail?: Anomaly to Prompt for Forecasting Future Anomalies in Time Series
di: Park, Min-Yeong, et al.
Pubblicazione: (2025)
di: Park, Min-Yeong, et al.
Pubblicazione: (2025)
I0T: Embedding Standardization Method Towards Zero Modality Gap
di: An, Na Min, et al.
Pubblicazione: (2024)
di: An, Na Min, et al.
Pubblicazione: (2024)
Influence of carcass mass on decomposition rate: A medico‐legal entomology perspective
di: Hyeon‐Seok Oh, et al.
Pubblicazione: (2024)
di: Hyeon‐Seok Oh, et al.
Pubblicazione: (2024)
Enhancing Robustness of Retrieval-Augmented Language Models with In-Context Learning
di: Park, Seong-Il, et al.
Pubblicazione: (2024)
di: Park, Seong-Il, et al.
Pubblicazione: (2024)
Flavor Lexicon and Testing for Kimchi Available in the United States
di: Jeehyun Lee, et al.
Pubblicazione: (2025)
di: Jeehyun Lee, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Diffusion Models Through a Global Lens: Are They Culturally Inclusive?
di: Bayramli, Zahra, et al.
Pubblicazione: (2025) -
CLIcK: A Benchmark Dataset of Cultural and Linguistic Intelligence in Korean
di: Kim, Eunsu, et al.
Pubblicazione: (2024) -
World in a Frame: Understanding Culture Mixing as a New Challenge for Vision-Language Models
di: Kim, Eunsu, et al.
Pubblicazione: (2025) -
Are they lovers or friends? Evaluating LLMs' Social Reasoning in English and Korean Dialogues
di: Kim, Eunsu, et al.
Pubblicazione: (2025) -
Flex-TravelPlanner: A Benchmark for Flexible Planning with Language Agents
di: Oh, Juhyun, et al.
Pubblicazione: (2025)