Multimodal Cultural Safety: Evaluation Framework and Alignment Strategies
Fuente:
arXiv
Salvato in:
| Autori principali: | Qiu, Haoyi, Huang, Kung-Hsiang, Zheng, Ruichen, Sun, Jiao, Peng, Nanyun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SafeWorld: Geo-Diverse Safety Alignment
di: Yin, Da, et al.
Pubblicazione: (2024)
di: Yin, Da, et al.
Pubblicazione: (2024)
AMRFact: Enhancing Summarization Factuality Evaluation with AMR-Driven Negative Samples Generation
di: Qiu, Haoyi, et al.
Pubblicazione: (2023)
di: Qiu, Haoyi, et al.
Pubblicazione: (2023)
MMPersuade: A Dataset and Evaluation Framework for Multimodal Persuasion
di: Qiu, Haoyi, et al.
Pubblicazione: (2025)
di: Qiu, Haoyi, et al.
Pubblicazione: (2025)
Evaluating Cultural and Social Awareness of LLM Web Agents
di: Qiu, Haoyi, et al.
Pubblicazione: (2024)
di: Qiu, Haoyi, et al.
Pubblicazione: (2024)
Evaluating Human Alignment and Model Faithfulness of LLM Rationale
di: Fayyaz, Mohsen, et al.
Pubblicazione: (2024)
di: Fayyaz, Mohsen, et al.
Pubblicazione: (2024)
VALOR-EVAL: Holistic Coverage and Faithfulness Evaluation of Large Vision-Language Models
di: Qiu, Haoyi, et al.
Pubblicazione: (2024)
di: Qiu, Haoyi, et al.
Pubblicazione: (2024)
GUI-KV: Efficient GUI Agents via KV Cache with Spatio-Temporal Awareness
di: Huang, Kung-Hsiang, et al.
Pubblicazione: (2025)
di: Huang, Kung-Hsiang, et al.
Pubblicazione: (2025)
SkillVerse : Assessing and Enhancing LLMs with Tree Evaluation
di: Tian, Yufei, et al.
Pubblicazione: (2025)
di: Tian, Yufei, et al.
Pubblicazione: (2025)
Decoupling Task-Solving and Output Formatting in LLM Generation
di: Deng, Haikang, et al.
Pubblicazione: (2025)
di: Deng, Haikang, et al.
Pubblicazione: (2025)
MMGR: Multi-Modal Generative Reasoning
di: Cai, Zefan, et al.
Pubblicazione: (2025)
di: Cai, Zefan, et al.
Pubblicazione: (2025)
From Preferences to Prejudice: The Role of Alignment Tuning in Shaping Social Bias in Video Diffusion Models
di: Cai, Zefan, et al.
Pubblicazione: (2025)
di: Cai, Zefan, et al.
Pubblicazione: (2025)
Model Extrapolation Expedites Alignment
di: Zheng, Chujie, et al.
Pubblicazione: (2024)
di: Zheng, Chujie, et al.
Pubblicazione: (2024)
Why Vision Language Models Struggle with Visual Arithmetic? Towards Enhanced Chart and Geometry Understanding
di: Huang, Kung-Hsiang, et al.
Pubblicazione: (2025)
di: Huang, Kung-Hsiang, et al.
Pubblicazione: (2025)
Are Akpans Trick or Treat: Unveiling Helpful Biases in Assistant Systems
di: Sun, Jiao, et al.
Pubblicazione: (2022)
di: Sun, Jiao, et al.
Pubblicazione: (2022)
MMA-ASIA: A Multilingual and Multimodal Alignment Framework for Culturally-Grounded Evaluation
di: Zheng, Weihua, et al.
Pubblicazione: (2025)
di: Zheng, Weihua, et al.
Pubblicazione: (2025)
Mind the Gesture: Evaluating AI Sensitivity to Culturally Offensive Non-Verbal Gestures
di: Yerukola, Akhila, et al.
Pubblicazione: (2025)
di: Yerukola, Akhila, et al.
Pubblicazione: (2025)
Adaptable Logical Control for Large Language Models
di: Zhang, Honghua, et al.
Pubblicazione: (2024)
di: Zhang, Honghua, et al.
Pubblicazione: (2024)
Think in Safety: Unveiling and Mitigating Safety Alignment Collapse in Multimodal Large Reasoning Model
di: Lou, Xinyue, et al.
Pubblicazione: (2025)
di: Lou, Xinyue, et al.
Pubblicazione: (2025)
From Pixels to Insights: A Survey on Automatic Chart Understanding in the Era of Large Foundation Models
di: Huang, Kung-Hsiang, et al.
Pubblicazione: (2024)
di: Huang, Kung-Hsiang, et al.
Pubblicazione: (2024)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
di: Vo, Truong, et al.
Pubblicazione: (2025)
di: Vo, Truong, et al.
Pubblicazione: (2025)
RLCD: Reinforcement Learning from Contrastive Distillation for Language Model Alignment
di: Yang, Kevin, et al.
Pubblicazione: (2023)
di: Yang, Kevin, et al.
Pubblicazione: (2023)
Multilingual Blending: LLM Safety Alignment Evaluation with Language Mixture
di: Song, Jiayang, et al.
Pubblicazione: (2024)
di: Song, Jiayang, et al.
Pubblicazione: (2024)
Assessing Socio-Cultural Alignment and Technical Safety of Sovereign LLMs
di: Chae, Kyubyung, et al.
Pubblicazione: (2025)
di: Chae, Kyubyung, et al.
Pubblicazione: (2025)
Utilizing Large Language Models for Event Deconstruction to Enhance Multimodal Aspect-Based Sentiment Analysis
di: Huang, Xiaoyong, et al.
Pubblicazione: (2024)
di: Huang, Xiaoyong, et al.
Pubblicazione: (2024)
Rethinking Creativity Evaluation: A Critical Analysis of Existing Creativity Evaluations
di: Lu, Li-Chun, et al.
Pubblicazione: (2025)
di: Lu, Li-Chun, et al.
Pubblicazione: (2025)
ManiTweet: A New Benchmark for Identifying Manipulation of News on Social Media
di: Huang, Kung-Hsiang, et al.
Pubblicazione: (2023)
di: Huang, Kung-Hsiang, et al.
Pubblicazione: (2023)
New Job, New Gender? Measuring the Social Bias in Image Generation Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
Understanding Multimodal Procedural Knowledge by Sequencing Multimodal Instructional Manuals
di: Wu, Te-Lin, et al.
Pubblicazione: (2021)
di: Wu, Te-Lin, et al.
Pubblicazione: (2021)
Open-Domain Text Evaluation via Contrastive Distribution Methods
di: Lu, Sidi, et al.
Pubblicazione: (2023)
di: Lu, Sidi, et al.
Pubblicazione: (2023)
Scientific Discourse Tagging for Evidence Extraction
di: Li, Xiangci, et al.
Pubblicazione: (2019)
di: Li, Xiangci, et al.
Pubblicazione: (2019)
A Paragraph-level Multi-task Learning Model for Scientific Fact-Verification
di: Li, Xiangci, et al.
Pubblicazione: (2020)
di: Li, Xiangci, et al.
Pubblicazione: (2020)
STAR: Boosting Low-Resource Information Extraction by Structure-to-Text Data Generation with Large Language Models
di: Ma, Mingyu Derek, et al.
Pubblicazione: (2023)
di: Ma, Mingyu Derek, et al.
Pubblicazione: (2023)
NewsEdits 2.0: Learning the Intentions Behind Updating News
di: Spangher, Alexander, et al.
Pubblicazione: (2024)
di: Spangher, Alexander, et al.
Pubblicazione: (2024)
Evaluating Cultural Knowledge Processing in Large Language Models: A Cognitive Benchmarking Framework Integrating Retrieval-Augmented Generation
di: Lee, Hung-Shin, et al.
Pubblicazione: (2025)
di: Lee, Hung-Shin, et al.
Pubblicazione: (2025)
CoKe: Customizable Fine-Grained Story Evaluation via Chain-of-Keyword Rationalization
di: Joshi, Brihi, et al.
Pubblicazione: (2025)
di: Joshi, Brihi, et al.
Pubblicazione: (2025)
MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models
di: Hu, Wenbo, et al.
Pubblicazione: (2024)
di: Hu, Wenbo, et al.
Pubblicazione: (2024)
VLDBench Evaluating Multimodal Disinformation with Regulatory Alignment
di: Raza, Shaina, et al.
Pubblicazione: (2025)
di: Raza, Shaina, et al.
Pubblicazione: (2025)
PhonologyBench: Evaluating Phonological Skills of Large Language Models
di: Suvarna, Ashima, et al.
Pubblicazione: (2024)
di: Suvarna, Ashima, et al.
Pubblicazione: (2024)
MSR-Align: Policy-Grounded Multimodal Alignment for Safety-Aware Reasoning in Vision-Language Models
di: Xia, Yinan, et al.
Pubblicazione: (2025)
di: Xia, Yinan, et al.
Pubblicazione: (2025)
Bridging the Multilingual Safety Divide: Efficient, Culturally-Aware Alignment for Global South Languages
di: Banerjee, Somnath, et al.
Pubblicazione: (2026)
di: Banerjee, Somnath, et al.
Pubblicazione: (2026)
Documenti analoghi
-
SafeWorld: Geo-Diverse Safety Alignment
di: Yin, Da, et al.
Pubblicazione: (2024) -
AMRFact: Enhancing Summarization Factuality Evaluation with AMR-Driven Negative Samples Generation
di: Qiu, Haoyi, et al.
Pubblicazione: (2023) -
MMPersuade: A Dataset and Evaluation Framework for Multimodal Persuasion
di: Qiu, Haoyi, et al.
Pubblicazione: (2025) -
Evaluating Cultural and Social Awareness of LLM Web Agents
di: Qiu, Haoyi, et al.
Pubblicazione: (2024) -
Evaluating Human Alignment and Model Faithfulness of LLM Rationale
di: Fayyaz, Mohsen, et al.
Pubblicazione: (2024)