Robust Adaptation of Large Multimodal Models for Retrieval Augmented Hateful Meme Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Mei, Jingbiao, Chen, Jinghong, Yang, Guangyu, Lin, Weizhe, Byrne, Bill |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Improving Hateful Meme Detection through Retrieval-Guided Contrastive Learning
di: Mei, Jingbiao, et al.
Pubblicazione: (2023)
di: Mei, Jingbiao, et al.
Pubblicazione: (2023)
Retrieval-Augmented Defense: Adaptive and Controllable Jailbreak Prevention for Large Language Models
di: Yang, Guangyu, et al.
Pubblicazione: (2025)
di: Yang, Guangyu, et al.
Pubblicazione: (2025)
According to Me: Long-Term Personalized Referential Memory QA
di: Mei, Jingbiao, et al.
Pubblicazione: (2026)
di: Mei, Jingbiao, et al.
Pubblicazione: (2026)
Few-Shot VQA with Frozen LLMs: A Tale of Two Approaches
di: Sterner, Igor, et al.
Pubblicazione: (2024)
di: Sterner, Igor, et al.
Pubblicazione: (2024)
BERAG: Bayesian Ensemble Retrieval-Augmented Generation for Knowledge-based Visual Question Answering
di: Chen, Jinghong, et al.
Pubblicazione: (2026)
di: Chen, Jinghong, et al.
Pubblicazione: (2026)
PreFLMR: Scaling Up Fine-Grained Late-Interaction Multi-modal Retrievers
di: Lin, Weizhe, et al.
Pubblicazione: (2024)
di: Lin, Weizhe, et al.
Pubblicazione: (2024)
On Extending Direct Preference Optimization to Accommodate Ties
di: Chen, Jinghong, et al.
Pubblicazione: (2024)
di: Chen, Jinghong, et al.
Pubblicazione: (2024)
ExPO-HM: Learning to Explain-then-Detect for Hateful Meme Detection
di: Mei, Jingbiao, et al.
Pubblicazione: (2025)
di: Mei, Jingbiao, et al.
Pubblicazione: (2025)
Evolver: Chain-of-Evolution Prompting to Boost Large Multimodal Models for Hateful Meme Detection
di: Huang, Jinfa, et al.
Pubblicazione: (2024)
di: Huang, Jinfa, et al.
Pubblicazione: (2024)
Control-DAG: Constrained Decoding for Non-Autoregressive Directed Acyclic T5 using Weighted Finite State Automata
di: Chen, Jinghong, et al.
Pubblicazione: (2024)
di: Chen, Jinghong, et al.
Pubblicazione: (2024)
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models
di: Van, Minh-Hao, et al.
Pubblicazione: (2025)
di: Van, Minh-Hao, et al.
Pubblicazione: (2025)
Modularized Networks for Few-shot Hateful Meme Detection
di: Cao, Rui, et al.
Pubblicazione: (2024)
di: Cao, Rui, et al.
Pubblicazione: (2024)
See, Explain, and Intervene: A Few-Shot Multimodal Agent Framework for Hateful Meme Moderation
di: Rizwan, Naquee, et al.
Pubblicazione: (2026)
di: Rizwan, Naquee, et al.
Pubblicazione: (2026)
Direct Preference Optimization for Neural Machine Translation with Minimum Bayes Risk Decoding
di: Yang, Guangyu, et al.
Pubblicazione: (2023)
di: Yang, Guangyu, et al.
Pubblicazione: (2023)
TRACE: Textual Relevance Augmentation and Contextual Encoding for Multimodal Hate Detection
di: Koushik, Girish A., et al.
Pubblicazione: (2025)
di: Koushik, Girish A., et al.
Pubblicazione: (2025)
MemeIntel: Explainable Detection of Propagandistic and Hateful Memes
di: Kmainasi, Mohamed Bayan, et al.
Pubblicazione: (2025)
di: Kmainasi, Mohamed Bayan, et al.
Pubblicazione: (2025)
Detecting Offensive Memes with Social Biases in Singapore Context Using Multimodal Large Language Models
di: Yuxuan, Cao, et al.
Pubblicazione: (2025)
di: Yuxuan, Cao, et al.
Pubblicazione: (2025)
FBHM: Functional Benchmarking and Steering of VLMs for Hateful Meme Detection
di: Bhaskar, Paramananda, et al.
Pubblicazione: (2026)
di: Bhaskar, Paramananda, et al.
Pubblicazione: (2026)
Exploring the Limits of Zero Shot Vision Language Models for Hate Meme Detection: The Vulnerabilities and their Interpretations
di: Rizwan, Naquee, et al.
Pubblicazione: (2024)
di: Rizwan, Naquee, et al.
Pubblicazione: (2024)
GatedCLIP: Gated Multimodal Fusion for Hateful Memes Detection
di: Guo, Yingying, et al.
Pubblicazione: (2026)
di: Guo, Yingying, et al.
Pubblicazione: (2026)
MemeMind: A Large-Scale Multimodal Dataset with Chain-of-Thought Reasoning for Harmful Meme Detection
di: Gu, Hexiang, et al.
Pubblicazione: (2025)
di: Gu, Hexiang, et al.
Pubblicazione: (2025)
Text or Image? What is More Important in Cross-Domain Generalization Capabilities of Hate Meme Detection Models?
di: Aggarwal, Piush, et al.
Pubblicazione: (2024)
di: Aggarwal, Piush, et al.
Pubblicazione: (2024)
Improving Multimodal Hateful Meme Detection Exploiting LMM-Generated Knowledge
di: Tzelepi, Maria, et al.
Pubblicazione: (2025)
di: Tzelepi, Maria, et al.
Pubblicazione: (2025)
MER-Bench: A Comprehensive Benchmark for Multimodal Meme Reappraisal
di: Nie, Yiqi, et al.
Pubblicazione: (2026)
di: Nie, Yiqi, et al.
Pubblicazione: (2026)
OSPC: Detecting Harmful Memes with Large Language Model as a Catalyst
di: Cao, Jingtao, et al.
Pubblicazione: (2024)
di: Cao, Jingtao, et al.
Pubblicazione: (2024)
Understanding Retrieval Robustness for Retrieval-Augmented Image Captioning
di: Li, Wenyan, et al.
Pubblicazione: (2024)
di: Li, Wenyan, et al.
Pubblicazione: (2024)
DeHate: A Stable Diffusion-based Multimodal Approach to Mitigate Hate Speech in Images
di: Dalal, Dwip, et al.
Pubblicazione: (2025)
di: Dalal, Dwip, et al.
Pubblicazione: (2025)
Multimodal Retrieval-Augmented Generation with Large Language Models for Medical VQA
di: Karim, A H M Rezaul, et al.
Pubblicazione: (2025)
di: Karim, A H M Rezaul, et al.
Pubblicazione: (2025)
Beyond Meme Templates: Limitations of Visual Similarity Measures in Meme Matching
di: Hazman, Muzhaffar, et al.
Pubblicazione: (2025)
di: Hazman, Muzhaffar, et al.
Pubblicazione: (2025)
RAVENEA: A Benchmark for Multimodal Retrieval-Augmented Visual Culture Understanding
di: Li, Jiaang, et al.
Pubblicazione: (2025)
di: Li, Jiaang, et al.
Pubblicazione: (2025)
Benchmarking Deflection and Hallucination in Large Vision-Language Models
di: Moratelli, Nicholas, et al.
Pubblicazione: (2026)
di: Moratelli, Nicholas, et al.
Pubblicazione: (2026)
VimRAG: Navigating Massive Visual Context in Retrieval-Augmented Generation via Multimodal Memory Graph
di: Wang, Qiuchen, et al.
Pubblicazione: (2026)
di: Wang, Qiuchen, et al.
Pubblicazione: (2026)
MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models
di: Hu, Wenbo, et al.
Pubblicazione: (2024)
di: Hu, Wenbo, et al.
Pubblicazione: (2024)
Scaling Beyond Context: A Survey of Multimodal Retrieval-Augmented Generation for Document Understanding
di: Gao, Sensen, et al.
Pubblicazione: (2025)
di: Gao, Sensen, et al.
Pubblicazione: (2025)
Reminding Multimodal Large Language Models of Object-aware Knowledge with Retrieved Tags
di: Qi, Daiqing, et al.
Pubblicazione: (2024)
di: Qi, Daiqing, et al.
Pubblicazione: (2024)
Specializing Large Models for Oracle Bone Script Interpretation via Component-Grounded Multimodal Knowledge Augmentation
di: Zhang, Jianing, et al.
Pubblicazione: (2026)
di: Zhang, Jianing, et al.
Pubblicazione: (2026)
Labels or Input? Rethinking Augmentation in Multimodal Hate Detection
di: Singh, Sahajpreet, et al.
Pubblicazione: (2025)
di: Singh, Sahajpreet, et al.
Pubblicazione: (2025)
RAP: Retrieval-Augmented Personalization for Multimodal Large Language Models
di: Hao, Haoran, et al.
Pubblicazione: (2024)
di: Hao, Haoran, et al.
Pubblicazione: (2024)
Alleviating Hallucination in Large Vision-Language Models with Active Retrieval Augmentation
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
AesBench: An Expert Benchmark for Multimodal Large Language Models on Image Aesthetics Perception
di: Huang, Yipo, et al.
Pubblicazione: (2024)
di: Huang, Yipo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Improving Hateful Meme Detection through Retrieval-Guided Contrastive Learning
di: Mei, Jingbiao, et al.
Pubblicazione: (2023) -
Retrieval-Augmented Defense: Adaptive and Controllable Jailbreak Prevention for Large Language Models
di: Yang, Guangyu, et al.
Pubblicazione: (2025) -
According to Me: Long-Term Personalized Referential Memory QA
di: Mei, Jingbiao, et al.
Pubblicazione: (2026) -
Few-Shot VQA with Frozen LLMs: A Tale of Two Approaches
di: Sterner, Igor, et al.
Pubblicazione: (2024) -
BERAG: Bayesian Ensemble Retrieval-Augmented Generation for Knowledge-based Visual Question Answering
di: Chen, Jinghong, et al.
Pubblicazione: (2026)