I Came, I Saw, I Explained: Benchmarking Multimodal LLMs on Figurative Meaning in Memes
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhou, Shijia, Mohammad, Saif M., Plank, Barbara, Frassinelli, Diego |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Do LLMs Give Psychometrically Plausible Responses in Educational Assessments?
por: Säuberli, Andreas, et al.
Publicado: (2025)
por: Säuberli, Andreas, et al.
Publicado: (2025)
Resource-Lean Lexicon Induction for German Dialects
por: Litschko, Robert, et al.
Publicado: (2026)
por: Litschko, Robert, et al.
Publicado: (2026)
What Media Frames Reveal About Stance: A Dataset and Study about Memes in Climate Change Discourse
por: Zhou, Shijia, et al.
Publicado: (2025)
por: Zhou, Shijia, et al.
Publicado: (2025)
Mind the Uncertainty in Human Disagreement: Evaluating Discrepancies between Model Predictions and Human Responses in VQA
por: Lan, Jian, et al.
Publicado: (2024)
por: Lan, Jian, et al.
Publicado: (2024)
Controlling Reading Ease with Gaze-Guided Text Generation
por: Säuberli, Andreas, et al.
Publicado: (2026)
por: Säuberli, Andreas, et al.
Publicado: (2026)
Make Every Letter Count: Building Dialect Variation Dictionaries from Monolingual Corpora
por: Litschko, Robert, et al.
Publicado: (2025)
por: Litschko, Robert, et al.
Publicado: (2025)
To Know or Not To Know? Analyzing Self-Consistency of Large Language Models under Ambiguity
por: Sedova, Anastasiia, et al.
Publicado: (2024)
por: Sedova, Anastasiia, et al.
Publicado: (2024)
CLIMATELI: Evaluating Entity Linking on Climate Change Data
por: Zhou, Shijia, et al.
Publicado: (2024)
por: Zhou, Shijia, et al.
Publicado: (2024)
MaiNLP at SemEval-2024 Task 1: Analyzing Source Language Selection in Cross-Lingual Textual Relatedness
por: Zhou, Shijia, et al.
Publicado: (2024)
por: Zhou, Shijia, et al.
Publicado: (2024)
Generalizable Sarcasm Detection Is Just Around The Corner, Of Course!
por: Jang, Hyewon, et al.
Publicado: (2024)
por: Jang, Hyewon, et al.
Publicado: (2024)
Liar, Liar, Logical Mire: A Benchmark for Suppositional Reasoning in Large Language Models
por: Mondorf, Philipp, et al.
Publicado: (2024)
por: Mondorf, Philipp, et al.
Publicado: (2024)
See, Explain, and Intervene: A Few-Shot Multimodal Agent Framework for Hateful Meme Moderation
por: Rizwan, Naquee, et al.
Publicado: (2026)
por: Rizwan, Naquee, et al.
Publicado: (2026)
MER-Bench: A Comprehensive Benchmark for Multimodal Meme Reappraisal
por: Nie, Yiqi, et al.
Publicado: (2026)
por: Nie, Yiqi, et al.
Publicado: (2026)
Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation
por: Chen, Beiduo, et al.
Publicado: (2025)
por: Chen, Beiduo, et al.
Publicado: (2025)
Investigating the Nature of Disagreements on Mid-Scale Ratings: A Case Study on the Abstractness-Concreteness Continuum
por: Knupleš, Urban, et al.
Publicado: (2023)
por: Knupleš, Urban, et al.
Publicado: (2023)
Decoding Memes: Benchmarking Narrative Role Classification across Multilingual and Multimodal Models
por: Sharma, Shivam, et al.
Publicado: (2025)
por: Sharma, Shivam, et al.
Publicado: (2025)
ExPO-HM: Learning to Explain-then-Detect for Hateful Meme Detection
por: Mei, Jingbiao, et al.
Publicado: (2025)
por: Mei, Jingbiao, et al.
Publicado: (2025)
Meme-ingful Analysis: Enhanced Understanding of Cyberbullying in Memes Through Multimodal Explanations
por: Jha, Prince, et al.
Publicado: (2024)
por: Jha, Prince, et al.
Publicado: (2024)
RAcQUEt: Unveiling the Dangers of Overlooked Referential Ambiguity in Visual LLMs
por: Testoni, Alberto, et al.
Publicado: (2024)
por: Testoni, Alberto, et al.
Publicado: (2024)
Figurative-cum-Commonsense Knowledge Infusion for Multimodal Mental Health Meme Classification
por: Mazhar, Abdullah, et al.
Publicado: (2025)
por: Mazhar, Abdullah, et al.
Publicado: (2025)
MemeMQA: Multimodal Question Answering for Memes via Rationale-Based Inferencing
por: Agarwal, Siddhant, et al.
Publicado: (2024)
por: Agarwal, Siddhant, et al.
Publicado: (2024)
"I Wrote, I Paused, I Rewrote" Teaching LLMs to Read Between the Lines of Student Writing
por: Zafar, Samra, et al.
Publicado: (2025)
por: Zafar, Samra, et al.
Publicado: (2025)
LeWiDi-2025 at NLPerspectives: Third Edition of the Learning with Disagreements Shared Task
por: Leonardelli, Elisa, et al.
Publicado: (2025)
por: Leonardelli, Elisa, et al.
Publicado: (2025)
Unveiling the Mystery of Visual Attributes of Concrete and Abstract Concepts: Variability, Nearest Neighbors, and Challenging Categories
por: Tater, Tarun, et al.
Publicado: (2024)
por: Tater, Tarun, et al.
Publicado: (2024)
MemeCLIP: Leveraging CLIP Representations for Multimodal Meme Classification
por: Shah, Siddhant Bikram, et al.
Publicado: (2024)
por: Shah, Siddhant Bikram, et al.
Publicado: (2024)
On VLMs for Diverse Tasks in Multimodal Meme Classification
por: Gavit, Deepesh, et al.
Publicado: (2025)
por: Gavit, Deepesh, et al.
Publicado: (2025)
Beyond Specialization: Benchmarking LLMs for Transliteration of Indian Languages
por: Azam, Gulfarogh, et al.
Publicado: (2025)
por: Azam, Gulfarogh, et al.
Publicado: (2025)
From Trial by Fire To Sleep Like a Baby: A Lexicon of Anxiety Associations for 20k English Multiword Expressions
por: Mohammad, Saif M.
Publicado: (2026)
por: Mohammad, Saif M.
Publicado: (2026)
When are We Worried? Temporal Trends of Anxiety and What They Reveal about Us
por: Mohammad, Saif M.
Publicado: (2026)
por: Mohammad, Saif M.
Publicado: (2026)
NRC VAD Lexicon v2: Norms for Valence, Arousal, and Dominance for over 55k English Terms
por: Mohammad, Saif M.
Publicado: (2025)
por: Mohammad, Saif M.
Publicado: (2025)
Breaking Bad: Norms for Valence, Arousal, and Dominance for over 10k English Multiword Expressions
por: Mohammad, Saif M.
Publicado: (2025)
por: Mohammad, Saif M.
Publicado: (2025)
WorryWords: Norms of Anxiety Association for over 44k English Words
por: Mohammad, Saif M.
Publicado: (2024)
por: Mohammad, Saif M.
Publicado: (2024)
Words of Warmth: Trust and Sociability Norms for over 26k English Words
por: Mohammad, Saif M.
Publicado: (2025)
por: Mohammad, Saif M.
Publicado: (2025)
LogicSkills: A Structured Benchmark for Formal Reasoning in Large Language Models
por: Rabern, Brian, et al.
Publicado: (2026)
por: Rabern, Brian, et al.
Publicado: (2026)
When Meanings Meet: Investigating the Emergence and Quality of Shared Concept Spaces during Multilingual Language Model Training
por: Körner, Felicia, et al.
Publicado: (2026)
por: Körner, Felicia, et al.
Publicado: (2026)
Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs
por: Camassa, Carolina, et al.
Publicado: (2026)
por: Camassa, Carolina, et al.
Publicado: (2026)
Unimodal Intermediate Training for Multimodal Meme Sentiment Classification
por: Hazman, Muzhaffar, et al.
Publicado: (2023)
por: Hazman, Muzhaffar, et al.
Publicado: (2023)
Multi-Granular Multimodal Clue Fusion for Meme Understanding
por: Zheng, Li, et al.
Publicado: (2025)
por: Zheng, Li, et al.
Publicado: (2025)
"I know myself better, but not really greatly": How Well Can LLMs Detect and Explain LLM-Generated Texts?
por: Ji, Jiazhou, et al.
Publicado: (2025)
por: Ji, Jiazhou, et al.
Publicado: (2025)
Agree, Disagree, Explain: Decomposing Human Label Variation in NLI through the Lens of Explanations
por: Hong, Pingjun, et al.
Publicado: (2025)
por: Hong, Pingjun, et al.
Publicado: (2025)
Ejemplares similares
-
Do LLMs Give Psychometrically Plausible Responses in Educational Assessments?
por: Säuberli, Andreas, et al.
Publicado: (2025) -
Resource-Lean Lexicon Induction for German Dialects
por: Litschko, Robert, et al.
Publicado: (2026) -
What Media Frames Reveal About Stance: A Dataset and Study about Memes in Climate Change Discourse
por: Zhou, Shijia, et al.
Publicado: (2025) -
Mind the Uncertainty in Human Disagreement: Evaluating Discrepancies between Model Predictions and Human Responses in VQA
por: Lan, Jian, et al.
Publicado: (2024) -
Controlling Reading Ease with Gaze-Guided Text Generation
por: Säuberli, Andreas, et al.
Publicado: (2026)