NYK-MS: A Well-annotated Multi-modal Metaphor and Sarcasm Understanding Benchmark on Cartoon-Caption Dataset
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chang, Ke, Li, Hao, Zhang, Junzhao, Wu, Yunfang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CFMS: Towards Explainable and Fine-Grained Chinese Multimodal Sarcasm Detection Benchmark
von: Zhang, Junzhao, et al.
Veröffentlicht: (2026)
von: Zhang, Junzhao, et al.
Veröffentlicht: (2026)
Multi-modal Semantic Understanding with Contrastive Cross-modal Feature Alignment
von: Zhang, Ming, et al.
Veröffentlicht: (2024)
von: Zhang, Ming, et al.
Veröffentlicht: (2024)
Mixture-of-Prompt-Experts for Multi-modal Semantic Understanding
von: Wu, Zichen, et al.
Veröffentlicht: (2024)
von: Wu, Zichen, et al.
Veröffentlicht: (2024)
Towards Multimodal Metaphor Understanding: A Chinese Dataset and Model for Metaphor Mapping Identification
von: Zhang, Dongyu, et al.
Veröffentlicht: (2025)
von: Zhang, Dongyu, et al.
Veröffentlicht: (2025)
Humor in AI: Massive Scale Crowd-Sourced Preferences and Benchmarks for Cartoon Captioning
von: Zhang, Jifan, et al.
Veröffentlicht: (2024)
von: Zhang, Jifan, et al.
Veröffentlicht: (2024)
Unlocking the Power of LLM Uncertainty for Active In-Context Example Selection
von: Huang, Hsiu-Yuan, et al.
Veröffentlicht: (2024)
von: Huang, Hsiu-Yuan, et al.
Veröffentlicht: (2024)
Metaphor Understanding Challenge Dataset for LLMs
von: Tong, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Tong, Xiaoyu, et al.
Veröffentlicht: (2024)
SarcasmBench: Towards Evaluating Large Language Models on Sarcasm Understanding
von: Zhang, Yazhou, et al.
Veröffentlicht: (2024)
von: Zhang, Yazhou, et al.
Veröffentlicht: (2024)
Cultural Bias Matters: A Cross-Cultural Benchmark Dataset and Sentiment-Enriched Model for Understanding Multimodal Metaphors
von: Yang, Senqi, et al.
Veröffentlicht: (2025)
von: Yang, Senqi, et al.
Veröffentlicht: (2025)
HAIC: Improving Human Action Understanding and Generation with Better Captions for Multi-modal Large Language Models
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
Integrating Feedback Loss from Bi-modal Sarcasm Detector for Sarcastic Speech Synthesis
von: Li, Zhu, et al.
Veröffentlicht: (2025)
von: Li, Zhu, et al.
Veröffentlicht: (2025)
Was that Sarcasm?: A Literature Survey on Sarcasm Detection
von: Bagga, Harleen Kaur, et al.
Veröffentlicht: (2024)
von: Bagga, Harleen Kaur, et al.
Veröffentlicht: (2024)
Evaluating Multimodal Large Language Models on Spoken Sarcasm Understanding
von: Li, Zhu, et al.
Veröffentlicht: (2025)
von: Li, Zhu, et al.
Veröffentlicht: (2025)
Unveiling the Invisible: Captioning Videos with Metaphors
von: Kalarani, Abisek Rajakumar, et al.
Veröffentlicht: (2024)
von: Kalarani, Abisek Rajakumar, et al.
Veröffentlicht: (2024)
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos
von: Zang, Yuan, et al.
Veröffentlicht: (2025)
von: Zang, Yuan, et al.
Veröffentlicht: (2025)
InterCLIP-MEP: Interactive CLIP and Memory-Enhanced Predictor for Multi-modal Sarcasm Detection
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
Multi-View Incongruity Learning for Multimodal Sarcasm Detection
von: Guo, Diandian, et al.
Veröffentlicht: (2024)
von: Guo, Diandian, et al.
Veröffentlicht: (2024)
Friends-MMC: A Dataset for Multi-modal Multi-party Conversation Understanding
von: Wang, Yueqian, et al.
Veröffentlicht: (2024)
von: Wang, Yueqian, et al.
Veröffentlicht: (2024)
MuVaC: A Variational Causal Framework for Multimodal Sarcasm Understanding in Dialogues
von: Guo, Diandian, et al.
Veröffentlicht: (2026)
von: Guo, Diandian, et al.
Veröffentlicht: (2026)
MetaphorShare: A Dynamic Collaborative Repository of Open Metaphor Datasets
von: Boisson, Joanne, et al.
Veröffentlicht: (2024)
von: Boisson, Joanne, et al.
Veröffentlicht: (2024)
Can Large Vision-Language Models Understand Multimodal Sarcasm?
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
MulCogBench: A Multi-modal Cognitive Benchmark Dataset for Evaluating Chinese and English Computational Language Models
von: Zhang, Yunhao, et al.
Veröffentlicht: (2024)
von: Zhang, Yunhao, et al.
Veröffentlicht: (2024)
Density Matrices for Metaphor Understanding
von: Owers, Jay, et al.
Veröffentlicht: (2024)
von: Owers, Jay, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models' Ability Using a Psychiatric Screening Tool Based on Metaphor and Sarcasm Scenarios
von: Yakura, Hiromu
Veröffentlicht: (2023)
von: Yakura, Hiromu
Veröffentlicht: (2023)
CMMU: A Benchmark for Chinese Multi-modal Multi-type Question Understanding and Reasoning
von: He, Zheqi, et al.
Veröffentlicht: (2024)
von: He, Zheqi, et al.
Veröffentlicht: (2024)
BESSTIE: A Benchmark for Sentiment and Sarcasm Classification for Varieties of English
von: Srirag, Dipankar, et al.
Veröffentlicht: (2024)
von: Srirag, Dipankar, et al.
Veröffentlicht: (2024)
Unsupervised Distractor Generation via Large Language Model Distilling and Counterfactual Contrastive Decoding
von: Qu, Fanyi, et al.
Veröffentlicht: (2024)
von: Qu, Fanyi, et al.
Veröffentlicht: (2024)
Multi-modal Retrieval Augmented Multi-modal Generation: Datasets, Evaluation Metrics and Strong Baselines
von: Ma, Zi-Ao, et al.
Veröffentlicht: (2024)
von: Ma, Zi-Ao, et al.
Veröffentlicht: (2024)
CMDAG: A Chinese Metaphor Dataset with Annotated Grounds as CoT for Boosting Metaphor Generation
von: Shao, Yujie, et al.
Veröffentlicht: (2024)
von: Shao, Yujie, et al.
Veröffentlicht: (2024)
Visual Language Tracking with Multi-modal Interaction: A Robust Benchmark
von: Li, Xuchen, et al.
Veröffentlicht: (2024)
von: Li, Xuchen, et al.
Veröffentlicht: (2024)
DanmakuTPPBench: A Multi-modal Benchmark for Temporal Point Process Modeling and Understanding
von: Jiang, Yue, et al.
Veröffentlicht: (2025)
von: Jiang, Yue, et al.
Veröffentlicht: (2025)
AMuSeD: An Attentive Deep Neural Network for Multimodal Sarcasm Detection Incorporating Bi-modal Data Augmentation
von: Gao, Xiyuan, et al.
Veröffentlicht: (2024)
von: Gao, Xiyuan, et al.
Veröffentlicht: (2024)
Lost in the Passage: Passage-level In-context Learning Does Not Necessarily Need a "Passage"
von: Sun, Hao, et al.
Veröffentlicht: (2025)
von: Sun, Hao, et al.
Veröffentlicht: (2025)
Hummus: A Dataset of Humorous Multimodal Metaphor Use
von: Tong, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Tong, Xiaoyu, et al.
Veröffentlicht: (2025)
A Survey of Multimodal Sarcasm Detection
von: Farabi, Shafkat, et al.
Veröffentlicht: (2024)
von: Farabi, Shafkat, et al.
Veröffentlicht: (2024)
KoCoSa: Korean Context-aware Sarcasm Detection Dataset
von: Kim, Yumin, et al.
Veröffentlicht: (2024)
von: Kim, Yumin, et al.
Veröffentlicht: (2024)
SarcasmMiner: A Dual-Track Post-Training Framework for Robust Audio-Visual Sarcasm Reasoning
von: Li, Zhu, et al.
Veröffentlicht: (2026)
von: Li, Zhu, et al.
Veröffentlicht: (2026)
BRACE: A Benchmark for Robust Audio Caption Quality Evaluation
von: Guo, Tianyu, et al.
Veröffentlicht: (2025)
von: Guo, Tianyu, et al.
Veröffentlicht: (2025)
Face-Human-Bench: A Comprehensive Benchmark of Face and Human Understanding for Multi-modal Assistants
von: Qin, Lixiong, et al.
Veröffentlicht: (2025)
von: Qin, Lixiong, et al.
Veröffentlicht: (2025)
Is Sarcasm Detection A Step-by-Step Reasoning Process in Large Language Models?
von: Yao, Ben, et al.
Veröffentlicht: (2024)
von: Yao, Ben, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CFMS: Towards Explainable and Fine-Grained Chinese Multimodal Sarcasm Detection Benchmark
von: Zhang, Junzhao, et al.
Veröffentlicht: (2026) -
Multi-modal Semantic Understanding with Contrastive Cross-modal Feature Alignment
von: Zhang, Ming, et al.
Veröffentlicht: (2024) -
Mixture-of-Prompt-Experts for Multi-modal Semantic Understanding
von: Wu, Zichen, et al.
Veröffentlicht: (2024) -
Towards Multimodal Metaphor Understanding: A Chinese Dataset and Model for Metaphor Mapping Identification
von: Zhang, Dongyu, et al.
Veröffentlicht: (2025) -
Humor in AI: Massive Scale Crowd-Sourced Preferences and Benchmarks for Cartoon Captioning
von: Zhang, Jifan, et al.
Veröffentlicht: (2024)