Harm or Humor: A Multimodal, Multilingual Benchmark for Overt and Covert Harmful Humor
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Sharshar, Ahmed, Elgendy, Hosam, Ahmed, Saad El Dine, Rohaim, Yasser, Wang, Yuxia |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Not Only Grey Matter: OmniBrain for Robust Multimodal Classification of Alzheimer's Disease
par: Sharshar, Ahmed, et autres
Publié: (2025)
par: Sharshar, Ahmed, et autres
Publié: (2025)
HALF: Harm-Aware LLM Fairness Evaluation Aligned with Deployment
par: Mekky, Ali, et autres
Publié: (2025)
par: Mekky, Ali, et autres
Publié: (2025)
Small But Funny: A Feedback-Driven Approach to Humor Distillation
par: Ravi, Sahithya, et autres
Publié: (2024)
par: Ravi, Sahithya, et autres
Publié: (2024)
GeoLLaVA: Efficient Fine-Tuned Vision-Language Models for Temporal Change Detection in Remote Sensing
par: Elgendy, Hosam, et autres
Publié: (2024)
par: Elgendy, Hosam, et autres
Publié: (2024)
Oogiri-Master: Benchmarking Humor Understanding via Oogiri
par: Murakami, Soichiro, et autres
Publié: (2025)
par: Murakami, Soichiro, et autres
Publié: (2025)
HarmMetric Eval: Benchmarking Metrics and Judges for LLM Harmfulness Assessment
par: Yang, Langqi, et autres
Publié: (2025)
par: Yang, Langqi, et autres
Publié: (2025)
AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents
par: Andriushchenko, Maksym, et autres
Publié: (2024)
par: Andriushchenko, Maksym, et autres
Publié: (2024)
Chumor 2.0: Towards Benchmarking Chinese Humor Understanding
par: He, Ruiqi, et autres
Publié: (2024)
par: He, Ruiqi, et autres
Publié: (2024)
Humor in Pixels: Benchmarking Large Multimodal Models Understanding of Online Comics
par: Ryan, Yuriel, et autres
Publié: (2025)
par: Ryan, Yuriel, et autres
Publié: (2025)
PluriHarms: Benchmarking the Full Spectrum of Human Judgments on AI Harm
par: Li, Jing-Jing, et autres
Publié: (2026)
par: Li, Jing-Jing, et autres
Publié: (2026)
Cards Against LLMs: Benchmarking Humor Alignment in Large Language Models
par: Fettach, Yousra, et autres
Publié: (2026)
par: Fettach, Yousra, et autres
Publié: (2026)
ChatENV: An Interactive Vision-Language Model for Sensor-Guided Environmental Monitoring and Scenario Simulation
par: Elgendy, Hosam, et autres
Publié: (2025)
par: Elgendy, Hosam, et autres
Publié: (2025)
HUMORCHAIN: Theory-Guided Multi-Stage Reasoning for Interpretable Multimodal Humor Generation
par: Zhang, Jiajun, et autres
Publié: (2025)
par: Zhang, Jiajun, et autres
Publié: (2025)
Getting Serious about Humor: Crafting Humor Datasets with Unfunny Large Language Models
par: Horvitz, Zachary, et autres
Publié: (2024)
par: Horvitz, Zachary, et autres
Publié: (2024)
"They are uncultured": Unveiling Covert Harms and Social Threats in LLM Generated Conversations
par: Dammu, Preetam Prabhu Srikar, et autres
Publié: (2024)
par: Dammu, Preetam Prabhu Srikar, et autres
Publié: (2024)
Self-HarmLLM: Can Large Language Model Harm Itself?
par: Kim, Heehwan, et autres
Publié: (2025)
par: Kim, Heehwan, et autres
Publié: (2025)
v-HUB: A Benchmark for Video Humor Understanding from Vision and Sound
par: Shi, Zhengpeng, et autres
Publié: (2025)
par: Shi, Zhengpeng, et autres
Publié: (2025)
SPACT18: Spiking Human Action Recognition Benchmark Dataset with Complementary RGB and Thermal Modalities
par: Ashraf, Yasser, et autres
Publié: (2025)
par: Ashraf, Yasser, et autres
Publié: (2025)
Learning to Think Like a Cartoon Captionist: Incongruity-Resolution Supervision for Multimodal Humor Understanding
par: Vural, Hatice Merve, et autres
Publié: (2026)
par: Vural, Hatice Merve, et autres
Publié: (2026)
ChineseHarm-Bench: A Chinese Harmful Content Detection Benchmark
par: Liu, Kangwei, et autres
Publié: (2025)
par: Liu, Kangwei, et autres
Publié: (2025)
Humor in AI: Massive Scale Crowd-Sourced Preferences and Benchmarks for Cartoon Captioning
par: Zhang, Jifan, et autres
Publié: (2024)
par: Zhang, Jifan, et autres
Publié: (2024)
HarmTransform: Transforming Explicit Harmful Queries into Stealthy via Multi-Agent Debate
par: Zhu, Shenzhe
Publié: (2025)
par: Zhu, Shenzhe
Publié: (2025)
Can Pre-trained Language Models Understand Chinese Humor?
par: Chen, Yuyan, et autres
Publié: (2024)
par: Chen, Yuyan, et autres
Publié: (2024)
The Multilingual Alignment Prism: Aligning Global and Local Preferences to Reduce Harm
par: Aakanksha, et autres
Publié: (2024)
par: Aakanksha, et autres
Publié: (2024)
Booster: Tackling Harmful Fine-tuning for Large Language Models via Attenuating Harmful Perturbation
par: Huang, Tiansheng, et autres
Publié: (2024)
par: Huang, Tiansheng, et autres
Publié: (2024)
Exploring Chinese Humor Generation: A Study on Two-Part Allegorical Sayings
par: Xu, Rongwu
Publié: (2024)
par: Xu, Rongwu
Publié: (2024)
One Joke to Rule them All? On the (Im)possibility of Generalizing Humor
par: Turgeman, Mor, et autres
Publié: (2025)
par: Turgeman, Mor, et autres
Publié: (2025)
`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts
par: Schoene, Annika M, et autres
Publié: (2025)
par: Schoene, Annika M, et autres
Publié: (2025)
VaccineRAG: Boosting Multimodal Large Language Models' Immunity to Harmful RAG Samples
par: Sun, Qixin, et autres
Publié: (2025)
par: Sun, Qixin, et autres
Publié: (2025)
SocialHarmBench: Revealing LLM Vulnerabilities to Socially Harmful Requests
par: Pandey, Punya Syon, et autres
Publié: (2025)
par: Pandey, Punya Syon, et autres
Publié: (2025)
Towards Generalizable Generic Harmful Speech Datasets for Implicit Hate Speech Detection
par: Almohaimeed, Saad, et autres
Publié: (2025)
par: Almohaimeed, Saad, et autres
Publié: (2025)
Deceptive Humor: A Synthetic Multilingual Benchmark Dataset for Bridging Fabricated Claims with Humorous Content
par: Kasu, Sai Kartheek Reddy, et autres
Publié: (2025)
par: Kasu, Sai Kartheek Reddy, et autres
Publié: (2025)
Towards Comprehensive Detection of Chinese Harmful Memes
par: Lu, Junyu, et autres
Publié: (2024)
par: Lu, Junyu, et autres
Publié: (2024)
Towards Explainable Harmful Meme Detection through Multimodal Debate between Large Language Models
par: Lin, Hongzhan, et autres
Publié: (2024)
par: Lin, Hongzhan, et autres
Publié: (2024)
Assessing the Capabilities of LLMs in Humor:A Multi-dimensional Analysis of Oogiri Generation and Evaluation
par: Sakabe, Ritsu, et autres
Publié: (2025)
par: Sakabe, Ritsu, et autres
Publié: (2025)
AdamMeme: Adaptively Probe the Reasoning Capacity of Multimodal Large Language Models on Harmfulness
par: Chen, Zixin, et autres
Publié: (2025)
par: Chen, Zixin, et autres
Publié: (2025)
Moderating Harm: Benchmarking Large Language Models for Cyberbullying Detection in YouTube Comments
par: Muminovic, Amel
Publié: (2025)
par: Muminovic, Amel
Publié: (2025)
Towards Inclusive NLP: Assessing Compressed Multilingual Transformers across Diverse Language Benchmarks
par: Alshehhi, Maitha, et autres
Publié: (2025)
par: Alshehhi, Maitha, et autres
Publié: (2025)
Human-Guided Harm Recovery for Computer Use Agents
par: Li, Christy, et autres
Publié: (2026)
par: Li, Christy, et autres
Publié: (2026)
Vision-Language Models for Edge Networks: A Comprehensive Survey
par: Sharshar, Ahmed, et autres
Publié: (2025)
par: Sharshar, Ahmed, et autres
Publié: (2025)
Documents similaires
-
Not Only Grey Matter: OmniBrain for Robust Multimodal Classification of Alzheimer's Disease
par: Sharshar, Ahmed, et autres
Publié: (2025) -
HALF: Harm-Aware LLM Fairness Evaluation Aligned with Deployment
par: Mekky, Ali, et autres
Publié: (2025) -
Small But Funny: A Feedback-Driven Approach to Humor Distillation
par: Ravi, Sahithya, et autres
Publié: (2024) -
GeoLLaVA: Efficient Fine-Tuned Vision-Language Models for Temporal Change Detection in Remote Sensing
par: Elgendy, Hosam, et autres
Publié: (2024) -
Oogiri-Master: Benchmarking Humor Understanding via Oogiri
par: Murakami, Soichiro, et autres
Publié: (2025)