MetaHarm: Harmful YouTube Video Dataset Annotated by Domain Experts, GPT-4-Turbo, and Crowdworkers
Fuente:
arXiv
Saved in:
| Main Authors: | Jo, Wonjeong, Wojcieszak, Magdalena |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Harmful YouTube Video Detection: A Taxonomy of Online Harm and MLLMs as Alternative Annotators
by: Jo, Claire Wonjeong, et al.
Published: (2024)
by: Jo, Claire Wonjeong, et al.
Published: (2024)
YouTube-Occ: Learning Indoor 3D Semantic Occupancy Prediction from YouTube Videos
by: Chen, Haoming, et al.
Published: (2025)
by: Chen, Haoming, et al.
Published: (2025)
YouTube SFV+HDR Quality Dataset
by: Wang, Yilin, et al.
Published: (2024)
by: Wang, Yilin, et al.
Published: (2024)
Can Language Models Laugh at YouTube Short-form Videos?
by: Ko, Dayoon, et al.
Published: (2023)
by: Ko, Dayoon, et al.
Published: (2023)
AHOY! Animatable Humans under Occlusion from YouTube Videos with Gaussian Splatting and Video Diffusion Priors
by: Mir, Aymen, et al.
Published: (2026)
by: Mir, Aymen, et al.
Published: (2026)
FinCap: Topic-Aligned Captions for Short-Form Financial YouTube Videos
by: Sukhani, Siddhant, et al.
Published: (2025)
by: Sukhani, Siddhant, et al.
Published: (2025)
MultiHateClip: A Multilingual Benchmark Dataset for Hateful Video Detection on YouTube and Bilibili
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
Unboxing Engagement in YouTube Influencer Videos: An Attention-Based Approach
by: Rajaram, Prashant, et al.
Published: (2020)
by: Rajaram, Prashant, et al.
Published: (2020)
YouTube Video Analytics for Patient Engagement: Evidence from Colonoscopy Preparation Videos
by: Guo, Yawen, et al.
Published: (2024)
by: Guo, Yawen, et al.
Published: (2024)
Failures to Surface Harmful Contents in Video Large Language Models
by: Cao, Yuxin, et al.
Published: (2025)
by: Cao, Yuxin, et al.
Published: (2025)
T2Vs Meet VLMs: A Scalable Multimodal Dataset for Visual Harmfulness Recognition
by: Yeh, Chen, et al.
Published: (2024)
by: Yeh, Chen, et al.
Published: (2024)
Residual Connections Harm Generative Representation Learning
by: Zhang, Xiao, et al.
Published: (2024)
by: Zhang, Xiao, et al.
Published: (2024)
The Bias of Harmful Label Associations in Vision-Language Models
by: Hazirbas, Caner, et al.
Published: (2024)
by: Hazirbas, Caner, et al.
Published: (2024)
Augmenting Chest X-ray Datasets with Non-Expert Annotations
by: Cheplygina, Veronika, et al.
Published: (2023)
by: Cheplygina, Veronika, et al.
Published: (2023)
Too Many Frames, Not All Useful: Efficient Strategies for Long-Form Video QA
by: Park, Jongwoo, et al.
Published: (2024)
by: Park, Jongwoo, et al.
Published: (2024)
Harmfully Manipulated Images Matter in Multimodal Misinformation Detection
by: Wang, Bing, et al.
Published: (2024)
by: Wang, Bing, et al.
Published: (2024)
SafeCFG: Controlling Harmful Features with Dynamic Safe Guidance for Safe Generation
by: Pan, Jiadong, et al.
Published: (2024)
by: Pan, Jiadong, et al.
Published: (2024)
TurboVSR: Fantastic Video Upscalers and Where to Find Them
by: Wang, Zhongdao, et al.
Published: (2025)
by: Wang, Zhongdao, et al.
Published: (2025)
Robust Harmful Meme Detection under Missing Modalities via Shared Representation Learning
by: Breiteneder, Felix, et al.
Published: (2026)
by: Breiteneder, Felix, et al.
Published: (2026)
MemeMind: A Large-Scale Multimodal Dataset with Chain-of-Thought Reasoning for Harmful Meme Detection
by: Gu, Hexiang, et al.
Published: (2025)
by: Gu, Hexiang, et al.
Published: (2025)
On the Generation and Mitigation of Harmful Geometry in Image-to-3D Models
by: Liu, Yule, et al.
Published: (2026)
by: Liu, Yule, et al.
Published: (2026)
Turbo-VAED: Fast and Stable Transfer of Video-VAEs to Mobile Devices
by: Zou, Ya, et al.
Published: (2025)
by: Zou, Ya, et al.
Published: (2025)
CableInspect-AD: An Expert-Annotated Anomaly Detection Dataset
by: Arodi, Akshatha, et al.
Published: (2024)
by: Arodi, Akshatha, et al.
Published: (2024)
TurboReg: TurboClique for Robust and Efficient Point Cloud Registration
by: Yan, Shaocheng, et al.
Published: (2025)
by: Yan, Shaocheng, et al.
Published: (2025)
SpatialVID: A Large-Scale Video Dataset with Spatial Annotations
by: Wang, Jiahao, et al.
Published: (2025)
by: Wang, Jiahao, et al.
Published: (2025)
Benign-to-Toxic Jailbreaking: Inducing Harmful Responses from Harmless Prompts
by: Kim, Hee-Seon, et al.
Published: (2025)
by: Kim, Hee-Seon, et al.
Published: (2025)
Recognition of Harmful Phytoplankton from Microscopic Images using Deep Learning
by: Khaldi, Aymane, et al.
Published: (2024)
by: Khaldi, Aymane, et al.
Published: (2024)
Transferable-guided Attention Is All You Need for Video Domain Adaptation
by: Sacilotti, André, et al.
Published: (2024)
by: Sacilotti, André, et al.
Published: (2024)
VideoCoT: A Video Chain-of-Thought Dataset with Active Annotation Tool
by: Wang, Yan, et al.
Published: (2024)
by: Wang, Yan, et al.
Published: (2024)
Don't Let Your Robot be Harmful: Responsible Robotic Manipulation via Safety-as-Policy
by: Ni, Minheng, et al.
Published: (2024)
by: Ni, Minheng, et al.
Published: (2024)
GeneVA: A Dataset of Human Annotations for Generative Text to Video Artifacts
by: Kang, Jenna, et al.
Published: (2025)
by: Kang, Jenna, et al.
Published: (2025)
VUDG: A Dataset for Video Understanding Domain Generalization
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
Stronger Semantic Encoders Can Harm Relighting Performance: Probing Visual Priors via Augmented Latent Intrinsics
by: Xing, Xiaoyan, et al.
Published: (2026)
by: Xing, Xiaoyan, et al.
Published: (2026)
Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization
by: Maljkovic, Igor, et al.
Published: (2026)
by: Maljkovic, Igor, et al.
Published: (2026)
Benchmarking Bias Mitigation Toward Fairness Without Harm from Vision to LVLMs
by: Tan, Xuwei, et al.
Published: (2026)
by: Tan, Xuwei, et al.
Published: (2026)
Robustness of Vision Language Models Against Split-Image Harmful Input Attacks
by: Rashid, Md Rafi Ur, et al.
Published: (2026)
by: Rashid, Md Rafi Ur, et al.
Published: (2026)
OSPC: Detecting Harmful Memes with Large Language Model as a Catalyst
by: Cao, Jingtao, et al.
Published: (2024)
by: Cao, Jingtao, et al.
Published: (2024)
From Shallow Humor to Metaphor: Towards Label-Free Harmful Meme Detection via LMM Agent Self-Improvement
by: Lang, Jian, et al.
Published: (2025)
by: Lang, Jian, et al.
Published: (2025)
A Fast Horizon Detector and a New Annotated Dataset for Maritime Video Processing
by: Zardoua, Yassir, et al.
Published: (2021)
by: Zardoua, Yassir, et al.
Published: (2021)
FineBio: A Fine-Grained Video Dataset of Biological Experiments with Hierarchical Annotation
by: Yagi, Takuma, et al.
Published: (2024)
by: Yagi, Takuma, et al.
Published: (2024)
Similar Items
-
Harmful YouTube Video Detection: A Taxonomy of Online Harm and MLLMs as Alternative Annotators
by: Jo, Claire Wonjeong, et al.
Published: (2024) -
YouTube-Occ: Learning Indoor 3D Semantic Occupancy Prediction from YouTube Videos
by: Chen, Haoming, et al.
Published: (2025) -
YouTube SFV+HDR Quality Dataset
by: Wang, Yilin, et al.
Published: (2024) -
Can Language Models Laugh at YouTube Short-form Videos?
by: Ko, Dayoon, et al.
Published: (2023) -
AHOY! Animatable Humans under Occlusion from YouTube Videos with Gaussian Splatting and Video Diffusion Priors
by: Mir, Aymen, et al.
Published: (2026)