DVD: A Comprehensive Dataset for Advancing Violence Detection in Real-World Scenarios

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Kollias, Dimitrios, Senadeera, Damith C., Zheng, Jianian, Yadav, Kaushal K. K., Slabaugh, Greg, Awais, Muhammad, Yang, Xiaoyun
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916781266829312
author Kollias, Dimitrios
Senadeera, Damith C.
Zheng, Jianian
Yadav, Kaushal K. K.
Slabaugh, Greg
Awais, Muhammad
Yang, Xiaoyun
author_facet Kollias, Dimitrios
Senadeera, Damith C.
Zheng, Jianian
Yadav, Kaushal K. K.
Slabaugh, Greg
Awais, Muhammad
Yang, Xiaoyun
contents Violence Detection (VD) has become an increasingly vital area of research. Existing automated VD efforts are hindered by the limited availability of diverse, well-annotated databases. Existing databases suffer from coarse video-level annotations, limited scale and diversity, and lack of metadata, restricting the generalization of models. To address these challenges, we introduce DVD, a large-scale (500 videos, 2.7M frames), frame-level annotated VD database with diverse environments, varying lighting conditions, multiple camera sources, complex social interactions, and rich metadata. DVD is designed to capture the complexities of real-world violent events.
format Preprint
id arxiv_https___arxiv_org_abs_2506_05372
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle DVD: A Comprehensive Dataset for Advancing Violence Detection in Real-World Scenarios
Kollias, Dimitrios
Senadeera, Damith C.
Zheng, Jianian
Yadav, Kaushal K. K.
Slabaugh, Greg
Awais, Muhammad
Yang, Xiaoyun
Computer Vision and Pattern Recognition
Violence Detection (VD) has become an increasingly vital area of research. Existing automated VD efforts are hindered by the limited availability of diverse, well-annotated databases. Existing databases suffer from coarse video-level annotations, limited scale and diversity, and lack of metadata, restricting the generalization of models. To address these challenges, we introduce DVD, a large-scale (500 videos, 2.7M frames), frame-level annotated VD database with diverse environments, varying lighting conditions, multiple camera sources, complex social interactions, and rich metadata. DVD is designed to capture the complexities of real-world violent events.
title DVD: A Comprehensive Dataset for Advancing Violence Detection in Real-World Scenarios
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2506.05372