DVD: A Comprehensive Dataset for Advancing Violence Detection in Real-World Scenarios
Fuente:
arXiv
Saved in:
| Main Authors: | , , , , , , |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
| _version_ | 1866916781266829312 |
|---|---|
| author | Kollias, Dimitrios Senadeera, Damith C. Zheng, Jianian Yadav, Kaushal K. K. Slabaugh, Greg Awais, Muhammad Yang, Xiaoyun |
| author_facet | Kollias, Dimitrios Senadeera, Damith C. Zheng, Jianian Yadav, Kaushal K. K. Slabaugh, Greg Awais, Muhammad Yang, Xiaoyun |
| contents | Violence Detection (VD) has become an increasingly vital area of research. Existing automated VD efforts are hindered by the limited availability of diverse, well-annotated databases. Existing databases suffer from coarse video-level annotations, limited scale and diversity, and lack of metadata, restricting the generalization of models. To address these challenges, we introduce DVD, a large-scale (500 videos, 2.7M frames), frame-level annotated VD database with diverse environments, varying lighting conditions, multiple camera sources, complex social interactions, and rich metadata. DVD is designed to capture the complexities of real-world violent events. |
| format | Preprint |
| id |
arxiv_https___arxiv_org_abs_2506_05372 |
| institution | arXiv |
| publishDate | 2025 |
| record_format | arxiv |
| spellingShingle | DVD: A Comprehensive Dataset for Advancing Violence Detection in Real-World Scenarios Kollias, Dimitrios Senadeera, Damith C. Zheng, Jianian Yadav, Kaushal K. K. Slabaugh, Greg Awais, Muhammad Yang, Xiaoyun Computer Vision and Pattern Recognition Violence Detection (VD) has become an increasingly vital area of research. Existing automated VD efforts are hindered by the limited availability of diverse, well-annotated databases. Existing databases suffer from coarse video-level annotations, limited scale and diversity, and lack of metadata, restricting the generalization of models. To address these challenges, we introduce DVD, a large-scale (500 videos, 2.7M frames), frame-level annotated VD database with diverse environments, varying lighting conditions, multiple camera sources, complex social interactions, and rich metadata. DVD is designed to capture the complexities of real-world violent events. |
| title | DVD: A Comprehensive Dataset for Advancing Violence Detection in Real-World Scenarios |
| topic | Computer Vision and Pattern Recognition |
| url | https://arxiv.org/abs/2506.05372 |