Human-Centric Anomaly Detection in Surveillance Videos Using YOLO-World and Spatio-Temporal Deep Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Naeen, Mohammad Ali Etemadi, Mohammadzade, Hoda, Shouraki, Saeed Bagheri |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
UAV-assisted Visual SLAM Generating Reconstructed 3D Scene Graphs in GPS-denied Environments
di: Radwan, Ahmed, et al.
Pubblicazione: (2024)
di: Radwan, Ahmed, et al.
Pubblicazione: (2024)
Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models
di: Seo, Huichan, et al.
Pubblicazione: (2025)
di: Seo, Huichan, et al.
Pubblicazione: (2025)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
di: Tourani, Ali, et al.
Pubblicazione: (2023)
di: Tourani, Ali, et al.
Pubblicazione: (2023)
SpectraNet: FFT-assisted Deep Learning Classifier for Deepfake Face Detection
di: Jayarathne, Nithira, et al.
Pubblicazione: (2025)
di: Jayarathne, Nithira, et al.
Pubblicazione: (2025)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
di: Raoufi, Behnam, et al.
Pubblicazione: (2025)
di: Raoufi, Behnam, et al.
Pubblicazione: (2025)
MaSC: A Masked Similarity Metric for Evaluating Concept-Driven Generation
di: Bartkowiak, Patryk, et al.
Pubblicazione: (2026)
di: Bartkowiak, Patryk, et al.
Pubblicazione: (2026)
Cortex 2.0: Grounding World Models in Real-World Industrial Deployment
di: Aida, Adriana, et al.
Pubblicazione: (2026)
di: Aida, Adriana, et al.
Pubblicazione: (2026)
Ego-Motion Aware Target Prediction Module for Robust Multi-Object Tracking
di: Mahdian, Navid, et al.
Pubblicazione: (2024)
di: Mahdian, Navid, et al.
Pubblicazione: (2024)
Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight
di: Romero, Angel, et al.
Pubblicazione: (2025)
di: Romero, Angel, et al.
Pubblicazione: (2025)
WayFASTER: a Self-Supervised Traversability Prediction for Increased Navigation Awareness
di: Gasparino, Mateus Valverde, et al.
Pubblicazione: (2024)
di: Gasparino, Mateus Valverde, et al.
Pubblicazione: (2024)
Closed-Loop Neural Activation Control in Vision-Language-Action Models
di: Babu, Abhijith, et al.
Pubblicazione: (2026)
di: Babu, Abhijith, et al.
Pubblicazione: (2026)
Advanced Long-term Earth System Forecasting
di: Wu, Hao, et al.
Pubblicazione: (2025)
di: Wu, Hao, et al.
Pubblicazione: (2025)
Towards Localizing Structural Elements: Merging Geometrical Detection with Semantic Verification in RGB-D Data
di: Tourani, Ali, et al.
Pubblicazione: (2024)
di: Tourani, Ali, et al.
Pubblicazione: (2024)
eStonefish-Scenes: A Sim-to-Real Validated and Robot-Centric Event-based Optical Flow Dataset for Underwater Vehicles
di: Mansour, Jad, et al.
Pubblicazione: (2025)
di: Mansour, Jad, et al.
Pubblicazione: (2025)
Circuit Mechanisms for Spatial Relation Generation in Diffusion Transformers
di: Wang, Binxu, et al.
Pubblicazione: (2026)
di: Wang, Binxu, et al.
Pubblicazione: (2026)
Systematic Comparison of Projection Methods for Monocular 3D Human Pose Estimation on Fisheye Images
di: Käs, Stephanie, et al.
Pubblicazione: (2025)
di: Käs, Stephanie, et al.
Pubblicazione: (2025)
Multimodal Ensemble with Conditional Feature Fusion for Dysgraphia Diagnosis in Children from Handwriting Samples
di: Kunhoth, Jayakanth, et al.
Pubblicazione: (2024)
di: Kunhoth, Jayakanth, et al.
Pubblicazione: (2024)
CLARE: Continual Learning for Vision-Language-Action Models via Autonomous Adapter Routing and Expansion
di: Römer, Ralf, et al.
Pubblicazione: (2026)
di: Römer, Ralf, et al.
Pubblicazione: (2026)
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
di: Shahin, Nada, et al.
Pubblicazione: (2025)
di: Shahin, Nada, et al.
Pubblicazione: (2025)
Accelerating Post-Tornado Disaster Assessment Using Advanced Deep Learning Models
di: Umeike, Robinson, et al.
Pubblicazione: (2024)
di: Umeike, Robinson, et al.
Pubblicazione: (2024)
Learn&Drop: Fast Learning of CNNs based on Layer Dropping
di: Cruciata, Giorgio, et al.
Pubblicazione: (2026)
di: Cruciata, Giorgio, et al.
Pubblicazione: (2026)
Knee Osteoarthritis Severity Grading Using Optimized Deep Learning and LLM-Driven Intelligent AI on Computationally Limited Systems
di: Nadeem, Dayam, et al.
Pubblicazione: (2026)
di: Nadeem, Dayam, et al.
Pubblicazione: (2026)
From Demonstrations to Safe Deployment: Path-Consistent Safety Filtering for Diffusion Policies
di: Römer, Ralf, et al.
Pubblicazione: (2025)
di: Römer, Ralf, et al.
Pubblicazione: (2025)
Curb Your Attention: Causal Attention Gating for Robust Trajectory Prediction in Autonomous Driving
di: Ahmadi, Ehsan, et al.
Pubblicazione: (2024)
di: Ahmadi, Ehsan, et al.
Pubblicazione: (2024)
Reference Dataset and Benchmark for Reconstructing Laser Parameters from On-axis Video in Powder Bed Fusion of Bulk Stainless Steel
di: Blanc, Cyril, et al.
Pubblicazione: (2024)
di: Blanc, Cyril, et al.
Pubblicazione: (2024)
eCARLA-scenes: A synthetically generated dataset for event-based optical flow prediction
di: Mansour, Jad, et al.
Pubblicazione: (2024)
di: Mansour, Jad, et al.
Pubblicazione: (2024)
ADAT: Time-Series-Aware Adaptive Transformer Architecture for Sign Language Translation
di: Shahin, Nada, et al.
Pubblicazione: (2025)
di: Shahin, Nada, et al.
Pubblicazione: (2025)
Positive Style Accumulation: A Style Screening and Continuous Utilization Framework for Federated DG-ReID
di: Xu, Xin, et al.
Pubblicazione: (2025)
di: Xu, Xin, et al.
Pubblicazione: (2025)
NumeriKontrol: Adding Numeric Control to Diffusion Transformers for Instruction-based Image Editing
di: Xu, Zhenyu, et al.
Pubblicazione: (2025)
di: Xu, Zhenyu, et al.
Pubblicazione: (2025)
Bridge Diffusion Model: Bridge Chinese Text-to-Image Diffusion Model with English Communities
di: Liu, Shanyuan, et al.
Pubblicazione: (2023)
di: Liu, Shanyuan, et al.
Pubblicazione: (2023)
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
di: Hu, X., et al.
Pubblicazione: (2025)
di: Hu, X., et al.
Pubblicazione: (2025)
Explainable Classifier for Malignant Lymphoma Subtyping via Cell Graph and Image Fusion
di: Nishiyama, Daiki, et al.
Pubblicazione: (2025)
di: Nishiyama, Daiki, et al.
Pubblicazione: (2025)
Convolutional Model Trees
di: Armstrong, William Ward, et al.
Pubblicazione: (2025)
di: Armstrong, William Ward, et al.
Pubblicazione: (2025)
RoboPack: Learning Tactile-Informed Dynamics Models for Dense Packing
di: Ai, Bo, et al.
Pubblicazione: (2024)
di: Ai, Bo, et al.
Pubblicazione: (2024)
Self-Supervised Polyp Re-Identification in Colonoscopy
di: Intrator, Yotam, et al.
Pubblicazione: (2023)
di: Intrator, Yotam, et al.
Pubblicazione: (2023)
Transformers for Image-Goal Navigation
di: Pelluri, Nikhilanj
Pubblicazione: (2024)
di: Pelluri, Nikhilanj
Pubblicazione: (2024)
HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos
di: Wang, Zhi, et al.
Pubblicazione: (2026)
di: Wang, Zhi, et al.
Pubblicazione: (2026)
FlexDoc: Parameterized Sampling for Diverse Multilingual Synthetic Documents for Training Document Understanding Models
di: Dua, Karan, et al.
Pubblicazione: (2025)
di: Dua, Karan, et al.
Pubblicazione: (2025)
APT: Adaptive Personalized Training for Diffusion Models with Limited Data
di: Chae, JungWoo, et al.
Pubblicazione: (2025)
di: Chae, JungWoo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
UAV-assisted Visual SLAM Generating Reconstructed 3D Scene Graphs in GPS-denied Environments
di: Radwan, Ahmed, et al.
Pubblicazione: (2024) -
Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models
di: Seo, Huichan, et al.
Pubblicazione: (2025) -
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
di: Tourani, Ali, et al.
Pubblicazione: (2023) -
SpectraNet: FFT-assisted Deep Learning Classifier for Deepfake Face Detection
di: Jayarathne, Nithira, et al.
Pubblicazione: (2025) -
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
di: Raoufi, Behnam, et al.
Pubblicazione: (2025)