GridVAD: Open-Set Video Anomaly Detection via Spatial Reasoning over Stratified Frame Grids
Fuente:
arXiv
Guardado en:
| Autores principales: | Eltahir, Mohamed, Ibrahim, Ahmed O., Siralkhatim, Obada, Abdallah, Tabarak, Mohamed, Sondos |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GridProbe: Posterior-Probing for Adaptive Test-Time Compute in Long-Video VLMs
por: Eltahir, Mohamed, et al.
Publicado: (2026)
por: Eltahir, Mohamed, et al.
Publicado: (2026)
ComplexVAD: Detecting Interaction Anomalies in Video
por: Mumcu, Furkan, et al.
Publicado: (2025)
por: Mumcu, Furkan, et al.
Publicado: (2025)
DUAL-VAD: Dual Benchmarks and Anomaly-Focused Sampling for Video Anomaly Detection
por: Jung, Seoik, et al.
Publicado: (2025)
por: Jung, Seoik, et al.
Publicado: (2025)
CoReVAD: A Contextual Reasoning Framework for Training-Free Video Anomaly Detection
por: Lim, Hyeongmuk, et al.
Publicado: (2026)
por: Lim, Hyeongmuk, et al.
Publicado: (2026)
EventVAD: Training-Free Event-Aware Video Anomaly Detection
por: Shao, Yihua, et al.
Publicado: (2025)
por: Shao, Yihua, et al.
Publicado: (2025)
GV-VAD : Exploring Video Generation for Weakly-Supervised Video Anomaly Detection
por: Cai, Suhang, et al.
Publicado: (2025)
por: Cai, Suhang, et al.
Publicado: (2025)
SphereVAD: Training-Free Video Anomaly Detection via Geodesic Inference on the Unit Hypersphere
por: Huang, Chao, et al.
Publicado: (2026)
por: Huang, Chao, et al.
Publicado: (2026)
Holmes-VAD: Towards Unbiased and Explainable Video Anomaly Detection via Multi-modal LLM
por: Zhang, Huaxin, et al.
Publicado: (2024)
por: Zhang, Huaxin, et al.
Publicado: (2024)
GlanceVAD: Exploring Glance Supervision for Label-efficient Video Anomaly Detection
por: Zhang, Huaxin, et al.
Publicado: (2024)
por: Zhang, Huaxin, et al.
Publicado: (2024)
RefineVAD: Semantic-Guided Feature Recalibration for Weakly Supervised Video Anomaly Detection
por: Lee, Junhee, et al.
Publicado: (2025)
por: Lee, Junhee, et al.
Publicado: (2025)
HeadHunt-VAD: Hunting Robust Anomaly-Sensitive Heads in MLLM for Tuning-Free Video Anomaly Detection
por: Cai, Zhaolin, et al.
Publicado: (2025)
por: Cai, Zhaolin, et al.
Publicado: (2025)
ProDisc-VAD: An Efficient System for Weakly-Supervised Anomaly Detection in Video Surveillance Applications
por: Zhu, Tao, et al.
Publicado: (2025)
por: Zhu, Tao, et al.
Publicado: (2025)
HyCoVAD: A Hybrid SSL-LLM Model for Complex Video Anomaly Detection
por: Hemmatyar, Mohammad Mahdi, et al.
Publicado: (2025)
por: Hemmatyar, Mohammad Mahdi, et al.
Publicado: (2025)
Bi-Grid Reconstruction for Image Anomaly Detection
por: Huang, Huichuan, et al.
Publicado: (2025)
por: Huang, Huichuan, et al.
Publicado: (2025)
HiProbe-VAD: Video Anomaly Detection via Hidden States Probing in Tuning-Free Multimodal LLMs
por: Cai, Zhaolin, et al.
Publicado: (2025)
por: Cai, Zhaolin, et al.
Publicado: (2025)
SlowFastVAD: Video Anomaly Detection via Integrating Simple Detector and RAG-Enhanced Vision-Language Model
por: Ding, Zongcan, et al.
Publicado: (2025)
por: Ding, Zongcan, et al.
Publicado: (2025)
VideoAtlas: Navigating Long-Form Video in Logarithmic Compute
por: Eltahir, Mohamed, et al.
Publicado: (2026)
por: Eltahir, Mohamed, et al.
Publicado: (2026)
Grid Spatial Understanding: A Dataset for Textual Spatial Reasoning over Grids, Embodied Settings, and Coordinate Structures
por: Sidhu, Risham, et al.
Publicado: (2026)
por: Sidhu, Risham, et al.
Publicado: (2026)
VAD4Space: Visual Anomaly Detection for Planetary Surface Imagery
por: Genilotti, Fabrizio, et al.
Publicado: (2026)
por: Genilotti, Fabrizio, et al.
Publicado: (2026)
Leveraging Digital Twin and Machine Learning Techniques for Anomaly Detection in Power Electronics Dominated Grid
por: Idrisov, Ildar N., et al.
Publicado: (2025)
por: Idrisov, Ildar N., et al.
Publicado: (2025)
UniVAD: A Training-free Unified Model for Few-shot Visual Anomaly Detection
por: Gu, Zhaopeng, et al.
Publicado: (2024)
por: Gu, Zhaopeng, et al.
Publicado: (2024)
Open-Vocabulary Video Anomaly Detection
por: Wu, Peng, et al.
Publicado: (2023)
por: Wu, Peng, et al.
Publicado: (2023)
AutoArabic: A Three-Stage Framework for Localizing Video-Text Retrieval Benchmarks
por: Eltahir, Mohamed, et al.
Publicado: (2025)
por: Eltahir, Mohamed, et al.
Publicado: (2025)
Multimodal Lengthy Videos Retrieval Framework and Evaluation Metric
por: Eltahir, Mohamed, et al.
Publicado: (2025)
por: Eltahir, Mohamed, et al.
Publicado: (2025)
Vote-in-Context: Turning VLMs into Zero-Shot Rank Fusers
por: Eltahir, Mohamed, et al.
Publicado: (2025)
por: Eltahir, Mohamed, et al.
Publicado: (2025)
Real-Time On-the-Go Annotation Framework Using YOLO for Automated Dataset Generation
por: Salem, Mohamed Abdallah, et al.
Publicado: (2025)
por: Salem, Mohamed Abdallah, et al.
Publicado: (2025)
Grid-LOGAT: Grid Based Local and Global Area Transcription for Video Question Answering
por: Chowdhury, Md Intisar, et al.
Publicado: (2025)
por: Chowdhury, Md Intisar, et al.
Publicado: (2025)
StreamSTGS: Streaming Spatial and Temporal Gaussian Grids for Real-Time Free-Viewpoint Video
por: Ke, Zhihui, et al.
Publicado: (2025)
por: Ke, Zhihui, et al.
Publicado: (2025)
Key-Grid: Unsupervised 3D Keypoints Detection using Grid Heatmap Features
por: Hou, Chengkai, et al.
Publicado: (2024)
por: Hou, Chengkai, et al.
Publicado: (2024)
Grid Diffusion Models for Text-to-Video Generation
por: Lee, Taegyeong, et al.
Publicado: (2024)
por: Lee, Taegyeong, et al.
Publicado: (2024)
A Data Efficiency Study of Synthetic Fog for Object Detection Using the Clear2Fog Pipeline
por: Mohamed, Mohamed Ahmed, et al.
Publicado: (2026)
por: Mohamed, Mohamed Ahmed, et al.
Publicado: (2026)
From Frames to Events: Rethinking Evaluation in Human-Centric Video Anomaly Detection
por: Rashvand, Narges, et al.
Publicado: (2026)
por: Rashvand, Narges, et al.
Publicado: (2026)
The Role of Deep Learning in Advancing Proactive Cybersecurity Measures for Smart Grid Networks: A Survey
por: Abdi, Nima, et al.
Publicado: (2024)
por: Abdi, Nima, et al.
Publicado: (2024)
FrameMind: Frame-Interleaved Video Reasoning via Reinforcement Learning
por: Ge, Haonan, et al.
Publicado: (2025)
por: Ge, Haonan, et al.
Publicado: (2025)
Anomize: Better Open Vocabulary Video Anomaly Detection
por: Li, Fei, et al.
Publicado: (2025)
por: Li, Fei, et al.
Publicado: (2025)
Enhancing Visual Token Representations for Video Large Language Models via Training-Free Spatial-Temporal Pooling and Gridding
por: Luo, Bingjun, et al.
Publicado: (2026)
por: Luo, Bingjun, et al.
Publicado: (2026)
Towards a Safer and Sustainable Manufacturing Process: Material classification in Laser Cutting Using Deep Learning
por: Salem, Mohamed Abdallah, et al.
Publicado: (2025)
por: Salem, Mohamed Abdallah, et al.
Publicado: (2025)
Artificial intelligence approaches for energy-efficient laser cutting machines
por: Salem, Mohamed Abdallah, et al.
Publicado: (2025)
por: Salem, Mohamed Abdallah, et al.
Publicado: (2025)
Qffusion: Controllable Portrait Video Editing via Quadrant-Grid Attention Learning
por: Li, Maomao, et al.
Publicado: (2025)
por: Li, Maomao, et al.
Publicado: (2025)
Chain-of-Frames: Advancing Video Understanding in Multimodal LLMs via Frame-Aware Reasoning
por: Ghazanfari, Sara, et al.
Publicado: (2025)
por: Ghazanfari, Sara, et al.
Publicado: (2025)
Ejemplares similares
-
GridProbe: Posterior-Probing for Adaptive Test-Time Compute in Long-Video VLMs
por: Eltahir, Mohamed, et al.
Publicado: (2026) -
ComplexVAD: Detecting Interaction Anomalies in Video
por: Mumcu, Furkan, et al.
Publicado: (2025) -
DUAL-VAD: Dual Benchmarks and Anomaly-Focused Sampling for Video Anomaly Detection
por: Jung, Seoik, et al.
Publicado: (2025) -
CoReVAD: A Contextual Reasoning Framework for Training-Free Video Anomaly Detection
por: Lim, Hyeongmuk, et al.
Publicado: (2026) -
EventVAD: Training-Free Event-Aware Video Anomaly Detection
por: Shao, Yihua, et al.
Publicado: (2025)