Evaluation of Vision-LLMs in Surveillance Video
Fuente:
arXiv
Salvato in:
| Autori principali: | Benschop, Pascal, Meo, Cristian, Dauwels, Justin, Mense, Jelte P. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Assessing Situational and Spatial Awareness of VLMs with Synthetically Generated Video
di: Benschop, Pascal, et al.
Pubblicazione: (2026)
di: Benschop, Pascal, et al.
Pubblicazione: (2026)
Compositional Scene Understanding through Inverse Generative Modeling
di: Wang, Yanbo, et al.
Pubblicazione: (2025)
di: Wang, Yanbo, et al.
Pubblicazione: (2025)
Slot-VAE: Object-Centric Scene Generation with Slot Attention
di: Wang, Yanbo, et al.
Pubblicazione: (2023)
di: Wang, Yanbo, et al.
Pubblicazione: (2023)
Identifying Ethical Biases in Action Recognition Models
di: Baltaretu, Ana, et al.
Pubblicazione: (2026)
di: Baltaretu, Ana, et al.
Pubblicazione: (2026)
Object-Centric Temporal Consistency via Conditional Autoregressive Inductive Biases
di: Meo, Cristian, et al.
Pubblicazione: (2024)
di: Meo, Cristian, et al.
Pubblicazione: (2024)
Retrieval, Refinement, and Ranking for Text-to-Video Generation via Prompt Optimization and Test-Time Scaling
di: Rahman, Zillur, et al.
Pubblicazione: (2026)
di: Rahman, Zillur, et al.
Pubblicazione: (2026)
Detection of Customer Interested Garments in Surveillance Video using Computer Vision
di: Ijjina, Earnest Paul, et al.
Pubblicazione: (2025)
di: Ijjina, Earnest Paul, et al.
Pubblicazione: (2025)
Detection of Object Throwing Behavior in Surveillance Videos
di: Kersten, Ivo P. C., et al.
Pubblicazione: (2024)
di: Kersten, Ivo P. C., et al.
Pubblicazione: (2024)
Large Language Models for Video Surveillance Applications
di: De Silva, Ulindu, et al.
Pubblicazione: (2025)
di: De Silva, Ulindu, et al.
Pubblicazione: (2025)
SurveillanceVQA-589K: A Benchmark for Comprehensive Surveillance Video-Language Understanding with Large Models
di: Liu, Bo, et al.
Pubblicazione: (2025)
di: Liu, Bo, et al.
Pubblicazione: (2025)
Customer Analytics using Surveillance Video
di: Ijjina, Earnest Paul, et al.
Pubblicazione: (2025)
di: Ijjina, Earnest Paul, et al.
Pubblicazione: (2025)
Grounding Foundational Vision Models with 3D Human Poses for Robust Action Recognition
di: Babey, Nicholas, et al.
Pubblicazione: (2025)
di: Babey, Nicholas, et al.
Pubblicazione: (2025)
ARGUS: Hallucination and Omission Evaluation in Video-LLMs
di: Rawal, Ruchit, et al.
Pubblicazione: (2025)
di: Rawal, Ruchit, et al.
Pubblicazione: (2025)
Zero-Shot Action Recognition in Surveillance Videos
di: Pereira, Joao, et al.
Pubblicazione: (2024)
di: Pereira, Joao, et al.
Pubblicazione: (2024)
A Benchmark for Crime Surveillance Video Analysis with Large Models
di: Chen, Haoran, et al.
Pubblicazione: (2025)
di: Chen, Haoran, et al.
Pubblicazione: (2025)
A Flying Bird Object Detection Method for Surveillance Video
di: Sun, Ziwei, et al.
Pubblicazione: (2024)
di: Sun, Ziwei, et al.
Pubblicazione: (2024)
AssistPDA: An Online Video Surveillance Assistant for Video Anomaly Prediction, Detection, and Analysis
di: Yang, Zhiwei, et al.
Pubblicazione: (2025)
di: Yang, Zhiwei, et al.
Pubblicazione: (2025)
Vision Transformer for Robust Occluded Person Reidentification in Complex Surveillance Scenes
di: Li, Bo, et al.
Pubblicazione: (2025)
di: Li, Bo, et al.
Pubblicazione: (2025)
Vision Technologies with Applications in Traffic Surveillance Systems: A Holistic Survey
di: Zhou, Wei, et al.
Pubblicazione: (2024)
di: Zhou, Wei, et al.
Pubblicazione: (2024)
CRCL: Causal Representation Consistency Learning for Anomaly Detection in Surveillance Videos
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
Interpretable Human Activity Recognition for Subtle Robbery Detection in Surveillance Videos
di: Leyva, Bryan Jhoan Cazáres, et al.
Pubblicazione: (2026)
di: Leyva, Bryan Jhoan Cazáres, et al.
Pubblicazione: (2026)
LiteFrame: Efficient Vision Encoders Unlock Frame Scaling in Video LLMs
di: Kim, Jihwan, et al.
Pubblicazione: (2026)
di: Kim, Jihwan, et al.
Pubblicazione: (2026)
MME-VideoOCR: Evaluating OCR-Based Capabilities of Multimodal LLMs in Video Scenarios
di: Shi, Yang, et al.
Pubblicazione: (2025)
di: Shi, Yang, et al.
Pubblicazione: (2025)
ForeSea: AI Forensic Search with Multi-modal Queries for Video Surveillance
di: Park, Hyojin, et al.
Pubblicazione: (2026)
di: Park, Hyojin, et al.
Pubblicazione: (2026)
Clustering Aided Weakly Supervised Training to Detect Anomalous Events in Surveillance Videos
di: Zaheer, Muhammad Zaigham, et al.
Pubblicazione: (2022)
di: Zaheer, Muhammad Zaigham, et al.
Pubblicazione: (2022)
Flying Bird Object Detection Algorithm in Surveillance Video Based on Motion Information
di: Sun, Ziwei, et al.
Pubblicazione: (2023)
di: Sun, Ziwei, et al.
Pubblicazione: (2023)
FBD-SV-2024: Flying Bird Object Detection Dataset in Surveillance Video
di: Sun, Zi-Wei, et al.
Pubblicazione: (2024)
di: Sun, Zi-Wei, et al.
Pubblicazione: (2024)
Enhancing Temporal Understanding in Video-LLMs through Stacked Temporal Attention in Vision Encoders
di: Rasekh, Ali, et al.
Pubblicazione: (2025)
di: Rasekh, Ali, et al.
Pubblicazione: (2025)
Recognition of Abnormal Events in Surveillance Videos using Weakly Supervised Dual-Encoder Models
di: Tsfaty, Noam, et al.
Pubblicazione: (2025)
di: Tsfaty, Noam, et al.
Pubblicazione: (2025)
Distributed Intelligent Video Surveillance for Early Armed Robbery Detection based on Deep Learning
di: Fernandez-Testa, Sergio, et al.
Pubblicazione: (2024)
di: Fernandez-Testa, Sergio, et al.
Pubblicazione: (2024)
MTFL: Multi-Timescale Feature Learning for Weakly-Supervised Anomaly Detection in Surveillance Videos
di: Zhang, Yiling, et al.
Pubblicazione: (2024)
di: Zhang, Yiling, et al.
Pubblicazione: (2024)
Accelerated Event-Based Feature Detection and Compression for Surveillance Video Systems
di: Freeman, Andrew C., et al.
Pubblicazione: (2023)
di: Freeman, Andrew C., et al.
Pubblicazione: (2023)
LVOmniBench: Pioneering Long Audio-Video Understanding Evaluation for Omnimodal LLMs
di: Tao, Keda, et al.
Pubblicazione: (2026)
di: Tao, Keda, et al.
Pubblicazione: (2026)
Video Forgery Detection for Surveillance Cameras: A Review
di: Tayfor, Noor B., et al.
Pubblicazione: (2025)
di: Tayfor, Noor B., et al.
Pubblicazione: (2025)
Few-shot Semantic Encoding and Decoding for Video Surveillance
di: Cheng, Baoping, et al.
Pubblicazione: (2025)
di: Cheng, Baoping, et al.
Pubblicazione: (2025)
ProDisc-VAD: An Efficient System for Weakly-Supervised Anomaly Detection in Video Surveillance Applications
di: Zhu, Tao, et al.
Pubblicazione: (2025)
di: Zhu, Tao, et al.
Pubblicazione: (2025)
Video-based Vehicle Surveillance in the Wild: License Plate, Make, and Model Recognition with Self Reflective Vision-Language Models
di: Parsa, Pouya, et al.
Pubblicazione: (2025)
di: Parsa, Pouya, et al.
Pubblicazione: (2025)
Two-Pass Zero-Shot Temporal-Spatial Grounding of Rare Traffic Events in Surveillance Video
di: Huang, Jiantang
Pubblicazione: (2026)
di: Huang, Jiantang
Pubblicazione: (2026)
How Important are Videos for Training Video LLMs?
di: Lydakis, George, et al.
Pubblicazione: (2025)
di: Lydakis, George, et al.
Pubblicazione: (2025)
Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
di: Fu, Chaoyou, et al.
Pubblicazione: (2024)
di: Fu, Chaoyou, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Assessing Situational and Spatial Awareness of VLMs with Synthetically Generated Video
di: Benschop, Pascal, et al.
Pubblicazione: (2026) -
Compositional Scene Understanding through Inverse Generative Modeling
di: Wang, Yanbo, et al.
Pubblicazione: (2025) -
Slot-VAE: Object-Centric Scene Generation with Slot Attention
di: Wang, Yanbo, et al.
Pubblicazione: (2023) -
Identifying Ethical Biases in Action Recognition Models
di: Baltaretu, Ana, et al.
Pubblicazione: (2026) -
Object-Centric Temporal Consistency via Conditional Autoregressive Inductive Biases
di: Meo, Cristian, et al.
Pubblicazione: (2024)