Video Anomaly Detection and Explanation via Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Lv, Hui, Sun, Qianru |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Harnessing Large Language Models for Training-free Video Anomaly Detection
di: Zanella, Luca, et al.
Pubblicazione: (2024)
di: Zanella, Luca, et al.
Pubblicazione: (2024)
Follow the Rules: Reasoning for Video Anomaly Detection with Large Language Models
di: Yang, Yuchen, et al.
Pubblicazione: (2024)
di: Yang, Yuchen, et al.
Pubblicazione: (2024)
MoniTor: Exploiting Large Language Models with Instruction for Online Video Anomaly Detection
di: Yang, Shengtian, et al.
Pubblicazione: (2025)
di: Yang, Shengtian, et al.
Pubblicazione: (2025)
Exploring Large Vision-Language Models for Robust and Efficient Industrial Anomaly Detection
di: Qian, Kun, et al.
Pubblicazione: (2024)
di: Qian, Kun, et al.
Pubblicazione: (2024)
Weakly-Supervised Semantic Segmentation with Image-Level Labels: from Traditional Models to Foundation Models
di: Chen, Zhaozheng, et al.
Pubblicazione: (2023)
di: Chen, Zhaozheng, et al.
Pubblicazione: (2023)
Vision-Language Models Assisted Unsupervised Video Anomaly Detection
di: Jiang, Yalong, et al.
Pubblicazione: (2024)
di: Jiang, Yalong, et al.
Pubblicazione: (2024)
Aligning Effective Tokens with Video Anomaly in Large Language Models
di: Chen, Yingxian, et al.
Pubblicazione: (2025)
di: Chen, Yingxian, et al.
Pubblicazione: (2025)
Frame-Voyager: Learning to Query Frames for Video Large Language Models
di: Yu, Sicheng, et al.
Pubblicazione: (2024)
di: Yu, Sicheng, et al.
Pubblicazione: (2024)
AF-CLIP: Zero-Shot Anomaly Detection via Anomaly-Focused CLIP Adaptation
di: Fang, Qingqing, et al.
Pubblicazione: (2025)
di: Fang, Qingqing, et al.
Pubblicazione: (2025)
Cerberus: Real-Time Video Anomaly Detection via Cascaded Vision-Language Models
di: Zheng, Yue, et al.
Pubblicazione: (2025)
di: Zheng, Yue, et al.
Pubblicazione: (2025)
Open-Vocabulary Video Anomaly Detection
di: Wu, Peng, et al.
Pubblicazione: (2023)
di: Wu, Peng, et al.
Pubblicazione: (2023)
3D Question Answering via only 2D Vision-Language Models
di: Wang, Fengyun, et al.
Pubblicazione: (2025)
di: Wang, Fengyun, et al.
Pubblicazione: (2025)
VADMamba++: Efficient Video Anomaly Detection via Hybrid Modeling in Grayscale Space
di: Lyu, Jihao, et al.
Pubblicazione: (2026)
di: Lyu, Jihao, et al.
Pubblicazione: (2026)
VADER: Towards Causal Video Anomaly Understanding with Relation-Aware Large Language Models
di: Cheng, Ying, et al.
Pubblicazione: (2025)
di: Cheng, Ying, et al.
Pubblicazione: (2025)
MVP: Enhancing Video Large Language Models via Self-supervised Masked Video Prediction
di: Sun, Xiaokun, et al.
Pubblicazione: (2026)
di: Sun, Xiaokun, et al.
Pubblicazione: (2026)
Forward Consistency Learning with Gated Context Aggregation for Video Anomaly Detection
di: Lyu, Jiahao, et al.
Pubblicazione: (2026)
di: Lyu, Jiahao, et al.
Pubblicazione: (2026)
Simplifying Traffic Anomaly Detection with Video Foundation Models
di: Orlova, Svetlana, et al.
Pubblicazione: (2025)
di: Orlova, Svetlana, et al.
Pubblicazione: (2025)
Unlocking Vision-Language Models for Video Anomaly Detection via Fine-Grained Prompting
di: Zou, Shu, et al.
Pubblicazione: (2025)
di: Zou, Shu, et al.
Pubblicazione: (2025)
SlowFastVAD: Video Anomaly Detection via Integrating Simple Detector and RAG-Enhanced Vision-Language Model
di: Ding, Zongcan, et al.
Pubblicazione: (2025)
di: Ding, Zongcan, et al.
Pubblicazione: (2025)
Generalized Visual Relation Detection with Diffusion Models
di: Gao, Kaifeng, et al.
Pubblicazione: (2025)
di: Gao, Kaifeng, et al.
Pubblicazione: (2025)
Can Multimodal Large Language Models be Guided to Improve Industrial Anomaly Detection?
di: Chen, Zhiling, et al.
Pubblicazione: (2025)
di: Chen, Zhiling, et al.
Pubblicazione: (2025)
ColorGPT: Leveraging Large Language Models for Multimodal Color Recommendation
di: Xia, Ding, et al.
Pubblicazione: (2025)
di: Xia, Ding, et al.
Pubblicazione: (2025)
VideoLLM-online: Online Video Large Language Model for Streaming Video
di: Chen, Joya, et al.
Pubblicazione: (2024)
di: Chen, Joya, et al.
Pubblicazione: (2024)
SmartHome-Bench: A Comprehensive Benchmark for Video Anomaly Detection in Smart Homes Using Multi-Modal Large Language Models
di: Zhao, Xinyi, et al.
Pubblicazione: (2025)
di: Zhao, Xinyi, et al.
Pubblicazione: (2025)
Sparse Reasoning is Enough: Biological-Inspired Framework for Video Anomaly Detection with Large Pre-trained Models
di: Huang, He, et al.
Pubblicazione: (2025)
di: Huang, He, et al.
Pubblicazione: (2025)
VMAD: Visual-enhanced Multimodal Large Language Model for Zero-Shot Anomaly Detection
di: Deng, Huilin, et al.
Pubblicazione: (2024)
di: Deng, Huilin, et al.
Pubblicazione: (2024)
Detect, Classify, Act: Categorizing Industrial Anomalies with Multi-Modal Large Language Models
di: Mokhtar, Sassan, et al.
Pubblicazione: (2025)
di: Mokhtar, Sassan, et al.
Pubblicazione: (2025)
Topo-R1: Detecting Topological Anomalies via Vision-Language Models
di: Xu, Meilong, et al.
Pubblicazione: (2026)
di: Xu, Meilong, et al.
Pubblicazione: (2026)
Language-guided Open-world Video Anomaly Detection under Weak Supervision
di: Liu, Zihao, et al.
Pubblicazione: (2025)
di: Liu, Zihao, et al.
Pubblicazione: (2025)
Towards Artwork Explanation in Large-scale Vision Language Models
di: Hayashi, Kazuki, et al.
Pubblicazione: (2024)
di: Hayashi, Kazuki, et al.
Pubblicazione: (2024)
Appearance Blur-driven AutoEncoder and Motion-guided Memory Module for Video Anomaly Detection
di: Lyu, Jiahao, et al.
Pubblicazione: (2024)
di: Lyu, Jiahao, et al.
Pubblicazione: (2024)
Unsupervised Visual Chain-of-Thought Reasoning via Preference Optimization
di: Zhao, Kesen, et al.
Pubblicazione: (2025)
di: Zhao, Kesen, et al.
Pubblicazione: (2025)
Learn Suspected Anomalies from Event Prompts for Video Anomaly Detection
di: Tao, Chenchen, et al.
Pubblicazione: (2024)
di: Tao, Chenchen, et al.
Pubblicazione: (2024)
AnyAnomaly: Zero-Shot Customizable Video Anomaly Detection with LVLM
di: Ahn, Sunghyun, et al.
Pubblicazione: (2025)
di: Ahn, Sunghyun, et al.
Pubblicazione: (2025)
TUBench: Benchmarking Large Vision-Language Models on Trustworthiness with Unanswerable Questions
di: He, Xingwei, et al.
Pubblicazione: (2024)
di: He, Xingwei, et al.
Pubblicazione: (2024)
Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
di: Xu, Jiacong, et al.
Pubblicazione: (2025)
di: Xu, Jiacong, et al.
Pubblicazione: (2025)
A Unified Anomaly Synthesis Strategy with Gradient Ascent for Industrial Anomaly Detection and Localization
di: Chen, Qiyu, et al.
Pubblicazione: (2024)
di: Chen, Qiyu, et al.
Pubblicazione: (2024)
Generalized Logit Adjustment: Calibrating Fine-tuned Models by Removing Label Bias in Foundation Models
di: Zhu, Beier, et al.
Pubblicazione: (2023)
di: Zhu, Beier, et al.
Pubblicazione: (2023)
Video Anomaly Detection with Motion and Appearance Guided Patch Diffusion Model
di: Zhou, Hang, et al.
Pubblicazione: (2024)
di: Zhou, Hang, et al.
Pubblicazione: (2024)
Bounding Boxes and Probabilistic Graphical Models: Video Anomaly Detection Simplified
di: Siemon, Mia, et al.
Pubblicazione: (2024)
di: Siemon, Mia, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Harnessing Large Language Models for Training-free Video Anomaly Detection
di: Zanella, Luca, et al.
Pubblicazione: (2024) -
Follow the Rules: Reasoning for Video Anomaly Detection with Large Language Models
di: Yang, Yuchen, et al.
Pubblicazione: (2024) -
MoniTor: Exploiting Large Language Models with Instruction for Online Video Anomaly Detection
di: Yang, Shengtian, et al.
Pubblicazione: (2025) -
Exploring Large Vision-Language Models for Robust and Efficient Industrial Anomaly Detection
di: Qian, Kun, et al.
Pubblicazione: (2024) -
Weakly-Supervised Semantic Segmentation with Image-Level Labels: from Traditional Models to Foundation Models
di: Chen, Zhaozheng, et al.
Pubblicazione: (2023)