Human-Centric Video Anomaly Detection Through Spatio-Temporal Pose Tokenization and Transformer
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Noghre, Ghazal Alinezhad, Pazho, Armin Danesh, Tabkhi, Hamed |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
An Exploratory Study on Human-Centric Video Anomaly Detection through Variational Autoencoders and Trajectory Prediction
par: Noghre, Ghazal Alinezhad, et autres
Publié: (2024)
par: Noghre, Ghazal Alinezhad, et autres
Publié: (2024)
A Survey on Video Anomaly Detection via Deep Learning: Human, Vehicle, and Environment
par: Noghre, Ghazal Alinezhad, et autres
Publié: (2025)
par: Noghre, Ghazal Alinezhad, et autres
Publié: (2025)
Evaluating the Effectiveness of Video Anomaly Detection in the Wild: Online Learning and Inference for Real-world Deployment
par: Yao, Shanle, et autres
Publié: (2024)
par: Yao, Shanle, et autres
Publié: (2024)
Exploring Pose-Based Anomaly Detection for Retail Security: A Real-World Shoplifting Dataset and Benchmark
par: Rashvand, Narges, et autres
Publié: (2025)
par: Rashvand, Narges, et autres
Publié: (2025)
Adversarially-Refined VQ-GAN with Dense Motion Tokenization for Spatio-Temporal Heatmaps
par: Maldonado, Gabriel, et autres
Publié: (2025)
par: Maldonado, Gabriel, et autres
Publié: (2025)
VT-Former: An Exploratory Study on Vehicle Trajectory Prediction for Highway Surveillance through Graph Isomorphism and Transformer
par: Pazho, Armin Danesh, et autres
Publié: (2023)
par: Pazho, Armin Danesh, et autres
Publié: (2023)
Towards Adaptive Human-centric Video Anomaly Detection: A Comprehensive Framework and A New Benchmark
par: Pazho, Armin Danesh, et autres
Publié: (2024)
par: Pazho, Armin Danesh, et autres
Publié: (2024)
ALFred: An Active Learning Framework for Real-world Semi-supervised Anomaly Detection with Adaptive Thresholds
par: Yao, Shanle, et autres
Publié: (2025)
par: Yao, Shanle, et autres
Publié: (2025)
Shopformer: Transformer-Based Framework for Detecting Shoplifting via Human Pose
par: Rashvand, Narges, et autres
Publié: (2025)
par: Rashvand, Narges, et autres
Publié: (2025)
MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation
par: Maldonado, Gabriel, et autres
Publié: (2025)
par: Maldonado, Gabriel, et autres
Publié: (2025)
MoFM: A Large-Scale Human Motion Foundation Model
par: Baharani, Mohammadreza, et autres
Publié: (2025)
par: Baharani, Mohammadreza, et autres
Publié: (2025)
From Lab to Field: Real-World Evaluation of an AI-Driven Smart Video Solution to Enhance Community Safety
par: Yao, Shanle, et autres
Publié: (2023)
par: Yao, Shanle, et autres
Publié: (2023)
Are Multimodal LLMs Ready for Surveillance? A Reality Check on Zero-Shot Anomaly Detection in the Wild
par: Yao, Shanle, et autres
Publié: (2026)
par: Yao, Shanle, et autres
Publié: (2026)
From Frames to Events: Rethinking Evaluation in Human-Centric Video Anomaly Detection
par: Rashvand, Narges, et autres
Publié: (2026)
par: Rashvand, Narges, et autres
Publié: (2026)
From Offline to Periodic Adaptation for Pose-Based Shoplifting Detection in Real-world Retail Security
par: Yao, Shanle, et autres
Publié: (2026)
par: Yao, Shanle, et autres
Publié: (2026)
Weakly Supervised Video Anomaly Detection and Localization with Spatio-Temporal Prompts
par: Wu, Peng, et autres
Publié: (2024)
par: Wu, Peng, et autres
Publié: (2024)
Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models
par: Cao, Tri, et autres
Publié: (2026)
par: Cao, Tri, et autres
Publié: (2026)
Video Anomaly Detection via Spatio-Temporal Pseudo-Anomaly Generation : A Unified Approach
par: Rai, Ayush K., et autres
Publié: (2023)
par: Rai, Ayush K., et autres
Publié: (2023)
Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs
par: Hyun, Jeongseok, et autres
Publié: (2025)
par: Hyun, Jeongseok, et autres
Publié: (2025)
Unified Spatio-Temporal Token Scoring for Efficient Video VLMs
par: Zhang, Jianrui, et autres
Publié: (2026)
par: Zhang, Jianrui, et autres
Publié: (2026)
Human-Centric Anomaly Detection in Surveillance Videos Using YOLO-World and Spatio-Temporal Deep Learning
par: Naeen, Mohammad Ali Etemadi, et autres
Publié: (2025)
par: Naeen, Mohammad Ali Etemadi, et autres
Publié: (2025)
Spatio-Temporal Token Pruning for Efficient High-Resolution GUI Agents
par: Xu, Zhou, et autres
Publié: (2026)
par: Xu, Zhou, et autres
Publié: (2026)
STAA: Spatio-Temporal Attention Attribution for Real-Time Interpreting Transformer-based Video Models
par: Wang, Zerui, et autres
Publié: (2024)
par: Wang, Zerui, et autres
Publié: (2024)
Knowing Your Target: Target-Aware Transformer Makes Better Spatio-Temporal Video Grounding
par: Gu, Xin, et autres
Publié: (2025)
par: Gu, Xin, et autres
Publié: (2025)
Refined Temporal Pyramidal Compression-and-Amplification Transformer for 3D Human Pose Estimation
par: Liu, Hanbing, et autres
Publié: (2023)
par: Liu, Hanbing, et autres
Publié: (2023)
SAM-PM: Enhancing Video Camouflaged Object Detection using Spatio-Temporal Attention
par: Meeran, Muhammad Nawfal, et autres
Publié: (2024)
par: Meeran, Muhammad Nawfal, et autres
Publié: (2024)
HumanVideo-MME: Benchmarking MLLMs for Human-Centric Video Understanding
par: Cai, Yuxuan, et autres
Publié: (2025)
par: Cai, Yuxuan, et autres
Publié: (2025)
OCT-SelfNet: A Self-Supervised Framework with Multi-Modal Datasets for Generalized and Robust Retinal Disease Detection
par: Jannat, Fatema-E, et autres
Publié: (2024)
par: Jannat, Fatema-E, et autres
Publié: (2024)
AvatarShield: Visual Reinforcement Learning for Human-Centric Synthetic Video Detection
par: Xu, Zhipei, et autres
Publié: (2025)
par: Xu, Zhipei, et autres
Publié: (2025)
H$_{2}$OT: Hierarchical Hourglass Tokenizer for Efficient Video Pose Transformers
par: Li, Wenhao, et autres
Publié: (2025)
par: Li, Wenhao, et autres
Publié: (2025)
NCSTR: Node-Centric Decoupled Spatio-Temporal Reasoning for Video-based Human Pose Estimation
par: Huynh, Quang Dang, et autres
Publié: (2026)
par: Huynh, Quang Dang, et autres
Publié: (2026)
B-VLLM: A Vision Large Language Model with Balanced Spatio-Temporal Tokens
par: Lu, Zhuqiang, et autres
Publié: (2024)
par: Lu, Zhuqiang, et autres
Publié: (2024)
Hourglass Tokenizer for Efficient Transformer-Based 3D Human Pose Estimation
par: Li, Wenhao, et autres
Publié: (2023)
par: Li, Wenhao, et autres
Publié: (2023)
TimeBlind: A Spatio-Temporal Compositionality Benchmark for Video LLMs
par: Li, Baiqi, et autres
Publié: (2026)
par: Li, Baiqi, et autres
Publié: (2026)
DropletVideo: A Dataset and Approach to Explore Integral Spatio-Temporal Consistent Video Generation
par: Zhang, Runze, et autres
Publié: (2025)
par: Zhang, Runze, et autres
Publié: (2025)
VideoStir: Understanding Long Videos via Spatio-Temporally Structured and Intent-Aware RAG
par: Fu, Honghao, et autres
Publié: (2026)
par: Fu, Honghao, et autres
Publié: (2026)
Spatio-Temporal Context Prompting for Zero-Shot Action Detection
par: Huang, Wei-Jhe, et autres
Publié: (2024)
par: Huang, Wei-Jhe, et autres
Publié: (2024)
ST-Prune: Training-Free Spatio-Temporal Token Pruning for Vision-Language Models in Autonomous Driving
par: Sha, Lin, et autres
Publié: (2026)
par: Sha, Lin, et autres
Publié: (2026)
Mining Multi-Modality Spatio-Temporal Cues for Video Important Person Identification
par: Wang, Xiao, et autres
Publié: (2026)
par: Wang, Xiao, et autres
Publié: (2026)
ToG-Bench: Task-Oriented Spatio-Temporal Grounding in Egocentric Videos
par: Xu, Qi'ao, et autres
Publié: (2025)
par: Xu, Qi'ao, et autres
Publié: (2025)
Documents similaires
-
An Exploratory Study on Human-Centric Video Anomaly Detection through Variational Autoencoders and Trajectory Prediction
par: Noghre, Ghazal Alinezhad, et autres
Publié: (2024) -
A Survey on Video Anomaly Detection via Deep Learning: Human, Vehicle, and Environment
par: Noghre, Ghazal Alinezhad, et autres
Publié: (2025) -
Evaluating the Effectiveness of Video Anomaly Detection in the Wild: Online Learning and Inference for Real-world Deployment
par: Yao, Shanle, et autres
Publié: (2024) -
Exploring Pose-Based Anomaly Detection for Retail Security: A Real-World Shoplifting Dataset and Benchmark
par: Rashvand, Narges, et autres
Publié: (2025) -
Adversarially-Refined VQ-GAN with Dense Motion Tokenization for Spatio-Temporal Heatmaps
par: Maldonado, Gabriel, et autres
Publié: (2025)