Adversarially-Refined VQ-GAN with Dense Motion Tokenization for Spatio-Temporal Heatmaps
Fuente:
arXiv
Saved in:
| Main Authors: | Maldonado, Gabriel, Rashvand, Narges, Pazho, Armin Danesh, Noghre, Ghazal Alinezhad, Katariya, Vinit, Tabkhi, Hamed |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation
by: Maldonado, Gabriel, et al.
Published: (2025)
by: Maldonado, Gabriel, et al.
Published: (2025)
Human-Centric Video Anomaly Detection Through Spatio-Temporal Pose Tokenization and Transformer
by: Noghre, Ghazal Alinezhad, et al.
Published: (2024)
by: Noghre, Ghazal Alinezhad, et al.
Published: (2024)
VT-Former: An Exploratory Study on Vehicle Trajectory Prediction for Highway Surveillance through Graph Isomorphism and Transformer
by: Pazho, Armin Danesh, et al.
Published: (2023)
by: Pazho, Armin Danesh, et al.
Published: (2023)
Exploring Pose-Based Anomaly Detection for Retail Security: A Real-World Shoplifting Dataset and Benchmark
by: Rashvand, Narges, et al.
Published: (2025)
by: Rashvand, Narges, et al.
Published: (2025)
Shopformer: Transformer-Based Framework for Detecting Shoplifting via Human Pose
by: Rashvand, Narges, et al.
Published: (2025)
by: Rashvand, Narges, et al.
Published: (2025)
A Survey on Video Anomaly Detection via Deep Learning: Human, Vehicle, and Environment
by: Noghre, Ghazal Alinezhad, et al.
Published: (2025)
by: Noghre, Ghazal Alinezhad, et al.
Published: (2025)
An Exploratory Study on Human-Centric Video Anomaly Detection through Variational Autoencoders and Trajectory Prediction
by: Noghre, Ghazal Alinezhad, et al.
Published: (2024)
by: Noghre, Ghazal Alinezhad, et al.
Published: (2024)
MoFM: A Large-Scale Human Motion Foundation Model
by: Baharani, Mohammadreza, et al.
Published: (2025)
by: Baharani, Mohammadreza, et al.
Published: (2025)
Towards Adaptive Human-centric Video Anomaly Detection: A Comprehensive Framework and A New Benchmark
by: Pazho, Armin Danesh, et al.
Published: (2024)
by: Pazho, Armin Danesh, et al.
Published: (2024)
ALFred: An Active Learning Framework for Real-world Semi-supervised Anomaly Detection with Adaptive Thresholds
by: Yao, Shanle, et al.
Published: (2025)
by: Yao, Shanle, et al.
Published: (2025)
Evaluating the Effectiveness of Video Anomaly Detection in the Wild: Online Learning and Inference for Real-world Deployment
by: Yao, Shanle, et al.
Published: (2024)
by: Yao, Shanle, et al.
Published: (2024)
Are Multimodal LLMs Ready for Surveillance? A Reality Check on Zero-Shot Anomaly Detection in the Wild
by: Yao, Shanle, et al.
Published: (2026)
by: Yao, Shanle, et al.
Published: (2026)
From Offline to Periodic Adaptation for Pose-Based Shoplifting Detection in Real-world Retail Security
by: Yao, Shanle, et al.
Published: (2026)
by: Yao, Shanle, et al.
Published: (2026)
From Frames to Events: Rethinking Evaluation in Human-Centric Video Anomaly Detection
by: Rashvand, Narges, et al.
Published: (2026)
by: Rashvand, Narges, et al.
Published: (2026)
From Lab to Field: Real-World Evaluation of an AI-Driven Smart Video Solution to Enhance Community Safety
by: Yao, Shanle, et al.
Published: (2023)
by: Yao, Shanle, et al.
Published: (2023)
Distributed learning for automatic modulation recognition in bandwidth-limited networks
by: Rashvand, Narges, et al.
Published: (2025)
by: Rashvand, Narges, et al.
Published: (2025)
Enhancing Automatic Modulation Recognition for IoT Applications Using Transformers
by: Rashvand, Narges, et al.
Published: (2024)
by: Rashvand, Narges, et al.
Published: (2024)
EdgeVTP: Exploration of Latency-efficient Trajectory Prediction for Edge-based Embedded Vision Applications
by: Kim, Seungjin, et al.
Published: (2026)
by: Kim, Seungjin, et al.
Published: (2026)
Intelligent CCTV for Urban Design: AI-Based Analysis of Soft Infrastructure at Intersections
by: Katariya, Vinit, et al.
Published: (2026)
by: Katariya, Vinit, et al.
Published: (2026)
Real-Time Bus Departure Prediction Using Neural Networks for Smart IoT Public Bus Transit
by: Rashvand, Narges, et al.
Published: (2025)
by: Rashvand, Narges, et al.
Published: (2025)
Real-Time Bus Arrival Prediction: A Deep Learning Approach for Enhanced Urban Mobility
by: Rashvand, Narges, et al.
Published: (2023)
by: Rashvand, Narges, et al.
Published: (2023)
Transformer RGBT Tracking with Spatio-Temporal Multimodal Tokens
by: Sun, Dengdi, et al.
Published: (2024)
by: Sun, Dengdi, et al.
Published: (2024)
Refining CNN-based Heatmap Regression with Gradient-based Corner Points for Electrode Localization
by: Wu, Lin
Published: (2024)
by: Wu, Lin
Published: (2024)
HRVGAN: High Resolution Video Generation using Spatio-Temporal GAN
by: Sagar, Abhinav
Published: (2020)
by: Sagar, Abhinav
Published: (2020)
TrackNetV5: Residual-Driven Spatio-Temporal Refinement and Motion Direction Decoupling for Fast Object Tracking
by: Tang, Haonan, et al.
Published: (2025)
by: Tang, Haonan, et al.
Published: (2025)
Spatio-Temporal Branching for Motion Prediction using Motion Increments
by: Wang, Jiexin, et al.
Published: (2023)
by: Wang, Jiexin, et al.
Published: (2023)
Latent Anomaly Detection: Masked VQ-GAN for Unsupervised Segmentation in Medical CBCT
by: Wang, Pengwei
Published: (2025)
by: Wang, Pengwei
Published: (2025)
ODTrack: Online Dense Temporal Token Learning for Visual Tracking
by: Zheng, Yaozong, et al.
Published: (2024)
by: Zheng, Yaozong, et al.
Published: (2024)
Cross-Task Affinity Learning for Multitask Dense Scene Predictions
by: Sinodinos, Dimitrios, et al.
Published: (2024)
by: Sinodinos, Dimitrios, et al.
Published: (2024)
RefineStyle: Dynamic Convolution Refinement for StyleGAN
by: Xia, Siwei, et al.
Published: (2024)
by: Xia, Siwei, et al.
Published: (2024)
Anatomy-Aware Unsupervised Detection and Localization of Retinal Abnormalities in Optical Coherence Tomography
by: Haghighi, Tania, et al.
Published: (2026)
by: Haghighi, Tania, et al.
Published: (2026)
SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer
by: Chen, Hao, et al.
Published: (2024)
by: Chen, Hao, et al.
Published: (2024)
MaskControl: Spatio-Temporal Control for Masked Motion Synthesis
by: Pinyoanuntapong, Ekkasit, et al.
Published: (2024)
by: Pinyoanuntapong, Ekkasit, et al.
Published: (2024)
MoGAN: Improving Motion Quality in Video Diffusion via Few-Step Motion Adversarial Post-Training
by: Xue, Haotian, et al.
Published: (2025)
by: Xue, Haotian, et al.
Published: (2025)
From Sparse to Dense: Spatio-Temporal Fusion for Multi-View 3D Human Pose Estimation with DenseWarper
by: Li, Ling, et al.
Published: (2026)
by: Li, Ling, et al.
Published: (2026)
Unified Spatio-Temporal Token Scoring for Efficient Video VLMs
by: Zhang, Jianrui, et al.
Published: (2026)
by: Zhang, Jianrui, et al.
Published: (2026)
Spatio-Temporal Token Pruning for Efficient High-Resolution GUI Agents
by: Xu, Zhou, et al.
Published: (2026)
by: Xu, Zhou, et al.
Published: (2026)
Spatio-Temporal Difference Guided Motion Deblurring with the Complementary Vision Sensor
by: Meng, Yapeng, et al.
Published: (2026)
by: Meng, Yapeng, et al.
Published: (2026)
Towards Universal Modal Tracking with Online Dense Temporal Token Learning
by: Zheng, Yaozong, et al.
Published: (2025)
by: Zheng, Yaozong, et al.
Published: (2025)
VQ-Seg: Vector-Quantized Token Perturbation for Semi-Supervised Medical Image Segmentation
by: Yang, Sicheng, et al.
Published: (2026)
by: Yang, Sicheng, et al.
Published: (2026)
Similar Items
-
MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation
by: Maldonado, Gabriel, et al.
Published: (2025) -
Human-Centric Video Anomaly Detection Through Spatio-Temporal Pose Tokenization and Transformer
by: Noghre, Ghazal Alinezhad, et al.
Published: (2024) -
VT-Former: An Exploratory Study on Vehicle Trajectory Prediction for Highway Surveillance through Graph Isomorphism and Transformer
by: Pazho, Armin Danesh, et al.
Published: (2023) -
Exploring Pose-Based Anomaly Detection for Retail Security: A Real-World Shoplifting Dataset and Benchmark
by: Rashvand, Narges, et al.
Published: (2025) -
Shopformer: Transformer-Based Framework for Detecting Shoplifting via Human Pose
by: Rashvand, Narges, et al.
Published: (2025)