STF: Spatio-Temporal Fusion Module for Improving Video Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Anwar, Noreen, Bilodeau, Guillaume-Alexandre, Bouachir, Wassim |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dual-Stream Attention with Multi-Modal Queries for Object Detection in Transportation Applications
by: Anwar, Noreen, et al.
Published: (2025)
by: Anwar, Noreen, et al.
Published: (2025)
Detection of Micromobility Vehicles in Urban Traffic Videos
by: Sabri, Khalil, et al.
Published: (2024)
by: Sabri, Khalil, et al.
Published: (2024)
ReL-SAR: Representation Learning for Skeleton Action Recognition with Convolutional Transformers and BYOL
by: Naimi, Safwen, et al.
Published: (2024)
by: Naimi, Safwen, et al.
Published: (2024)
Suicide Risk Assessment from AI-powered Video Surveillance: An Interpretable Framework for Prevention in Metro Stations
by: Naimi, Safwen, et al.
Published: (2026)
by: Naimi, Safwen, et al.
Published: (2026)
InceptoFormer: A Multi-Signal Neural Framework for Parkinson's Disease Severity Evaluation from Gait
by: Naimi, Safwen, et al.
Published: (2025)
by: Naimi, Safwen, et al.
Published: (2025)
STDiff: Spatio-temporal Diffusion for Continuous Stochastic Video Prediction
by: Ye, Xi, et al.
Published: (2023)
by: Ye, Xi, et al.
Published: (2023)
Learning Data Association for Multi-Object Tracking using Only Coordinates
by: Miah, Mehdi, et al.
Published: (2024)
by: Miah, Mehdi, et al.
Published: (2024)
STF: Spatial Temporal Fusion for Trajectory Prediction
by: Han, Pengqian, et al.
Published: (2023)
by: Han, Pengqian, et al.
Published: (2023)
VideoFusion: A Spatio-Temporal Collaborative Network for Multi-modal Video Fusion
by: Tang, Linfeng, et al.
Published: (2025)
by: Tang, Linfeng, et al.
Published: (2025)
CenterDisks: Real-time instance segmentation with disk covering
by: Litto, Katia Jodogne-Del, et al.
Published: (2024)
by: Litto, Katia Jodogne-Del, et al.
Published: (2024)
How good are deep learning methods for automated road safety analysis using video data? An experimental study
by: Liu, Qingwu, et al.
Published: (2025)
by: Liu, Qingwu, et al.
Published: (2025)
PMMA: The Polytechnique Montreal Mobility Aids Dataset
by: Liu, Qingwu, et al.
Published: (2026)
by: Liu, Qingwu, et al.
Published: (2026)
Beyond Boxes: Mask-Guided Spatio-Temporal Feature Aggregation for Video Object Detection
by: Hashmi, Khurram Azeem, et al.
Published: (2024)
by: Hashmi, Khurram Azeem, et al.
Published: (2024)
Detection of Autonomous Shuttles in Urban Traffic Images Using Adaptive Residual Context
by: Younes, Mohamed Aziz, et al.
Published: (2026)
by: Younes, Mohamed Aziz, et al.
Published: (2026)
OmniSTVG: Toward Spatio-Temporal Omni-Object Video Grounding
by: Yao, Jiali, et al.
Published: (2025)
by: Yao, Jiali, et al.
Published: (2025)
Beyond Pixels: Leveraging the Language of Soccer to Improve Spatio-Temporal Action Detection in Broadcast Videos
by: Ochin, Jeremie, et al.
Published: (2025)
by: Ochin, Jeremie, et al.
Published: (2025)
Patch Spatio-Temporal Relation Prediction for Video Anomaly Detection
by: Shen, Hao, et al.
Published: (2024)
by: Shen, Hao, et al.
Published: (2024)
Equivariant Spatio-Temporal Self-Supervision for LiDAR Object Detection
by: Hegde, Deepti, et al.
Published: (2024)
by: Hegde, Deepti, et al.
Published: (2024)
SAM-PM: Enhancing Video Camouflaged Object Detection using Spatio-Temporal Attention
by: Meeran, Muhammad Nawfal, et al.
Published: (2024)
by: Meeran, Muhammad Nawfal, et al.
Published: (2024)
Standardization for improved Spatio-Temporal Image Fusion
by: Goyena, Harkaitz, et al.
Published: (2025)
by: Goyena, Harkaitz, et al.
Published: (2025)
Vulnerability-Aware Spatio-Temporal Learning for Generalizable Deepfake Video Detection
by: Nguyen, Dat, et al.
Published: (2025)
by: Nguyen, Dat, et al.
Published: (2025)
JSTR: Joint Spatio-Temporal Reasoning for Event-based Moving Object Detection
by: Zhou, Hanyu, et al.
Published: (2024)
by: Zhou, Hanyu, et al.
Published: (2024)
STAF: 3D Human Mesh Recovery from Video with Spatio-Temporal Alignment Fusion
by: Yao, Wei, et al.
Published: (2024)
by: Yao, Wei, et al.
Published: (2024)
Student Classroom Behavior Detection based on Spatio-Temporal Network and Multi-Model Fusion
by: Yang, Fan, et al.
Published: (2023)
by: Yang, Fan, et al.
Published: (2023)
Mamba-based Spatio-Frequency Motion Perception for Video Camouflaged Object Detection
by: Li, Xin, et al.
Published: (2025)
by: Li, Xin, et al.
Published: (2025)
Context-Guided Spatio-Temporal Video Grounding
by: Gu, Xin, et al.
Published: (2024)
by: Gu, Xin, et al.
Published: (2024)
DVFL-Net: A Lightweight Distilled Video Focal Modulation Network for Spatio-Temporal Action Recognition
by: Ullah, Hayat, et al.
Published: (2025)
by: Ullah, Hayat, et al.
Published: (2025)
Improving Token-based Object Detection with Video
by: Singh, Abhineet, et al.
Published: (2025)
by: Singh, Abhineet, et al.
Published: (2025)
AttentiveGRU: Recurrent Spatio-Temporal Modeling for Advanced Radar-Based BEV Object Detection
by: Saini, Loveneet, et al.
Published: (2025)
by: Saini, Loveneet, et al.
Published: (2025)
Multimodal Spatio-temporal Graph Learning for Alignment-free RGBT Video Object Detection
by: Wang, Qishun, et al.
Published: (2025)
by: Wang, Qishun, et al.
Published: (2025)
Rethinking Early-Fusion Strategies for Improved Multispectral Object Detection
by: Zhang, Xue, et al.
Published: (2024)
by: Zhang, Xue, et al.
Published: (2024)
ACTrack: Adding Spatio-Temporal Condition for Visual Object Tracking
by: Han, Yushan, et al.
Published: (2024)
by: Han, Yushan, et al.
Published: (2024)
Towards Long-Form Spatio-Temporal Video Grounding
by: Gu, Xin, et al.
Published: (2026)
by: Gu, Xin, et al.
Published: (2026)
VISTA: Video Interaction Spatio-Temporal Analysis Benchmark
by: Aparcedo, Alejandro, et al.
Published: (2026)
by: Aparcedo, Alejandro, et al.
Published: (2026)
VideoMolmo: Spatio-Temporal Grounding Meets Pointing
by: Ahmad, Ghazi Shazan, et al.
Published: (2025)
by: Ahmad, Ghazi Shazan, et al.
Published: (2025)
Learning Spatio-Temporal Feature Representations for Video-Based Gaze Estimation
by: Personnic, Alexandre, et al.
Published: (2025)
by: Personnic, Alexandre, et al.
Published: (2025)
CRT-Fusion: Camera, Radar, Temporal Fusion Using Motion Information for 3D Object Detection
by: Kim, Jisong, et al.
Published: (2024)
by: Kim, Jisong, et al.
Published: (2024)
Open-Vocabulary Spatio-Temporal Action Detection
by: Wu, Tao, et al.
Published: (2024)
by: Wu, Tao, et al.
Published: (2024)
Deepfake Detection with Spatio-Temporal Consistency and Attention
by: Chen, Yunzhuo, et al.
Published: (2025)
by: Chen, Yunzhuo, et al.
Published: (2025)
Co-Fusion4D: Spatio-temporal Collaborative Fusion for Robust 3D Object Detection
by: Li, Wenxuan, et al.
Published: (2026)
by: Li, Wenxuan, et al.
Published: (2026)
Similar Items
-
Dual-Stream Attention with Multi-Modal Queries for Object Detection in Transportation Applications
by: Anwar, Noreen, et al.
Published: (2025) -
Detection of Micromobility Vehicles in Urban Traffic Videos
by: Sabri, Khalil, et al.
Published: (2024) -
ReL-SAR: Representation Learning for Skeleton Action Recognition with Convolutional Transformers and BYOL
by: Naimi, Safwen, et al.
Published: (2024) -
Suicide Risk Assessment from AI-powered Video Surveillance: An Interpretable Framework for Prevention in Metro Stations
by: Naimi, Safwen, et al.
Published: (2026) -
InceptoFormer: A Multi-Signal Neural Framework for Parkinson's Disease Severity Evaluation from Gait
by: Naimi, Safwen, et al.
Published: (2025)