MambaVT: Spatio-Temporal Contextual Modeling for robust RGB-T Tracking
Fuente:
arXiv
Saved in:
| Main Authors: | Lai, Simiao, Liu, Chang, Zhu, Jiawen, Kang, Ben, Liu, Yang, Wang, Dong, Lu, Huchuan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring Enhanced Contextual Information for Video-Level Object Tracking
by: Kang, Ben, et al.
Published: (2024)
by: Kang, Ben, et al.
Published: (2024)
Unified Sequence-to-Sequence Learning for Single- and Multi-Modal Visual Object Tracking
by: Chen, Xin, et al.
Published: (2023)
by: Chen, Xin, et al.
Published: (2023)
SUTrack: Towards Simple and Unified Single Object Tracking
by: Chen, Xin, et al.
Published: (2024)
by: Chen, Xin, et al.
Published: (2024)
Two-stream Beats One-stream: Asymmetric Siamese Network for Efficient Visual Tracking
by: Zhu, Jiawen, et al.
Published: (2025)
by: Zhu, Jiawen, et al.
Published: (2025)
AVS-Mamba: Exploring Temporal and Multi-modal Mamba for Audio-Visual Segmentation
by: Gong, Sitong, et al.
Published: (2025)
by: Gong, Sitong, et al.
Published: (2025)
SRRT: Exploring Search Region Regulation for Visual Object Tracking
by: Zhu, Jiawen, et al.
Published: (2022)
by: Zhu, Jiawen, et al.
Published: (2022)
Dynamic Subframe Splitting and Spatio-Temporal Motion Entangled Sparse Attention for RGB-E Tracking
by: Shao, Pengcheng, et al.
Published: (2024)
by: Shao, Pengcheng, et al.
Published: (2024)
CADTrack: Learning Contextual Aggregation with Deformable Alignment for Robust RGBT Tracking
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
DefMamba: Deformable Visual State Space Model
by: Liu, Leiye, et al.
Published: (2025)
by: Liu, Leiye, et al.
Published: (2025)
InterMamba: Efficient Human-Human Interaction Generation with Adaptive Spatio-Temporal Mamba
by: Wu, Zizhao, et al.
Published: (2025)
by: Wu, Zizhao, et al.
Published: (2025)
MambaPro: Multi-Modal Object Re-Identification with Mamba Aggregation and Synergistic Prompt
by: Wang, Yuhao, et al.
Published: (2024)
by: Wang, Yuhao, et al.
Published: (2024)
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding
by: Xiong, Haomiao, et al.
Published: (2025)
by: Xiong, Haomiao, et al.
Published: (2025)
UETrack: A Unified and Efficient Framework for Single Object Tracking
by: Kang, Ben, et al.
Published: (2026)
by: Kang, Ben, et al.
Published: (2026)
Exploiting Lightweight Hierarchical ViT and Dynamic Framework for Efficient Visual Tracking
by: Kang, Ben, et al.
Published: (2025)
by: Kang, Ben, et al.
Published: (2025)
ColorMamba: Towards High-quality NIR-to-RGB Spectral Translation with Mamba
by: Zhai, Huiyu, et al.
Published: (2024)
by: Zhai, Huiyu, et al.
Published: (2024)
VideoMamba: Spatio-Temporal Selective State Space Model
by: Park, Jinyoung, et al.
Published: (2024)
by: Park, Jinyoung, et al.
Published: (2024)
Exploring Dynamic Transformer for Efficient Object Tracking
by: Zhu, Jiawen, et al.
Published: (2024)
by: Zhu, Jiawen, et al.
Published: (2024)
Parameter Aware Mamba Model for Multi-task Dense Prediction
by: Yu, Xinzhuo, et al.
Published: (2025)
by: Yu, Xinzhuo, et al.
Published: (2025)
SMTrack: State-Aware Mamba for Efficient Temporal Modeling in Visual Tracking
by: Ma, Yinchao, et al.
Published: (2026)
by: Ma, Yinchao, et al.
Published: (2026)
Tracking with Human-Intent Reasoning
by: Zhu, Jiawen, et al.
Published: (2023)
by: Zhu, Jiawen, et al.
Published: (2023)
WaveMamba: Wavelet-Driven Mamba Fusion for RGB-Infrared Object Detection
by: Zhu, Haodong, et al.
Published: (2025)
by: Zhu, Haodong, et al.
Published: (2025)
Transformer RGBT Tracking with Spatio-Temporal Multimodal Tokens
by: Sun, Dengdi, et al.
Published: (2024)
by: Sun, Dengdi, et al.
Published: (2024)
GMSR:Gradient-Guided Mamba for Spectral Reconstruction from RGB Images
by: Wang, Xinying, et al.
Published: (2024)
by: Wang, Xinying, et al.
Published: (2024)
DMTrack: Spatio-Temporal Multimodal Tracking via Dual-Adapter
by: Li, Weihong, et al.
Published: (2025)
by: Li, Weihong, et al.
Published: (2025)
Modality-Guided Dynamic Graph Fusion and Temporal Diffusion for Self-Supervised RGB-T Tracking
by: Li, Shenglan, et al.
Published: (2025)
by: Li, Shenglan, et al.
Published: (2025)
High-Performance Few-Shot Segmentation with Foundation Models: An Empirical Study
by: Chang, Shijie, et al.
Published: (2024)
by: Chang, Shijie, et al.
Published: (2024)
PoseMamba: Monocular 3D Human Pose Estimation with Bidirectional Global-Local Spatio-Temporal State Space Model
by: Huang, Yunlong, et al.
Published: (2024)
by: Huang, Yunlong, et al.
Published: (2024)
DCPT: Darkness Clue-Prompted Tracking in Nighttime UAVs
by: Zhu, Jiawen, et al.
Published: (2023)
by: Zhu, Jiawen, et al.
Published: (2023)
Interactive Spatial-Frequency Fusion Mamba for Multi-Modal Image Fusion
by: Zhu, Yixin, et al.
Published: (2026)
by: Zhu, Yixin, et al.
Published: (2026)
MVAFormer: RGB-based Multi-View Spatio-Temporal Action Recognition with Transformer
by: Yamane, Taiga, et al.
Published: (2025)
by: Yamane, Taiga, et al.
Published: (2025)
TrackingMiM: Efficient Mamba-in-Mamba Serialization for Real-time UAV Object Tracking
by: Liu, Bingxi, et al.
Published: (2025)
by: Liu, Bingxi, et al.
Published: (2025)
Efficient Motion Prompt Learning for Robust Visual Tracking
by: Zhao, Jie, et al.
Published: (2025)
by: Zhao, Jie, et al.
Published: (2025)
Autoregressive Queries for Adaptive Tracking with Spatio-TemporalTransformers
by: Xie, Jinxia, et al.
Published: (2024)
by: Xie, Jinxia, et al.
Published: (2024)
SPMTrack: Spatio-Temporal Parameter-Efficient Fine-Tuning with Mixture of Experts for Scalable Visual Tracking
by: Cai, Wenrui, et al.
Published: (2025)
by: Cai, Wenrui, et al.
Published: (2025)
MambaTrack3D: A State Space Model Framework for LiDAR-Based Object Tracking under High Temporal Variation
by: Tian, Shengjing, et al.
Published: (2025)
by: Tian, Shengjing, et al.
Published: (2025)
RAGTrack: Language-aware RGBT Tracking with Retrieval-Augmented Generation
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Hybrid-SORT: Weak Cues Matter for Online Multi-Object Tracking
by: Yang, Mingzhan, et al.
Published: (2023)
by: Yang, Mingzhan, et al.
Published: (2023)
EventMamba: Enhancing Spatio-Temporal Locality with State Space Models for Event-Based Video Reconstruction
by: Ge, Chengjie, et al.
Published: (2025)
by: Ge, Chengjie, et al.
Published: (2025)
DynSTG-Mamba: Dynamic Spatio-Temporal Graph Mamba with Cross-Graph Knowledge Distillation for Gait Disorders Recognition
by: Zrimek, Zakariae, et al.
Published: (2025)
by: Zrimek, Zakariae, et al.
Published: (2025)
UBATrack: Spatio-Temporal State Space Model for General Multi-Modal Tracking
by: Liang, Qihua, et al.
Published: (2026)
by: Liang, Qihua, et al.
Published: (2026)
Similar Items
-
Exploring Enhanced Contextual Information for Video-Level Object Tracking
by: Kang, Ben, et al.
Published: (2024) -
Unified Sequence-to-Sequence Learning for Single- and Multi-Modal Visual Object Tracking
by: Chen, Xin, et al.
Published: (2023) -
SUTrack: Towards Simple and Unified Single Object Tracking
by: Chen, Xin, et al.
Published: (2024) -
Two-stream Beats One-stream: Asymmetric Siamese Network for Efficient Visual Tracking
by: Zhu, Jiawen, et al.
Published: (2025) -
AVS-Mamba: Exploring Temporal and Multi-modal Mamba for Audio-Visual Segmentation
by: Gong, Sitong, et al.
Published: (2025)