MAMBA4D: Efficient Long-Sequence Point Cloud Video Understanding with Disentangled Spatial-Temporal State Space Models
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Jiuming, Han, Jinru, Liu, Lihao, Aviles-Rivero, Angelica I., Jiang, Chaokang, Liu, Zhe, Wang, Hesheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DifFlow3D: Toward Robust Uncertainty-Aware Scene Flow Estimation with Diffusion Model
by: Liu, Jiuming, et al.
Published: (2023)
by: Liu, Jiuming, et al.
Published: (2023)
EADReg: Probabilistic Correspondence Generation with Efficient Autoregressive Diffusion Model for Outdoor Point Cloud Registration
by: Gong, Linrui, et al.
Published: (2024)
by: Gong, Linrui, et al.
Published: (2024)
RegFormer++: An Efficient Large-Scale 3D LiDAR Point Registration Network with Projection-Aware 2D Transformer
by: Liu, Jiuming, et al.
Published: (2026)
by: Liu, Jiuming, et al.
Published: (2026)
VectorWorld: Efficient Streaming World Model via Diffusion Flow on Vector Graphs
by: Jiang, Chaokang, et al.
Published: (2026)
by: Jiang, Chaokang, et al.
Published: (2026)
Point Mamba: A Novel Point Cloud Backbone Based on State Space Model with Octree-Based Ordering Strategy
by: Liu, Jiuming, et al.
Published: (2024)
by: Liu, Jiuming, et al.
Published: (2024)
NeuroGauss4D-PCI: 4D Neural Fields and Gaussian Deformation Fields for Point Cloud Interpolation
by: Jiang, Chaokang, et al.
Published: (2024)
by: Jiang, Chaokang, et al.
Published: (2024)
Spherical Frustum Sparse Convolution Network for LiDAR Point Cloud Semantic Segmentation
by: Zheng, Yu, et al.
Published: (2023)
by: Zheng, Yu, et al.
Published: (2023)
Optimised ProPainter for Video Diminished Reality Inpainting
by: Li, Pengze, et al.
Published: (2024)
by: Li, Pengze, et al.
Published: (2024)
Mamba-VGGT: Persistent Long-Sequence Video Geometry Grounded Transformer via External Sliding Window Mamba Memory
by: Deng, Tianchen, et al.
Published: (2026)
by: Deng, Tianchen, et al.
Published: (2026)
3DSFLabelling: Boosting 3D Scene Flow Estimation by Pseudo Auto-labelling
by: Jiang, Chaokang, et al.
Published: (2024)
by: Jiang, Chaokang, et al.
Published: (2024)
TopoLiDM: Topology-Aware LiDAR Diffusion Models for Interpretable and Realistic LiDAR Point Cloud Generation
by: Liu, Jiuming, et al.
Published: (2025)
by: Liu, Jiuming, et al.
Published: (2025)
UST-SSM: Unified Spatio-Temporal State Space Models for Point Cloud Video Modeling
by: Li, Peiming, et al.
Published: (2025)
by: Li, Peiming, et al.
Published: (2025)
DVLO: Deep Visual-LiDAR Odometry with Local-to-Global Feature Fusion and Bi-Directional Structure Alignment
by: Liu, Jiuming, et al.
Published: (2024)
by: Liu, Jiuming, et al.
Published: (2024)
Biophysics Informed Pathological Regularisation for Brain Tumour Segmentation
by: Zhang, Lipei, et al.
Published: (2024)
by: Zhang, Lipei, et al.
Published: (2024)
GMF-Drive: Gated Mamba Fusion with Spatial-Aware BEV Representation for End-to-End Autonomous Driving
by: Wang, Jian, et al.
Published: (2025)
by: Wang, Jian, et al.
Published: (2025)
SemGauss-SLAM: Dense Semantic Gaussian Splatting SLAM
by: Zhu, Siting, et al.
Published: (2024)
by: Zhu, Siting, et al.
Published: (2024)
Spatial Visibility and Temporal Dynamics: Revolutionizing Field of View Prediction in Adaptive Point Cloud Video Streaming
by: Li, Chen, et al.
Published: (2024)
by: Li, Chen, et al.
Published: (2024)
VCBench: A Streaming Counting Benchmark for Spatial-Temporal State Maintenance in Long Videos
by: Liu, Pengyiang, et al.
Published: (2026)
by: Liu, Pengyiang, et al.
Published: (2026)
Leveraging Joint Spectral and Spatial Learning with MAMBA for Multichannel Speech Enhancement
by: Ren, Wenze, et al.
Published: (2024)
by: Ren, Wenze, et al.
Published: (2024)
TAPTRv3: Spatial and Temporal Context Foster Robust Tracking of Any Point in Long Video
by: Qu, Jinyuan, et al.
Published: (2024)
by: Qu, Jinyuan, et al.
Published: (2024)
ESP-PCT: Enhanced VR Semantic Performance through Efficient Compression of Temporal and Spatial Redundancies in Point Cloud Transformers
by: Mei, Luoyu, et al.
Published: (2024)
by: Mei, Luoyu, et al.
Published: (2024)
VideoCompressa: Data-Efficient Video Understanding via Joint Temporal Compression and Spatial Reconstruction
by: Wang, Shaobo, et al.
Published: (2025)
by: Wang, Shaobo, et al.
Published: (2025)
STS-Mixer: Spatio-Temporal-Spectral Mixer for 4D Point Cloud Video Understanding
by: Li, Wenhao, et al.
Published: (2026)
by: Li, Wenhao, et al.
Published: (2026)
State Space Prompting via Gathering and Spreading Spatio-Temporal Information for Video Understanding
by: Zhou, Jiahuan, et al.
Published: (2025)
by: Zhou, Jiahuan, et al.
Published: (2025)
CrossVideo: Self-supervised Cross-modal Contrastive Learning for Point Cloud Video Understanding
by: Liu, Yunze, et al.
Published: (2024)
by: Liu, Yunze, et al.
Published: (2024)
SDE: A Simplified and Disentangled Dependency Encoding Framework for State Space Models in Time Series Forecasting
by: Weng, Zixuan, et al.
Published: (2024)
by: Weng, Zixuan, et al.
Published: (2024)
Wonder Wins Ways: Curiosity-Driven Exploration through Multi-Agent Contextual Calibration
by: Pan, Yiyuan, et al.
Published: (2025)
by: Pan, Yiyuan, et al.
Published: (2025)
STOP: Integrated Spatial-Temporal Dynamic Prompting for Video Understanding
by: Liu, Zichen, et al.
Published: (2025)
by: Liu, Zichen, et al.
Published: (2025)
Efficient Extractive Summarization with MAMBA-Transformer Hybrids for Low-Resource Scenarios
by: Khayi, Nisrine Ait
Published: (2026)
by: Khayi, Nisrine Ait
Published: (2026)
End-to-end 2D-3D Registration between Image and LiDAR Point Cloud for Vehicle Localization
by: Wang, Guangming, et al.
Published: (2023)
by: Wang, Guangming, et al.
Published: (2023)
MAMBA: Multi-level Aggregation via Memory Bank for Video Object Detection
by: Sun, Guanxiong, et al.
Published: (2024)
by: Sun, Guanxiong, et al.
Published: (2024)
STORM: Token-Efficient Long Video Understanding for Multimodal LLMs
by: Jiang, Jindong, et al.
Published: (2025)
by: Jiang, Jindong, et al.
Published: (2025)
SM3D: Mitigating Spectral Bias and Semantic Dilution in Point Cloud State Space Models
by: Liu, Bin, et al.
Published: (2025)
by: Liu, Bin, et al.
Published: (2025)
4DSTR: Advancing Generative 4D Gaussians with Spatial-Temporal Rectification for High-Quality and Consistent 4D Generation
by: Liu, Mengmeng, et al.
Published: (2025)
by: Liu, Mengmeng, et al.
Published: (2025)
Enhancing Exploratory Capability of Visual Navigation Using Uncertainty of Implicit Scene Representation
by: Wang, Yichen, et al.
Published: (2024)
by: Wang, Yichen, et al.
Published: (2024)
Exploiting Temporal State Space Sharing for Video Semantic Segmentation
by: Hesham, Syed Ariff Syed, et al.
Published: (2025)
by: Hesham, Syed Ariff Syed, et al.
Published: (2025)
D$^2$GSLAM: 4D Dynamic Gaussian Splatting SLAM
by: Zhu, Siting, et al.
Published: (2025)
by: Zhu, Siting, et al.
Published: (2025)
Do Neural Operators Forget Geometry? The Forgetting Hypothesis in Deep Operator Learning
by: Xia, Yanming, et al.
Published: (2026)
by: Xia, Yanming, et al.
Published: (2026)
Don't Fix the Basis -- Learn It: Spectral Representation with Adaptive Basis Learning for PDEs
by: Zhao, Xuxiang, et al.
Published: (2026)
by: Zhao, Xuxiang, et al.
Published: (2026)
VideoMamba: State Space Model for Efficient Video Understanding
by: Li, Kunchang, et al.
Published: (2024)
by: Li, Kunchang, et al.
Published: (2024)
Similar Items
-
DifFlow3D: Toward Robust Uncertainty-Aware Scene Flow Estimation with Diffusion Model
by: Liu, Jiuming, et al.
Published: (2023) -
EADReg: Probabilistic Correspondence Generation with Efficient Autoregressive Diffusion Model for Outdoor Point Cloud Registration
by: Gong, Linrui, et al.
Published: (2024) -
RegFormer++: An Efficient Large-Scale 3D LiDAR Point Registration Network with Projection-Aware 2D Transformer
by: Liu, Jiuming, et al.
Published: (2026) -
VectorWorld: Efficient Streaming World Model via Diffusion Flow on Vector Graphs
by: Jiang, Chaokang, et al.
Published: (2026) -
Point Mamba: A Novel Point Cloud Backbone Based on State Space Model with Octree-Based Ordering Strategy
by: Liu, Jiuming, et al.
Published: (2024)