Saved in:
| Main Authors: | Mei, Xiaodong, Wang, Sheng, Cheng, Jie, Chen, Yingbing, Xu, Dan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.15703 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FlowMamba: Learning Point Cloud Scene Flow with Global Motion Propagation
by: Lin, Min, et al.
Published: (2024)
by: Lin, Min, et al.
Published: (2024)
Learning Context-Adaptive Motion Priors for Masked Motion Diffusion Models with Efficient Kinematic Attention Aggregation
by: Jiang, Junkun, et al.
Published: (2026)
by: Jiang, Junkun, et al.
Published: (2026)
MoReFun: Past-Movement Guided Motion Representation Learning for Future Motion Prediction and Understanding
by: Shi, Junyu, et al.
Published: (2024)
by: Shi, Junyu, et al.
Published: (2024)
MambaRain: Multi-Scale Mamba-Attention Framework for 0-3 Hour Precipitation Nowcasting
by: Shi, Chunlei, et al.
Published: (2026)
by: Shi, Chunlei, et al.
Published: (2026)
Hybrid Primal Sketch: Combining Analogy, Qualitative Representations, and Computer Vision for Scene Understanding
by: Forbus, Kenneth D., et al.
Published: (2024)
by: Forbus, Kenneth D., et al.
Published: (2024)
LHPF: Look back the History and Plan for the Future in Autonomous Driving
by: Wang, Sheng, et al.
Published: (2024)
by: Wang, Sheng, et al.
Published: (2024)
HexPlane Representation for 3D Semantic Scene Understanding
by: Chen, Zeren, et al.
Published: (2025)
by: Chen, Zeren, et al.
Published: (2025)
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation
by: Chen, Wenchao, et al.
Published: (2024)
by: Chen, Wenchao, et al.
Published: (2024)
Establishing Reality-Virtuality Interconnections in Urban Digital Twins for Superior Intelligent Road Inspection and Simulation
by: Zhang, Yikang, et al.
Published: (2024)
by: Zhang, Yikang, et al.
Published: (2024)
A Structure-aware and Motion-adaptive Framework for 3D Human Pose Estimation with Mamba
by: Lu, Ye, et al.
Published: (2025)
by: Lu, Ye, et al.
Published: (2025)
LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model
by: Mei, Xiaodong, et al.
Published: (2026)
by: Mei, Xiaodong, et al.
Published: (2026)
Vamba: Understanding Hour-Long Videos with Hybrid Mamba-Transformers
by: Ren, Weiming, et al.
Published: (2025)
by: Ren, Weiming, et al.
Published: (2025)
Parts-Mamba: Augmenting Joint Context with Part-Level Scanning for Occluded Human Skeleton
by: Shen, Tianyi, et al.
Published: (2025)
by: Shen, Tianyi, et al.
Published: (2025)
Flow-NeRF: Joint Learning of Geometry, Poses, and Dense Flow within Unified Neural Representations
by: Zheng, Xunzhi, et al.
Published: (2025)
by: Zheng, Xunzhi, et al.
Published: (2025)
Hybrid Mamba for Few-Shot Segmentation
by: Xu, Qianxiong, et al.
Published: (2024)
by: Xu, Qianxiong, et al.
Published: (2024)
Hybrid Mesh-Gaussian Representation for Efficient Indoor Scene Reconstruction
by: Huang, Binxiao, et al.
Published: (2025)
by: Huang, Binxiao, et al.
Published: (2025)
Trajectory Mamba: Efficient Attention-Mamba Forecasting Model Based on Selective SSM
by: Huang, Yizhou, et al.
Published: (2025)
by: Huang, Yizhou, et al.
Published: (2025)
Matten: Video Generation with Mamba-Attention
by: Gao, Yu, et al.
Published: (2024)
by: Gao, Yu, et al.
Published: (2024)
Incremental Joint Learning of Depth, Pose and Implicit Scene Representation on Monocular Camera in Large-scale Scenes
by: Deng, Tianchen, et al.
Published: (2024)
by: Deng, Tianchen, et al.
Published: (2024)
MTMamba: Enhancing Multi-Task Dense Scene Understanding by Mamba-Based Decoders
by: Lin, Baijiong, et al.
Published: (2024)
by: Lin, Baijiong, et al.
Published: (2024)
PoinTramba: A Hybrid Transformer-Mamba Framework for Point Cloud Analysis
by: Wang, Zicheng, et al.
Published: (2024)
by: Wang, Zicheng, et al.
Published: (2024)
PyGS: Large-scale Scene Representation with Pyramidal 3D Gaussian Splatting
by: Wang, Zipeng, et al.
Published: (2024)
by: Wang, Zipeng, et al.
Published: (2024)
Bridging Scene Generation and Planning: Driving with World Model via Unifying Vision and Motion Representation
by: Gui, Xingtai, et al.
Published: (2026)
by: Gui, Xingtai, et al.
Published: (2026)
Jointly Learning Structured Representations and Stabilized Affinity for Human Motion Segmentation
by: Meng, Xianghan, et al.
Published: (2026)
by: Meng, Xianghan, et al.
Published: (2026)
Scene Change Detection with Vision-Language Representation Learning
by: Sheng, Diwei, et al.
Published: (2026)
by: Sheng, Diwei, et al.
Published: (2026)
Motion4D: Learning 3D-Consistent Motion and Semantics for 4D Scene Understanding
by: Zhou, Haoran, et al.
Published: (2025)
by: Zhou, Haoran, et al.
Published: (2025)
Learning Human Motion with Temporally Conditional Mamba
by: Nguyen, Quang, et al.
Published: (2025)
by: Nguyen, Quang, et al.
Published: (2025)
Hybrid CNN-ViT Framework for Motion-Blurred Scene Text Restoration
by: Rashid, Umar, et al.
Published: (2025)
by: Rashid, Umar, et al.
Published: (2025)
Robust Egocentric Visual Attention Prediction Through Language-guided Scene Context-aware Learning
by: Park, Sungjune, et al.
Published: (2026)
by: Park, Sungjune, et al.
Published: (2026)
SEAL: Semantic Attention Learning for Long Video Representation
by: Wang, Lan, et al.
Published: (2024)
by: Wang, Lan, et al.
Published: (2024)
MTMamba++: Enhancing Multi-Task Dense Scene Understanding via Mamba-Based Decoders
by: Lin, Baijiong, et al.
Published: (2024)
by: Lin, Baijiong, et al.
Published: (2024)
Mamba Learns in Context: Structure-Aware Domain Generalization for Multi-Task Point Cloud Understanding
by: Jiang, Jincen, et al.
Published: (2026)
by: Jiang, Jincen, et al.
Published: (2026)
UAM: A Unified Attention-Mamba Backbone of Multimodal Framework for Tumor Cell Classification
by: Chen, Taixi, et al.
Published: (2025)
by: Chen, Taixi, et al.
Published: (2025)
Scenes as Tokens: Multi-Scale Normal Distributions Transform Tokenizer for General 3D Vision-Language Understanding
by: Tang, Yutao, et al.
Published: (2025)
by: Tang, Yutao, et al.
Published: (2025)
V2M: Visual 2-Dimensional Mamba for Image Representation Learning
by: Wang, Chengkun, et al.
Published: (2024)
by: Wang, Chengkun, et al.
Published: (2024)
MambaCAFU: Hybrid Multi-Scale and Multi-Attention Model with Mamba-Based Fusion for Medical Image Segmentation
by: Bui, T-Mai, et al.
Published: (2025)
by: Bui, T-Mai, et al.
Published: (2025)
MambaVesselNet++: A Hybrid CNN-Mamba Architecture for Medical Image Segmentation
by: Xu, Qing, et al.
Published: (2025)
by: Xu, Qing, et al.
Published: (2025)
SpectroMotion: Dynamic 3D Reconstruction of Specular Scenes
by: Fan, Cheng-De, et al.
Published: (2024)
by: Fan, Cheng-De, et al.
Published: (2024)
TAFormer: A Unified Target-Aware Transformer for Video and Motion Joint Prediction in Aerial Scenes
by: Xu, Liangyu, et al.
Published: (2024)
by: Xu, Liangyu, et al.
Published: (2024)
A Unified Framework for 3D Scene Understanding
by: Xu, Wei, et al.
Published: (2024)
by: Xu, Wei, et al.
Published: (2024)
Similar Items
-
FlowMamba: Learning Point Cloud Scene Flow with Global Motion Propagation
by: Lin, Min, et al.
Published: (2024) -
Learning Context-Adaptive Motion Priors for Masked Motion Diffusion Models with Efficient Kinematic Attention Aggregation
by: Jiang, Junkun, et al.
Published: (2026) -
MoReFun: Past-Movement Guided Motion Representation Learning for Future Motion Prediction and Understanding
by: Shi, Junyu, et al.
Published: (2024) -
MambaRain: Multi-Scale Mamba-Attention Framework for 0-3 Hour Precipitation Nowcasting
by: Shi, Chunlei, et al.
Published: (2026) -
Hybrid Primal Sketch: Combining Analogy, Qualitative Representations, and Computer Vision for Scene Understanding
by: Forbus, Kenneth D., et al.
Published: (2024)