Snakes and Ladders: Two Steps Up for VideoMamba
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lu, Hui, Salah, Albert Ali, Poppe, Ronald |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TCNet: Continuous Sign Language Recognition from Trajectories and Correlated Regions
von: Lu, Hui, et al.
Veröffentlicht: (2024)
von: Lu, Hui, et al.
Veröffentlicht: (2024)
Enhancing Video Transformers for Action Understanding with VLM-aided Training
von: Lu, Hui, et al.
Veröffentlicht: (2024)
von: Lu, Hui, et al.
Veröffentlicht: (2024)
VideoMamba: State Space Model for Efficient Video Understanding
von: Li, Kunchang, et al.
Veröffentlicht: (2024)
von: Li, Kunchang, et al.
Veröffentlicht: (2024)
VideoMamba: Spatio-Temporal Selective State Space Model
von: Park, Jinyoung, et al.
Veröffentlicht: (2024)
von: Park, Jinyoung, et al.
Veröffentlicht: (2024)
Dual Branch VideoMamba with Gated Class Token Fusion for Violence Detection
von: Senadeera, Damith Chamalke, et al.
Veröffentlicht: (2025)
von: Senadeera, Damith Chamalke, et al.
Veröffentlicht: (2025)
Self-Supervised Partial Cycle-Consistency for Multi-View Matching
von: Taggenbrock, Fedor, et al.
Veröffentlicht: (2025)
von: Taggenbrock, Fedor, et al.
Veröffentlicht: (2025)
About Time: Advances, Challenges, and Outlooks of Action Understanding
von: Stergiou, Alexandros, et al.
Veröffentlicht: (2024)
von: Stergiou, Alexandros, et al.
Veröffentlicht: (2024)
Representation Learning and Identity Adversarial Training for Facial Behavior Understanding
von: Ning, Mang, et al.
Veröffentlicht: (2024)
von: Ning, Mang, et al.
Veröffentlicht: (2024)
MVQA: Mamba with Unified Sampling for Efficient Video Quality Assessment
von: Mi, Yachun, et al.
Veröffentlicht: (2025)
von: Mi, Yachun, et al.
Veröffentlicht: (2025)
MambaVision: A Hybrid Mamba-Transformer Vision Backbone
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2024)
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2024)
Matten: Video Generation with Mamba-Attention
von: Gao, Yu, et al.
Veröffentlicht: (2024)
von: Gao, Yu, et al.
Veröffentlicht: (2024)
HieraMamba: Video Temporal Grounding via Hierarchical Anchor-Mamba Pooling
von: An, Joungbin, et al.
Veröffentlicht: (2025)
von: An, Joungbin, et al.
Veröffentlicht: (2025)
Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding
von: Chen, Guo, et al.
Veröffentlicht: (2024)
von: Chen, Guo, et al.
Veröffentlicht: (2024)
MambaPainter: Neural Stroke-Based Rendering in a Single Step
von: Sawada, Tomoya, et al.
Veröffentlicht: (2024)
von: Sawada, Tomoya, et al.
Veröffentlicht: (2024)
Vivim: a Video Vision Mamba for Medical Video Segmentation
von: Yang, Yijun, et al.
Veröffentlicht: (2024)
von: Yang, Yijun, et al.
Veröffentlicht: (2024)
MambaVideo for Discrete Video Tokenization with Channel-Split Quantization
von: Argaw, Dawit Mureja, et al.
Veröffentlicht: (2025)
von: Argaw, Dawit Mureja, et al.
Veröffentlicht: (2025)
LadderMoE: Ladder-Side Mixture of Experts Adapters for Bronze Inscription Recognition
von: Zhou, Rixin, et al.
Veröffentlicht: (2025)
von: Zhou, Rixin, et al.
Veröffentlicht: (2025)
GlobalMamba: Global Image Serialization for Vision Mamba
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
Take A Step Back: Rethinking the Two Stages in Visual Reasoning
von: Zhang, Mingyu, et al.
Veröffentlicht: (2024)
von: Zhang, Mingyu, et al.
Veröffentlicht: (2024)
VideoMAP: Toward Scalable Mamba-based Video Autoregressive Pretraining
von: Liu, Yunze, et al.
Veröffentlicht: (2025)
von: Liu, Yunze, et al.
Veröffentlicht: (2025)
Aligning Step-by-Step Instructional Diagrams to Video Demonstrations
von: Zhang, Jiahao, et al.
Veröffentlicht: (2023)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2023)
AVS-Mamba: Exploring Temporal and Multi-modal Mamba for Audio-Visual Segmentation
von: Gong, Sitong, et al.
Veröffentlicht: (2025)
von: Gong, Sitong, et al.
Veröffentlicht: (2025)
Pop-Up Distractions Reveal Bag-of-Events Behavior in Video Large Language Models
von: Chew, Oscar, et al.
Veröffentlicht: (2026)
von: Chew, Oscar, et al.
Veröffentlicht: (2026)
Vision Mamba for Permeability Prediction of Porous Media
von: Kashefi, Ali, et al.
Veröffentlicht: (2025)
von: Kashefi, Ali, et al.
Veröffentlicht: (2025)
Geometric Transformation-Embedded Mamba for Learned Video Compression
von: Wei, Hao, et al.
Veröffentlicht: (2026)
von: Wei, Hao, et al.
Veröffentlicht: (2026)
Step Differences in Instructional Video
von: Nagarajan, Tushar, et al.
Veröffentlicht: (2024)
von: Nagarajan, Tushar, et al.
Veröffentlicht: (2024)
Spatial-Frequency Enhanced Mamba for Multi-Modal Image Fusion
von: Sun, Hui, et al.
Veröffentlicht: (2025)
von: Sun, Hui, et al.
Veröffentlicht: (2025)
Mamba-OTR: a Mamba-based Solution for Online Take and Release Detection from Untrimmed Egocentric Video
von: Catinello, Alessandro Sebastiano, et al.
Veröffentlicht: (2025)
von: Catinello, Alessandro Sebastiano, et al.
Veröffentlicht: (2025)
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation
von: Chen, Wenchao, et al.
Veröffentlicht: (2024)
von: Chen, Wenchao, et al.
Veröffentlicht: (2024)
NeuroMamba: Multi-Perspective Feature Interaction with Visual Mamba for Neuron Segmentation
von: Jiang, Liuyun, et al.
Veröffentlicht: (2026)
von: Jiang, Liuyun, et al.
Veröffentlicht: (2026)
StepAL: Step-aware Active Learning for Cataract Surgical Videos
von: Shah, Nisarg A., et al.
Veröffentlicht: (2025)
von: Shah, Nisarg A., et al.
Veröffentlicht: (2025)
DeMamba: AI-Generated Video Detection on Million-Scale GenVideo Benchmark
von: Chen, Haoxing, et al.
Veröffentlicht: (2024)
von: Chen, Haoxing, et al.
Veröffentlicht: (2024)
MambaVF: State Space Model for Efficient Video Fusion
von: Zhao, Zixiang, et al.
Veröffentlicht: (2026)
von: Zhao, Zixiang, et al.
Veröffentlicht: (2026)
Vamba: Understanding Hour-Long Videos with Hybrid Mamba-Transformers
von: Ren, Weiming, et al.
Veröffentlicht: (2025)
von: Ren, Weiming, et al.
Veröffentlicht: (2025)
DAPointMamba: Domain Adaptive Point Mamba for Point Cloud Completion
von: Li, Yinghui, et al.
Veröffentlicht: (2025)
von: Li, Yinghui, et al.
Veröffentlicht: (2025)
Phased One-Step Adversarial Equilibrium for Video Diffusion Models
von: Cheng, Jiaxiang, et al.
Veröffentlicht: (2025)
von: Cheng, Jiaxiang, et al.
Veröffentlicht: (2025)
Mamba-VGGT: Persistent Long-Sequence Video Geometry Grounded Transformer via External Sliding Window Mamba Memory
von: Deng, Tianchen, et al.
Veröffentlicht: (2026)
von: Deng, Tianchen, et al.
Veröffentlicht: (2026)
Vision Mamba for Classification of Breast Ultrasound Images
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2024)
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2024)
VP-Hype: A Hybrid Mamba-Transformer Framework with Visual-Textual Prompting for Hyperspectral Image Classification
von: Sellam, Abdellah Zakaria, et al.
Veröffentlicht: (2026)
von: Sellam, Abdellah Zakaria, et al.
Veröffentlicht: (2026)
Mamba-in-Mamba: Centralized Mamba-Cross-Scan in Tokenized Mamba Model for Hyperspectral Image Classification
von: Zhou, Weilian, et al.
Veröffentlicht: (2024)
von: Zhou, Weilian, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
TCNet: Continuous Sign Language Recognition from Trajectories and Correlated Regions
von: Lu, Hui, et al.
Veröffentlicht: (2024) -
Enhancing Video Transformers for Action Understanding with VLM-aided Training
von: Lu, Hui, et al.
Veröffentlicht: (2024) -
VideoMamba: State Space Model for Efficient Video Understanding
von: Li, Kunchang, et al.
Veröffentlicht: (2024) -
VideoMamba: Spatio-Temporal Selective State Space Model
von: Park, Jinyoung, et al.
Veröffentlicht: (2024) -
Dual Branch VideoMamba with Gated Class Token Fusion for Violence Detection
von: Senadeera, Damith Chamalke, et al.
Veröffentlicht: (2025)