Vision-Proprioception Fusion with Mamba2 in End-to-End Reinforcement Learning for Motion Control
Fuente:
arXiv
Saved in:
| Main Authors: | Tao, Xiaowen, Wang, Yinuo, Zhou, Jinzhao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LocoMamba: Vision-Driven Locomotion via End-to-End Deep Reinforcement Learning with Mamba
by: Wang, Yinuo, et al.
Published: (2025)
by: Wang, Yinuo, et al.
Published: (2025)
QuadKAN: KAN-Enhanced Quadruped Motion Control via End-to-End Reinforcement Learning
by: Wang, Yinuo, et al.
Published: (2025)
by: Wang, Yinuo, et al.
Published: (2025)
DeepIPCv3: Event-Aware Multi-Modal Sensor Fusion for Sudden Pedestrian Crossing Avoidance
by: Natan, Oskar, et al.
Published: (2026)
by: Natan, Oskar, et al.
Published: (2026)
Recent Advances in Transformer and Large Language Models for UAV Applications
by: Kheddar, Hamza, et al.
Published: (2025)
by: Kheddar, Hamza, et al.
Published: (2025)
Enhancing Synthetic CT from CBCT via Multimodal Fusion and End-To-End Registration
by: Tschuchnig, Maximilian, et al.
Published: (2025)
by: Tschuchnig, Maximilian, et al.
Published: (2025)
Vision Transformers for End-to-End Vision-Based Quadrotor Obstacle Avoidance
by: Bhattacharya, Anish, et al.
Published: (2024)
by: Bhattacharya, Anish, et al.
Published: (2024)
VoxelPrompt: A Vision Agent for End-to-End Medical Image Analysis
by: Hoopes, Andrew, et al.
Published: (2024)
by: Hoopes, Andrew, et al.
Published: (2024)
ASMA: An Adaptive Safety Margin Algorithm for Vision-Language Drone Navigation via Scene-Aware Control Barrier Functions
by: Sanyal, Sourav, et al.
Published: (2024)
by: Sanyal, Sourav, et al.
Published: (2024)
End-to-End Deep Learning for Structural Brain Imaging: A Unified Framework
by: Su, Yao, et al.
Published: (2025)
by: Su, Yao, et al.
Published: (2025)
The Era of End-to-End Autonomy: Transitioning from Rule-Based Driving to Large Driving Models
by: Nebot, Eduardo, et al.
Published: (2026)
by: Nebot, Eduardo, et al.
Published: (2026)
Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery
by: Wang, Guankun, et al.
Published: (2024)
by: Wang, Guankun, et al.
Published: (2024)
ClinicalFMamba: Advancing Clinical Assessment using Mamba-based Multimodal Neuroimaging Fusion
by: Zhou, Meng, et al.
Published: (2025)
by: Zhou, Meng, et al.
Published: (2025)
LangMamba: A Language-driven Mamba Framework for Low-dose CT Denoising with Vision-language Models
by: Chen, Zhihao, et al.
Published: (2025)
by: Chen, Zhihao, et al.
Published: (2025)
See Silhouettes in Motion with Neuromorphic Vision
by: Zhang, Pei, et al.
Published: (2026)
by: Zhang, Pei, et al.
Published: (2026)
CFMW: Cross-modality Fusion Mamba for Robust Object Detection under Adverse Weather
by: Li, Haoyuan, et al.
Published: (2024)
by: Li, Haoyuan, et al.
Published: (2024)
EndoIR: Degradation-Agnostic All-in-One Endoscopic Image Restoration via Noise-Aware Routing Diffusion
by: Chen, Tong, et al.
Published: (2025)
by: Chen, Tong, et al.
Published: (2025)
SiMBA: Simplified Mamba-Based Architecture for Vision and Multivariate Time series
by: Patro, Badri N., et al.
Published: (2024)
by: Patro, Badri N., et al.
Published: (2024)
SDCM: Simulated Densifying and Compensatory Modeling Fusion for Radar-Vision 3-D Object Detection in Internet of Vehicles
by: Li, Shucong, et al.
Published: (2026)
by: Li, Shucong, et al.
Published: (2026)
Vision Controlled Orthotic Hand Exoskeleton
by: Blais, Connor, et al.
Published: (2025)
by: Blais, Connor, et al.
Published: (2025)
F2PASeg: Feature Fusion for Pituitary Anatomy Segmentation in Endoscopic Surgery
by: Chen, Lumin, et al.
Published: (2025)
by: Chen, Lumin, et al.
Published: (2025)
UAV-Based Intelligent Traffic Surveillance System: Real-Time Vehicle Detection, Classification, Tracking, and Behavioral Analysis
by: Khanpour, Ali, et al.
Published: (2025)
by: Khanpour, Ali, et al.
Published: (2025)
DiffuseRAW: End-to-End Generative RAW Image Processing for Low-Light Images
by: Dagli, Rishit
Published: (2023)
by: Dagli, Rishit
Published: (2023)
Automatic laminectomy cutting plane planning based on artificial intelligence in robot assisted laminectomy surgery
by: Li, Zhuofu, et al.
Published: (2023)
by: Li, Zhuofu, et al.
Published: (2023)
Exploring Few-Shot Adaptation for Activity Recognition on Diverse Domains
by: Peng, Kunyu, et al.
Published: (2023)
by: Peng, Kunyu, et al.
Published: (2023)
AffordTissue: Dense Affordance Prediction for Tool-Action Specific Tissue Interaction
by: Maksutova, Aiza, et al.
Published: (2026)
by: Maksutova, Aiza, et al.
Published: (2026)
Deep Transformer Network for Monocular Pose Estimation of Shipborne Unmanned Aerial Vehicle
by: Wickramasuriya, Maneesha, et al.
Published: (2024)
by: Wickramasuriya, Maneesha, et al.
Published: (2024)
Surgical SAM 2: Real-time Segment Anything in Surgical Video by Efficient Frame Pruning
by: Liu, Haofeng, et al.
Published: (2024)
by: Liu, Haofeng, et al.
Published: (2024)
Pose Estimation for Intra-cardiac Echocardiography Catheter via AI-Based Anatomical Understanding
by: Huh, Jaeyoung, et al.
Published: (2025)
by: Huh, Jaeyoung, et al.
Published: (2025)
FlatLands: Generative Floormap Completion From a Single Egocentric View
by: Bhattacharjee, Subhransu S., et al.
Published: (2026)
by: Bhattacharjee, Subhransu S., et al.
Published: (2026)
SurgiSR4K: A High-Resolution Endoscopic Video Dataset for Robotic-Assisted Minimally Invasive Procedures
by: Jiang, Fengyi, et al.
Published: (2025)
by: Jiang, Fengyi, et al.
Published: (2025)
Cardiac Copilot: Automatic Probe Guidance for Echocardiography with World Model
by: Jiang, Haojun, et al.
Published: (2024)
by: Jiang, Haojun, et al.
Published: (2024)
MIST: A Simple and Scalable End-To-End 3D Medical Imaging Segmentation Framework
by: Celaya, Adrian, et al.
Published: (2024)
by: Celaya, Adrian, et al.
Published: (2024)
MTCNet: Motion and Topology Consistency Guided Learning for Mitral Valve Segmentationin 4D Ultrasound
by: Chen, Rusi, et al.
Published: (2025)
by: Chen, Rusi, et al.
Published: (2025)
Deep-Learning-Assisted Highly-Accurate COVID-19 Diagnosis on Lung Computed Tomography Images
by: Wang, Yinuo, et al.
Published: (2025)
by: Wang, Yinuo, et al.
Published: (2025)
GLFC: Unified Global-Local Feature and Contrast Learning with Mamba-Enhanced UNet for Synthetic CT Generation from CBCT
by: Zhou, Xianhao, et al.
Published: (2025)
by: Zhou, Xianhao, et al.
Published: (2025)
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model
by: Zeng, Kang, et al.
Published: (2024)
by: Zeng, Kang, et al.
Published: (2024)
The Machine Vision Iceberg Explained: Advancing Dynamic Testing by Considering Holistic Environmental Relations
by: Padusinski, Hubert, et al.
Published: (2024)
by: Padusinski, Hubert, et al.
Published: (2024)
EndoControlMag: Robust Endoscopic Vascular Motion Magnification with Periodic Reference Resetting and Hierarchical Tissue-aware Dual-Mask Control
by: Wang, An, et al.
Published: (2025)
by: Wang, An, et al.
Published: (2025)
UNet with Self-Adaptive Mamba-Like Attention and Causal-Resonance Learning for Medical Image Segmentation
by: Qamar, Saqib, et al.
Published: (2025)
by: Qamar, Saqib, et al.
Published: (2025)
Control and Automation for Industrial Production Storage Zone: Generation of Optimal Route Using Image Processing
by: Huerfano, Bejamin A., et al.
Published: (2024)
by: Huerfano, Bejamin A., et al.
Published: (2024)
Similar Items
-
LocoMamba: Vision-Driven Locomotion via End-to-End Deep Reinforcement Learning with Mamba
by: Wang, Yinuo, et al.
Published: (2025) -
QuadKAN: KAN-Enhanced Quadruped Motion Control via End-to-End Reinforcement Learning
by: Wang, Yinuo, et al.
Published: (2025) -
DeepIPCv3: Event-Aware Multi-Modal Sensor Fusion for Sudden Pedestrian Crossing Avoidance
by: Natan, Oskar, et al.
Published: (2026) -
Recent Advances in Transformer and Large Language Models for UAV Applications
by: Kheddar, Hamza, et al.
Published: (2025) -
Enhancing Synthetic CT from CBCT via Multimodal Fusion and End-To-End Registration
by: Tschuchnig, Maximilian, et al.
Published: (2025)