Learning Surgical Robotic Manipulation with 3D Spatial Priors
Fuente:
arXiv
Saved in:
| Main Authors: | Sheng, Yu, Wang, Lidian, Chu, Xiaomeng, Deng, Jiajun, Cheng, Min, Zhang, Yanyong, Hua, Bei, Li, Houqiang, Ji, Jianmin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MSGField: A Unified Scene Representation Integrating Motion, Semantics, and Geometry for Robotic Manipulation
by: Sheng, Yu, et al.
Published: (2024)
by: Sheng, Yu, et al.
Published: (2024)
OA-DET3D: Embedding Object Awareness as a General Plug-in for Multi-Camera 3D Object Detection
by: Chu, Xiaomeng, et al.
Published: (2023)
by: Chu, Xiaomeng, et al.
Published: (2023)
GraspCoT: Integrating Physical Property Reasoning for 6-DoF Grasping under Flexible Language Instructions
by: Chu, Xiaomeng, et al.
Published: (2025)
by: Chu, Xiaomeng, et al.
Published: (2025)
SpatialSplat: Efficient Semantic 3D from Sparse Unposed Images
by: Sheng, Yu, et al.
Published: (2025)
by: Sheng, Yu, et al.
Published: (2025)
CELLmap: Enhancing LiDAR SLAM through Elastic and Lightweight Spherical Map Representation
by: Duan, Yifan, et al.
Published: (2024)
by: Duan, Yifan, et al.
Published: (2024)
STDArm: Transferring Visuomotor Policies From Static Data Training to Dynamic Robot Manipulation
by: Duan, Yifan, et al.
Published: (2025)
by: Duan, Yifan, et al.
Published: (2025)
MM-Gaussian: 3D Gaussian-based Multi-modal Fusion for Localization and Reconstruction in Unbounded Scenes
by: Wu, Chenyang, et al.
Published: (2024)
by: Wu, Chenyang, et al.
Published: (2024)
MHRC: Closed-loop Decentralized Multi-Heterogeneous Robot Collaboration with Large Language Models
by: Yu, Wenhao, et al.
Published: (2024)
by: Yu, Wenhao, et al.
Published: (2024)
FAVLA: A Force-Adaptive Fast-Slow VLA model for Contact-Rich Robotic Manipulation
by: Li, Yao, et al.
Published: (2026)
by: Li, Yao, et al.
Published: (2026)
OCC-VO: Dense Mapping via 3D Occupancy-Based Visual Odometry for Autonomous Driving
by: Li, Heng, et al.
Published: (2023)
by: Li, Heng, et al.
Published: (2023)
Rendering-Enhanced Automatic Image-to-Point Cloud Registration for Roadside Scenes
by: Sheng, Yu, et al.
Published: (2024)
by: Sheng, Yu, et al.
Published: (2024)
RaCFormer: Towards High-Quality 3D Object Detection via Query-based Radar-Camera Fusion
by: Chu, Xiaomeng, et al.
Published: (2024)
by: Chu, Xiaomeng, et al.
Published: (2024)
S3R-GS: Streamlining the Pipeline for Large-Scale Street Scene Reconstruction
by: Zheng, Guangting, et al.
Published: (2025)
by: Zheng, Guangting, et al.
Published: (2025)
LDP: A Local Diffusion Planner for Efficient Robot Navigation and Collision Avoidance
by: Yu, Wenhao, et al.
Published: (2024)
by: Yu, Wenhao, et al.
Published: (2024)
Rotation Initialization and Stepwise Refinement for Universal LiDAR Calibration
by: Duan, Yifan, et al.
Published: (2024)
by: Duan, Yifan, et al.
Published: (2024)
Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control
by: Gao, Yuxuan, et al.
Published: (2026)
by: Gao, Yuxuan, et al.
Published: (2026)
Perception Helps Planning: Facilitating Multi-Stage Lane-Level Integration via Double-Edge Structures
by: You, Guoliang, et al.
Published: (2024)
by: You, Guoliang, et al.
Published: (2024)
Structural Action Transformer for 3D Dexterous Manipulation
by: Lei, Xiaohan, et al.
Published: (2026)
by: Lei, Xiaohan, et al.
Published: (2026)
D3RoMa: Disparity Diffusion-based Depth Sensing for Material-Agnostic Robotic Manipulation
by: Wei, Songlin, et al.
Published: (2024)
by: Wei, Songlin, et al.
Published: (2024)
Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation
by: Xie, Yifan, et al.
Published: (2026)
by: Xie, Yifan, et al.
Published: (2026)
Generalizable Geometric Prior and Recurrent Spiking Feature Learning for Humanoid Robot Manipulation
by: Li, Xuetao, et al.
Published: (2026)
by: Li, Xuetao, et al.
Published: (2026)
Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation
by: Xiao, Junjin, et al.
Published: (2026)
by: Xiao, Junjin, et al.
Published: (2026)
3D Affordance Keypoint Detection for Robotic Manipulation
by: Liu, Zhiyang, et al.
Published: (2025)
by: Liu, Zhiyang, et al.
Published: (2025)
CAFE-AD: Cross-Scenario Adaptive Feature Enhancement for Trajectory Planning in Autonomous Driving
by: Zhang, Junrui, et al.
Published: (2025)
by: Zhang, Junrui, et al.
Published: (2025)
GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions
by: Huang, Helong, et al.
Published: (2025)
by: Huang, Helong, et al.
Published: (2025)
Learning Geometrically-Grounded 3D Visual Representations for View-Generalizable Robotic Manipulation
by: Zhang, Di, et al.
Published: (2026)
by: Zhang, Di, et al.
Published: (2026)
Surgical Robot Transformer (SRT): Imitation Learning for Surgical Tasks
by: Kim, Ji Woong, et al.
Published: (2024)
by: Kim, Ji Woong, et al.
Published: (2024)
S3D: A Spatial Steerable Surgical Drilling Framework for Robotic Spinal Fixation Procedures
by: Maroufi, Daniyal, et al.
Published: (2025)
by: Maroufi, Daniyal, et al.
Published: (2025)
Diffusion Stabilizer Policy for Automated Surgical Robot Manipulations
by: Ho, Chonlam, et al.
Published: (2025)
by: Ho, Chonlam, et al.
Published: (2025)
LLM-based Interactive Imitation Learning for Robotic Manipulation
by: Werner, Jonas, et al.
Published: (2025)
by: Werner, Jonas, et al.
Published: (2025)
CRPlace: Camera-Radar Fusion with BEV Representation for Place Recognition
by: Fu, Shaowei, et al.
Published: (2024)
by: Fu, Shaowei, et al.
Published: (2024)
UrgenGo: Urgency-Aware Transparent GPU Kernel Launching for Autonomous Driving
by: Zhu, Hanqi, et al.
Published: (2025)
by: Zhu, Hanqi, et al.
Published: (2025)
MT-PCR: Leveraging Modality Transformation for Large-Scale Point Cloud Registration with Limited Overlap
by: Wu, Yilong, et al.
Published: (2025)
by: Wu, Yilong, et al.
Published: (2025)
High-Precision Surgical Robotic System for Intraocular Procedures
by: Lai, Yu-Ting, et al.
Published: (2025)
by: Lai, Yu-Ting, et al.
Published: (2025)
Robot Tape Manipulation for 3D Printing
by: Tushar, Nahid, et al.
Published: (2024)
by: Tushar, Nahid, et al.
Published: (2024)
GaussianDream: A Feed-Forward 3D Gaussian World Model for Robotic Manipulation
by: Zhang, Zijian, et al.
Published: (2026)
by: Zhang, Zijian, et al.
Published: (2026)
Manipulate-to-Navigate: Reinforcement Learning with Visual Affordances and Manipulability Priors
by: Zhang, Yuying, et al.
Published: (2025)
by: Zhang, Yuying, et al.
Published: (2025)
CalibFormer: A Transformer-based Automatic LiDAR-Camera Calibration Network
by: Xiao, Yuxuan, et al.
Published: (2023)
by: Xiao, Yuxuan, et al.
Published: (2023)
Beyond the Majority: Long-tail Imitation Learning for Robotic Manipulation
by: Zhu, Junhong, et al.
Published: (2026)
by: Zhu, Junhong, et al.
Published: (2026)
DexRepNet++: Learning Dexterous Robotic Manipulation with Geometric and Spatial Hand-Object Representations
by: Liu, Qingtao, et al.
Published: (2026)
by: Liu, Qingtao, et al.
Published: (2026)
Similar Items
-
MSGField: A Unified Scene Representation Integrating Motion, Semantics, and Geometry for Robotic Manipulation
by: Sheng, Yu, et al.
Published: (2024) -
OA-DET3D: Embedding Object Awareness as a General Plug-in for Multi-Camera 3D Object Detection
by: Chu, Xiaomeng, et al.
Published: (2023) -
GraspCoT: Integrating Physical Property Reasoning for 6-DoF Grasping under Flexible Language Instructions
by: Chu, Xiaomeng, et al.
Published: (2025) -
SpatialSplat: Efficient Semantic 3D from Sparse Unposed Images
by: Sheng, Yu, et al.
Published: (2025) -
CELLmap: Enhancing LiDAR SLAM through Elastic and Lightweight Spherical Map Representation
by: Duan, Yifan, et al.
Published: (2024)