Depth Helps: Improving Pre-trained RGB-based Policy with Depth Information Injection
Fuente:
arXiv
Saved in:
| Main Authors: | Pang, Xincheng, Xia, Wenke, Wang, Zhigang, Zhao, Bin, Hu, Di, Wang, Dong, Li, Xuelong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Kinematic-aware Prompting for Generalizable Articulated Object Manipulation with LLMs
by: Xia, Wenke, et al.
Published: (2023)
by: Xia, Wenke, et al.
Published: (2023)
KOI: Accelerating Online Imitation Learning via Hybrid Key-state Guidance
by: Lu, Jingxian, et al.
Published: (2024)
by: Lu, Jingxian, et al.
Published: (2024)
Closed-Loop Action Chunks with Dynamic Corrections for Training-Free Diffusion Policy
by: Wu, Pengyuan, et al.
Published: (2026)
by: Wu, Pengyuan, et al.
Published: (2026)
SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation
by: Zhang, Junjie, et al.
Published: (2024)
by: Zhang, Junjie, et al.
Published: (2024)
Play to the Score: Stage-Guided Dynamic Multi-Sensory Fusion for Robotic Manipulation
by: Feng, Ruoxuan, et al.
Published: (2024)
by: Feng, Ruoxuan, et al.
Published: (2024)
Phoenix: A Motion-based Self-Reflection Framework for Fine-grained Robotic Action Correction
by: Xia, Wenke, et al.
Published: (2025)
by: Xia, Wenke, et al.
Published: (2025)
COHERENT: Collaboration of Heterogeneous Multi-Robot System with Large Language Models
by: Liu, Kehui, et al.
Published: (2024)
by: Liu, Kehui, et al.
Published: (2024)
When would Vision-Proprioception Policies Fail in Robotic Manipulation?
by: Lu, Jingxian, et al.
Published: (2026)
by: Lu, Jingxian, et al.
Published: (2026)
RPG: Robust Policy Gating for Smooth Multi-Skill Transitions in Humanoid Fighting
by: Xin, Yucheng, et al.
Published: (2026)
by: Xin, Yucheng, et al.
Published: (2026)
GeCo-SRT: Geometry-aware Continual Adaptation for Robotic Cross-Task Sim-to-Real Transfer
by: Yu, Wenbo, et al.
Published: (2026)
by: Yu, Wenbo, et al.
Published: (2026)
Human-assisted Robotic Policy Refinement via Action Preference Optimization
by: Xia, Wenke, et al.
Published: (2025)
by: Xia, Wenke, et al.
Published: (2025)
Geometry-Aware Sparse Depth Sampling for High-Fidelity RGB-D Depth Completion in Robotic Systems
by: Salloom, Tony, et al.
Published: (2025)
by: Salloom, Tony, et al.
Published: (2025)
A Depth Control Method for Full Ocean Depth AUV
by: Yueming Li, et al.
Published: (2026)
by: Yueming Li, et al.
Published: (2026)
Learning Manipulation by Predicting Interaction
by: Zeng, Jia, et al.
Published: (2024)
by: Zeng, Jia, et al.
Published: (2024)
Dropping the D: RGB-D SLAM Without the Depth Sensor
by: Kiray, Mert, et al.
Published: (2025)
by: Kiray, Mert, et al.
Published: (2025)
Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-Training
by: He, Haoran, et al.
Published: (2024)
by: He, Haoran, et al.
Published: (2024)
Self-Aligning Depth-regularized Radiance Fields for Asynchronous RGB-D Sequences
by: Huang, Yuxin, et al.
Published: (2022)
by: Huang, Yuxin, et al.
Published: (2022)
InternData-A1: Pioneering High-Fidelity Synthetic Data for Pre-training Generalist Policy
by: Tian, Yang, et al.
Published: (2025)
by: Tian, Yang, et al.
Published: (2025)
Floor Plan-Guided Visual Navigation Incorporating Depth and Directional Cues
by: Huang, Weiqi, et al.
Published: (2025)
by: Huang, Weiqi, et al.
Published: (2025)
UniLACT: Depth-Aware RGB Latent Action Learning for Vision-Language-Action Models
by: Govind, Manish Kumar, et al.
Published: (2026)
by: Govind, Manish Kumar, et al.
Published: (2026)
Visual Robotic Manipulation with Depth-Aware Pretraining
by: Wang, Wanying, et al.
Published: (2024)
by: Wang, Wanying, et al.
Published: (2024)
Boosting Omnidirectional Stereo Matching with a Pre-trained Depth Foundation Model
by: Endres, Jannik, et al.
Published: (2025)
by: Endres, Jannik, et al.
Published: (2025)
DPL: Depth-only Perceptive Humanoid Locomotion via Realistic Depth Synthesis and Cross-Attention Terrain Reconstruction
by: Sun, Jingkai, et al.
Published: (2025)
by: Sun, Jingkai, et al.
Published: (2025)
RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation
by: Wang, Sheng
Published: (2025)
by: Wang, Sheng
Published: (2025)
IDLS: Inverse Depth Line based Visual-Inertial SLAM
by: Li, Wanting, et al.
Published: (2023)
by: Li, Wanting, et al.
Published: (2023)
Monocular One-Shot Metric-Depth Alignment for RGB-Based Robot Grasping
by: Guo, Teng, et al.
Published: (2025)
by: Guo, Teng, et al.
Published: (2025)
The RoboDepth Challenge: Methods and Advancements Towards Robust Depth Estimation
by: Kong, Lingdong, et al.
Published: (2023)
by: Kong, Lingdong, et al.
Published: (2023)
Depth Matters: Multimodal RGB-D Perception for Robust Autonomous Agents
by: Clement, Mihaela-Larisa, et al.
Published: (2025)
by: Clement, Mihaela-Larisa, et al.
Published: (2025)
M3Depth: Wavelet-Enhanced Depth Estimation on Mars via Mutual Boosting of Dual-Modal Data
by: Li, Junjie, et al.
Published: (2025)
by: Li, Junjie, et al.
Published: (2025)
Evo-Depth: A Lightweight Depth-Enhanced Vision-Language-Action Model
by: Lin, Tao, et al.
Published: (2026)
by: Lin, Tao, et al.
Published: (2026)
X-Loco: Towards Generalist Humanoid Locomotion Control via Synergetic Policy Distillation
by: Wang, Dewei, et al.
Published: (2026)
by: Wang, Dewei, et al.
Published: (2026)
MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation
by: Zhang, Pingrui, et al.
Published: (2025)
by: Zhang, Pingrui, et al.
Published: (2025)
Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation
by: Zhang, Pingrui, et al.
Published: (2025)
by: Zhang, Pingrui, et al.
Published: (2025)
RoboBrain 2.5: Depth in Sight, Time in Mind
by: Tan, Huajie, et al.
Published: (2026)
by: Tan, Huajie, et al.
Published: (2026)
Depth Jitter: Seeing through the Depth
by: Rahman, Md Sazidur, et al.
Published: (2025)
by: Rahman, Md Sazidur, et al.
Published: (2025)
D3RoMa: Disparity Diffusion-based Depth Sensing for Material-Agnostic Robotic Manipulation
by: Wei, Songlin, et al.
Published: (2024)
by: Wei, Songlin, et al.
Published: (2024)
DepthCache: Depth-Guided Training-Free Visual Token Merging for Vision-Language-Action Model Inference
by: Li, Yuquan, et al.
Published: (2026)
by: Li, Yuquan, et al.
Published: (2026)
Think Small, Act Big: Primitive Prompt Learning for Lifelong Robot Manipulation
by: Yao, Yuanqi, et al.
Published: (2025)
by: Yao, Yuanqi, et al.
Published: (2025)
GUIDES: Guidance Using Instructor-Distilled Embeddings for Pre-trained Robot Policy Enhancement
by: Gao, Minquan, et al.
Published: (2025)
by: Gao, Minquan, et al.
Published: (2025)
LungDepth: Self‐Supervised Multi‐Frame Monocular Depth Estimation for Bronchoscopy
by: Jingsheng Xu, et al.
Published: (2025)
by: Jingsheng Xu, et al.
Published: (2025)
Similar Items
-
Kinematic-aware Prompting for Generalizable Articulated Object Manipulation with LLMs
by: Xia, Wenke, et al.
Published: (2023) -
KOI: Accelerating Online Imitation Learning via Hybrid Key-state Guidance
by: Lu, Jingxian, et al.
Published: (2024) -
Closed-Loop Action Chunks with Dynamic Corrections for Training-Free Diffusion Policy
by: Wu, Pengyuan, et al.
Published: (2026) -
SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation
by: Zhang, Junjie, et al.
Published: (2024) -
Play to the Score: Stage-Guided Dynamic Multi-Sensory Fusion for Robotic Manipulation
by: Feng, Ruoxuan, et al.
Published: (2024)