R3DP: Real-Time 3D-Aware Policy for Embodied Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yuhao, Dong, Wanxi, Shi, Yue, Liang, Yi, Gao, Jingnan, Yang, Qiaochu, Lyu, Yaxing, Liang, Zhixuan, Liu, Yibin, Xu, Congsheng, Guo, Xianda, Sui, Wei, Jin, Yaohui, Yang, Xiaokang, Xu, Yanyan, Mu, Yao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HyCodePolicy: Hybrid Language Controllers for Multimodal Monitoring and Decision in Embodied Agents
by: Liu, Yibin, et al.
Published: (2025)
by: Liu, Yibin, et al.
Published: (2025)
From Passive Observer to Active Critic: Reinforcement Learning Elicits Process Reasoning for Robotic Manipulation
by: Liu, Yibin, et al.
Published: (2026)
by: Liu, Yibin, et al.
Published: (2026)
HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System
by: Yang, Tianshuo, et al.
Published: (2026)
by: Yang, Tianshuo, et al.
Published: (2026)
UniVTAC: A Unified Simulation Platform for Visuo-Tactile Manipulation Data Generation, Learning, and Benchmarking
by: Chen, Baijun, et al.
Published: (2026)
by: Chen, Baijun, et al.
Published: (2026)
Rein3D: Reinforced 3D Indoor Scene Generation with Panoramic Video Diffusion Models
by: Wang, Dehui, et al.
Published: (2026)
by: Wang, Dehui, et al.
Published: (2026)
AnchorDP3: 3D Affordance Guided Sparse Diffusion Policy for Robotic Manipulation
by: Zhao, Ziyan, et al.
Published: (2025)
by: Zhao, Ziyan, et al.
Published: (2025)
Learning Chemical Reaction Representation with Reactant-Product Alignment
by: Zeng, Kaipeng, et al.
Published: (2024)
by: Zeng, Kaipeng, et al.
Published: (2024)
AniSDF: Fused-Granularity Neural Surfaces with Anisotropic Encoding for High-Fidelity 3D Reconstruction
by: Gao, Jingnan, et al.
Published: (2024)
by: Gao, Jingnan, et al.
Published: (2024)
ManiDP: Manipulability-Aware Diffusion Policy for Posture-Dependent Bimanual Manipulation
by: Li, Zhuo, et al.
Published: (2025)
by: Li, Zhuo, et al.
Published: (2025)
$π_0$-EqM: Equilibrium Matching for Closed-Loop Vision-Language-Action Control
by: Liu, Huanming, et al.
Published: (2026)
by: Liu, Huanming, et al.
Published: (2026)
Transferable Human Mobility Network Reconstruction with neuroGravity
by: Yang, Jinming, et al.
Published: (2026)
by: Yang, Jinming, et al.
Published: (2026)
Directional Texture Editing for 3D Models
by: Liu, Shengqi, et al.
Published: (2023)
by: Liu, Shengqi, et al.
Published: (2023)
EvaSurf: Efficient View-Aware Implicit Textured Surface Reconstruction
by: Gao, Jingnan, et al.
Published: (2023)
by: Gao, Jingnan, et al.
Published: (2023)
MoRE: 3D Visual Geometry Reconstruction Meets Mixture-of-Experts
by: Gao, Jingnan, et al.
Published: (2025)
by: Gao, Jingnan, et al.
Published: (2025)
MoE-DP: An MoE-Enhanced Diffusion Policy for Robust Long-Horizon Robotic Manipulation with Skill Decomposition and Failure Recovery
by: Cheng, Baiye, et al.
Published: (2025)
by: Cheng, Baiye, et al.
Published: (2025)
Dens3R: A Foundation Model for 3D Geometry Prediction
by: Fang, Xianze, et al.
Published: (2025)
by: Fang, Xianze, et al.
Published: (2025)
G3Flow: Generative 3D Semantic Flow for Pose-aware and Generalizable Object Manipulation
by: Chen, Tianxing, et al.
Published: (2024)
by: Chen, Tianxing, et al.
Published: (2024)
VO-DP: Semantic-Geometric Adaptive Diffusion Policy for Vision-Only Robotic Manipulation
by: Ni, Zehao, et al.
Published: (2025)
by: Ni, Zehao, et al.
Published: (2025)
ChemActor: Enhancing Automated Extraction of Chemical Synthesis Actions with LLM-Generated Data
by: Zhang, Yu, et al.
Published: (2025)
by: Zhang, Yu, et al.
Published: (2025)
Text-Augmented Multimodal LLMs for Chemical Reaction Condition Recommendation
by: Zhang, Yu, et al.
Published: (2024)
by: Zhang, Yu, et al.
Published: (2024)
UAlign: Pushing the Limit of Template-free Retrosynthesis Prediction with Unsupervised SMILES Alignment
by: Zeng, Kaipeng, et al.
Published: (2024)
by: Zeng, Kaipeng, et al.
Published: (2024)
HIMO: A New Benchmark for Full-Body Human Interacting with Multiple Objects
by: Lv, Xintao, et al.
Published: (2024)
by: Lv, Xintao, et al.
Published: (2024)
Learning Instruction-Guided Manipulation Affordance via Large Models for Embodied Robotic Tasks
by: Li, Dayou, et al.
Published: (2024)
by: Li, Dayou, et al.
Published: (2024)
SA-LSPL:Sequence-Aware Long- and Short- Term Preference Learning for next POI recommendation
by: Wang, Bin, et al.
Published: (2024)
by: Wang, Bin, et al.
Published: (2024)
RoboScript: Code Generation for Free-Form Manipulation Tasks across Real and Simulation
by: Chen, Junting, et al.
Published: (2024)
by: Chen, Junting, et al.
Published: (2024)
DexHiL: A Human-in-the-Loop Framework for Vision-Language-Action Model Post-Training in Dexterous Manipulation
by: Han, Yifan, et al.
Published: (2026)
by: Han, Yifan, et al.
Published: (2026)
Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
by: Liang, Zhixuan, et al.
Published: (2025)
by: Liang, Zhixuan, et al.
Published: (2025)
One-Policy-Fits-All: Geometry-Aware Action Latents for Cross-Embodiment Manipulation
by: Mu, Juncheng, et al.
Published: (2026)
by: Mu, Juncheng, et al.
Published: (2026)
OmniDP: Beyond-FOV Large-Workspace Humanoid Manipulation with Omnidirectional 3D Perception
by: Qu, Pei, et al.
Published: (2026)
by: Qu, Pei, et al.
Published: (2026)
Learning Tactile-Aware Quadrupedal Loco-Manipulation Policies
by: Zhou, Pokuang, et al.
Published: (2026)
by: Zhou, Pokuang, et al.
Published: (2026)
H$^3$DP: Triply-Hierarchical Diffusion Policy for Visuomotor Learning
by: Lu, Yiyang, et al.
Published: (2025)
by: Lu, Yiyang, et al.
Published: (2025)
RiEMann: Near Real-Time SE(3)-Equivariant Robot Manipulation without Point Cloud Segmentation
by: Gao, Chongkai, et al.
Published: (2024)
by: Gao, Chongkai, et al.
Published: (2024)
Expertise need not monopolize: Action-Specialized Mixture of Experts for Vision-Language-Action Learning
by: Shen, Weijie, et al.
Published: (2025)
by: Shen, Weijie, et al.
Published: (2025)
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025)
by: Yuan, Yifu, et al.
Published: (2025)
TrajGEOS: Trajectory Graph Enhanced Orientation-based Sequential Network for Mobility Prediction
by: Hu, Zhaoping, et al.
Published: (2024)
by: Hu, Zhaoping, et al.
Published: (2024)
BORA: Bridging Offline Reinforcement Learning and Online Residual Adaptation for Real-World Dexterous VLA Models
by: Chen, Zhongxi, et al.
Published: (2026)
by: Chen, Zhongxi, et al.
Published: (2026)
Perceiving and Acting in First-Person: A Dataset and Benchmark for Egocentric Human-Object-Human Interactions
by: Xu, Liang, et al.
Published: (2025)
by: Xu, Liang, et al.
Published: (2025)
Purging the Gray Zone: Latent-Geometric Denoising for Precise Knowledge Boundary Awareness
by: An, Hao, et al.
Published: (2026)
by: An, Hao, et al.
Published: (2026)
VolumeDP: Modeling Volumetric Representation for Manipulation Policy Learning
by: Zhou, Tianxing, et al.
Published: (2026)
by: Zhou, Tianxing, et al.
Published: (2026)
Hyper-DP3: Frequency-Aware Right-Sizing of 3D Diffusion Policies for Visuomotor Control
by: Zhang, Jinhao, et al.
Published: (2026)
by: Zhang, Jinhao, et al.
Published: (2026)
Similar Items
-
HyCodePolicy: Hybrid Language Controllers for Multimodal Monitoring and Decision in Embodied Agents
by: Liu, Yibin, et al.
Published: (2025) -
From Passive Observer to Active Critic: Reinforcement Learning Elicits Process Reasoning for Robotic Manipulation
by: Liu, Yibin, et al.
Published: (2026) -
HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System
by: Yang, Tianshuo, et al.
Published: (2026) -
UniVTAC: A Unified Simulation Platform for Visuo-Tactile Manipulation Data Generation, Learning, and Benchmarking
by: Chen, Baijun, et al.
Published: (2026) -
Rein3D: Reinforced 3D Indoor Scene Generation with Panoramic Video Diffusion Models
by: Wang, Dehui, et al.
Published: (2026)