ManiVID-3D: Generalizable View-Invariant Reinforcement Learning for Robotic Manipulation via Disentangled 3D Representations
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Zheng, Qu, Pei, Jia, Yufei, Zhou, Shihui, Ge, Haizhou, Cao, Jiahang, Zhou, Jinni, Zhou, Guyue, Ma, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Arm-Constrained Curriculum Learning for Loco-Manipulation of the Wheel-Legged Robot
by: Wang, Zifan, et al.
Published: (2024)
by: Wang, Zifan, et al.
Published: (2024)
OmniDP: Beyond-FOV Large-Workspace Humanoid Manipulation with Omnidirectional 3D Perception
by: Qu, Pei, et al.
Published: (2026)
by: Qu, Pei, et al.
Published: (2026)
FILIC: Dual-Loop Force-Guided Imitation Learning with Impedance Torque Control for Contact-Rich Manipulation Tasks
by: Ge, Haizhou, et al.
Published: (2025)
by: Ge, Haizhou, et al.
Published: (2025)
Learning Geometrically-Grounded 3D Visual Representations for View-Generalizable Robotic Manipulation
by: Zhang, Di, et al.
Published: (2026)
by: Zhang, Di, et al.
Published: (2026)
An Real-Sim-Real (RSR) Loop Framework for Generalizable Robotic Policy Transfer with Differentiable Simulation
by: Shi, Lu, et al.
Published: (2025)
by: Shi, Lu, et al.
Published: (2025)
Merging and Disentangling Views in Visual Reinforcement Learning for Robotic Manipulation
by: Almuzairee, Abdulaziz, et al.
Published: (2025)
by: Almuzairee, Abdulaziz, et al.
Published: (2025)
Embracing Bulky Objects with Humanoid Robots: Whole-Body Manipulation with Reinforcement Learning
by: Zheng, Chunxin, et al.
Published: (2025)
by: Zheng, Chunxin, et al.
Published: (2025)
MoViD: View-Invariant 3D Human Pose Estimation via Motion-View Disentanglement
by: Liu, Yejia, et al.
Published: (2026)
by: Liu, Yejia, et al.
Published: (2026)
Robot Tape Manipulation for 3D Printing
by: Tushar, Nahid, et al.
Published: (2024)
by: Tushar, Nahid, et al.
Published: (2024)
VLBiMan: Vision-Language Anchored One-Shot Demonstration Enables Generalizable Bimanual Robotic Manipulation
by: Zhou, Huayi, et al.
Published: (2025)
by: Zhou, Huayi, et al.
Published: (2025)
StyleDecoupler: Generalizable Artistic Style Disentanglement
by: Jia, Zexi, et al.
Published: (2026)
by: Jia, Zexi, et al.
Published: (2026)
MVISTA-4D: View-Consistent 4D World Model with Test-Time Action Inference for Robotic Manipulation
by: Wang, Jiaxu, et al.
Published: (2026)
by: Wang, Jiaxu, et al.
Published: (2026)
Dexterity-BEV: Aligning 3D World and Actions for Generalizable Robot Policies Learning
by: Zhou, Huayi, et al.
Published: (2026)
by: Zhou, Huayi, et al.
Published: (2026)
Integrating LMM Planners and 3D Skill Policies for Generalizable Manipulation
by: Li, Yuelei, et al.
Published: (2025)
by: Li, Yuelei, et al.
Published: (2025)
HiSplat: Hierarchical 3D Gaussian Splatting for Generalizable Sparse-View Reconstruction
by: Tang, Shengji, et al.
Published: (2024)
by: Tang, Shengji, et al.
Published: (2024)
OmniD: Generalizable Robot Manipulation Policy via Image-Based BEV Representation
by: Mao, Jilei, et al.
Published: (2025)
by: Mao, Jilei, et al.
Published: (2025)
3D Affordance Keypoint Detection for Robotic Manipulation
by: Liu, Zhiyang, et al.
Published: (2025)
by: Liu, Zhiyang, et al.
Published: (2025)
Generalizable Sparse-View 3D Reconstruction from Unconstrained Images
by: Gupta, Vinayak, et al.
Published: (2026)
by: Gupta, Vinayak, et al.
Published: (2026)
Generalizable Coarse-to-Fine Robot Manipulation via Language-Aligned 3D Keypoints
by: Hu, Jianshu, et al.
Published: (2025)
by: Hu, Jianshu, et al.
Published: (2025)
3D-Printed Hydraulic Fluidic Logic Circuitry for Soft Robots
by: Lin, Yuxin, et al.
Published: (2024)
by: Lin, Yuxin, et al.
Published: (2024)
ViewCraft3D: High-Fidelity and View-Consistent 3D Vector Graphics Synthesis
by: Wang, Chuang, et al.
Published: (2025)
by: Wang, Chuang, et al.
Published: (2025)
A Generalizable 3D Diffusion Framework for Low-Dose and Few-View Cardiac SPECT
by: Xie, Huidong, et al.
Published: (2024)
by: Xie, Huidong, et al.
Published: (2024)
Human Demonstrations are Generalizable Knowledge for Robots
by: Cui, Te, et al.
Published: (2023)
by: Cui, Te, et al.
Published: (2023)
ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping
by: Pang, Youxin, et al.
Published: (2024)
by: Pang, Youxin, et al.
Published: (2024)
Generalizable Humanoid Manipulation with 3D Diffusion Policies
by: Ze, Yanjie, et al.
Published: (2024)
by: Ze, Yanjie, et al.
Published: (2024)
Learning Generalizable 3D Manipulation With 10 Demonstrations
by: Ren, Yu, et al.
Published: (2024)
by: Ren, Yu, et al.
Published: (2024)
3DGR-CT: Sparse-View CT Reconstruction with a 3D Gaussian Representation
by: Li, Yingtai, et al.
Published: (2023)
by: Li, Yingtai, et al.
Published: (2023)
O$^3$Afford: One-Shot 3D Object-to-Object Affordance Grounding for Generalizable Robotic Manipulation
by: Tian, Tongxuan, et al.
Published: (2025)
by: Tian, Tongxuan, et al.
Published: (2025)
Closed-Loop Visuomotor Control with Generative Expectation for Robotic Manipulation
by: Bu, Qingwen, et al.
Published: (2024)
by: Bu, Qingwen, et al.
Published: (2024)
Forecasting Future Videos from Novel Views via Disentangled 3D Scene Representation
by: Yarram, Sudhir, et al.
Published: (2024)
by: Yarram, Sudhir, et al.
Published: (2024)
Disentangled Object-Centric Image Representation for Robotic Manipulation
by: Emukpere, David, et al.
Published: (2025)
by: Emukpere, David, et al.
Published: (2025)
Multi-Scale Accent Modeling and Disentangling for Multi-Speaker Multi-Accent Text-to-Speech Synthesis
by: Zhou, Xuehao, et al.
Published: (2024)
by: Zhou, Xuehao, et al.
Published: (2024)
UniLegs: Universal Multi-Legged Robot Control through Morphology-Agnostic Policy Distillation
by: Xi, Weijie, et al.
Published: (2025)
by: Xi, Weijie, et al.
Published: (2025)
Envision3D: One Image to 3D with Anchor Views Interpolation
by: Pang, Yatian, et al.
Published: (2024)
by: Pang, Yatian, et al.
Published: (2024)
No Need for Real 3D: Fusing 2D Vision with Pseudo 3D Representations for Robotic Manipulation Learning
by: Yu, Run, et al.
Published: (2025)
by: Yu, Run, et al.
Published: (2025)
Local Reactive Control for Mobile Manipulators with Whole-Body Safety in Complex Environments
by: Zheng, Chunxin, et al.
Published: (2025)
by: Zheng, Chunxin, et al.
Published: (2025)
Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images
by: Zhou, Bo, et al.
Published: (2026)
by: Zhou, Bo, et al.
Published: (2026)
A High-Fidelity Digital Twin for Robotic Manipulation Based on 3D Gaussian Splatting
by: Sun, Ziyang, et al.
Published: (2026)
by: Sun, Ziyang, et al.
Published: (2026)
GUIDE: A Diffusion-Based Autonomous Robot Exploration Framework Using Global Graph Inference
by: Che, Zijun, et al.
Published: (2025)
by: Che, Zijun, et al.
Published: (2025)
Towards Generalizable Vision-Language Robotic Manipulation: A Benchmark and LLM-guided 3D Policy
by: Garcia, Ricardo, et al.
Published: (2024)
by: Garcia, Ricardo, et al.
Published: (2024)
Similar Items
-
Arm-Constrained Curriculum Learning for Loco-Manipulation of the Wheel-Legged Robot
by: Wang, Zifan, et al.
Published: (2024) -
OmniDP: Beyond-FOV Large-Workspace Humanoid Manipulation with Omnidirectional 3D Perception
by: Qu, Pei, et al.
Published: (2026) -
FILIC: Dual-Loop Force-Guided Imitation Learning with Impedance Torque Control for Contact-Rich Manipulation Tasks
by: Ge, Haizhou, et al.
Published: (2025) -
Learning Geometrically-Grounded 3D Visual Representations for View-Generalizable Robotic Manipulation
by: Zhang, Di, et al.
Published: (2026) -
An Real-Sim-Real (RSR) Loop Framework for Generalizable Robotic Policy Transfer with Differentiable Simulation
by: Shi, Lu, et al.
Published: (2025)