OmniD: Generalizable Robot Manipulation Policy via Image-Based BEV Representation
Fuente:
arXiv
Saved in:
| Main Authors: | Mao, Jilei, Guan, Jiarui, Tang, Yingjuan, Hu, Qirui, Li, Zhihang, Yu, Junjie, Mao, Yongjie, Sun, Yunzhe, Liu, Shuang, Ju, Xiaozhu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dexterity-BEV: Aligning 3D World and Actions for Generalizable Robot Policies Learning
by: Zhou, Huayi, et al.
Published: (2026)
by: Zhou, Huayi, et al.
Published: (2026)
RAM: Retrieval-Based Affordance Transfer for Generalizable Zero-Shot Robotic Manipulation
by: Kuang, Yuxuan, et al.
Published: (2024)
by: Kuang, Yuxuan, et al.
Published: (2024)
VO-DP: Semantic-Geometric Adaptive Diffusion Policy for Vision-Only Robotic Manipulation
by: Ni, Zehao, et al.
Published: (2025)
by: Ni, Zehao, et al.
Published: (2025)
Octopus-inspired Distributed Control for Soft Robotic Arms: A Graph Neural Network-Based Attention Policy with Environmental Interaction
by: Hou, Linxin, et al.
Published: (2026)
by: Hou, Linxin, et al.
Published: (2026)
Efficient Training of Generalizable Visuomotor Policies via Control-Aware Augmentation
by: Zhao, Yinuo, et al.
Published: (2024)
by: Zhao, Yinuo, et al.
Published: (2024)
RGMP: Recurrent Geometric-prior Multimodal Policy for Generalizable Humanoid Robot Manipulation
by: Li, Xuetao, et al.
Published: (2025)
by: Li, Xuetao, et al.
Published: (2025)
Discrete Policy: Learning Disentangled Action Space for Multi-Task Robotic Manipulation
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
BEV-ODOM2: Enhanced BEV-based Monocular Visual Odometry with PV-BEV Fusion and Dense Flow Supervision for Ground Robots
by: Wei, Yufei, et al.
Published: (2025)
by: Wei, Yufei, et al.
Published: (2025)
Language-Guided Object-Centric Diffusion Policy for Generalizable and Collision-Aware Robotic Manipulation
by: Li, Hang, et al.
Published: (2024)
by: Li, Hang, et al.
Published: (2024)
Learning Geometrically-Grounded 3D Visual Representations for View-Generalizable Robotic Manipulation
by: Zhang, Di, et al.
Published: (2026)
by: Zhang, Di, et al.
Published: (2026)
ARM: Advantage Reward Modeling for Long-Horizon Manipulation
by: Mao, Yiming, et al.
Published: (2026)
by: Mao, Yiming, et al.
Published: (2026)
Skill-Aware Diffusion for Generalizable Robotic Manipulation
by: Huang, Aoshen, et al.
Published: (2026)
by: Huang, Aoshen, et al.
Published: (2026)
BEV-ODOM: Reducing Scale Drift in Monocular Visual Odometry with BEV Representation
by: Wei, Yufei, et al.
Published: (2024)
by: Wei, Yufei, et al.
Published: (2024)
Scaling Diffusion Policy in Transformer to 1 Billion Parameters for Robotic Manipulation
by: Zhu, Minjie, et al.
Published: (2024)
by: Zhu, Minjie, et al.
Published: (2024)
ImaginationPolicy: Towards Generalizable, Precise and Reliable End-to-End Policy for Robotic Manipulation
by: Lu, Dekun, et al.
Published: (2025)
by: Lu, Dekun, et al.
Published: (2025)
A Quantitative Comparison of Centralised and Distributed Reinforcement Learning-Based Control for Soft Robotic Arms
by: Hou, Linxin, et al.
Published: (2025)
by: Hou, Linxin, et al.
Published: (2025)
RoboOmni: Proactive Robot Manipulation in Omni-modal Context
by: Wang, Siyin, et al.
Published: (2025)
by: Wang, Siyin, et al.
Published: (2025)
Dual-BEV Nav: Dual-layer BEV-based Heuristic Path Planning for Robotic Navigation in Unstructured Outdoor Environments
by: Zhang, Jianfeng, et al.
Published: (2025)
by: Zhang, Jianfeng, et al.
Published: (2025)
WristWorld: Generating Wrist-Views via 4D World Models for Robotic Manipulation
by: Qian, Zezhong, et al.
Published: (2025)
by: Qian, Zezhong, et al.
Published: (2025)
DSPv2: Improved Dense Policy for Effective and Generalizable Whole-body Mobile Manipulation
by: Su, Yue, et al.
Published: (2025)
by: Su, Yue, et al.
Published: (2025)
DexGarmentLab: Dexterous Garment Manipulation Environment with Generalizable Policy
by: Wang, Yuran, et al.
Published: (2025)
by: Wang, Yuran, et al.
Published: (2025)
Transferring Foundation Models for Generalizable Robotic Manipulation
by: Yang, Jiange, et al.
Published: (2023)
by: Yang, Jiange, et al.
Published: (2023)
PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation
by: Li, Yutai, et al.
Published: (2026)
by: Li, Yutai, et al.
Published: (2026)
OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation
by: Zheng, Yuhang, et al.
Published: (2026)
by: Zheng, Yuhang, et al.
Published: (2026)
Learning Efficient Robotic Garment Manipulation with Standardization
by: Zhou, Changshi, et al.
Published: (2025)
by: Zhou, Changshi, et al.
Published: (2025)
FMB: a Functional Manipulation Benchmark for Generalizable Robotic Learning
by: Luo, Jianlan, et al.
Published: (2024)
by: Luo, Jianlan, et al.
Published: (2024)
Learning the Generalizable Manipulation Skills on Soft-body Tasks via Guided Self-attention Behavior Cloning Policy
by: Li, Xuetao, et al.
Published: (2024)
by: Li, Xuetao, et al.
Published: (2024)
Generalizable Coarse-to-Fine Robot Manipulation via Language-Aligned 3D Keypoints
by: Hu, Jianshu, et al.
Published: (2025)
by: Hu, Jianshu, et al.
Published: (2025)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
by: Zhang, Jesse, et al.
Published: (2025)
by: Zhang, Jesse, et al.
Published: (2025)
Diffusion Stabilizer Policy for Automated Surgical Robot Manipulations
by: Ho, Chonlam, et al.
Published: (2025)
by: Ho, Chonlam, et al.
Published: (2025)
TC-IDM: Grounding Video Generation for Executable Zero-shot Robot Motion
by: Mi, Weishi, et al.
Published: (2026)
by: Mi, Weishi, et al.
Published: (2026)
YOR: Your Own Mobile Manipulator for Generalizable Robotics
by: Anjaria, Manan H, et al.
Published: (2026)
by: Anjaria, Manan H, et al.
Published: (2026)
Veo-Act: How Far Can Frontier Video Models Advance Generalizable Robot Manipulation?
by: Zhang, Zhongru, et al.
Published: (2026)
by: Zhang, Zhongru, et al.
Published: (2026)
Two by Two: Learning Multi-Task Pairwise Objects Assembly for Generalizable Robot Manipulation
by: Qi, Yu, et al.
Published: (2025)
by: Qi, Yu, et al.
Published: (2025)
RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
Generalizable Image Repair for Robust Visual Control
by: Sobolewski, Carson, et al.
Published: (2025)
by: Sobolewski, Carson, et al.
Published: (2025)
DYMO-Hair: Generalizable Volumetric Dynamics Modeling for Robot Hair Manipulation
by: Zhao, Chengyang, et al.
Published: (2025)
by: Zhao, Chengyang, et al.
Published: (2025)
CAGE: Causal Attention Enables Data-Efficient Generalizable Robotic Manipulation
by: Xia, Shangning, et al.
Published: (2024)
by: Xia, Shangning, et al.
Published: (2024)
CABTO: Context-Aware Behavior Tree Grounding for Robot Manipulation
by: Cai, Yishuai, et al.
Published: (2026)
by: Cai, Yishuai, et al.
Published: (2026)
AffordGen: Generating Diverse Demonstrations for Generalizable Object Manipulation with Afford Correspondence
by: Zhang, Jiawei, et al.
Published: (2026)
by: Zhang, Jiawei, et al.
Published: (2026)
Similar Items
-
Dexterity-BEV: Aligning 3D World and Actions for Generalizable Robot Policies Learning
by: Zhou, Huayi, et al.
Published: (2026) -
RAM: Retrieval-Based Affordance Transfer for Generalizable Zero-Shot Robotic Manipulation
by: Kuang, Yuxuan, et al.
Published: (2024) -
VO-DP: Semantic-Geometric Adaptive Diffusion Policy for Vision-Only Robotic Manipulation
by: Ni, Zehao, et al.
Published: (2025) -
Octopus-inspired Distributed Control for Soft Robotic Arms: A Graph Neural Network-Based Attention Policy with Environmental Interaction
by: Hou, Linxin, et al.
Published: (2026) -
Efficient Training of Generalizable Visuomotor Policies via Control-Aware Augmentation
by: Zhao, Yinuo, et al.
Published: (2024)