Generalizable Coarse-to-Fine Robot Manipulation via Language-Aligned 3D Keypoints
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Jianshu, Wang, Lidi, Li, Shujia, Jiang, Yunpeng, Li, Xiao, Weng, Paul, Ban, Yutong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Time Reversal Symmetry for Efficient Robotic Manipulations in Deep Reinforcement Learning
by: Jiang, Yunpeng, et al.
Published: (2025)
by: Jiang, Yunpeng, et al.
Published: (2025)
DSSP: Diffusion State Space Policy with Full-History Encoding
by: Guan, Zhiyuan, et al.
Published: (2026)
by: Guan, Zhiyuan, et al.
Published: (2026)
Diffusion Stabilizer Policy for Automated Surgical Robot Manipulations
by: Ho, Chonlam, et al.
Published: (2025)
by: Ho, Chonlam, et al.
Published: (2025)
SKIL: Semantic Keypoint Imitation Learning for Generalizable Data-efficient Manipulation
by: Wang, Shengjie, et al.
Published: (2025)
by: Wang, Shengjie, et al.
Published: (2025)
Language-Guided Grasp Detection with Coarse-to-Fine Learning for Robotic Manipulation
by: Jiang, Zebin, et al.
Published: (2025)
by: Jiang, Zebin, et al.
Published: (2025)
3D Affordance Keypoint Detection for Robotic Manipulation
by: Liu, Zhiyang, et al.
Published: (2025)
by: Liu, Zhiyang, et al.
Published: (2025)
AFFORD2ACT: Affordance-Guided Automatic Keypoint Selection for Generalizable and Lightweight Robotic Manipulation
by: Singh, Anukriti, et al.
Published: (2025)
by: Singh, Anukriti, et al.
Published: (2025)
OmniD: Generalizable Robot Manipulation Policy via Image-Based BEV Representation
by: Mao, Jilei, et al.
Published: (2025)
by: Mao, Jilei, et al.
Published: (2025)
OMP: One-step Meanflow Policy with Directional Alignment
by: Fang, Han, et al.
Published: (2025)
by: Fang, Han, et al.
Published: (2025)
SSP: Safety-guaranteed Surgical Policy via Joint Optimization of Behavioral and Spatial Constraints
by: Hu, Jianshu, et al.
Published: (2026)
by: Hu, Jianshu, et al.
Published: (2026)
Dexterity-BEV: Aligning 3D World and Actions for Generalizable Robot Policies Learning
by: Zhou, Huayi, et al.
Published: (2026)
by: Zhou, Huayi, et al.
Published: (2026)
SKT: Integrating State-Aware Keypoint Trajectories with Vision-Language Models for Robotic Garment Manipulation
by: Li, Xin, et al.
Published: (2024)
by: Li, Xin, et al.
Published: (2024)
State-Novelty Guided Action Persistence in Deep Reinforcement Learning
by: Hu, Jianshu, et al.
Published: (2024)
by: Hu, Jianshu, et al.
Published: (2024)
Distilling and Retrieving Generalizable Knowledge for Robot Manipulation via Language Corrections
by: Zha, Lihan, et al.
Published: (2023)
by: Zha, Lihan, et al.
Published: (2023)
Robotic Manipulation Framework Based on Semantic Keypoints for Packing Shoes of Different Sizes, Shapes, and Softness
by: Dong, Yi, et al.
Published: (2025)
by: Dong, Yi, et al.
Published: (2025)
Utilizing Inpainting for Keypoint Detection for Vision-Based Control of Robotic Manipulators
by: Chatterjee, Sreejani, et al.
Published: (2026)
by: Chatterjee, Sreejani, et al.
Published: (2026)
KUDA: Keypoints to Unify Dynamics Learning and Visual Prompting for Open-Vocabulary Robotic Manipulation
by: Liu, Zixian, et al.
Published: (2025)
by: Liu, Zixian, et al.
Published: (2025)
ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
by: Huang, Wenlong, et al.
Published: (2024)
by: Huang, Wenlong, et al.
Published: (2024)
Language-Guided Object-Centric Diffusion Policy for Generalizable and Collision-Aware Robotic Manipulation
by: Li, Hang, et al.
Published: (2024)
by: Li, Hang, et al.
Published: (2024)
PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation
by: Li, Yutai, et al.
Published: (2026)
by: Li, Yutai, et al.
Published: (2026)
RGMP: Recurrent Geometric-prior Multimodal Policy for Generalizable Humanoid Robot Manipulation
by: Li, Xuetao, et al.
Published: (2025)
by: Li, Xuetao, et al.
Published: (2025)
Coarse-to-Fine 3D Keyframe Transporter
by: Zhu, Xupeng, et al.
Published: (2025)
by: Zhu, Xupeng, et al.
Published: (2025)
Towards Generalizable Robotic Manipulation in Dynamic Environments
by: Fang, Heng, et al.
Published: (2026)
by: Fang, Heng, et al.
Published: (2026)
Learning Geometrically-Grounded 3D Visual Representations for View-Generalizable Robotic Manipulation
by: Zhang, Di, et al.
Published: (2026)
by: Zhang, Di, et al.
Published: (2026)
Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation
by: Chen, Shizhe, et al.
Published: (2025)
by: Chen, Shizhe, et al.
Published: (2025)
ManiVID-3D: Generalizable View-Invariant Reinforcement Learning for Robotic Manipulation via Disentangled 3D Representations
by: Li, Zheng, et al.
Published: (2025)
by: Li, Zheng, et al.
Published: (2025)
Skill-Aware Diffusion for Generalizable Robotic Manipulation
by: Huang, Aoshen, et al.
Published: (2026)
by: Huang, Aoshen, et al.
Published: (2026)
RoboKeyGen: Robot Pose and Joint Angles Estimation via Diffusion-based 3D Keypoint Generation
by: Tian, Yang, et al.
Published: (2024)
by: Tian, Yang, et al.
Published: (2024)
Integrating LMM Planners and 3D Skill Policies for Generalizable Manipulation
by: Li, Yuelei, et al.
Published: (2025)
by: Li, Yuelei, et al.
Published: (2025)
Generalizable Motion Policies through Keypoint Parameterization and Transportation Maps
by: Franzese, Giovanni, et al.
Published: (2024)
by: Franzese, Giovanni, et al.
Published: (2024)
Bridging VLM and KMP: Enabling Fine-grained robotic manipulation via Semantic Keypoints Representation
by: Zhu, Junjie, et al.
Published: (2025)
by: Zhu, Junjie, et al.
Published: (2025)
CLASP: General-Purpose Clothes Manipulation with Semantic Keypoints
by: Deng, Yuhong, et al.
Published: (2025)
by: Deng, Yuhong, et al.
Published: (2025)
Transferring Foundation Models for Generalizable Robotic Manipulation
by: Yang, Jiange, et al.
Published: (2023)
by: Yang, Jiange, et al.
Published: (2023)
Visuo-Tactile Keypoint Correspondences for Object Manipulation
by: Kim, Jeong-Jung, et al.
Published: (2024)
by: Kim, Jeong-Jung, et al.
Published: (2024)
Understanding and Reducing the Class-Dependent Effects of Data Augmentation with A Two-Player Game Approach
by: Jiang, Yunpeng, et al.
Published: (2024)
by: Jiang, Yunpeng, et al.
Published: (2024)
FoldNet: Learning Generalizable Closed-Loop Policy for Garment Folding via Keypoint-Driven Asset and Demonstration Synthesis
by: Chen, Yuxing, et al.
Published: (2025)
by: Chen, Yuxing, et al.
Published: (2025)
Learning Generalizable Language-Conditioned Cloth Manipulation from Long Demonstrations
by: Zhao, Hanyi, et al.
Published: (2025)
by: Zhao, Hanyi, et al.
Published: (2025)
Keypoint Detection Technique for Image-Based Visual Servoing of Manipulators
by: Amiri, Niloufar, et al.
Published: (2024)
by: Amiri, Niloufar, et al.
Published: (2024)
Towards Generalizable Vision-Language Robotic Manipulation: A Benchmark and LLM-guided 3D Policy
by: Garcia, Ricardo, et al.
Published: (2024)
by: Garcia, Ricardo, et al.
Published: (2024)
$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills
by: Xiao, Siyao, et al.
Published: (2026)
by: Xiao, Siyao, et al.
Published: (2026)
Similar Items
-
Time Reversal Symmetry for Efficient Robotic Manipulations in Deep Reinforcement Learning
by: Jiang, Yunpeng, et al.
Published: (2025) -
DSSP: Diffusion State Space Policy with Full-History Encoding
by: Guan, Zhiyuan, et al.
Published: (2026) -
Diffusion Stabilizer Policy for Automated Surgical Robot Manipulations
by: Ho, Chonlam, et al.
Published: (2025) -
SKIL: Semantic Keypoint Imitation Learning for Generalizable Data-efficient Manipulation
by: Wang, Shengjie, et al.
Published: (2025) -
Language-Guided Grasp Detection with Coarse-to-Fine Learning for Robotic Manipulation
by: Jiang, Zebin, et al.
Published: (2025)