SKT: Integrating State-Aware Keypoint Trajectories with Vision-Language Models for Robotic Garment Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Xin, Huang, Siyuan, Yu, Qiaojun, Jiang, Zhengkai, Hao, Ce, Zhu, Yimeng, Li, Hongsheng, Gao, Peng, Lu, Cewu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UniAff: A Unified Representation of Affordances for Tool Usage and Articulation with Vision-Language Models
by: Yu, Qiaojun, et al.
Published: (2024)
by: Yu, Qiaojun, et al.
Published: (2024)
ManipVQA: Injecting Robotic Affordance and Physically Grounded Information into Multi-Modal Large Language Models
by: Huang, Siyuan, et al.
Published: (2024)
by: Huang, Siyuan, et al.
Published: (2024)
A3VLM: Actionable Articulation-Aware Vision Language Model
by: Huang, Siyuan, et al.
Published: (2024)
by: Huang, Siyuan, et al.
Published: (2024)
VFP: Variational Flow-Matching Policy for Multi-Modal Robot Manipulation
by: Zhai, Xuanran, et al.
Published: (2025)
by: Zhai, Xuanran, et al.
Published: (2025)
ManiPose: A Comprehensive Benchmark for Pose-aware Object Manipulation in Robotics
by: Yu, Qiaojun, et al.
Published: (2024)
by: Yu, Qiaojun, et al.
Published: (2024)
GAMMA: Generalizable Articulation Modeling and Manipulation for Articulated Objects
by: Yu, Qiaojun, et al.
Published: (2023)
by: Yu, Qiaojun, et al.
Published: (2023)
ForceVLA2: Unleashing Hybrid Force-Position Control with Force Awareness for Contact-Rich Manipulation
by: Li, Yang, et al.
Published: (2026)
by: Li, Yang, et al.
Published: (2026)
CoFreeVLA: Collision-Free Dual-Arm Manipulation via Vision-Language-Action Model and Risk Estimation
by: Zhai, Xuanran, et al.
Published: (2026)
by: Zhai, Xuanran, et al.
Published: (2026)
EnerVerse: Envisioning Embodied Future Space for Robotics Manipulation
by: Huang, Siyuan, et al.
Published: (2025)
by: Huang, Siyuan, et al.
Published: (2025)
FoAR: Force-Aware Reactive Policy for Contact-Rich Robotic Manipulation
by: He, Zihao, et al.
Published: (2024)
by: He, Zihao, et al.
Published: (2024)
Towards Human-Like Manipulation through RL-Augmented Teleoperation and Mixture-of-Dexterous-Experts VLA
by: Tang, Tutian, et al.
Published: (2026)
by: Tang, Tutian, et al.
Published: (2026)
ArtGS:3D Gaussian Splatting for Interactive Visual-Physical Modeling and Manipulation of Articulated Objects
by: Yu, Qiaojun, et al.
Published: (2025)
by: Yu, Qiaojun, et al.
Published: (2025)
CAGE: Causal Attention Enables Data-Efficient Generalizable Robotic Manipulation
by: Xia, Shangning, et al.
Published: (2024)
by: Xia, Shangning, et al.
Published: (2024)
Utilizing Inpainting for Keypoint Detection for Vision-Based Control of Robotic Manipulators
by: Chatterjee, Sreejani, et al.
Published: (2026)
by: Chatterjee, Sreejani, et al.
Published: (2026)
Hybrid Consistency Policy: Decoupling Multi-Modal Diversity and Real-Time Efficiency in Robotic Manipulation
by: Zhao, Qianyou, et al.
Published: (2025)
by: Zhao, Qianyou, et al.
Published: (2025)
GarmentTracking: Category-Level Garment Pose Tracking
by: Xue, Han, et al.
Published: (2023)
by: Xue, Han, et al.
Published: (2023)
Generalizable Coarse-to-Fine Robot Manipulation via Language-Aligned 3D Keypoints
by: Hu, Jianshu, et al.
Published: (2025)
by: Hu, Jianshu, et al.
Published: (2025)
Towards Effective Utilization of Mixed-Quality Demonstrations in Robotic Manipulation via Segment-Level Selection and Optimization
by: Chen, Jingjing, et al.
Published: (2024)
by: Chen, Jingjing, et al.
Published: (2024)
Learning Efficient Robotic Garment Manipulation with Standardization
by: Zhou, Changshi, et al.
Published: (2025)
by: Zhou, Changshi, et al.
Published: (2025)
Real Garment Benchmark (RGBench): A Comprehensive Benchmark for Robotic Garment Manipulation featuring a High-Fidelity Scalable Simulator
by: Hu, Wenkang, et al.
Published: (2025)
by: Hu, Wenkang, et al.
Published: (2025)
3D Affordance Keypoint Detection for Robotic Manipulation
by: Liu, Zhiyang, et al.
Published: (2025)
by: Liu, Zhiyang, et al.
Published: (2025)
UniPlan: Vision-Language Task Planning for Mobile Manipulation with Unified PDDL Formulation
by: Ye, Haoming, et al.
Published: (2026)
by: Ye, Haoming, et al.
Published: (2026)
GarmentPile++: Affordance-Driven Cluttered Garments Retrieval with Vision-Language Reasoning
by: Li, Mingleyang, et al.
Published: (2026)
by: Li, Mingleyang, et al.
Published: (2026)
RPMArt: Towards Robust Perception and Manipulation for Articulated Objects
by: Wang, Junbo, et al.
Published: (2024)
by: Wang, Junbo, et al.
Published: (2024)
GarmentLab: A Unified Simulation and Benchmark for Garment Manipulation
by: Lu, Haoran, et al.
Published: (2024)
by: Lu, Haoran, et al.
Published: (2024)
ForceVLA: Enhancing VLA Models with a Force-aware MoE for Contact-rich Manipulation
by: Yu, Jiawen, et al.
Published: (2025)
by: Yu, Jiawen, et al.
Published: (2025)
Self-Wearing Adaptive Garments via Soft Robotic Unfurling
by: Kim, Nam Gyun, et al.
Published: (2025)
by: Kim, Nam Gyun, et al.
Published: (2025)
Hierarchical Audio-Visual-Proprioceptive Fusion for Precise Robotic Manipulation
by: Li, Siyuan, et al.
Published: (2026)
by: Li, Siyuan, et al.
Published: (2026)
AgentWorld: An Interactive Simulation Platform for Scene Construction and Mobile Robotic Manipulation
by: Zhang, Yizheng, et al.
Published: (2025)
by: Zhang, Yizheng, et al.
Published: (2025)
Learning Multi-Modal Trajectory Policies for Data-Efficient Robotic Manipulation
by: Chen, Zijia, et al.
Published: (2026)
by: Chen, Zijia, et al.
Published: (2026)
GraphGarment: Learning Garment Dynamics for Bimanual Cloth Manipulation Tasks
by: Chen, Wei, et al.
Published: (2025)
by: Chen, Wei, et al.
Published: (2025)
SkillVLA: Tackling Combinatorial Diversity in Dual-Arm Manipulation via Skill Reuse
by: Zhai, Xuanran, et al.
Published: (2026)
by: Zhai, Xuanran, et al.
Published: (2026)
Robotic Manipulation Framework Based on Semantic Keypoints for Packing Shoes of Different Sizes, Shapes, and Softness
by: Dong, Yi, et al.
Published: (2025)
by: Dong, Yi, et al.
Published: (2025)
KUDA: Keypoints to Unify Dynamics Learning and Visual Prompting for Open-Vocabulary Robotic Manipulation
by: Liu, Zixian, et al.
Published: (2025)
by: Liu, Zixian, et al.
Published: (2025)
DISCO: Language-Guided Manipulation with Diffusion Policies and Constrained Inpainting
by: Hao, Ce, et al.
Published: (2024)
by: Hao, Ce, et al.
Published: (2024)
Confusion-Aware In-Context-Learning for Vision-Language Models in Robotic Manipulation
by: He, Yayun, et al.
Published: (2026)
by: He, Yayun, et al.
Published: (2026)
Abstracting Robot Manipulation Skills via Mixture-of-Experts Diffusion Policies
by: Hao, Ce, et al.
Published: (2026)
by: Hao, Ce, et al.
Published: (2026)
SKIL: Semantic Keypoint Imitation Learning for Generalizable Data-efficient Manipulation
by: Wang, Shengjie, et al.
Published: (2025)
by: Wang, Shengjie, et al.
Published: (2025)
ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
by: Huang, Wenlong, et al.
Published: (2024)
by: Huang, Wenlong, et al.
Published: (2024)
AFFORD2ACT: Affordance-Guided Automatic Keypoint Selection for Generalizable and Lightweight Robotic Manipulation
by: Singh, Anukriti, et al.
Published: (2025)
by: Singh, Anukriti, et al.
Published: (2025)
Similar Items
-
UniAff: A Unified Representation of Affordances for Tool Usage and Articulation with Vision-Language Models
by: Yu, Qiaojun, et al.
Published: (2024) -
ManipVQA: Injecting Robotic Affordance and Physically Grounded Information into Multi-Modal Large Language Models
by: Huang, Siyuan, et al.
Published: (2024) -
A3VLM: Actionable Articulation-Aware Vision Language Model
by: Huang, Siyuan, et al.
Published: (2024) -
VFP: Variational Flow-Matching Policy for Multi-Modal Robot Manipulation
by: Zhai, Xuanran, et al.
Published: (2025) -
ManiPose: A Comprehensive Benchmark for Pose-aware Object Manipulation in Robotics
by: Yu, Qiaojun, et al.
Published: (2024)