Learning Manipulation by Predicting Interaction
Fuente:
arXiv
Saved in:
| Main Authors: | Zeng, Jia, Bu, Qingwen, Wang, Bangjun, Xia, Wenke, Chen, Li, Dong, Hao, Song, Haoming, Wang, Dong, Hu, Di, Luo, Ping, Cui, Heming, Zhao, Bin, Li, Xuelong, Qiao, Yu, Li, Hongyang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation
by: Bu, Qingwen, et al.
Published: (2024)
by: Bu, Qingwen, et al.
Published: (2024)
Closed-Loop Visuomotor Control with Generative Expectation for Robotic Manipulation
by: Bu, Qingwen, et al.
Published: (2024)
by: Bu, Qingwen, et al.
Published: (2024)
Kinematic-aware Prompting for Generalizable Articulated Object Manipulation with LLMs
by: Xia, Wenke, et al.
Published: (2023)
by: Xia, Wenke, et al.
Published: (2023)
CodeCoT: Tackling Code Syntax Errors in CoT Reasoning for Code Generation
by: Huang, Dong, et al.
Published: (2023)
by: Huang, Dong, et al.
Published: (2023)
KOI: Accelerating Online Imitation Learning via Hybrid Key-state Guidance
by: Lu, Jingxian, et al.
Published: (2024)
by: Lu, Jingxian, et al.
Published: (2024)
Depth Helps: Improving Pre-trained RGB-based Policy with Depth Information Injection
by: Pang, Xincheng, et al.
Published: (2024)
by: Pang, Xincheng, et al.
Published: (2024)
Play to the Score: Stage-Guided Dynamic Multi-Sensory Fusion for Robotic Manipulation
by: Feng, Ruoxuan, et al.
Published: (2024)
by: Feng, Ruoxuan, et al.
Published: (2024)
SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation
by: Zhang, Junjie, et al.
Published: (2024)
by: Zhang, Junjie, et al.
Published: (2024)
Think Small, Act Big: Primitive Prompt Learning for Lifelong Robot Manipulation
by: Yao, Yuanqi, et al.
Published: (2025)
by: Yao, Yuanqi, et al.
Published: (2025)
AgentCoder: Multi-Agent-based Code Generation with Iterative Testing and Optimisation
by: Huang, Dong, et al.
Published: (2023)
by: Huang, Dong, et al.
Published: (2023)
Bias Testing and Mitigation in LLM-based Code Generation
by: Huang, Dong, et al.
Published: (2023)
by: Huang, Dong, et al.
Published: (2023)
Adversarial Feature Map Pruning for Backdoor
by: Huang, Dong, et al.
Published: (2023)
by: Huang, Dong, et al.
Published: (2023)
FastUMI-100K: Advancing Data-driven Robotic Manipulation with a Large-scale UMI-style Dataset
by: Liu, Kehui, et al.
Published: (2025)
by: Liu, Kehui, et al.
Published: (2025)
Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning
by: Gao, Xianqiang, et al.
Published: (2026)
by: Gao, Xianqiang, et al.
Published: (2026)
LaneSegNet: Map Learning with Lane Segment Perception for Autonomous Driving
by: Li, Tianyu, et al.
Published: (2023)
by: Li, Tianyu, et al.
Published: (2023)
Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation
by: Tian, Yang, et al.
Published: (2024)
by: Tian, Yang, et al.
Published: (2024)
ZeroWBC: Learning Natural Visuomotor Humanoid Control Directly from Human Egocentric Video
by: Yang, Haoran, et al.
Published: (2026)
by: Yang, Haoran, et al.
Published: (2026)
Phoenix: A Motion-based Self-Reflection Framework for Fine-grained Robotic Action Correction
by: Xia, Wenke, et al.
Published: (2025)
by: Xia, Wenke, et al.
Published: (2025)
Agility Meets Stability: Versatile Humanoid Control with Heterogeneous Data
by: Pan, Yixuan, et al.
Published: (2025)
by: Pan, Yixuan, et al.
Published: (2025)
Embodied Understanding of Driving Scenarios
by: Zhou, Yunsong, et al.
Published: (2024)
by: Zhou, Yunsong, et al.
Published: (2024)
Two Heads are Better than One: Robust Learning Meets Multi-branch Models
by: Zhang, Zongyuan, et al.
Published: (2022)
by: Zhang, Zongyuan, et al.
Published: (2022)
WholeBodyVLA: Towards Unified Latent VLA for Whole-Body Loco-Manipulation Control
by: Jiang, Haoran, et al.
Published: (2025)
by: Jiang, Haoran, et al.
Published: (2025)
UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
by: Bu, Qingwen, et al.
Published: (2025)
by: Bu, Qingwen, et al.
Published: (2025)
Themis: Automatic and Efficient Deep Learning System Testing with Strong Fault Detection Capability
by: Huang, Dong, et al.
Published: (2024)
by: Huang, Dong, et al.
Published: (2024)
FastUMI: A Scalable and Hardware-Independent Universal Manipulation Interface with Dataset
by: Zhaxizhuoma, et al.
Published: (2024)
by: Zhaxizhuoma, et al.
Published: (2024)
Physics in Next-token Prediction
by: An, Hongjun, et al.
Published: (2024)
by: An, Hongjun, et al.
Published: (2024)
$χ_{0}$: Resource-Aware Robust Manipulation via Taming Distributional Inconsistencies
by: Yu, Checheng, et al.
Published: (2026)
by: Yu, Checheng, et al.
Published: (2026)
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
by: Qu, Delin, et al.
Published: (2025)
by: Qu, Delin, et al.
Published: (2025)
Characterized Diffusion Networks for Enhanced Autonomous Driving Trajectory Prediction
by: Li, Haoming
Published: (2024)
by: Li, Haoming
Published: (2024)
LiveScene: Language Embedding Interactive Radiance Fields for Physical Scene Rendering and Control
by: Qu, Delin, et al.
Published: (2024)
by: Qu, Delin, et al.
Published: (2024)
Improving Transferable Targeted Attacks with Feature Tuning Mixup
by: Liang, Kaisheng, et al.
Published: (2024)
by: Liang, Kaisheng, et al.
Published: (2024)
EgoHumanoid: Unlocking In-the-Wild Loco-Manipulation with Robot-Free Egocentric Demonstration
by: Shi, Modi, et al.
Published: (2026)
by: Shi, Modi, et al.
Published: (2026)
MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation
by: Zhang, Pingrui, et al.
Published: (2025)
by: Zhang, Pingrui, et al.
Published: (2025)
When would Vision-Proprioception Policies Fail in Robotic Manipulation?
by: Lu, Jingxian, et al.
Published: (2026)
by: Lu, Jingxian, et al.
Published: (2026)
Night-to-Day Translation via Illumination Degradation Disentanglement
by: Lan, Guanzhou, et al.
Published: (2024)
by: Lan, Guanzhou, et al.
Published: (2024)
COHERENT: Collaboration of Heterogeneous Multi-Robot System with Large Language Models
by: Liu, Kehui, et al.
Published: (2024)
by: Liu, Kehui, et al.
Published: (2024)
Closed-Loop Action Chunks with Dynamic Corrections for Training-Free Diffusion Policy
by: Wu, Pengyuan, et al.
Published: (2026)
by: Wu, Pengyuan, et al.
Published: (2026)
Open-Vocabulary Octree-Graph for 3D Scene Understanding
by: Wang, Zhigang, et al.
Published: (2024)
by: Wang, Zhigang, et al.
Published: (2024)
Is Diversity All You Need for Scalable Robotic Manipulation?
by: Shi, Modi, et al.
Published: (2025)
by: Shi, Modi, et al.
Published: (2025)
AIC MLLM: Autonomous Interactive Correction MLLM for Robust Robotic Manipulation
by: Xiong, Chuyan, et al.
Published: (2024)
by: Xiong, Chuyan, et al.
Published: (2024)
Similar Items
-
Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation
by: Bu, Qingwen, et al.
Published: (2024) -
Closed-Loop Visuomotor Control with Generative Expectation for Robotic Manipulation
by: Bu, Qingwen, et al.
Published: (2024) -
Kinematic-aware Prompting for Generalizable Articulated Object Manipulation with LLMs
by: Xia, Wenke, et al.
Published: (2023) -
CodeCoT: Tackling Code Syntax Errors in CoT Reasoning for Code Generation
by: Huang, Dong, et al.
Published: (2023) -
KOI: Accelerating Online Imitation Learning via Hybrid Key-state Guidance
by: Lu, Jingxian, et al.
Published: (2024)