VLMPC: Vision-Language Model Predictive Control for Robotic Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Wentao, Chen, Jiaming, Meng, Ziyu, Mao, Donghui, Song, Ran, Zhang, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vision-Language Model Predictive Control for Manipulation Planning and Trajectory Generation
by: Chen, Jiaming, et al.
Published: (2025)
by: Chen, Jiaming, et al.
Published: (2025)
Skill-Aware Diffusion for Generalizable Robotic Manipulation
by: Huang, Aoshen, et al.
Published: (2026)
by: Huang, Aoshen, et al.
Published: (2026)
SafeFall: Learning Protective Control for Humanoid Robots
by: Meng, Ziyu, et al.
Published: (2025)
by: Meng, Ziyu, et al.
Published: (2025)
VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation
by: Zhao, Wei, et al.
Published: (2025)
by: Zhao, Wei, et al.
Published: (2025)
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey
by: Shao, Rui, et al.
Published: (2025)
by: Shao, Rui, et al.
Published: (2025)
ManiSoft: Towards Vision-Language Manipulation for Soft Continuum Robotics
by: Wei, Ziyu, et al.
Published: (2026)
by: Wei, Ziyu, et al.
Published: (2026)
ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation
by: Zhao, Enyu, et al.
Published: (2025)
by: Zhao, Enyu, et al.
Published: (2025)
CrayonRobo: Object-Centric Prompt-Driven Vision-Language-Action Model for Robotic Manipulation
by: Li, Xiaoqi, et al.
Published: (2025)
by: Li, Xiaoqi, et al.
Published: (2025)
AHA: A Vision-Language-Model for Detecting and Reasoning Over Failures in Robotic Manipulation
by: Duan, Jiafei, et al.
Published: (2024)
by: Duan, Jiafei, et al.
Published: (2024)
Confusion-Aware In-Context-Learning for Vision-Language Models in Robotic Manipulation
by: He, Yayun, et al.
Published: (2026)
by: He, Yayun, et al.
Published: (2026)
Task-oriented Robotic Manipulation with Vision Language Models
by: Guran, Nurhan Bulus, et al.
Published: (2024)
by: Guran, Nurhan Bulus, et al.
Published: (2024)
Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation
by: Feng, Yunhai, et al.
Published: (2025)
by: Feng, Yunhai, et al.
Published: (2025)
Manipulate-Anything: Automating Real-World Robots using Vision-Language Models
by: Duan, Jiafei, et al.
Published: (2024)
by: Duan, Jiafei, et al.
Published: (2024)
Reactive Model Predictive Contouring Control for Robot Manipulators
by: Yoon, Junheon, et al.
Published: (2025)
by: Yoon, Junheon, et al.
Published: (2025)
Integrating Controllable Motion Skills from Demonstrations
by: Liao, Honghao, et al.
Published: (2024)
by: Liao, Honghao, et al.
Published: (2024)
VLATest: Testing and Evaluating Vision-Language-Action Models for Robotic Manipulation
by: Wang, Zhijie, et al.
Published: (2024)
by: Wang, Zhijie, et al.
Published: (2024)
Long-VLA: Unleashing Long-Horizon Capability of Vision Language Action Model for Robot Manipulation
by: Fan, Yiguo, et al.
Published: (2025)
by: Fan, Yiguo, et al.
Published: (2025)
Online Robot Navigation and Manipulation with Distilled Vision-Language Models
by: Liu, Kangcheng
Published: (2024)
by: Liu, Kangcheng
Published: (2024)
OmniVIC: A Self-Improving Variable Impedance Controller with Vision-Language In-Context Learning for Safe Robotic Manipulation
by: Zhang, Heng, et al.
Published: (2025)
by: Zhang, Heng, et al.
Published: (2025)
Robot Trajectron: Trajectory Prediction-based Shared Control for Robot Manipulation
by: Song, Pinhao, et al.
Published: (2024)
by: Song, Pinhao, et al.
Published: (2024)
Robotic Manipulation is Vision-to-Geometry Mapping ($f(v) \rightarrow G$): Vision-Geometry Backbones over Language and Video Models
by: Song, Zijian, et al.
Published: (2026)
by: Song, Zijian, et al.
Published: (2026)
DAM-VLA: A Dynamic Action Model-Based Vision-Language-Action Framework for Robot Manipulation
by: Peng, Xiongfeng, et al.
Published: (2026)
by: Peng, Xiongfeng, et al.
Published: (2026)
LaST$_{0}$: Latent Spatio-Temporal Chain-of-Thought for Robotic Vision-Language-Action Model
by: Liu, Zhuoyang, et al.
Published: (2026)
by: Liu, Zhuoyang, et al.
Published: (2026)
StyleLoco: Generative Adversarial Distillation for Natural Humanoid Robot Locomotion
by: Ma, Le, et al.
Published: (2025)
by: Ma, Le, et al.
Published: (2025)
Experiences from Benchmarking Vision-Language-Action Models for Robotic Manipulation
by: Zhang, Yihao, et al.
Published: (2025)
by: Zhang, Yihao, et al.
Published: (2025)
Audio-VLA: Adding Contact Audio Perception to Vision-Language-Action Model for Robotic Manipulation
by: Wei, Xiangyi, et al.
Published: (2025)
by: Wei, Xiangyi, et al.
Published: (2025)
TIC-VLA: A Think-in-Control Vision-Language-Action Model for Robot Navigation in Dynamic Environments
by: Huang, Zhiyu, et al.
Published: (2026)
by: Huang, Zhiyu, et al.
Published: (2026)
VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation
by: Zhao, Han, et al.
Published: (2025)
by: Zhao, Han, et al.
Published: (2025)
MLA: A Multisensory Language-Action Model for Multimodal Understanding and Forecasting in Robotic Manipulation
by: Liu, Zhuoyang, et al.
Published: (2025)
by: Liu, Zhuoyang, et al.
Published: (2025)
LADEV: A Language-Driven Testing and Evaluation Platform for Vision-Language-Action Models in Robotic Manipulation
by: Wang, Zhijie, et al.
Published: (2024)
by: Wang, Zhijie, et al.
Published: (2024)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
by: Huang, Haifeng, et al.
Published: (2025)
by: Huang, Haifeng, et al.
Published: (2025)
Survey of Vision-Language-Action Models for Embodied Manipulation
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
RedVLA: Physical Red Teaming for Vision-Language-Action Models
by: Zhang, Yuhao, et al.
Published: (2026)
by: Zhang, Yuhao, et al.
Published: (2026)
VIP: Vision Instructed Pre-training for Robotic Manipulation
by: Li, Zhuoling, et al.
Published: (2024)
by: Li, Zhuoling, et al.
Published: (2024)
CollaBot: Vision-Language Guided Simultaneous Collaborative Manipulation
by: Song, Kun, et al.
Published: (2025)
by: Song, Kun, et al.
Published: (2025)
Scaffolding Dexterous Manipulation with Vision-Language Models
by: de Bakker, Vincent, et al.
Published: (2025)
by: de Bakker, Vincent, et al.
Published: (2025)
DySL-VLA: Efficient Vision-Language-Action Model Inference via Dynamic-Static Layer-Skipping for Robot Manipulation
by: Yang, Zebin, et al.
Published: (2026)
by: Yang, Zebin, et al.
Published: (2026)
Utilizing Inpainting for Keypoint Detection for Vision-Based Control of Robotic Manipulators
by: Chatterjee, Sreejani, et al.
Published: (2026)
by: Chatterjee, Sreejani, et al.
Published: (2026)
Image-Based Roadmaps for Vision-Only Planning and Control of Robotic Manipulators
by: Chatterjee, Sreejani, et al.
Published: (2025)
by: Chatterjee, Sreejani, et al.
Published: (2025)
World Models for Robotic Manipulation: A Survey
by: Wang, Fangyuan, et al.
Published: (2026)
by: Wang, Fangyuan, et al.
Published: (2026)
Similar Items
-
Vision-Language Model Predictive Control for Manipulation Planning and Trajectory Generation
by: Chen, Jiaming, et al.
Published: (2025) -
Skill-Aware Diffusion for Generalizable Robotic Manipulation
by: Huang, Aoshen, et al.
Published: (2026) -
SafeFall: Learning Protective Control for Humanoid Robots
by: Meng, Ziyu, et al.
Published: (2025) -
VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation
by: Zhao, Wei, et al.
Published: (2025) -
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey
by: Shao, Rui, et al.
Published: (2025)