VLM-DEWM: Dynamic External World Model for Verifiable and Resilient Vision-Language Planning in Manufacturing
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Guoqin, Jia, Qingxuan, Chen, Gang, Li, Tong, Huang, Zeyuan, Lv, Zihang, Ji, Ning |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
3D-Grounded Vision-Language Framework for Robotic Task Planning: Automated Prompt Synthesis and Supervised Reasoning
by: Tang, Guoqin, et al.
Published: (2025)
by: Tang, Guoqin, et al.
Published: (2025)
Perceiving, Reasoning, Adapting: A Dual-Layer Framework for VLM-Guided Precision Robotic Manipulation
by: Jia, Qingxuan, et al.
Published: (2025)
by: Jia, Qingxuan, et al.
Published: (2025)
VitaTouch: Property-Aware Vision-Tactile-Language Model for Robotic Quality Inspection in Manufacturing
by: Zong, Junyi, et al.
Published: (2026)
by: Zong, Junyi, et al.
Published: (2026)
WorldVLM: Combining World Model Forecasting and Vision-Language Reasoning
by: Englmeier, Stefan, et al.
Published: (2026)
by: Englmeier, Stefan, et al.
Published: (2026)
ExploreVLM: Closed-Loop Robot Exploration Task Planning with Vision-Language Models
by: Lou, Zhichen, et al.
Published: (2025)
by: Lou, Zhichen, et al.
Published: (2025)
VLM-SAFE: Vision-Language Model-Guided Safety-Aware Reinforcement Learning with World Models for Autonomous Driving
by: Qu, Yansong, et al.
Published: (2025)
by: Qu, Yansong, et al.
Published: (2025)
DyNaVLM: Zero-Shot Vision-Language Navigation System with Dynamic Viewpoints and Self-Refining Graph Memory
by: Ji, Zihe, et al.
Published: (2025)
by: Ji, Zihe, et al.
Published: (2025)
A3VLM: Actionable Articulation-Aware Vision Language Model
by: Huang, Siyuan, et al.
Published: (2024)
by: Huang, Siyuan, et al.
Published: (2024)
Reinforced Embodied Planning with Verifiable Reward for Real-World Robotic Manipulation
by: Bo, Zitong, et al.
Published: (2025)
by: Bo, Zitong, et al.
Published: (2025)
VLM-MPC: Vision Language Foundation Model (VLM)-Guided Model Predictive Controller (MPC) for Autonomous Driving
by: Long, Keke, et al.
Published: (2024)
by: Long, Keke, et al.
Published: (2024)
AppleVLM: End-to-end Autonomous Driving with Advanced Perception and Planning-Enhanced Vision-Language Models
by: Han, Yuxuan, et al.
Published: (2026)
by: Han, Yuxuan, et al.
Published: (2026)
World-Value-Action Model: Implicit Planning for Vision-Language-Action Systems
by: Li, Runze, et al.
Published: (2026)
by: Li, Runze, et al.
Published: (2026)
DreamPlan: Efficient Reinforcement Fine-Tuning of Vision-Language Planners via Video World Models
by: Jia, Emily Yue-Ting, et al.
Published: (2026)
by: Jia, Emily Yue-Ting, et al.
Published: (2026)
A Hierarchical Test Platform for Vision Language Model (VLM)-Integrated Real-World Autonomous Driving
by: Zhou, Yupeng, et al.
Published: (2025)
by: Zhou, Yupeng, et al.
Published: (2025)
ImagineUAV: Aerial Vision-Language Navigation via World-Action Modeling and Kinodynamic Planning
by: Liu, Xuchen, et al.
Published: (2026)
by: Liu, Xuchen, et al.
Published: (2026)
Imagine, Verify, Execute: Memory-guided Agentic Exploration with Vision-Language Models
by: Lee, Seungjae, et al.
Published: (2025)
by: Lee, Seungjae, et al.
Published: (2025)
Action Draft and Verify: A Self-Verifying Framework for Vision-Language-Action Model
by: Zhao, Chen, et al.
Published: (2026)
by: Zhao, Chen, et al.
Published: (2026)
AERMANI-VLM: Structured Prompting and Reasoning for Aerial Manipulation with Vision Language Models
by: Mishra, Sarthak, et al.
Published: (2025)
by: Mishra, Sarthak, et al.
Published: (2025)
BEV-VLM: Trajectory Planning via Unified BEV Abstraction
by: Chen, Guancheng, et al.
Published: (2025)
by: Chen, Guancheng, et al.
Published: (2025)
HybridGen: VLM-Guided Hybrid Planning for Scalable Data Generation of Imitation Learning
by: Wang, Wensheng, et al.
Published: (2025)
by: Wang, Wensheng, et al.
Published: (2025)
Open-World Task and Motion Planning via Vision-Language Model Generated Constraints
by: Kumar, Nishanth, et al.
Published: (2024)
by: Kumar, Nishanth, et al.
Published: (2024)
NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation
by: Zhang, Jiazhao, et al.
Published: (2024)
by: Zhang, Jiazhao, et al.
Published: (2024)
VLA-REPLICA: A Low-Cost, Reproducible Benchmark for Real-World Evaluation of Vision-Language-Action Models
by: Huang, Alex S., et al.
Published: (2026)
by: Huang, Alex S., et al.
Published: (2026)
Vision-Language-Policy Model for Dynamic Robot Task Planning
by: Wang, Jin, et al.
Published: (2025)
by: Wang, Jin, et al.
Published: (2025)
Rethinking Intermediate Representation for VLM-based Robot Manipulation
by: Tang, Weiliang, et al.
Published: (2025)
by: Tang, Weiliang, et al.
Published: (2025)
DKPROMPT: Domain Knowledge Prompting Vision-Language Models for Open-World Planning
by: Zhang, Xiaohan, et al.
Published: (2024)
by: Zhang, Xiaohan, et al.
Published: (2024)
World-aware Planning Narratives Enhance Large Vision-Language Model Planner
by: Shi, Junhao, et al.
Published: (2025)
by: Shi, Junhao, et al.
Published: (2025)
Language-Augmented Symbolic Planner for Open-World Task Planning
by: Chen, Guanqi, et al.
Published: (2024)
by: Chen, Guanqi, et al.
Published: (2024)
Is Your VLM for Autonomous Driving Safety-Ready? A Comprehensive Benchmark for Evaluating External and In-Cabin Risks
by: Meng, Xianhui, et al.
Published: (2025)
by: Meng, Xianhui, et al.
Published: (2025)
VLM-Social-Nav: Socially Aware Robot Navigation through Scoring using Vision-Language Models
by: Song, Daeun, et al.
Published: (2024)
by: Song, Daeun, et al.
Published: (2024)
LocoVLM: Grounding Vision and Language for Adapting Versatile Legged Locomotion Policies
by: Nahrendra, I Made Aswin, et al.
Published: (2026)
by: Nahrendra, I Made Aswin, et al.
Published: (2026)
RoVer: Robot Reward Model as Test-Time Verifier for Vision-Language-Action Model
by: Dai, Mingtong, et al.
Published: (2025)
by: Dai, Mingtong, et al.
Published: (2025)
iFlyBot-VLM Technical Report
by: Nie, Xin, et al.
Published: (2025)
by: Nie, Xin, et al.
Published: (2025)
RoboDexVLM: Visual Language Model-Enabled Task Planning and Motion Control for Dexterous Robot Manipulation
by: Liu, Haichao, et al.
Published: (2025)
by: Liu, Haichao, et al.
Published: (2025)
VLM-RRT: Vision Language Model Guided RRT Search for Autonomous UAV Navigation
by: Ye, Jianlin, et al.
Published: (2025)
by: Ye, Jianlin, et al.
Published: (2025)
Dynam3D: Dynamic Layered 3D Tokens Empower VLM for Vision-and-Language Navigation
by: Wang, Zihan, et al.
Published: (2025)
by: Wang, Zihan, et al.
Published: (2025)
LLM-Drone: Aerial Additive Manufacturing with Drones Planned Using Large Language Models
by: Raman, Akshay, et al.
Published: (2025)
by: Raman, Akshay, et al.
Published: (2025)
SIMPACT: Simulation-Enabled Action Planning using Vision-Language Models
by: Liu, Haowen, et al.
Published: (2025)
by: Liu, Haowen, et al.
Published: (2025)
HumanoidVLM: Vision-Language-Guided Impedance Control for Contact-Rich Humanoid Manipulation
by: Mahmoud, Yara, et al.
Published: (2026)
by: Mahmoud, Yara, et al.
Published: (2026)
TIC-VLA: A Think-in-Control Vision-Language-Action Model for Robot Navigation in Dynamic Environments
by: Huang, Zhiyu, et al.
Published: (2026)
by: Huang, Zhiyu, et al.
Published: (2026)
Similar Items
-
3D-Grounded Vision-Language Framework for Robotic Task Planning: Automated Prompt Synthesis and Supervised Reasoning
by: Tang, Guoqin, et al.
Published: (2025) -
Perceiving, Reasoning, Adapting: A Dual-Layer Framework for VLM-Guided Precision Robotic Manipulation
by: Jia, Qingxuan, et al.
Published: (2025) -
VitaTouch: Property-Aware Vision-Tactile-Language Model for Robotic Quality Inspection in Manufacturing
by: Zong, Junyi, et al.
Published: (2026) -
WorldVLM: Combining World Model Forecasting and Vision-Language Reasoning
by: Englmeier, Stefan, et al.
Published: (2026) -
ExploreVLM: Closed-Loop Robot Exploration Task Planning with Vision-Language Models
by: Lou, Zhichen, et al.
Published: (2025)