Saved in:
| Main Authors: | Liu, Fengkai, Su, Hao, Chi, Haozhuang, Geng, Rui, Ren, Congzhi, Liu, Xuqing, Xu, Yucheng, Ohsita, Yuichi, Zhang, Liyun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2603.23950 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SG-VLA: Learning Spatially-Grounded Vision-Language-Action Models for Mobile Manipulation
by: Tu, Ruisen, et al.
Published: (2026)
by: Tu, Ruisen, et al.
Published: (2026)
Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation
by: Chen, Shizhe, et al.
Published: (2025)
by: Chen, Shizhe, et al.
Published: (2025)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
by: Huang, Haifeng, et al.
Published: (2025)
by: Huang, Haifeng, et al.
Published: (2025)
Casper: Inferring Diverse Intents for Assistive Teleoperation with Vision Language Models
by: Liu, Huihan, et al.
Published: (2025)
by: Liu, Huihan, et al.
Published: (2025)
Cooperative Modular Manipulation with Numerous Cable-Driven Robots for Assistive Construction and Gap Crossing
by: Murphy, Kevin, et al.
Published: (2024)
by: Murphy, Kevin, et al.
Published: (2024)
Long-Horizon Manipulation via Trace-Conditioned VLA Planning
by: Liu, Isabella, et al.
Published: (2026)
by: Liu, Isabella, et al.
Published: (2026)
CrayonRobo: Object-Centric Prompt-Driven Vision-Language-Action Model for Robotic Manipulation
by: Li, Xiaoqi, et al.
Published: (2025)
by: Li, Xiaoqi, et al.
Published: (2025)
ProDrive: Proactive Planning for Autonomous Driving via Ego-Environment Co-Evolution
by: Fu, Chuyao, et al.
Published: (2026)
by: Fu, Chuyao, et al.
Published: (2026)
PLanAR: Planning-Language-Grounded Agentic Reasoning for Robot Manipulation
by: Guo, Pengyuan, et al.
Published: (2026)
by: Guo, Pengyuan, et al.
Published: (2026)
Driver-WM: A Driver-Centric Traffic-Conditioned Latent World Model for In-Cabin Dynamics Rollout
by: Chi, Haozhuang, et al.
Published: (2026)
by: Chi, Haozhuang, et al.
Published: (2026)
FloorPlan-VLN: A New Paradigm for Floor Plan Guided Vision-Language Navigation
by: Chen, Kehan, et al.
Published: (2026)
by: Chen, Kehan, et al.
Published: (2026)
Hierarchical Vision-Language Planning for Multi-Step Humanoid Manipulation
by: Schakkal, André, et al.
Published: (2025)
by: Schakkal, André, et al.
Published: (2025)
A Novel Planning Framework for Complex Flipping Manipulation of Multiple Mobile Manipulators
by: Liu, Wenhang, et al.
Published: (2023)
by: Liu, Wenhang, et al.
Published: (2023)
UniPlan: Vision-Language Task Planning for Mobile Manipulation with Unified PDDL Formulation
by: Ye, Haoming, et al.
Published: (2026)
by: Ye, Haoming, et al.
Published: (2026)
AssistDLO: Assistive Teleoperation for Deformable Linear Object Manipulation
by: Guler, Berk, et al.
Published: (2026)
by: Guler, Berk, et al.
Published: (2026)
Online Robot Navigation and Manipulation with Distilled Vision-Language Models
by: Liu, Kangcheng
Published: (2024)
by: Liu, Kangcheng
Published: (2024)
SVLL: Staged Vision-Language Learning for Physically Grounded Embodied Task Planning
by: Yang, Yuyuan, et al.
Published: (2026)
by: Yang, Yuyuan, et al.
Published: (2026)
Grounding Sim-to-Real Generalization in Dexterous Manipulation: An Empirical Study with Vision-Language-Action Models
by: Jin, Ruixing, et al.
Published: (2026)
by: Jin, Ruixing, et al.
Published: (2026)
Event-Grounded Sparse Autoencoders for Vision-Language-Action Policies
by: Jin, Xinchen, et al.
Published: (2026)
by: Jin, Xinchen, et al.
Published: (2026)
Dynamic Planning for Sequential Whole-body Mobile Manipulation
by: Li, Zhitian, et al.
Published: (2024)
by: Li, Zhitian, et al.
Published: (2024)
Task-oriented Robotic Manipulation with Vision Language Models
by: Guran, Nurhan Bulus, et al.
Published: (2024)
by: Guran, Nurhan Bulus, et al.
Published: (2024)
Thinking in Text and Images: Interleaved Vision--Language Reasoning Traces for Long-Horizon Robot Manipulation
by: Liu, Jinkun, et al.
Published: (2026)
by: Liu, Jinkun, et al.
Published: (2026)
Vision-Language Model Predictive Control for Manipulation Planning and Trajectory Generation
by: Chen, Jiaming, et al.
Published: (2025)
by: Chen, Jiaming, et al.
Published: (2025)
Physically Grounded Vision-Language Models for Robotic Manipulation
by: Gao, Jensen, et al.
Published: (2023)
by: Gao, Jensen, et al.
Published: (2023)
Inclusion in Assistive Haircare Robotics: Practical and Ethical Considerations in Hair Manipulation
by: Yoo, Uksang, et al.
Published: (2024)
by: Yoo, Uksang, et al.
Published: (2024)
Dynamic Environment Adaptive Path Planning for Mobile Robots: A Hybrid Enhanced Path‐Planning Approach
by: Junxu Hou, et al.
Published: (2026)
by: Junxu Hou, et al.
Published: (2026)
Grounded Vision-Language Interpreter for Integrated Task and Motion Planning
by: Siburian, Jeremy, et al.
Published: (2025)
by: Siburian, Jeremy, et al.
Published: (2025)
CubeRobot: Grounding Language in Rubik's Cube Manipulation via Vision-Language Model
by: Wang, Feiyang, et al.
Published: (2025)
by: Wang, Feiyang, et al.
Published: (2025)
Configuration Space Distance Fields for Manipulation Planning
by: Li, Yiming, et al.
Published: (2024)
by: Li, Yiming, et al.
Published: (2024)
Proactive Route Planning for Electric Vehicles
by: Nasehi, Saeed, et al.
Published: (2024)
by: Nasehi, Saeed, et al.
Published: (2024)
DRAGON: A Dialogue-Based Robot for Assistive Navigation with Visual Language Grounding
by: Liu, Shuijing, et al.
Published: (2023)
by: Liu, Shuijing, et al.
Published: (2023)
Manip4Care: Robotic Manipulation of Human Limbs for Solving Assistive Tasks
by: Koh, Yubin, et al.
Published: (2025)
by: Koh, Yubin, et al.
Published: (2025)
Reflective VLM Planning for Dual-Arm Desktop Cleaning: Bridging Open-Vocabulary Perception and Precise Manipulation
by: Liu, Yufan, et al.
Published: (2025)
by: Liu, Yufan, et al.
Published: (2025)
FSR-VLN: Fast and Slow Reasoning for Vision-Language Navigation with Hierarchical Multi-modal Scene Graph
by: Zhou, Xiaolin, et al.
Published: (2025)
by: Zhou, Xiaolin, et al.
Published: (2025)
SPARK: Safe Protective and Assistive Robot Kit
by: Sun, Yifan, et al.
Published: (2025)
by: Sun, Yifan, et al.
Published: (2025)
Survey of Vision-Language-Action Models for Embodied Manipulation
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
VTLA: Vision-Tactile-Language-Action Model with Preference Learning for Insertion Manipulation
by: Zhang, Chaofan, et al.
Published: (2025)
by: Zhang, Chaofan, et al.
Published: (2025)
CoFreeVLA: Collision-Free Dual-Arm Manipulation via Vision-Language-Action Model and Risk Estimation
by: Zhai, Xuanran, et al.
Published: (2026)
by: Zhai, Xuanran, et al.
Published: (2026)
WheelArm-Sim: A Manipulation and Navigation Combined Multimodal Synthetic Data Generation Simulator for Unified Control in Assistive Robotics
by: Liu, Guangping, et al.
Published: (2026)
by: Liu, Guangping, et al.
Published: (2026)
High-density Electromyography for Effective Gesture-based Control of Physically Assistive Mobile Manipulators
by: Yang, Jehan, et al.
Published: (2023)
by: Yang, Jehan, et al.
Published: (2023)
Similar Items
-
SG-VLA: Learning Spatially-Grounded Vision-Language-Action Models for Mobile Manipulation
by: Tu, Ruisen, et al.
Published: (2026) -
Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation
by: Chen, Shizhe, et al.
Published: (2025) -
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
by: Huang, Haifeng, et al.
Published: (2025) -
Casper: Inferring Diverse Intents for Assistive Teleoperation with Vision Language Models
by: Liu, Huihan, et al.
Published: (2025) -
Cooperative Modular Manipulation with Numerous Cable-Driven Robots for Assistive Construction and Gap Crossing
by: Murphy, Kevin, et al.
Published: (2024)