Score the Steps, Not Just the Goal: VLM-Based Subgoal Evaluation for Robotic Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | ElMallah, Ramy, Chhajer, Krish, Lee, Chi-Guhn |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Offline Discovery of Interpretable Skills from Multi-Task Trajectories
by: Zhu, Chongyu, et al.
Published: (2026)
by: Zhu, Chongyu, et al.
Published: (2026)
Enhancing Cognitive Robotics with Commonsense through LLM-Generated Preconditions and Subgoals
by: Bachner, Ohad, et al.
Published: (2025)
by: Bachner, Ohad, et al.
Published: (2025)
Robot Collapse: Supply Chain Backdoor Attacks Against VLM-based Robotic Manipulation
by: Wang, Xianlong, et al.
Published: (2024)
by: Wang, Xianlong, et al.
Published: (2024)
Perceiving, Reasoning, Adapting: A Dual-Layer Framework for VLM-Guided Precision Robotic Manipulation
by: Jia, Qingxuan, et al.
Published: (2025)
by: Jia, Qingxuan, et al.
Published: (2025)
HybridFlow: A Two-Step Generative Policy for Robotic Manipulation
by: Dong, Zhenchen, et al.
Published: (2026)
by: Dong, Zhenchen, et al.
Published: (2026)
On the Vulnerability of LLM/VLM-Controlled Robotics
by: Wu, Xiyang, et al.
Published: (2024)
by: Wu, Xiyang, et al.
Published: (2024)
DM1: MeanFlow with Dispersive Regularization for 1-Step Robotic Manipulation
by: Zou, Guowei, et al.
Published: (2025)
by: Zou, Guowei, et al.
Published: (2025)
From Obstacles to Etiquette: Robot Social Navigation with VLM-Informed Path Selection
by: Fang, Zilin, et al.
Published: (2026)
by: Fang, Zilin, et al.
Published: (2026)
Hyper-GoalNet: Goal-Conditioned Manipulation Policy Learning with HyperNetworks
by: Zhou, Pei, et al.
Published: (2025)
by: Zhou, Pei, et al.
Published: (2025)
Think, Remember, Navigate: Zero-Shot Object-Goal Navigation with VLM-Powered Reasoning
by: Habibpour, Mobin, et al.
Published: (2025)
by: Habibpour, Mobin, et al.
Published: (2025)
Executable Analytic Concepts as the Missing Link Between VLM Insight and Precise Manipulation
by: Sun, Mingyang, et al.
Published: (2025)
by: Sun, Mingyang, et al.
Published: (2025)
SimpleVSF: VLM-Scoring Fusion for Trajectory Prediction of End-to-End Autonomous Driving
by: Zheng, Peiru, et al.
Published: (2025)
by: Zheng, Peiru, et al.
Published: (2025)
CTSAC: Curriculum-Based Transformer Soft Actor-Critic for Goal-Oriented Robot Exploration
by: Yang, Chunyu, et al.
Published: (2025)
by: Yang, Chunyu, et al.
Published: (2025)
Evaluating VLMs' Spatial Reasoning Over Robot Motion: A Step Towards Robot Planning with Motion Preferences
by: Wu, Wenxi, et al.
Published: (2026)
by: Wu, Wenxi, et al.
Published: (2026)
AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World
by: Zhou, Zhiyuan, et al.
Published: (2025)
by: Zhou, Zhiyuan, et al.
Published: (2025)
RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation
by: Jiang, Feng, et al.
Published: (2026)
by: Jiang, Feng, et al.
Published: (2026)
MetaWorld-X: Hierarchical World Modeling via VLM-Orchestrated Experts for Humanoid Loco-Manipulation
by: Shen, Yutong, et al.
Published: (2026)
by: Shen, Yutong, et al.
Published: (2026)
Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics
by: Chiche, Hendrik, et al.
Published: (2026)
by: Chiche, Hendrik, et al.
Published: (2026)
A Real-to-Sim-to-Real Approach to Robotic Manipulation with VLM-Generated Iterative Keypoint Rewards
by: Patel, Shivansh, et al.
Published: (2025)
by: Patel, Shivansh, et al.
Published: (2025)
GameVLM: A Decision-making Framework for Robotic Task Planning Based on Visual Language Models and Zero-sum Games
by: Mei, Aoran, et al.
Published: (2024)
by: Mei, Aoran, et al.
Published: (2024)
Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training
by: Liu, Yi, et al.
Published: (2025)
by: Liu, Yi, et al.
Published: (2025)
INTENTION: Inferring Tendencies of Humanoid Robot Motion Through Interactive Intuition and Grounded VLM
by: Wang, Jin, et al.
Published: (2025)
by: Wang, Jin, et al.
Published: (2025)
The Robot's Inner Critic: Self-Refinement of Social Behaviors through VLM-based Replanning
by: Lim, Jiyu, et al.
Published: (2026)
by: Lim, Jiyu, et al.
Published: (2026)
Phys2Real: Fusing VLM Priors with Interactive Online Adaptation for Uncertainty-Aware Sim-to-Real Manipulation
by: Wang, Maggie, et al.
Published: (2025)
by: Wang, Maggie, et al.
Published: (2025)
ManipDreamer: Boosting Robotic Manipulation World Model with Action Tree and Visual Guidance
by: Li, Ying, et al.
Published: (2025)
by: Li, Ying, et al.
Published: (2025)
MOKA: Open-World Robotic Manipulation through Mark-Based Visual Prompting
by: Liu, Fangchen, et al.
Published: (2024)
by: Liu, Fangchen, et al.
Published: (2024)
A Taxonomy for Evaluating Generalist Robot Manipulation Policies
by: Gao, Jensen, et al.
Published: (2025)
by: Gao, Jensen, et al.
Published: (2025)
THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation
by: Pumacay, Wilbert, et al.
Published: (2024)
by: Pumacay, Wilbert, et al.
Published: (2024)
A Semantic Autonomy Framework for VLM-Integrated Indoor Mobile Robots: Hybrid Deterministic Reasoning and Cross-Robot Adaptive Memory
by: Abaza, Bogdan Felician, et al.
Published: (2026)
by: Abaza, Bogdan Felician, et al.
Published: (2026)
Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets
by: Chen, Kaiyuan, et al.
Published: (2025)
by: Chen, Kaiyuan, et al.
Published: (2025)
Thinking in Text and Images: Interleaved Vision--Language Reasoning Traces for Long-Horizon Robot Manipulation
by: Liu, Jinkun, et al.
Published: (2026)
by: Liu, Jinkun, et al.
Published: (2026)
SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
Learning to Transfer Human Hand Skills for Robot Manipulations
by: Park, Sungjae, et al.
Published: (2025)
by: Park, Sungjae, et al.
Published: (2025)
LADEV: A Language-Driven Testing and Evaluation Platform for Vision-Language-Action Models in Robotic Manipulation
by: Wang, Zhijie, et al.
Published: (2024)
by: Wang, Zhijie, et al.
Published: (2024)
Transferring Foundation Models for Generalizable Robotic Manipulation
by: Yang, Jiange, et al.
Published: (2023)
by: Yang, Jiange, et al.
Published: (2023)
Affordance-based Robot Manipulation with Flow Matching
by: Zhang, Fan, et al.
Published: (2024)
by: Zhang, Fan, et al.
Published: (2024)
Lifelong Language-Conditioned Robotic Manipulation Learning
by: Wang, Xudong, et al.
Published: (2026)
by: Wang, Xudong, et al.
Published: (2026)
Bridging Speech, Emotion, and Motion: a VLM-based Multimodal Edge-deployable Framework for Humanoid Robots
by: Yang, Songhua, et al.
Published: (2026)
by: Yang, Songhua, et al.
Published: (2026)
Open-Ended Goal Inference through Actions and Language for Human-Robot Collaboration
by: Ghose, Debasmita, et al.
Published: (2025)
by: Ghose, Debasmita, et al.
Published: (2025)
Hierarchical Reinforcement Learning in Multi-Goal Spatial Navigation with Autonomous Mobile Robots
by: Johnson, Brendon, et al.
Published: (2025)
by: Johnson, Brendon, et al.
Published: (2025)
Similar Items
-
Offline Discovery of Interpretable Skills from Multi-Task Trajectories
by: Zhu, Chongyu, et al.
Published: (2026) -
Enhancing Cognitive Robotics with Commonsense through LLM-Generated Preconditions and Subgoals
by: Bachner, Ohad, et al.
Published: (2025) -
Robot Collapse: Supply Chain Backdoor Attacks Against VLM-based Robotic Manipulation
by: Wang, Xianlong, et al.
Published: (2024) -
Perceiving, Reasoning, Adapting: A Dual-Layer Framework for VLM-Guided Precision Robotic Manipulation
by: Jia, Qingxuan, et al.
Published: (2025) -
HybridFlow: A Two-Step Generative Policy for Robotic Manipulation
by: Dong, Zhenchen, et al.
Published: (2026)