STAR: A Foundation Model-driven Framework for Robust Task Planning and Failure Recovery in Robotic Systems
Fuente:
arXiv
Guardado en:
| Autores principales: | Sakib, Md Sadman, Sun, Yu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Consolidating Trees of Robotic Plans Generated Using Large Language Models to Improve Reliability
por: Sakib, Md Sadman, et al.
Publicado: (2024)
por: Sakib, Md Sadman, et al.
Publicado: (2024)
A Diver Attention Estimation Framework for Effective Underwater Human-Robot Interaction
por: Enan, Sadman Sakib, et al.
Publicado: (2022)
por: Enan, Sadman Sakib, et al.
Publicado: (2022)
Robix: A Unified Model for Robot Interaction, Reasoning and Planning
por: Fang, Huang, et al.
Publicado: (2025)
por: Fang, Huang, et al.
Publicado: (2025)
SDA-PLANNER: State-Dependency Aware Adaptive Planner for Embodied Task Planning
por: Shen, Zichao, et al.
Publicado: (2025)
por: Shen, Zichao, et al.
Publicado: (2025)
VLM-driven Behavior Tree for Context-aware Task Planning
por: Wake, Naoki, et al.
Publicado: (2025)
por: Wake, Naoki, et al.
Publicado: (2025)
PALMS+: Modular Image-Based Floor Plan Localization Leveraging Depth Foundation Model
por: Cheng, Yunqian, et al.
Publicado: (2025)
por: Cheng, Yunqian, et al.
Publicado: (2025)
KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis
por: Hosseinzadeh, Mehdi, et al.
Publicado: (2026)
por: Hosseinzadeh, Mehdi, et al.
Publicado: (2026)
General Flow as Foundation Affordance for Scalable Robot Learning
por: Yuan, Chengbo, et al.
Publicado: (2024)
por: Yuan, Chengbo, et al.
Publicado: (2024)
Robust Driving QA through Metadata-Grounded Context and Task-Specific Prompts
por: Yu, Seungjun, et al.
Publicado: (2025)
por: Yu, Seungjun, et al.
Publicado: (2025)
Explainable Adversarial-Robust Vision-Language-Action Model for Robotic Manipulation
por: Kim, Ju-Young, et al.
Publicado: (2025)
por: Kim, Ju-Young, et al.
Publicado: (2025)
Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks
por: Dalal, Murtaza, et al.
Publicado: (2024)
por: Dalal, Murtaza, et al.
Publicado: (2024)
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
por: Bhatt, Neel P., et al.
Publicado: (2024)
por: Bhatt, Neel P., et al.
Publicado: (2024)
This&That: Language-Gesture Controlled Video Generation for Robot Planning
por: Wang, Boyang, et al.
Publicado: (2024)
por: Wang, Boyang, et al.
Publicado: (2024)
NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks
por: Hung, Chia-Yu, et al.
Publicado: (2025)
por: Hung, Chia-Yu, et al.
Publicado: (2025)
Work Zones challenge VLM Trajectory Planning: Toward Mitigation and Robust Autonomous Driving
por: Liao, Yifan, et al.
Publicado: (2025)
por: Liao, Yifan, et al.
Publicado: (2025)
Prospective Role of Foundation Models in Advancing Autonomous Vehicles
por: Wu, Jianhua, et al.
Publicado: (2023)
por: Wu, Jianhua, et al.
Publicado: (2023)
Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation
por: Chen, Shizhe, et al.
Publicado: (2025)
por: Chen, Shizhe, et al.
Publicado: (2025)
Universal Actions for Enhanced Embodied Foundation Models
por: Zheng, Jinliang, et al.
Publicado: (2025)
por: Zheng, Jinliang, et al.
Publicado: (2025)
SEM: Enhancing Spatial Understanding for Robust Robot Manipulation
por: Lin, Xuewu, et al.
Publicado: (2025)
por: Lin, Xuewu, et al.
Publicado: (2025)
Can-Do! A Dataset and Neuro-Symbolic Grounded Framework for Embodied Planning with Large Multimodal Models
por: Chia, Yew Ken, et al.
Publicado: (2024)
por: Chia, Yew Ken, et al.
Publicado: (2024)
Efficient Robotic Policy Learning via Latent Space Backward Planning
por: Liu, Dongxiu, et al.
Publicado: (2025)
por: Liu, Dongxiu, et al.
Publicado: (2025)
Real-World Robot Applications of Foundation Models: A Review
por: Kawaharazuka, Kento, et al.
Publicado: (2024)
por: Kawaharazuka, Kento, et al.
Publicado: (2024)
Robotic State Recognition with Image-to-Text Retrieval Task of Pre-Trained Vision-Language Model and Black-Box Optimization
por: Kawaharazuka, Kento, et al.
Publicado: (2024)
por: Kawaharazuka, Kento, et al.
Publicado: (2024)
Humanoid Occupancy: Enabling A Generalized Multimodal Occupancy Perception System on Humanoid Robots
por: Cui, Wei, et al.
Publicado: (2025)
por: Cui, Wei, et al.
Publicado: (2025)
Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets
por: Jiang, Guangqi, et al.
Publicado: (2024)
por: Jiang, Guangqi, et al.
Publicado: (2024)
S.T.A.R.-Track: Latent Motion Models for End-to-End 3D Object Tracking with Adaptive Spatio-Temporal Appearance Representations
por: Doll, Simon, et al.
Publicado: (2023)
por: Doll, Simon, et al.
Publicado: (2023)
AIC MLLM: Autonomous Interactive Correction MLLM for Robust Robotic Manipulation
por: Xiong, Chuyan, et al.
Publicado: (2024)
por: Xiong, Chuyan, et al.
Publicado: (2024)
SocialNav: Training Human-Inspired Foundation Model for Socially-Aware Embodied Navigation
por: Chen, Ziyi, et al.
Publicado: (2025)
por: Chen, Ziyi, et al.
Publicado: (2025)
Theia: Distilling Diverse Vision Foundation Models for Robot Learning
por: Shang, Jinghuan, et al.
Publicado: (2024)
por: Shang, Jinghuan, et al.
Publicado: (2024)
SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning
por: Liu, Yuecheng, et al.
Publicado: (2025)
por: Liu, Yuecheng, et al.
Publicado: (2025)
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy
por: Chen, Xinyi, et al.
Publicado: (2025)
por: Chen, Xinyi, et al.
Publicado: (2025)
Robot Learning from a Physical World Model
por: Mao, Jiageng, et al.
Publicado: (2025)
por: Mao, Jiageng, et al.
Publicado: (2025)
PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation
por: Huang, Wenlong, et al.
Publicado: (2026)
por: Huang, Wenlong, et al.
Publicado: (2026)
Scene Graph-Guided Proactive Replanning for Failure-Resilient Embodied Agent
por: Yu, Che Rin, et al.
Publicado: (2025)
por: Yu, Che Rin, et al.
Publicado: (2025)
VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks
por: Zhang, Shiduo, et al.
Publicado: (2024)
por: Zhang, Shiduo, et al.
Publicado: (2024)
RADAR: Closed-Loop Robotic Data Generation via Semantic Planning and Autonomous Causal Environment Reset
por: Wang, Yongzhong, et al.
Publicado: (2026)
por: Wang, Yongzhong, et al.
Publicado: (2026)
Playing to Vision Foundation Model's Strengths in Stereo Matching
por: Liu, Chuang-Wei, et al.
Publicado: (2024)
por: Liu, Chuang-Wei, et al.
Publicado: (2024)
VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models
por: Si, Shengyu, et al.
Publicado: (2026)
por: Si, Shengyu, et al.
Publicado: (2026)
Phoenix: A Motion-based Self-Reflection Framework for Fine-grained Robotic Action Correction
por: Xia, Wenke, et al.
Publicado: (2025)
por: Xia, Wenke, et al.
Publicado: (2025)
DriveDreamer-Policy: A Geometry-Grounded World-Action Model for Unified Generation and Planning
por: Zhou, Yang, et al.
Publicado: (2026)
por: Zhou, Yang, et al.
Publicado: (2026)
Ejemplares similares
-
Consolidating Trees of Robotic Plans Generated Using Large Language Models to Improve Reliability
por: Sakib, Md Sadman, et al.
Publicado: (2024) -
A Diver Attention Estimation Framework for Effective Underwater Human-Robot Interaction
por: Enan, Sadman Sakib, et al.
Publicado: (2022) -
Robix: A Unified Model for Robot Interaction, Reasoning and Planning
por: Fang, Huang, et al.
Publicado: (2025) -
SDA-PLANNER: State-Dependency Aware Adaptive Planner for Embodied Task Planning
por: Shen, Zichao, et al.
Publicado: (2025) -
VLM-driven Behavior Tree for Context-aware Task Planning
por: Wake, Naoki, et al.
Publicado: (2025)