CookBench: A Long-Horizon Embodied Planning Benchmark for Complex Cooking Scenarios
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Muzhen, Chen, Xiubo, An, Yining, Zhang, Jiaxin, Wang, Xuesong, Xu, Wang, Zhang, Weinan, Liu, Ting |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cooking Task Planning using LLM and Verified by Graph Network
by: Takebayashi, Ryunosuke, et al.
Published: (2025)
by: Takebayashi, Ryunosuke, et al.
Published: (2025)
ET-Plan-Bench: Embodied Task-level Planning Benchmark Towards Spatial-Temporal Cognition with Foundation Models
by: Zhang, Lingfeng, et al.
Published: (2024)
by: Zhang, Lingfeng, et al.
Published: (2024)
Cook2LTL: Translating Cooking Recipes to LTL Formulae using Large Language Models
by: Mavrogiannis, Angelos, et al.
Published: (2023)
by: Mavrogiannis, Angelos, et al.
Published: (2023)
ContextFlow: Hierarchical Task-State Alignment for Long-Horizon Embodied Agents
by: Guo, Shuhan, et al.
Published: (2026)
by: Guo, Shuhan, et al.
Published: (2026)
CleanUpBench: Embodied Sweeping and Grasping Benchmark
by: Li, Wenbo, et al.
Published: (2025)
by: Li, Wenbo, et al.
Published: (2025)
EmboCoach-Bench: Benchmarking AI Agents on Developing Embodied Robots
by: Lei, Zixing, et al.
Published: (2026)
by: Lei, Zixing, et al.
Published: (2026)
Instruction-Augmented Long-Horizon Planning: Embedding Grounding Mechanisms in Embodied Mobile Manipulation
by: Wang, Fangyuan, et al.
Published: (2025)
by: Wang, Fangyuan, et al.
Published: (2025)
MOSAIC: Modular Foundation Models for Assistive and Interactive Cooking
by: Wang, Huaxiaoyue, et al.
Published: (2024)
by: Wang, Huaxiaoyue, et al.
Published: (2024)
LongBench: Evaluating Robotic Manipulation Policies on Real-World Long-Horizon Tasks
by: Chen, Xueyao, et al.
Published: (2026)
by: Chen, Xueyao, et al.
Published: (2026)
SafeAgentBench: A Benchmark for Safe Task Planning of Embodied LLM Agents
by: Yin, Sheng, et al.
Published: (2024)
by: Yin, Sheng, et al.
Published: (2024)
CausalNav: A Long-term Embodied Navigation System for Autonomous Mobile Robots in Dynamic Outdoor Scenarios
by: Duan, Hongbo, et al.
Published: (2026)
by: Duan, Hongbo, et al.
Published: (2026)
Compositional Diffusion with Guided Search for Long-Horizon Planning
by: Mishra, Utkarsh A, et al.
Published: (2025)
by: Mishra, Utkarsh A, et al.
Published: (2025)
V-CAGE: Context-Aware Generation and Verification for Scalable Long-Horizon Embodied Tasks
by: Liu, Yaru, et al.
Published: (2026)
by: Liu, Yaru, et al.
Published: (2026)
Long-Horizon Manipulation via Trace-Conditioned VLA Planning
by: Liu, Isabella, et al.
Published: (2026)
by: Liu, Isabella, et al.
Published: (2026)
Multi-Dimensional AGV Path Planning in 3D Warehouses Using Ant Colony Optimization and Advanced Neural Networks
by: Zhang, Bo, et al.
Published: (2025)
by: Zhang, Bo, et al.
Published: (2025)
World-Ego Modeling for Long-Horizon Evolution in Hybrid Embodied Tasks
by: Lin, Zuyao, et al.
Published: (2026)
by: Lin, Zuyao, et al.
Published: (2026)
PEPA: a Persistently Autonomous Embodied Agent with Personalities
by: Liu, Kaige, et al.
Published: (2026)
by: Liu, Kaige, et al.
Published: (2026)
Language-Guided Long Horizon Manipulation with LLM-based Planning and Visual Perception
by: Zhou, Changshi, et al.
Published: (2025)
by: Zhou, Changshi, et al.
Published: (2025)
Goal2Skill: Long-Horizon Manipulation with Adaptive Planning and Reflection
by: Liu, Zhen, et al.
Published: (2026)
by: Liu, Zhen, et al.
Published: (2026)
RiskBench: A Scenario-based Benchmark for Risk Identification
by: Kung, Chi-Hsi, et al.
Published: (2023)
by: Kung, Chi-Hsi, et al.
Published: (2023)
Characteristics Analysis of Autonomous Vehicle Pre-crash Scenarios
by: Li, Yixuan, et al.
Published: (2025)
by: Li, Yixuan, et al.
Published: (2025)
Anticipation-VLA: Solving Long-Horizon Embodied Tasks via Anticipation-based Subgoal Generation
by: Zhang, Zhilong, et al.
Published: (2026)
by: Zhang, Zhilong, et al.
Published: (2026)
Generative Skill Chaining: Long-Horizon Skill Planning with Diffusion Models
by: Mishra, Utkarsh A., et al.
Published: (2023)
by: Mishra, Utkarsh A., et al.
Published: (2023)
Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks
by: Liu, Zhihong, et al.
Published: (2026)
by: Liu, Zhihong, et al.
Published: (2026)
Follow-Bench: A Unified Motion Planning Benchmark for Socially-Aware Robot Person Following
by: Ye, Hanjing, et al.
Published: (2025)
by: Ye, Hanjing, et al.
Published: (2025)
RoboJailBench: Benchmarking Adversarial Attacks and Defenses in Embodied Robotic Agents
by: Yeke, Doguhuan, et al.
Published: (2026)
by: Yeke, Doguhuan, et al.
Published: (2026)
FLAF: Focal Line and Feature-constrained Active View Planning for Visual Teach and Repeat
by: Fu, Changfei, et al.
Published: (2024)
by: Fu, Changfei, et al.
Published: (2024)
Long-horizon Embodied Planning with Implicit Logical Inference and Hallucination Mitigation
by: Liu, Siyuan, et al.
Published: (2024)
by: Liu, Siyuan, et al.
Published: (2024)
SWITCH: Benchmarking Modeling and Handling of Tangible Interfaces in Long-horizon Embodied Scenarios
by: Lin, Jieru, et al.
Published: (2025)
by: Lin, Jieru, et al.
Published: (2025)
Mem2Ego: Empowering Vision-Language Models with Global-to-Ego Memory for Long-Horizon Embodied Navigation
by: Zhang, Lingfeng, et al.
Published: (2025)
by: Zhang, Lingfeng, et al.
Published: (2025)
TPT-Bench: A Large-Scale, Long-Term and Robot-Egocentric Dataset for Benchmarking Target Person Tracking
by: Ye, Hanjing, et al.
Published: (2025)
by: Ye, Hanjing, et al.
Published: (2025)
LoHoVLA: A Unified Vision-Language-Action Model for Long-Horizon Embodied Tasks
by: Yang, Yi, et al.
Published: (2025)
by: Yang, Yi, et al.
Published: (2025)
MLDT: Multi-Level Decomposition for Complex Long-Horizon Robotic Task Planning with Open-Source Large Language Model
by: Wu, Yike, et al.
Published: (2024)
by: Wu, Yike, et al.
Published: (2024)
Beyond Short-Horizon: VQ-Memory for Robust Long-Horizon Manipulation in Non-Markovian Simulation Benchmarks
by: Wang, Honghui, et al.
Published: (2026)
by: Wang, Honghui, et al.
Published: (2026)
Human-centered In-building Embodied Delivery Benchmark
by: Xu, Zhuoqun, et al.
Published: (2024)
by: Xu, Zhuoqun, et al.
Published: (2024)
Explicit-Implicit Subgoal Planning for Long-Horizon Tasks with Sparse Reward
by: Wang, Fangyuan, et al.
Published: (2023)
by: Wang, Fangyuan, et al.
Published: (2023)
Embodied Tree of Thoughts: Deliberate Manipulation Planning with Embodied World Model
by: Xu, Wenjiang, et al.
Published: (2025)
by: Xu, Wenjiang, et al.
Published: (2025)
EmbodiedBrain: Expanding Performance Boundaries of Task Planning for Embodied Intelligence
by: Zou, Ding, et al.
Published: (2025)
by: Zou, Ding, et al.
Published: (2025)
ESARBench: A Benchmark for Agentic UAV Embodied Search and Rescue
by: Zhang, Daoxuan, et al.
Published: (2026)
by: Zhang, Daoxuan, et al.
Published: (2026)
FloNa: Floor Plan Guided Embodied Visual Navigation
by: Li, Jiaxin, et al.
Published: (2024)
by: Li, Jiaxin, et al.
Published: (2024)
Similar Items
-
Cooking Task Planning using LLM and Verified by Graph Network
by: Takebayashi, Ryunosuke, et al.
Published: (2025) -
ET-Plan-Bench: Embodied Task-level Planning Benchmark Towards Spatial-Temporal Cognition with Foundation Models
by: Zhang, Lingfeng, et al.
Published: (2024) -
Cook2LTL: Translating Cooking Recipes to LTL Formulae using Large Language Models
by: Mavrogiannis, Angelos, et al.
Published: (2023) -
ContextFlow: Hierarchical Task-State Alignment for Long-Horizon Embodied Agents
by: Guo, Shuhan, et al.
Published: (2026) -
CleanUpBench: Embodied Sweeping and Grasping Benchmark
by: Li, Wenbo, et al.
Published: (2025)