LoTa-Bench: Benchmarking Language-oriented Task Planners for Embodied Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Choi, Jae-Woo, Yoon, Youngwoo, Ong, Hyobin, Kim, Jaehong, Jang, Minsu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ReAcTree: Hierarchical LLM Agent Trees with Control Flow for Long-Horizon Task Planning
by: Choi, Jae-Woo, et al.
Published: (2025)
by: Choi, Jae-Woo, et al.
Published: (2025)
Learning Dexterous Bimanual Catch Skills through Adversarial-Cooperative Heterogeneous-Agent Reinforcement Learning
by: Kim, Taewoo, et al.
Published: (2025)
by: Kim, Taewoo, et al.
Published: (2025)
World Model Implanting for Test-time Adaptation of Embodied Agents
by: Yoo, Minjong, et al.
Published: (2025)
by: Yoo, Minjong, et al.
Published: (2025)
Test-Time Mixture of World Models for Embodied Agents in Dynamic Environments
by: Jang, Jinwoo, et al.
Published: (2026)
by: Jang, Jinwoo, et al.
Published: (2026)
HELP: Hierarchical Embodied Language Planner for Household Tasks
by: Korchemnyi, Alexandr V., et al.
Published: (2025)
by: Korchemnyi, Alexandr V., et al.
Published: (2025)
Embodied CoT Distillation From LLM To Off-the-shelf Agents
by: Choi, Wonje, et al.
Published: (2024)
by: Choi, Wonje, et al.
Published: (2024)
SafeAgentBench: A Benchmark for Safe Task Planning of Embodied LLM Agents
by: Yin, Sheng, et al.
Published: (2024)
by: Yin, Sheng, et al.
Published: (2024)
EnvGen: Generating and Adapting Environments via LLMs for Training Embodied Agents
by: Zala, Abhay, et al.
Published: (2024)
by: Zala, Abhay, et al.
Published: (2024)
Efficient Policy Adaptation with Contrastive Prompt Ensemble for Embodied Agents
by: Choi, Wonje, et al.
Published: (2024)
by: Choi, Wonje, et al.
Published: (2024)
LLMServingSim: A HW/SW Co-Simulation Infrastructure for LLM Inference Serving at Scale
by: Cho, Jaehong, et al.
Published: (2024)
by: Cho, Jaehong, et al.
Published: (2024)
SPARK: Self-Play with Asymmetric Reward from Knowledge Graphs
by: Park, Hyobin, et al.
Published: (2026)
by: Park, Hyobin, et al.
Published: (2026)
NeSyPr: Neurosymbolic Proceduralization For Efficient Embodied Reasoning
by: Choi, Wonje, et al.
Published: (2025)
by: Choi, Wonje, et al.
Published: (2025)
EmboCoach-Bench: Benchmarking AI Agents on Developing Embodied Robots
by: Lei, Zixing, et al.
Published: (2026)
by: Lei, Zixing, et al.
Published: (2026)
PILOT-Bench: A Benchmark for Legal Reasoning in the Patent Domain with IRAC-Aligned Classification Tasks
by: Jang, Yehoon, et al.
Published: (2026)
by: Jang, Yehoon, et al.
Published: (2026)
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
by: Yang, Rui, et al.
Published: (2025)
by: Yang, Rui, et al.
Published: (2025)
The Great March 100: 100 Detail-oriented Tasks for Evaluating Embodied AI Agents
by: Wang, Ziyu, et al.
Published: (2026)
by: Wang, Ziyu, et al.
Published: (2026)
Semantic Skill Grounding for Embodied Instruction-Following in Cross-Domain Environments
by: Shin, Sangwoo, et al.
Published: (2024)
by: Shin, Sangwoo, et al.
Published: (2024)
Towards Reliable Code-as-Policies: A Neuro-Symbolic Framework for Embodied Task Planning
by: Ahn, Sanghyun, et al.
Published: (2025)
by: Ahn, Sanghyun, et al.
Published: (2025)
NeSyC: A Neuro-symbolic Continual Learner For Complex Embodied Tasks In Open Domains
by: Choi, Wonje, et al.
Published: (2025)
by: Choi, Wonje, et al.
Published: (2025)
EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models
by: Du, Mengfei, et al.
Published: (2024)
by: Du, Mengfei, et al.
Published: (2024)
TaskBench: Benchmarking Large Language Models for Task Automation
by: Shen, Yongliang, et al.
Published: (2023)
by: Shen, Yongliang, et al.
Published: (2023)
REI-Bench: Can Embodied Agents Understand Vague Human Instructions in Task Planning?
by: Jiang, Chenxi, et al.
Published: (2025)
by: Jiang, Chenxi, et al.
Published: (2025)
Hindsight Planner: A Closed-Loop Few-Shot Planner for Embodied Instruction Following
by: Yang, Yuxiao, et al.
Published: (2024)
by: Yang, Yuxiao, et al.
Published: (2024)
A Framework for Benchmarking and Aligning Task-Planning Safety in LLM-Based Embodied Agents
by: Huang, Yuting, et al.
Published: (2025)
by: Huang, Yuting, et al.
Published: (2025)
Confidence-guided Refinement Reasoning for Zero-shot Question Answering
by: Jang, Youwon, et al.
Published: (2025)
by: Jang, Youwon, et al.
Published: (2025)
ET-Plan-Bench: Embodied Task-level Planning Benchmark Towards Spatial-Temporal Cognition with Foundation Models
by: Zhang, Lingfeng, et al.
Published: (2024)
by: Zhang, Lingfeng, et al.
Published: (2024)
EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems
by: Qin, Xue, et al.
Published: (2026)
by: Qin, Xue, et al.
Published: (2026)
The Robot's Inner Critic: Self-Refinement of Social Behaviors through VLM-based Replanning
by: Lim, Jiyu, et al.
Published: (2026)
by: Lim, Jiyu, et al.
Published: (2026)
ReALFRED: An Embodied Instruction Following Benchmark in Photo-Realistic Environments
by: Kim, Taewoong, et al.
Published: (2024)
by: Kim, Taewoong, et al.
Published: (2024)
LoCoBench-Agent: An Interactive Benchmark for LLM Agents in Long-Context Software Engineering
by: Qiu, Jielin, et al.
Published: (2025)
by: Qiu, Jielin, et al.
Published: (2025)
SDA-PLANNER: State-Dependency Aware Adaptive Planner for Embodied Task Planning
by: Shen, Zichao, et al.
Published: (2025)
by: Shen, Zichao, et al.
Published: (2025)
S3LoRA: Safe Spectral Sharpness-Guided Pruning in Adaptation of Agent Planner
by: Ao, Shuang, et al.
Published: (2025)
by: Ao, Shuang, et al.
Published: (2025)
Exploratory Retrieval-Augmented Planning For Continual Embodied Instruction Following
by: Yoo, Minjong, et al.
Published: (2025)
by: Yoo, Minjong, et al.
Published: (2025)
Multi-Modal Grounded Planning and Efficient Replanning For Learning Embodied Agents with A Few Examples
by: Kim, Taewoong, et al.
Published: (2024)
by: Kim, Taewoong, et al.
Published: (2024)
Spatially-Aware Transformer for Embodied Agents
by: Cho, Junmo, et al.
Published: (2024)
by: Cho, Junmo, et al.
Published: (2024)
Context-Aware Planning and Environment-Aware Memory for Instruction Following Embodied Agents
by: Kim, Byeonghwi, et al.
Published: (2023)
by: Kim, Byeonghwi, et al.
Published: (2023)
Analysis of the Memorization and Generalization Capabilities of AI Agents: Are Continual Learners Robust?
by: Kim, Minsu, et al.
Published: (2023)
by: Kim, Minsu, et al.
Published: (2023)
SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
by: Li, Xiangyi, et al.
Published: (2026)
by: Li, Xiangyi, et al.
Published: (2026)
HWE-Bench: Benchmarking LLM Agents on Real-World Hardware Bug Repair Tasks
by: Cui, Fan, et al.
Published: (2026)
by: Cui, Fan, et al.
Published: (2026)
ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs
by: Wu, Xin, et al.
Published: (2026)
by: Wu, Xin, et al.
Published: (2026)
Similar Items
-
ReAcTree: Hierarchical LLM Agent Trees with Control Flow for Long-Horizon Task Planning
by: Choi, Jae-Woo, et al.
Published: (2025) -
Learning Dexterous Bimanual Catch Skills through Adversarial-Cooperative Heterogeneous-Agent Reinforcement Learning
by: Kim, Taewoo, et al.
Published: (2025) -
World Model Implanting for Test-time Adaptation of Embodied Agents
by: Yoo, Minjong, et al.
Published: (2025) -
Test-Time Mixture of World Models for Embodied Agents in Dynamic Environments
by: Jang, Jinwoo, et al.
Published: (2026) -
HELP: Hierarchical Embodied Language Planner for Household Tasks
by: Korchemnyi, Alexandr V., et al.
Published: (2025)