Robotouille: An Asynchronous Planning Benchmark for LLM Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gonzalez-Pumariega, Gonzalo, Yean, Leong Su, Sunkara, Neha, Choudhury, Sanjiban |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Query-Efficient Planning with Language Models
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2024)
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2024)
APRICOT: Active Preference Learning and Constraint-Aware Task Planning with LLMs
von: Wang, Huaxiaoyue, et al.
Veröffentlicht: (2024)
von: Wang, Huaxiaoyue, et al.
Veröffentlicht: (2024)
Multi-Turn Code Generation Through Single-Step Rewards
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2025)
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2025)
LLM-Empowered Embodied Agent for Memory-Augmented Task Planning in Household Robotics
von: Glocker, Marc, et al.
Veröffentlicht: (2025)
von: Glocker, Marc, et al.
Veröffentlicht: (2025)
Seeing Isn't Believing: Mitigating Belief Inertia via Active Intervention in Embodied Agents
von: Wang, Hanlin, et al.
Veröffentlicht: (2026)
von: Wang, Hanlin, et al.
Veröffentlicht: (2026)
Process Reward Models for LLM Agents: Practical Framework and Directions
von: Choudhury, Sanjiban
Veröffentlicht: (2025)
von: Choudhury, Sanjiban
Veröffentlicht: (2025)
LLM-A*: Large Language Model Enhanced Incremental Heuristic Search on Path Planning
von: Meng, Silin, et al.
Veröffentlicht: (2024)
von: Meng, Silin, et al.
Veröffentlicht: (2024)
REI-Bench: Can Embodied Agents Understand Vague Human Instructions in Task Planning?
von: Jiang, Chenxi, et al.
Veröffentlicht: (2025)
von: Jiang, Chenxi, et al.
Veröffentlicht: (2025)
Guide-LLM: An Embodied LLM Agent and Text-Based Topological Map for Robotic Guidance of People with Visual Impairments
von: Song, Sangmim, et al.
Veröffentlicht: (2024)
von: Song, Sangmim, et al.
Veröffentlicht: (2024)
ADAPT: Benchmarking Commonsense Planning under Unspecified Affordance Constraints
von: Chen, Pei-An, et al.
Veröffentlicht: (2026)
von: Chen, Pei-An, et al.
Veröffentlicht: (2026)
Can an Embodied Agent Find Your "Cat-shaped Mug"? LLM-Guided Exploration for Zero-Shot Object Navigation
von: Dorbala, Vishnu Sashank, et al.
Veröffentlicht: (2023)
von: Dorbala, Vishnu Sashank, et al.
Veröffentlicht: (2023)
Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making
von: Li, Manling, et al.
Veröffentlicht: (2024)
von: Li, Manling, et al.
Veröffentlicht: (2024)
AgentRefine: Enhancing Agent Generalization through Refinement Tuning
von: Fu, Dayuan, et al.
Veröffentlicht: (2025)
von: Fu, Dayuan, et al.
Veröffentlicht: (2025)
Vision-Language Interpreter for Robot Task Planning
von: Shirai, Keisuke, et al.
Veröffentlicht: (2023)
von: Shirai, Keisuke, et al.
Veröffentlicht: (2023)
Tool-Planner: Task Planning with Clusters across Multiple Tools
von: Liu, Yanming, et al.
Veröffentlicht: (2024)
von: Liu, Yanming, et al.
Veröffentlicht: (2024)
Automotive innovation landscaping using LLM
von: Gorain, Raju, et al.
Veröffentlicht: (2024)
von: Gorain, Raju, et al.
Veröffentlicht: (2024)
A Pragmatist Robot: Learning to Plan Tasks by Experiencing the Real World
von: Qu, Kaixian, et al.
Veröffentlicht: (2025)
von: Qu, Kaixian, et al.
Veröffentlicht: (2025)
Policy-Guided World Model Planning for Language-Conditioned Visual Navigation
von: Chahe, Amirhosein, et al.
Veröffentlicht: (2026)
von: Chahe, Amirhosein, et al.
Veröffentlicht: (2026)
HiAgent: Hierarchical Working Memory Management for Solving Long-Horizon Agent Tasks with Large Language Model
von: Hu, Mengkang, et al.
Veröffentlicht: (2024)
von: Hu, Mengkang, et al.
Veröffentlicht: (2024)
SELP: Generating Safe and Efficient Task Plans for Robot Agents with Large Language Models
von: Wu, Yi, et al.
Veröffentlicht: (2024)
von: Wu, Yi, et al.
Veröffentlicht: (2024)
SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models
von: Torres-Fonseca, Josue, et al.
Veröffentlicht: (2026)
von: Torres-Fonseca, Josue, et al.
Veröffentlicht: (2026)
Details Make a Difference: Object State-Sensitive Neurorobotic Task Planning
von: Sun, Xiaowen, et al.
Veröffentlicht: (2024)
von: Sun, Xiaowen, et al.
Veröffentlicht: (2024)
IPPON: Common Sense Guided Informative Path Planning for Object Goal Navigation
von: Qu, Kaixian, et al.
Veröffentlicht: (2024)
von: Qu, Kaixian, et al.
Veröffentlicht: (2024)
RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins
von: Mu, Yao, et al.
Veröffentlicht: (2025)
von: Mu, Yao, et al.
Veröffentlicht: (2025)
Simulating User Agents for Embodied Conversational-AI
von: Philipov, Daniel, et al.
Veröffentlicht: (2024)
von: Philipov, Daniel, et al.
Veröffentlicht: (2024)
Consolidating Trees of Robotic Plans Generated Using Large Language Models to Improve Reliability
von: Sakib, Md Sadman, et al.
Veröffentlicht: (2024)
von: Sakib, Md Sadman, et al.
Veröffentlicht: (2024)
Code-as-Symbolic-Planner: Foundation Model-Based Robot Planning via Symbolic Code Generation
von: Chen, Yongchao, et al.
Veröffentlicht: (2025)
von: Chen, Yongchao, et al.
Veröffentlicht: (2025)
A Prompt-driven Task Planning Method for Multi-drones based on Large Language Model
von: Liu, Yaohua
Veröffentlicht: (2024)
von: Liu, Yaohua
Veröffentlicht: (2024)
RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins (early version)
von: Mu, Yao, et al.
Veröffentlicht: (2024)
von: Mu, Yao, et al.
Veröffentlicht: (2024)
More Than Meets the Eye? Uncovering the Reasoning-Planning Disconnect in Training Vision-Language Driving Models
von: Song, Xurui, et al.
Veröffentlicht: (2025)
von: Song, Xurui, et al.
Veröffentlicht: (2025)
LLM4AD: Large Language Models for Autonomous Driving -- Concept, Review, Benchmark, Experiments, and Future Trends
von: Cui, Can, et al.
Veröffentlicht: (2024)
von: Cui, Can, et al.
Veröffentlicht: (2024)
From Novice to Expert: LLM Agent Policy Optimization via Step-wise Reinforcement Learning
von: Deng, Zhirui, et al.
Veröffentlicht: (2024)
von: Deng, Zhirui, et al.
Veröffentlicht: (2024)
HyCodePolicy: Hybrid Language Controllers for Multimodal Monitoring and Decision in Embodied Agents
von: Liu, Yibin, et al.
Veröffentlicht: (2025)
von: Liu, Yibin, et al.
Veröffentlicht: (2025)
Diffusion-ES: Gradient-free Planning with Diffusion for Autonomous Driving and Zero-Shot Instruction Following
von: Yang, Brian, et al.
Veröffentlicht: (2024)
von: Yang, Brian, et al.
Veröffentlicht: (2024)
Characterizing the Robustness of Black-Box LLM Planners Under Perturbed Observations with Adaptive Stress Testing
von: Chakraborty, Neeloy, et al.
Veröffentlicht: (2025)
von: Chakraborty, Neeloy, et al.
Veröffentlicht: (2025)
From Words to Collisions: LLM-Guided Evaluation and Adversarial Generation of Safety-Critical Driving Scenarios
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
Libra-VLA: Achieving Learning Equilibrium via Asynchronous Coarse-to-Fine Dual-System
von: Wei, Yifei, et al.
Veröffentlicht: (2026)
von: Wei, Yifei, et al.
Veröffentlicht: (2026)
Better than Your Teacher: LLM Agents that learn from Privileged AI Feedback
von: Choudhury, Sanjiban, et al.
Veröffentlicht: (2024)
von: Choudhury, Sanjiban, et al.
Veröffentlicht: (2024)
Motion Tracks: A Unified Representation for Human-Robot Transfer in Few-Shot Imitation Learning
von: Ren, Juntao, et al.
Veröffentlicht: (2025)
von: Ren, Juntao, et al.
Veröffentlicht: (2025)
InteRACT: Transformer Models for Human Intent Prediction Conditioned on Robot Actions
von: Kedia, Kushal, et al.
Veröffentlicht: (2023)
von: Kedia, Kushal, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Query-Efficient Planning with Language Models
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2024) -
APRICOT: Active Preference Learning and Constraint-Aware Task Planning with LLMs
von: Wang, Huaxiaoyue, et al.
Veröffentlicht: (2024) -
Multi-Turn Code Generation Through Single-Step Rewards
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2025) -
LLM-Empowered Embodied Agent for Memory-Augmented Task Planning in Household Robotics
von: Glocker, Marc, et al.
Veröffentlicht: (2025) -
Seeing Isn't Believing: Mitigating Belief Inertia via Active Intervention in Embodied Agents
von: Wang, Hanlin, et al.
Veröffentlicht: (2026)