PARC: An Autonomous Self-Reflective Coding Agent for Robust Execution of Long-Horizon Tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Orimo, Yuki, Kurata, Iori, Mori, Hodaka, Okuno, Ryuhei, Sawada, Ryohto, Okanohara, Daisuke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
When Robots Do the Chores: A Benchmark and Agent for Long-Horizon Household Task Execution
von: Zhu, Zilin, et al.
Veröffentlicht: (2026)
von: Zhu, Zilin, et al.
Veröffentlicht: (2026)
The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Execution
von: Li, Junlong, et al.
Veröffentlicht: (2025)
von: Li, Junlong, et al.
Veröffentlicht: (2025)
SAGE: Scene Graph-Aware Guidance and Execution for Long-Horizon Manipulation Tasks
von: Li, Jialiang, et al.
Veröffentlicht: (2025)
von: Li, Jialiang, et al.
Veröffentlicht: (2025)
Learning on the Job: An Experience-Driven Self-Evolving Agent for Long-Horizon Tasks
von: Yang, Cheng, et al.
Veröffentlicht: (2025)
von: Yang, Cheng, et al.
Veröffentlicht: (2025)
Memory as Action: Autonomous Context Curation for Long-Horizon Agentic Tasks
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2025)
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2025)
Learning Agent-Compatible Context Management for Long-Horizon Tasks
von: Yi, Lu, et al.
Veröffentlicht: (2026)
von: Yi, Lu, et al.
Veröffentlicht: (2026)
NEMO: Execution-Aware Optimization Modeling via Autonomous Coding Agents
von: Song, Yang, et al.
Veröffentlicht: (2026)
von: Song, Yang, et al.
Veröffentlicht: (2026)
SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks
von: Orlanski, Gabriel, et al.
Veröffentlicht: (2026)
von: Orlanski, Gabriel, et al.
Veröffentlicht: (2026)
Beyond Entangled Planning: Task-Decoupled Planning for Long-Horizon Agents
von: Li, Yunfan, et al.
Veröffentlicht: (2026)
von: Li, Yunfan, et al.
Veröffentlicht: (2026)
ARC: Active and Reflection-driven Context Management for Long-Horizon Information Seeking Agents
von: Yao, Yilun, et al.
Veröffentlicht: (2026)
von: Yao, Yilun, et al.
Veröffentlicht: (2026)
REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation
von: Yuan, Puzhen, et al.
Veröffentlicht: (2025)
von: Yuan, Puzhen, et al.
Veröffentlicht: (2025)
GTA: Generating Long-Horizon Tasks for Web Agents at Scale
von: Huang, Tenghao, et al.
Veröffentlicht: (2026)
von: Huang, Tenghao, et al.
Veröffentlicht: (2026)
OmegaUse: Building a General-Purpose GUI Agent for Autonomous Task Execution
von: Zhang, Le, et al.
Veröffentlicht: (2026)
von: Zhang, Le, et al.
Veröffentlicht: (2026)
The Illusion of Diminishing Returns: Measuring Long Horizon Execution in LLMs
von: Sinha, Akshit, et al.
Veröffentlicht: (2025)
von: Sinha, Akshit, et al.
Veröffentlicht: (2025)
ELHPlan: Efficient Long-Horizon Task Planning for Multi-Agent Collaboration
von: Ling, Shaobin, et al.
Veröffentlicht: (2025)
von: Ling, Shaobin, et al.
Veröffentlicht: (2025)
Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Tasks
von: Wu, Xiyang, et al.
Veröffentlicht: (2026)
von: Wu, Xiyang, et al.
Veröffentlicht: (2026)
AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
FCRF: Flexible Constructivism Reflection for Long-Horizon Robotic Task Planning with Large Language Models
von: Song, Yufan, et al.
Veröffentlicht: (2025)
von: Song, Yufan, et al.
Veröffentlicht: (2025)
MemPO: Self-Memory Policy Optimization for Long-Horizon Agents
von: Li, Ruoran, et al.
Veröffentlicht: (2026)
von: Li, Ruoran, et al.
Veröffentlicht: (2026)
Orchestrator: Active Inference for Multi-Agent Systems in Long-Horizon Tasks
von: Beckenbauer, Lukas, et al.
Veröffentlicht: (2025)
von: Beckenbauer, Lukas, et al.
Veröffentlicht: (2025)
Fault-Tolerant Sandboxing for AI Coding Agents: A Transactional Approach to Safe Autonomous Execution
von: Yan, Boyang
Veröffentlicht: (2025)
von: Yan, Boyang
Veröffentlicht: (2025)
The Illusion of Procedural Reasoning: Measuring Long-Horizon FSM Execution in LLMs
von: Samiei, Mahdi, et al.
Veröffentlicht: (2025)
von: Samiei, Mahdi, et al.
Veröffentlicht: (2025)
LH-Bench: Skill-Grounded Evaluation of Long-Horizon Agents on Subjective Enterprise Tasks
von: Chandwani, Abhishek, et al.
Veröffentlicht: (2026)
von: Chandwani, Abhishek, et al.
Veröffentlicht: (2026)
STMA: A Spatio-Temporal Memory Agent for Long-Horizon Embodied Task Planning
von: Lei, Mingcong, et al.
Veröffentlicht: (2025)
von: Lei, Mingcong, et al.
Veröffentlicht: (2025)
ContextFlow: Hierarchical Task-State Alignment for Long-Horizon Embodied Agents
von: Guo, Shuhan, et al.
Veröffentlicht: (2026)
von: Guo, Shuhan, et al.
Veröffentlicht: (2026)
Heterogeneous Multi-Expert Reinforcement Learning for Long-Horizon Multi-Goal Tasks in Autonomous Forklifts
von: Chen, Yun, et al.
Veröffentlicht: (2026)
von: Chen, Yun, et al.
Veröffentlicht: (2026)
Tree-of-Code: A Hybrid Approach for Robust Complex Task Planning and Execution
von: Ni, Ziyi, et al.
Veröffentlicht: (2024)
von: Ni, Ziyi, et al.
Veröffentlicht: (2024)
WebPilot: A Versatile and Autonomous Multi-Agent System for Web Task Execution with Strategic Exploration
von: Zhang, Yao, et al.
Veröffentlicht: (2024)
von: Zhang, Yao, et al.
Veröffentlicht: (2024)
Intrinsic Stability Limits of Autoregressive Reasoning: Structural Consequences for Long-Horizon Execution
von: Liao, Hsien-Jyh
Veröffentlicht: (2026)
von: Liao, Hsien-Jyh
Veröffentlicht: (2026)
Interaction as Intelligence Part II: Asynchronous Human-Agent Rollout for Long-Horizon Task Training
von: Fu, Dayuan, et al.
Veröffentlicht: (2025)
von: Fu, Dayuan, et al.
Veröffentlicht: (2025)
InternAgent-1.5: A Unified Agentic Framework for Long-Horizon Autonomous Scientific Discovery
von: Feng, Shiyang, et al.
Veröffentlicht: (2026)
von: Feng, Shiyang, et al.
Veröffentlicht: (2026)
Curriculum Guided Massive Multi Agent System Solving For Robust Long Horizon Tasks
von: Kar, Indrajit, et al.
Veröffentlicht: (2025)
von: Kar, Indrajit, et al.
Veröffentlicht: (2025)
When the Specification Emerges: Benchmarking Faithfulness Loss in Long-Horizon Coding Agents
von: Yan, Lu, et al.
Veröffentlicht: (2026)
von: Yan, Lu, et al.
Veröffentlicht: (2026)
The Conversations Beneath the Code: Triadic Data for Long-Horizon Software Engineering Agents
von: Kim, Yelin
Veröffentlicht: (2026)
von: Kim, Yelin
Veröffentlicht: (2026)
MineEvolve: Self-Evolution with Accumulated Knowledge for Long-Horizon Embodied Minecraft Agents
von: Xie, Zhengwei, et al.
Veröffentlicht: (2026)
von: Xie, Zhengwei, et al.
Veröffentlicht: (2026)
ConceptAgent: LLM-Driven Precondition Grounding and Tree Search for Robust Task Planning and Execution
von: Rivera, Corban, et al.
Veröffentlicht: (2024)
von: Rivera, Corban, et al.
Veröffentlicht: (2024)
STRUCTUREDAGENT: Planning with AND/OR Trees for Long-Horizon Web Tasks
von: Lobo, ELita, et al.
Veröffentlicht: (2026)
von: Lobo, ELita, et al.
Veröffentlicht: (2026)
ReAcTree: Hierarchical LLM Agent Trees with Control Flow for Long-Horizon Task Planning
von: Choi, Jae-Woo, et al.
Veröffentlicht: (2025)
von: Choi, Jae-Woo, et al.
Veröffentlicht: (2025)
SpecBench: Measuring Reward Hacking in Long-Horizon Coding Agents
von: Zhao, Bingchen, et al.
Veröffentlicht: (2026)
von: Zhao, Bingchen, et al.
Veröffentlicht: (2026)
Optimus-1: Hybrid Multimodal Memory Empowered Agents Excel in Long-Horizon Tasks
von: Li, Zaijing, et al.
Veröffentlicht: (2024)
von: Li, Zaijing, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
When Robots Do the Chores: A Benchmark and Agent for Long-Horizon Household Task Execution
von: Zhu, Zilin, et al.
Veröffentlicht: (2026) -
The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Execution
von: Li, Junlong, et al.
Veröffentlicht: (2025) -
SAGE: Scene Graph-Aware Guidance and Execution for Long-Horizon Manipulation Tasks
von: Li, Jialiang, et al.
Veröffentlicht: (2025) -
Learning on the Job: An Experience-Driven Self-Evolving Agent for Long-Horizon Tasks
von: Yang, Cheng, et al.
Veröffentlicht: (2025) -
Memory as Action: Autonomous Context Curation for Long-Horizon Agentic Tasks
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2025)