Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Youwei, Wang, Jian, Wang, Hanlin, Guo, Beichen, Li, Wenjie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
von: Liu, Youwei, et al.
Veröffentlicht: (2026)
von: Liu, Youwei, et al.
Veröffentlicht: (2026)
STeCa: Step-level Trajectory Calibration for LLM Agent Learning
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2026)
DeepImagine: Learning Biomedical Reasoning via Successive Counterfactual Imagining
von: Zheng, Youze, et al.
Veröffentlicht: (2026)
von: Zheng, Youze, et al.
Veröffentlicht: (2026)
E2CL: Exploration-based Error Correction Learning for Embodied Agents
von: Wang, Hanlin, et al.
Veröffentlicht: (2024)
von: Wang, Hanlin, et al.
Veröffentlicht: (2024)
Parallel Test-Time Scaling for Latent Reasoning Models
von: You, Runyang, et al.
Veröffentlicht: (2025)
von: You, Runyang, et al.
Veröffentlicht: (2025)
Scaling over Scaling: Exploring Test-Time Scaling Plateau in Large Reasoning Models
von: Wang, Jian, et al.
Veröffentlicht: (2025)
von: Wang, Jian, et al.
Veröffentlicht: (2025)
Adaptive Social Learning via Mode Policy Optimization for Language Agents
von: Wang, Minzheng, et al.
Veröffentlicht: (2025)
von: Wang, Minzheng, et al.
Veröffentlicht: (2025)
Self-Imagine: Effective Unimodal Reasoning with Multimodal Models using Self-Imagination
von: Akter, Syeda Nahida, et al.
Veröffentlicht: (2024)
von: Akter, Syeda Nahida, et al.
Veröffentlicht: (2024)
LiveClawBench: Benchmarking LLM Agents on Complex, Real-World Assistant Tasks
von: Long, Xiang, et al.
Veröffentlicht: (2026)
von: Long, Xiang, et al.
Veröffentlicht: (2026)
MCP-AgentBench: Evaluating Real-World Language Agent Performance with MCP-Mediated Tools
von: Guo, Zikang, et al.
Veröffentlicht: (2025)
von: Guo, Zikang, et al.
Veröffentlicht: (2025)
FutureSim: Replaying World Events to Evaluate Adaptive Agents
von: Goel, Shashwat, et al.
Veröffentlicht: (2026)
von: Goel, Shashwat, et al.
Veröffentlicht: (2026)
Reducing Credit Assignment Variance via Counterfactual Reasoning Paths
von: Ding, Fei, et al.
Veröffentlicht: (2026)
von: Ding, Fei, et al.
Veröffentlicht: (2026)
ToolACE-R: Model-aware Iterative Training and Adaptive Refinement for Tool Learning
von: Zeng, Xingshan, et al.
Veröffentlicht: (2025)
von: Zeng, Xingshan, et al.
Veröffentlicht: (2025)
TodoEvolve: Learning to Architect Agent Planning Systems
von: Liu, Jiaxi, et al.
Veröffentlicht: (2026)
von: Liu, Jiaxi, et al.
Veröffentlicht: (2026)
Evaluating Mathematical Reasoning of Large Language Models: A Focus on Error Identification and Correction
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2024)
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2024)
Large Language Models Meet Text-Attributed Graphs: A Survey of Integration Frameworks and Applications
von: Su, Guangxin, et al.
Veröffentlicht: (2025)
von: Su, Guangxin, et al.
Veröffentlicht: (2025)
Adaptive Milestone Reward for GUI Agents
von: Zheng, Congmin, et al.
Veröffentlicht: (2026)
von: Zheng, Congmin, et al.
Veröffentlicht: (2026)
ModelingAgent: Bridging LLMs and Mathematical Modeling for Real-World Challenges
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
Discovering Hierarchical Latent Capabilities of Language Models via Causal Representation Learning
von: Jin, Jikai, et al.
Veröffentlicht: (2025)
von: Jin, Jikai, et al.
Veröffentlicht: (2025)
Why Reasoning Fails to Plan: A Planning-Centric Analysis of Long-Horizon Decision Making in LLM Agents
von: Wang, Zehong, et al.
Veröffentlicht: (2026)
von: Wang, Zehong, et al.
Veröffentlicht: (2026)
PhoneWorld: Scaling Phone-Use Agent Environments
von: Tang, Zhengyang, et al.
Veröffentlicht: (2026)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2026)
MPO: Boosting LLM Agents with Meta Plan Optimization
von: Xiong, Weimin, et al.
Veröffentlicht: (2025)
von: Xiong, Weimin, et al.
Veröffentlicht: (2025)
A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis
von: Gur, Izzeddin, et al.
Veröffentlicht: (2023)
von: Gur, Izzeddin, et al.
Veröffentlicht: (2023)
Generalizable and Stable Finetuning of Pretrained Language Models on Low-Resource Texts
von: Somayajula, Sai Ashish, et al.
Veröffentlicht: (2024)
von: Somayajula, Sai Ashish, et al.
Veröffentlicht: (2024)
BoostStep: Boosting mathematical capability of Large Language Models via improved single-step reasoning
von: Zhang, Beichen, et al.
Veröffentlicht: (2025)
von: Zhang, Beichen, et al.
Veröffentlicht: (2025)
GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation
von: Li, Sijia, et al.
Veröffentlicht: (2026)
von: Li, Sijia, et al.
Veröffentlicht: (2026)
Current Agents Fail to Leverage World Model as Tool for Foresight
von: Qian, Cheng, et al.
Veröffentlicht: (2026)
von: Qian, Cheng, et al.
Veröffentlicht: (2026)
AR-Omni: A Unified Autoregressive Model for Any-to-Any Generation
von: Cheng, Dongjie, et al.
Veröffentlicht: (2026)
von: Cheng, Dongjie, et al.
Veröffentlicht: (2026)
Planning with Multi-Constraints via Collaborative Language Agents
von: Zhang, Cong, et al.
Veröffentlicht: (2024)
von: Zhang, Cong, et al.
Veröffentlicht: (2024)
Enhancing Hepatopathy Clinical Trial Efficiency: A Secure, Large Language Model-Powered Pre-Screening Pipeline
von: Gui, Xiongbin, et al.
Veröffentlicht: (2025)
von: Gui, Xiongbin, et al.
Veröffentlicht: (2025)
MLR-Copilot: Autonomous Machine Learning Research based on Large Language Models Agents
von: Li, Ruochen, et al.
Veröffentlicht: (2024)
von: Li, Ruochen, et al.
Veröffentlicht: (2024)
SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation
von: Yang, Wenjie, et al.
Veröffentlicht: (2025)
von: Yang, Wenjie, et al.
Veröffentlicht: (2025)
Seeing Isn't Believing: Mitigating Belief Inertia via Active Intervention in Embodied Agents
von: Wang, Hanlin, et al.
Veröffentlicht: (2026)
von: Wang, Hanlin, et al.
Veröffentlicht: (2026)
Adversarial Reinforcement Learning for Large Language Model Agent Safety
von: Wang, Zizhao, et al.
Veröffentlicht: (2025)
von: Wang, Zizhao, et al.
Veröffentlicht: (2025)
ALPINE: Unveiling the Planning Capability of Autoregressive Learning in Language Models
von: Wang, Siwei, et al.
Veröffentlicht: (2024)
von: Wang, Siwei, et al.
Veröffentlicht: (2024)
Agent Planning with World Knowledge Model
von: Qiao, Shuofei, et al.
Veröffentlicht: (2024)
von: Qiao, Shuofei, et al.
Veröffentlicht: (2024)
Benefits and Pitfalls of Reinforcement Learning for Language Model Planning: A Theoretical Perspective
von: Wang, Siwei, et al.
Veröffentlicht: (2025)
von: Wang, Siwei, et al.
Veröffentlicht: (2025)
OmniJARVIS: Unified Vision-Language-Action Tokenization Enables Open-World Instruction Following Agents
von: Wang, Zihao, et al.
Veröffentlicht: (2024)
von: Wang, Zihao, et al.
Veröffentlicht: (2024)
LiPUP-MA: A Residential Experience-centric Multi-Agent Framework for Living-in-the-loop Participatory Urban Planning
von: Ni, Hang, et al.
Veröffentlicht: (2024)
von: Ni, Hang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
von: Liu, Youwei, et al.
Veröffentlicht: (2026) -
STeCa: Step-level Trajectory Calibration for LLM Agent Learning
von: Wang, Hanlin, et al.
Veröffentlicht: (2025) -
Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2026) -
DeepImagine: Learning Biomedical Reasoning via Successive Counterfactual Imagining
von: Zheng, Youze, et al.
Veröffentlicht: (2026) -
E2CL: Exploration-based Error Correction Learning for Embodied Agents
von: Wang, Hanlin, et al.
Veröffentlicht: (2024)