MPO: Boosting LLM Agents with Meta Plan Optimization
Fuente:
arXiv
Salvato in:
| Autori principali: | Xiong, Weimin, Song, Yifan, Dong, Qingxiu, Zhao, Bingchan, Song, Feifan, Wang, Xun, Li, Sujian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
di: Xiong, Weimin, et al.
Pubblicazione: (2024)
di: Xiong, Weimin, et al.
Pubblicazione: (2024)
Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents
di: Song, Yifan, et al.
Pubblicazione: (2024)
di: Song, Yifan, et al.
Pubblicazione: (2024)
Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining
di: Xiong, Weimin, et al.
Pubblicazione: (2026)
di: Xiong, Weimin, et al.
Pubblicazione: (2026)
More Vulnerable than You Think: On the Stability of Tool-Integrated LLM Agents
di: Xiong, Weimin, et al.
Pubblicazione: (2025)
di: Xiong, Weimin, et al.
Pubblicazione: (2025)
AgentBank: Towards Generalized LLM Agents via Fine-Tuning on 50000+ Interaction Trajectories
di: Song, Yifan, et al.
Pubblicazione: (2024)
di: Song, Yifan, et al.
Pubblicazione: (2024)
Multiplex Thinking: Reasoning via Token-wise Branch-and-Merge
di: Tang, Yao, et al.
Pubblicazione: (2026)
di: Tang, Yao, et al.
Pubblicazione: (2026)
JudgeRLVR: Judge First, Generate Second for Efficient Reasoning
di: Duo, Jiangshan, et al.
Pubblicazione: (2026)
di: Duo, Jiangshan, et al.
Pubblicazione: (2026)
MASPO: Joint Prompt Optimization for LLM-based Multi-Agent Systems
di: Wang, Zhexuan, et al.
Pubblicazione: (2026)
di: Wang, Zhexuan, et al.
Pubblicazione: (2026)
Boosting Protein Language Models with Negative Sample Mining
di: Xu, Yaoyao, et al.
Pubblicazione: (2024)
di: Xu, Yaoyao, et al.
Pubblicazione: (2024)
Why Reasoning Fails to Plan: A Planning-Centric Analysis of Long-Horizon Decision Making in LLM Agents
di: Wang, Zehong, et al.
Pubblicazione: (2026)
di: Wang, Zehong, et al.
Pubblicazione: (2026)
EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis
di: Song, Xiaoshuai, et al.
Pubblicazione: (2026)
di: Song, Xiaoshuai, et al.
Pubblicazione: (2026)
LiPUP-MA: A Residential Experience-centric Multi-Agent Framework for Living-in-the-loop Participatory Urban Planning
di: Ni, Hang, et al.
Pubblicazione: (2024)
di: Ni, Hang, et al.
Pubblicazione: (2024)
Reinforcement Learning for Chain of Thought Compression with One-Domain-to-All Generalization
di: Li, Hanyu, et al.
Pubblicazione: (2025)
di: Li, Hanyu, et al.
Pubblicazione: (2025)
Planning In Natural Language Improves LLM Search For Code Generation
di: Wang, Evan, et al.
Pubblicazione: (2024)
di: Wang, Evan, et al.
Pubblicazione: (2024)
MetaGreen: Meta-Learning Inspired Transformer Selection for Green Semantic Communication
di: Mukherjee, Shubhabrata, et al.
Pubblicazione: (2024)
di: Mukherjee, Shubhabrata, et al.
Pubblicazione: (2024)
P-Aligner: Enabling Pre-Alignment of Language Models via Principled Instruction Synthesis
di: Song, Feifan, et al.
Pubblicazione: (2025)
di: Song, Feifan, et al.
Pubblicazione: (2025)
HealthFlow: A Self-Evolving AI Agent with Meta Planning for Autonomous Healthcare Research
di: Zhu, Yinghao, et al.
Pubblicazione: (2025)
di: Zhu, Yinghao, et al.
Pubblicazione: (2025)
Finetune Once: Decoupling General & Domain Learning with Dynamic Boosted Annealing
di: Tang, Yang, et al.
Pubblicazione: (2025)
di: Tang, Yang, et al.
Pubblicazione: (2025)
Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents
di: Zambrano, Alejandra, et al.
Pubblicazione: (2026)
di: Zambrano, Alejandra, et al.
Pubblicazione: (2026)
Automatic Curriculum Expert Iteration for Reliable LLM Reasoning
di: Zhao, Zirui, et al.
Pubblicazione: (2024)
di: Zhao, Zirui, et al.
Pubblicazione: (2024)
Dynamic Evaluation of Large Language Models by Meta Probing Agents
di: Zhu, Kaijie, et al.
Pubblicazione: (2024)
di: Zhu, Kaijie, et al.
Pubblicazione: (2024)
RAP: Retrieval-Augmented Planning with Contextual Memory for Multimodal LLM Agents
di: Kagaya, Tomoyuki, et al.
Pubblicazione: (2024)
di: Kagaya, Tomoyuki, et al.
Pubblicazione: (2024)
FreeKV: Boosting KV Cache Retrieval for Efficient LLM Inference
di: Liu, Guangda, et al.
Pubblicazione: (2025)
di: Liu, Guangda, et al.
Pubblicazione: (2025)
The Good, The Bad, and The Greedy: Evaluation of LLMs Should Not Ignore Non-Determinism
di: Song, Yifan, et al.
Pubblicazione: (2024)
di: Song, Yifan, et al.
Pubblicazione: (2024)
Optimizing Chain-of-Thought Reasoners via Gradient Variance Minimization in Rejection Sampling and RL
di: Yao, Jiarui, et al.
Pubblicazione: (2025)
di: Yao, Jiarui, et al.
Pubblicazione: (2025)
How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse
di: Deng, Yichuan, et al.
Pubblicazione: (2024)
di: Deng, Yichuan, et al.
Pubblicazione: (2024)
SABER: Switchable and Balanced Training for Efficient LLM Reasoning
di: Zhao, Kai, et al.
Pubblicazione: (2025)
di: Zhao, Kai, et al.
Pubblicazione: (2025)
AutoPDL: Automatic Prompt Optimization for LLM Agents
di: Spiess, Claudio, et al.
Pubblicazione: (2025)
di: Spiess, Claudio, et al.
Pubblicazione: (2025)
Your Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents
di: Shao, Shuai, et al.
Pubblicazione: (2025)
di: Shao, Shuai, et al.
Pubblicazione: (2025)
Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory Framework
di: Zhang, Zeyu, et al.
Pubblicazione: (2025)
di: Zhang, Zeyu, et al.
Pubblicazione: (2025)
Simplicity Prevails: Rethinking Negative Preference Optimization for LLM Unlearning
di: Fan, Chongyu, et al.
Pubblicazione: (2024)
di: Fan, Chongyu, et al.
Pubblicazione: (2024)
EERPD: Leveraging Emotion and Emotion Regulation for Improving Personality Detection
di: Li, Zheng, et al.
Pubblicazione: (2024)
di: Li, Zheng, et al.
Pubblicazione: (2024)
CreditAudit: 2$^\text{nd}$ Dimension for LLM Evaluation and Selection
di: Song, Yiliang, et al.
Pubblicazione: (2026)
di: Song, Yiliang, et al.
Pubblicazione: (2026)
Efficiency-Effectiveness Reranking FLOPs for LLM-based Rerankers
di: Peng, Zhiyuan, et al.
Pubblicazione: (2025)
di: Peng, Zhiyuan, et al.
Pubblicazione: (2025)
Adversarial Preference Optimization: Enhancing Your Alignment via RM-LLM Game
di: Cheng, Pengyu, et al.
Pubblicazione: (2023)
di: Cheng, Pengyu, et al.
Pubblicazione: (2023)
TrajAgent: An LLM-Agent Framework for Trajectory Modeling via Large-and-Small Model Collaboration
di: Du, Yuwei, et al.
Pubblicazione: (2024)
di: Du, Yuwei, et al.
Pubblicazione: (2024)
MetaTool: Facilitating Large Language Models to Master Tools with Meta-task Augmentation
di: Wang, Xiaohan, et al.
Pubblicazione: (2024)
di: Wang, Xiaohan, et al.
Pubblicazione: (2024)
Survey on Evaluation of LLM-based Agents
di: Yehudai, Asaf, et al.
Pubblicazione: (2025)
di: Yehudai, Asaf, et al.
Pubblicazione: (2025)
DataSciBench: An LLM Agent Benchmark for Data Science
di: Zhang, Dan, et al.
Pubblicazione: (2025)
di: Zhang, Dan, et al.
Pubblicazione: (2025)
LiteSearch: Efficacious Tree Search for LLM
di: Wang, Ante, et al.
Pubblicazione: (2024)
di: Wang, Ante, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
di: Xiong, Weimin, et al.
Pubblicazione: (2024) -
Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents
di: Song, Yifan, et al.
Pubblicazione: (2024) -
Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining
di: Xiong, Weimin, et al.
Pubblicazione: (2026) -
More Vulnerable than You Think: On the Stability of Tool-Integrated LLM Agents
di: Xiong, Weimin, et al.
Pubblicazione: (2025) -
AgentBank: Towards Generalized LLM Agents via Fine-Tuning on 50000+ Interaction Trajectories
di: Song, Yifan, et al.
Pubblicazione: (2024)