RAPO: Expanding Exploration for LLM Agents via Retrieval-Augmented Policy Optimization
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Siwei, Xiong, Yun, Chen, Xi, Jia, Zi'an, Huang, Renhong, Xu, Jiarong, Zhang, Jiawei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Unifying Text Semantics and Graph Structures for Temporal Text-attributed Graphs with Large Language Models
di: Zhang, Siwei, et al.
Pubblicazione: (2025)
di: Zhang, Siwei, et al.
Pubblicazione: (2025)
PolicySim: An LLM-Based Agent Social Simulation Sandbox for Proactive Policy Optimization
di: Huang, Renhong, et al.
Pubblicazione: (2026)
di: Huang, Renhong, et al.
Pubblicazione: (2026)
Rethinking Time Encoding via Learnable Transformation Functions
di: Chen, Xi, et al.
Pubblicazione: (2025)
di: Chen, Xi, et al.
Pubblicazione: (2025)
Rethinking Soft Compression in Retrieval-Augmented Generation: A Query-Conditioned Selector Perspective
di: Liu, Yunhao, et al.
Pubblicazione: (2026)
di: Liu, Yunhao, et al.
Pubblicazione: (2026)
RAPO: Risk-Aware Preference Optimization for Generalizable Safe Reasoning
di: Wei, Zeming, et al.
Pubblicazione: (2026)
di: Wei, Zeming, et al.
Pubblicazione: (2026)
Divergence-Augmented Policy Optimization
di: Wang, Qing, et al.
Pubblicazione: (2025)
di: Wang, Qing, et al.
Pubblicazione: (2025)
Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy Optimization
di: Liu, Zeyuan, et al.
Pubblicazione: (2026)
di: Liu, Zeyuan, et al.
Pubblicazione: (2026)
Towards Adaptive Neighborhood for Advancing Temporal Interaction Graph Modeling
di: Zhang, Siwei, et al.
Pubblicazione: (2024)
di: Zhang, Siwei, et al.
Pubblicazione: (2024)
NeuroGenPoisoning: Neuron-Guided Attacks on Retrieval-Augmented Generation of LLM via Genetic Optimization of External Knowledge
di: Zhu, Hanyu, et al.
Pubblicazione: (2025)
di: Zhu, Hanyu, et al.
Pubblicazione: (2025)
Enhancing LLM-based Search Agents via Contribution Weighted Group Relative Policy Optimization
di: Wang, Junzhe, et al.
Pubblicazione: (2026)
di: Wang, Junzhe, et al.
Pubblicazione: (2026)
LLM Agents Should Employ Security Principles
di: Zhang, Kaiyuan, et al.
Pubblicazione: (2025)
di: Zhang, Kaiyuan, et al.
Pubblicazione: (2025)
Customized Retrieval-Augmented Generation with LLM for Debiasing Recommendation Unlearning
di: Zhang, Haichao, et al.
Pubblicazione: (2025)
di: Zhang, Haichao, et al.
Pubblicazione: (2025)
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective
di: Jin, Bowen, et al.
Pubblicazione: (2025)
di: Jin, Bowen, et al.
Pubblicazione: (2025)
Retrieval Augmented Generation-Enhanced Distributed LLM Agents for Generalizable Traffic Signal Control with Emergency Vehicles
di: Li, Xinhang, et al.
Pubblicazione: (2025)
di: Li, Xinhang, et al.
Pubblicazione: (2025)
Retrieval-Augmented Generation for Large Language Models: A Survey
di: Gao, Yunfan, et al.
Pubblicazione: (2023)
di: Gao, Yunfan, et al.
Pubblicazione: (2023)
PTCG-Bench: Can LLM Agents Master Pokémon Trading Card Game?
di: Hua, Dongdong, et al.
Pubblicazione: (2026)
di: Hua, Dongdong, et al.
Pubblicazione: (2026)
Training LLM Agents for Spontaneous, Reward-Free Self-Evolution via World Knowledge Exploration
di: Zhang, Qifan, et al.
Pubblicazione: (2026)
di: Zhang, Qifan, et al.
Pubblicazione: (2026)
Prompt Learning on Temporal Interaction Graphs
di: Chen, Xi, et al.
Pubblicazione: (2024)
di: Chen, Xi, et al.
Pubblicazione: (2024)
LUMINA: LLM-Guided GPU Architecture Exploration via Bottleneck Analysis
di: Zhang, Tao, et al.
Pubblicazione: (2026)
di: Zhang, Tao, et al.
Pubblicazione: (2026)
Tree of Preferences for Diversified Recommendation
di: Yuan, Hanyang, et al.
Pubblicazione: (2025)
di: Yuan, Hanyang, et al.
Pubblicazione: (2025)
Efficient Agent: Optimizing Planning Capability for Multimodal Retrieval Augmented Generation
di: Wang, Yuechen, et al.
Pubblicazione: (2025)
di: Wang, Yuechen, et al.
Pubblicazione: (2025)
WESE: Weak Exploration to Strong Exploitation for LLM Agents
di: Huang, Xu, et al.
Pubblicazione: (2024)
di: Huang, Xu, et al.
Pubblicazione: (2024)
MMUEChange: A Generalized LLM Agent Framework for Intelligent Multi-Modal Urban Environment Change Analysis
di: Xiao, Zixuan, et al.
Pubblicazione: (2026)
di: Xiao, Zixuan, et al.
Pubblicazione: (2026)
Relative Policy-Transition Optimization for Fast Policy Transfer
di: Xu, Jiawei, et al.
Pubblicazione: (2022)
di: Xu, Jiawei, et al.
Pubblicazione: (2022)
From Novice to Expert: LLM Agent Policy Optimization via Step-wise Reinforcement Learning
di: Deng, Zhirui, et al.
Pubblicazione: (2024)
di: Deng, Zhirui, et al.
Pubblicazione: (2024)
APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents
di: Li, Yibo, et al.
Pubblicazione: (2026)
di: Li, Yibo, et al.
Pubblicazione: (2026)
METASYMBO: Multi-Agent Language-Guided Metamaterial Discovery via Symbolic Latent Evolution
di: Chen, Jianpeng, et al.
Pubblicazione: (2026)
di: Chen, Jianpeng, et al.
Pubblicazione: (2026)
Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents
di: Song, Yifan, et al.
Pubblicazione: (2024)
di: Song, Yifan, et al.
Pubblicazione: (2024)
Anchored Policy Optimization: Mitigating Exploration Collapse Via Support-Constrained Rectification
di: Wang, Tianyi, et al.
Pubblicazione: (2026)
di: Wang, Tianyi, et al.
Pubblicazione: (2026)
MAXS: Meta-Adaptive Exploration with LLM Agents
di: Zhang, Jian, et al.
Pubblicazione: (2026)
di: Zhang, Jian, et al.
Pubblicazione: (2026)
Agent-Based Decentralized Energy Management of EV Charging Station with Solar Photovoltaics via Multi-Agent Reinforcement Learning
di: Fan, Jiarong, et al.
Pubblicazione: (2025)
di: Fan, Jiarong, et al.
Pubblicazione: (2025)
Improved LLM Agents for Financial Document Question Answering
di: Tan, Nelvin, et al.
Pubblicazione: (2025)
di: Tan, Nelvin, et al.
Pubblicazione: (2025)
iDSE: Navigating Design Space Exploration in High-Level Synthesis Using LLMs
di: Li, Runkai, et al.
Pubblicazione: (2025)
di: Li, Runkai, et al.
Pubblicazione: (2025)
Benchmarking Retrieval-Augmented Generation for Medicine
di: Xiong, Guangzhi, et al.
Pubblicazione: (2024)
di: Xiong, Guangzhi, et al.
Pubblicazione: (2024)
AtomicRAG: Atom-Entity Graphs for Retrieval-Augmented Generation
di: Hou, Yanning, et al.
Pubblicazione: (2026)
di: Hou, Yanning, et al.
Pubblicazione: (2026)
Disentangling Policy from Offline Task Representation Learning via Adversarial Data Augmentation
di: Jia, Chengxing, et al.
Pubblicazione: (2024)
di: Jia, Chengxing, et al.
Pubblicazione: (2024)
Reinforcement Learning-Augmented LLM Agents for Collaborative Decision Making and Performance Optimization
di: Qiu, Dong, et al.
Pubblicazione: (2025)
di: Qiu, Dong, et al.
Pubblicazione: (2025)
Beyond Individual Mimicry: Constructing Human-Like Social network with Graph-Augmented LLM Agents
di: Bu, Haoran, et al.
Pubblicazione: (2026)
di: Bu, Haoran, et al.
Pubblicazione: (2026)
Retrieval-Augmented LLM Agents: Learning to Learn from Experience
di: Ferraz, Thomas Palmeira, et al.
Pubblicazione: (2026)
di: Ferraz, Thomas Palmeira, et al.
Pubblicazione: (2026)
Retrieval-Augmented Robots via Retrieve-Reason-Act
di: Temiraliev, Izat, et al.
Pubblicazione: (2026)
di: Temiraliev, Izat, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Unifying Text Semantics and Graph Structures for Temporal Text-attributed Graphs with Large Language Models
di: Zhang, Siwei, et al.
Pubblicazione: (2025) -
PolicySim: An LLM-Based Agent Social Simulation Sandbox for Proactive Policy Optimization
di: Huang, Renhong, et al.
Pubblicazione: (2026) -
Rethinking Time Encoding via Learnable Transformation Functions
di: Chen, Xi, et al.
Pubblicazione: (2025) -
Rethinking Soft Compression in Retrieval-Augmented Generation: A Query-Conditioned Selector Perspective
di: Liu, Yunhao, et al.
Pubblicazione: (2026) -
RAPO: Risk-Aware Preference Optimization for Generalizable Safe Reasoning
di: Wei, Zeming, et al.
Pubblicazione: (2026)