CAPF: Guiding Search-Agent Rollouts with Credit-Attenuated Privileged Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Bin, Liao, Xinye, Liu, Yiming, Liao, Xin, Liu, Chonghan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CuSearch: Curriculum Rollout Sampling via Search Depth for Agentic RAG
by: Shen, Jianghan, et al.
Published: (2026)
by: Shen, Jianghan, et al.
Published: (2026)
Rollout Cards: A Reproducibility Standard for Agent Research
by: Masters, Charlie, et al.
Published: (2026)
by: Masters, Charlie, et al.
Published: (2026)
RTMC: Step-Level Credit Assignment via Rollout Trees
by: Wang, Tao, et al.
Published: (2026)
by: Wang, Tao, et al.
Published: (2026)
Agentic Flow Steering and Parallel Rollout Search for Spatially Grounded Text-to-Image Generation
by: Chen, Ping, et al.
Published: (2026)
by: Chen, Ping, et al.
Published: (2026)
Better than Your Teacher: LLM Agents that learn from Privileged AI Feedback
by: Choudhury, Sanjiban, et al.
Published: (2024)
by: Choudhury, Sanjiban, et al.
Published: (2024)
TSR: Trajectory-Search Rollouts for Multi-Turn RL of LLM Agents
by: Djuhera, Aladin, et al.
Published: (2026)
by: Djuhera, Aladin, et al.
Published: (2026)
Do Coding Agents Understand Least-Privilege Authorization?
by: Yan, Zheng, et al.
Published: (2026)
by: Yan, Zheng, et al.
Published: (2026)
FORTIS: Benchmarking Over-Privilege in Agent Skills
by: Li, Shawn, et al.
Published: (2026)
by: Li, Shawn, et al.
Published: (2026)
Counterfactual Credit Policy Optimization for Multi-Agent Collaboration
by: Li, Zhongyi, et al.
Published: (2026)
by: Li, Zhongyi, et al.
Published: (2026)
ProRL Agent: Rollout-as-a-Service for RL Training of Multi-Turn LLM Agents
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
ASTRO: Adaptive Stitching via Dynamics-Guided Trajectory Rollouts
by: Yu, Hang, et al.
Published: (2025)
by: Yu, Hang, et al.
Published: (2025)
Debate, Reflect, and Distill: Multi-Agent Feedback with Tree-Structured Preference Optimization for Efficient Language Model Enhancement
by: Zhou, Xiaofeng, et al.
Published: (2025)
by: Zhou, Xiaofeng, et al.
Published: (2025)
Progent: Securing AI Agents with Privilege Control
by: Shi, Tianneng, et al.
Published: (2025)
by: Shi, Tianneng, et al.
Published: (2025)
Guided Cooperation in Hierarchical Reinforcement Learning via Model-based Rollout
by: Wang, Haoran, et al.
Published: (2023)
by: Wang, Haoran, et al.
Published: (2023)
DARTS: Distribution-Aware Active Rollout Trajectory Shaping for Accelerating LLM Reinforcement Learning
by: Wang, Yujie, et al.
Published: (2026)
by: Wang, Yujie, et al.
Published: (2026)
EDGE-OPD: Internalizing Privileged Context with Evidence Guided On-Policy Distillation
by: Lazaridis, Aristotelis, et al.
Published: (2026)
by: Lazaridis, Aristotelis, et al.
Published: (2026)
PiCA: Pivot-Based Credit Assignment for Search Agentic Reinforcement Learning
by: Liu, Dongyi, et al.
Published: (2026)
by: Liu, Dongyi, et al.
Published: (2026)
Emergence: Overcoming Privileged Information Bias in Asymmetric Embodied Agents via Active Querying
by: Baek, Shaun, et al.
Published: (2025)
by: Baek, Shaun, et al.
Published: (2025)
AnchorDrive: LLM Scenario Rollout with Anchor-Guided Diffusion Regeneration for Safety-Critical Scenario Generation
by: Jiang, Zhulin, et al.
Published: (2026)
by: Jiang, Zhulin, et al.
Published: (2026)
Interaction as Intelligence Part II: Asynchronous Human-Agent Rollout for Long-Horizon Task Training
by: Fu, Dayuan, et al.
Published: (2025)
by: Fu, Dayuan, et al.
Published: (2025)
Evaluating Privilege Usage of Agents with Real-World Tools
by: Zhang, Quan, et al.
Published: (2026)
by: Zhang, Quan, et al.
Published: (2026)
Compass: Navigating Global Marine Lead Data Integration through Expert-Guided LLM Agent
by: Liu, Yiming, et al.
Published: (2026)
by: Liu, Yiming, et al.
Published: (2026)
Beyond Trajectory Rewards: Step-level Credit Assignment for Agentic Search via Graph Modeling
by: Liu, Yuchen, et al.
Published: (2026)
by: Liu, Yuchen, et al.
Published: (2026)
Thinking on Maps: How Foundation Model Agents Explore, Remember, and Reason Map Environments
by: Wei, Zhiwei, et al.
Published: (2025)
by: Wei, Zhiwei, et al.
Published: (2025)
Agent-as-a-Service based on Agent Network
by: Zhu, Yuhan, et al.
Published: (2025)
by: Zhu, Yuhan, et al.
Published: (2025)
From Patches to Trajectories: Privileged Process Supervision for Software-Engineering Agents
by: Ma, Murong, et al.
Published: (2026)
by: Ma, Murong, et al.
Published: (2026)
Does RLVR Extend Reasoning Boundaries? Investigating Capability Expansion in Vision-Language Models
by: Shen, Minghe, et al.
Published: (2025)
by: Shen, Minghe, et al.
Published: (2025)
AutoFeedback: An LLM-based Framework for Efficient and Accurate API Request Generation
by: Liu, Huanxi, et al.
Published: (2024)
by: Liu, Huanxi, et al.
Published: (2024)
Feedback-MPPI: Fast Sampling-Based MPC via Rollout Differentiation -- Adios low-level controllers
by: Belvedere, Tommaso, et al.
Published: (2025)
by: Belvedere, Tommaso, et al.
Published: (2025)
Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent
by: Huang, Ziyang, et al.
Published: (2025)
by: Huang, Ziyang, et al.
Published: (2025)
Search, Do not Guess: Teaching Small Language Models to Be Effective Search Agents
by: Liu, Yizhou, et al.
Published: (2026)
by: Liu, Yizhou, et al.
Published: (2026)
SmartSearch: Process Reward-Guided Query Refinement for Search Agents
by: Wen, Tongyu, et al.
Published: (2026)
by: Wen, Tongyu, et al.
Published: (2026)
An Agent-Oriented Pluggable Experience-RAG Skill for Experience-Driven Retrieval Strategy Orchestration
by: Zhang, Dutao, et al.
Published: (2026)
by: Zhang, Dutao, et al.
Published: (2026)
AutoAgent: Evolving Cognition and Elastic Memory Orchestration for Adaptive Agents
by: Wang, Xiaoxing, et al.
Published: (2026)
by: Wang, Xiaoxing, et al.
Published: (2026)
DiffusionRollout: Uncertainty-Aware Rollout Planning in Long-Horizon PDE Solving
by: Yoo, Seungwoo, et al.
Published: (2026)
by: Yoo, Seungwoo, et al.
Published: (2026)
EGL-SCA: Structural Credit Assignment for Co-Evolving Instructions and Tools in Graph Reasoning Agents
by: Yuan, Zike, et al.
Published: (2026)
by: Yuan, Zike, et al.
Published: (2026)
EconAgent: Large Language Model-Empowered Agents for Simulating Macroeconomic Activities
by: Li, Nian, et al.
Published: (2023)
by: Li, Nian, et al.
Published: (2023)
ARROW: An Adaptive Rollout and Routing Method for Global Weather Forecasting
by: Tian, Jindong, et al.
Published: (2025)
by: Tian, Jindong, et al.
Published: (2025)
COMPASS: Cognitive MCTS-Guided Process Alignment for Safe Search Agents
by: Shen, Wenkai, et al.
Published: (2026)
by: Shen, Wenkai, et al.
Published: (2026)
Enhancing Linux Privilege Escalation Attack Capabilities of Local LLM Agents
by: Probst, Benjamin, et al.
Published: (2026)
by: Probst, Benjamin, et al.
Published: (2026)
Similar Items
-
CuSearch: Curriculum Rollout Sampling via Search Depth for Agentic RAG
by: Shen, Jianghan, et al.
Published: (2026) -
Rollout Cards: A Reproducibility Standard for Agent Research
by: Masters, Charlie, et al.
Published: (2026) -
RTMC: Step-Level Credit Assignment via Rollout Trees
by: Wang, Tao, et al.
Published: (2026) -
Agentic Flow Steering and Parallel Rollout Search for Spatially Grounded Text-to-Image Generation
by: Chen, Ping, et al.
Published: (2026) -
Better than Your Teacher: LLM Agents that learn from Privileged AI Feedback
by: Choudhury, Sanjiban, et al.
Published: (2024)