AgentRM: Enhancing Agent Generalization with Reward Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xia, Yu, Fan, Jingru, Chen, Weize, Yan, Siyu, Cong, Xin, Zhang, Zhong, Lu, Yaxi, Lin, Yankai, Liu, Zhiyuan, Sun, Maosong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning to Generate Structured Output with Schema Reinforcement Learning
von: Lu, Yaxi, et al.
Veröffentlicht: (2025)
von: Lu, Yaxi, et al.
Veröffentlicht: (2025)
RepoAgent: An LLM-Powered Open-Source Framework for Repository-level Code Documentation Generation
von: Luo, Qinyu, et al.
Veröffentlicht: (2024)
von: Luo, Qinyu, et al.
Veröffentlicht: (2024)
Rational Decision-Making Agent with Internalized Utility Judgment
von: Ye, Yining, et al.
Veröffentlicht: (2023)
von: Ye, Yining, et al.
Veröffentlicht: (2023)
Proactive Agent: Shifting LLM Agents from Reactive Responses to Active Assistance
von: Lu, Yaxi, et al.
Veröffentlicht: (2024)
von: Lu, Yaxi, et al.
Veröffentlicht: (2024)
AgentCPM-Report: Interleaving Drafting and Deepening for Open-Ended Deep Research
von: Li, Yishan, et al.
Veröffentlicht: (2026)
von: Li, Yishan, et al.
Veröffentlicht: (2026)
Investigate-Consolidate-Exploit: A General Strategy for Inter-Task Agent Self-Evolution
von: Qian, Cheng, et al.
Veröffentlicht: (2024)
von: Qian, Cheng, et al.
Veröffentlicht: (2024)
AgentRM: An OS-Inspired Resource Manager for LLM Agent Systems
von: She, Jianshu
Veröffentlicht: (2026)
von: She, Jianshu
Veröffentlicht: (2026)
Tell Me More! Towards Implicit User Intention Understanding of Language Model Driven Agents
von: Qian, Cheng, et al.
Veröffentlicht: (2024)
von: Qian, Cheng, et al.
Veröffentlicht: (2024)
WorkflowLLM: Enhancing Workflow Orchestration Capability of Large Language Models
von: Fan, Shengda, et al.
Veröffentlicht: (2024)
von: Fan, Shengda, et al.
Veröffentlicht: (2024)
Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System
von: Chen, Weize, et al.
Veröffentlicht: (2024)
von: Chen, Weize, et al.
Veröffentlicht: (2024)
ConsistRM: Improving Generative Reward Models via Consistency-Aware Self-Training
von: Liang, Yu, et al.
Veröffentlicht: (2026)
von: Liang, Yu, et al.
Veröffentlicht: (2026)
ProgRM: Build Better GUI Agents with Progress Rewards
von: Zhang, Danyang, et al.
Veröffentlicht: (2025)
von: Zhang, Danyang, et al.
Veröffentlicht: (2025)
Internet of Agents: Weaving a Web of Heterogeneous Agents for Collaborative Intelligence
von: Chen, Weize, et al.
Veröffentlicht: (2024)
von: Chen, Weize, et al.
Veröffentlicht: (2024)
Representation Learning for Natural Language Processing
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2020)
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2020)
Multi-Agent Collaboration via Evolving Orchestration
von: Dang, Yufan, et al.
Veröffentlicht: (2025)
von: Dang, Yufan, et al.
Veröffentlicht: (2025)
ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework
von: Qin, Kai, et al.
Veröffentlicht: (2026)
von: Qin, Kai, et al.
Veröffentlicht: (2026)
Experiential Co-Learning of Software-Developing Agents
von: Qian, Chen, et al.
Veröffentlicht: (2023)
von: Qian, Chen, et al.
Veröffentlicht: (2023)
ChatDev: Communicative Agents for Software Development
von: Qian, Chen, et al.
Veröffentlicht: (2023)
von: Qian, Chen, et al.
Veröffentlicht: (2023)
LLM$\times$MapReduce-V3: Enabling Interactive In-Depth Survey Generation through a MCP-Driven Hierarchically Modular Agent System
von: Chao, Yu, et al.
Veröffentlicht: (2025)
von: Chao, Yu, et al.
Veröffentlicht: (2025)
SteerRM: Debiasing Reward Models via Sparse Autoencoders
von: Sun, Mengyuan, et al.
Veröffentlicht: (2026)
von: Sun, Mengyuan, et al.
Veröffentlicht: (2026)
Enhancing Open-Domain Task-Solving Capability of LLMs via Autonomous Tool Integration from GitHub
von: Lyu, Bohan, et al.
Veröffentlicht: (2023)
von: Lyu, Bohan, et al.
Veröffentlicht: (2023)
Scaling Large Language Model-based Multi-Agent Collaboration
von: Qian, Chen, et al.
Veröffentlicht: (2024)
von: Qian, Chen, et al.
Veröffentlicht: (2024)
Proof-RM: A Scalable and Generalizable Reward Model for Math Proof
von: Yang, Haotong, et al.
Veröffentlicht: (2026)
von: Yang, Haotong, et al.
Veröffentlicht: (2026)
GUICourse: From General Vision Language Models to Versatile GUI Agents
von: Chen, Wentong, et al.
Veröffentlicht: (2024)
von: Chen, Wentong, et al.
Veröffentlicht: (2024)
RM-Distiller: Exploiting Generative LLM for Reward Model Distillation
von: Zhou, Hongli, et al.
Veröffentlicht: (2026)
von: Zhou, Hongli, et al.
Veröffentlicht: (2026)
MatPlotAgent: Method and Evaluation for LLM-Based Agentic Scientific Data Visualization
von: Yang, Zhiyu, et al.
Veröffentlicht: (2024)
von: Yang, Zhiyu, et al.
Veröffentlicht: (2024)
Iterative Experience Refinement of Software-Developing Agents
von: Qian, Chen, et al.
Veröffentlicht: (2024)
von: Qian, Chen, et al.
Veröffentlicht: (2024)
Fast-Slow Thinking RM: Efficient Integration of Scalar and Generative Reward Models
von: Wu, Jiayun, et al.
Veröffentlicht: (2026)
von: Wu, Jiayun, et al.
Veröffentlicht: (2026)
ToLeaP: Rethinking Development of Tool Learning with Large Language Models
von: Chen, Haotian, et al.
Veröffentlicht: (2025)
von: Chen, Haotian, et al.
Veröffentlicht: (2025)
RM-R1: Reward Modeling as Reasoning
von: Chen, Xiusi, et al.
Veröffentlicht: (2025)
von: Chen, Xiusi, et al.
Veröffentlicht: (2025)
DebugBench: Evaluating Debugging Capability of Large Language Models
von: Tian, Runchu, et al.
Veröffentlicht: (2024)
von: Tian, Runchu, et al.
Veröffentlicht: (2024)
AgentCPM-GUI: Building Mobile-Use Agents with Reinforcement Fine-Tuning
von: Zhang, Zhong, et al.
Veröffentlicht: (2025)
von: Zhang, Zhong, et al.
Veröffentlicht: (2025)
Cross-Task Experiential Learning on LLM-based Multi-Agent Collaboration
von: Li, Yilong, et al.
Veröffentlicht: (2025)
von: Li, Yilong, et al.
Veröffentlicht: (2025)
The Overthinker's DIET: Cutting Token Calories with DIfficulty-AwarE Training
von: Chen, Weize, et al.
Veröffentlicht: (2025)
von: Chen, Weize, et al.
Veröffentlicht: (2025)
AgentCPM-Explore: Realizing Long-Horizon Deep Exploration for Edge-Scale Agents
von: Chen, Haotian, et al.
Veröffentlicht: (2026)
von: Chen, Haotian, et al.
Veröffentlicht: (2026)
AgentMonitor: A Plug-and-Play Framework for Predictive and Secure Multi-Agent Systems
von: Chan, Chi-Min, et al.
Veröffentlicht: (2024)
von: Chan, Chi-Min, et al.
Veröffentlicht: (2024)
RAG-DDR: Optimizing Retrieval-Augmented Generation Using Differentiable Data Rewards
von: Li, Xinze, et al.
Veröffentlicht: (2024)
von: Li, Xinze, et al.
Veröffentlicht: (2024)
DogeRM: Equipping Reward Models with Domain Knowledge through Model Merging
von: Lin, Tzu-Han, et al.
Veröffentlicht: (2024)
von: Lin, Tzu-Han, et al.
Veröffentlicht: (2024)
Beyond Natural Language: LLMs Leveraging Alternative Formats for Enhanced Reasoning and Communication
von: Chen, Weize, et al.
Veröffentlicht: (2024)
von: Chen, Weize, et al.
Veröffentlicht: (2024)
P-GenRM: Personalized Generative Reward Model with Test-time User-based Scaling
von: Zhang, Pinyi, et al.
Veröffentlicht: (2026)
von: Zhang, Pinyi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Learning to Generate Structured Output with Schema Reinforcement Learning
von: Lu, Yaxi, et al.
Veröffentlicht: (2025) -
RepoAgent: An LLM-Powered Open-Source Framework for Repository-level Code Documentation Generation
von: Luo, Qinyu, et al.
Veröffentlicht: (2024) -
Rational Decision-Making Agent with Internalized Utility Judgment
von: Ye, Yining, et al.
Veröffentlicht: (2023) -
Proactive Agent: Shifting LLM Agents from Reactive Responses to Active Assistance
von: Lu, Yaxi, et al.
Veröffentlicht: (2024) -
AgentCPM-Report: Interleaving Drafting and Deepening for Open-Ended Deep Research
von: Li, Yishan, et al.
Veröffentlicht: (2026)