ProgAgent:A Continual RL Agent with Progress-Aware Rewards
Fuente:
arXiv
Saved in:
| Main Authors: | Tan, Jinzhou, Adineera, Gabriel, Kim, Jinoh |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ResWM: Residual-Action World Model for Visual RL
by: Zhang, Jseen, et al.
Published: (2026)
by: Zhang, Jseen, et al.
Published: (2026)
ProgRM: Build Better GUI Agents with Progress Rewards
by: Zhang, Danyang, et al.
Published: (2025)
by: Zhang, Danyang, et al.
Published: (2025)
ProgFed: Effective, Communication, and Computation Efficient Federated Learning by Progressive Training
by: Wang, Hui-Po, et al.
Published: (2021)
by: Wang, Hui-Po, et al.
Published: (2021)
Debate as Reward: A Multi-Agent Reward System for Scientific Ideation via RL Post-Training
by: Salimi, Moein, et al.
Published: (2026)
by: Salimi, Moein, et al.
Published: (2026)
Towards Improving Reward Design in RL: A Reward Alignment Metric for RL Practitioners
by: Muslimani, Calarina, et al.
Published: (2025)
by: Muslimani, Calarina, et al.
Published: (2025)
Capacity-Aware Planning and Scheduling in Budget-Constrained Multi-Agent MDPs: A Meta-RL Approach
by: Vora, Manav, et al.
Published: (2024)
by: Vora, Manav, et al.
Published: (2024)
Meta-RL Induces Exploration in Language Agents
by: Jiang, Yulun, et al.
Published: (2025)
by: Jiang, Yulun, et al.
Published: (2025)
Accelerating Diffusion Planners in Offline RL via Reward-Aware Consistency Trajectory Distillation
by: Duan, Xintong, et al.
Published: (2025)
by: Duan, Xintong, et al.
Published: (2025)
RDAR: Reward-Driven Agent Relevance Estimation for Autonomous Driving
by: Bosio, Carlo, et al.
Published: (2025)
by: Bosio, Carlo, et al.
Published: (2025)
RF-Agent: Automated Reward Function Design via Language Agent Tree Search
by: Gao, Ning, et al.
Published: (2026)
by: Gao, Ning, et al.
Published: (2026)
ArenaRL: Scaling RL for Open-Ended Agents via Tournament-based Relative Ranking
by: Zhang, Qiang, et al.
Published: (2026)
by: Zhang, Qiang, et al.
Published: (2026)
ProgCo: Program Helps Self-Correction of Large Language Models
by: Song, Xiaoshuai, et al.
Published: (2025)
by: Song, Xiaoshuai, et al.
Published: (2025)
AgentRM: Enhancing Agent Generalization with Reward Modeling
by: Xia, Yu, et al.
Published: (2025)
by: Xia, Yu, et al.
Published: (2025)
Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
by: Cherepanov, Egor, et al.
Published: (2024)
by: Cherepanov, Egor, et al.
Published: (2024)
Mind the Model, Not the Agent: The Primacy Bias in Model-based RL
by: Qiao, Zhongjian, et al.
Published: (2023)
by: Qiao, Zhongjian, et al.
Published: (2023)
Proto Successor Measure: Representing the Behavior Space of an RL Agent
by: Agarwal, Siddhant, et al.
Published: (2024)
by: Agarwal, Siddhant, et al.
Published: (2024)
RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large Language Models
by: Feng, Xiao, et al.
Published: (2026)
by: Feng, Xiao, et al.
Published: (2026)
FlickerFusion: Intra-trajectory Domain Generalizing Multi-Agent RL
by: Koh, Woosung, et al.
Published: (2024)
by: Koh, Woosung, et al.
Published: (2024)
Analysis of the Memorization and Generalization Capabilities of AI Agents: Are Continual Learners Robust?
by: Kim, Minsu, et al.
Published: (2023)
by: Kim, Minsu, et al.
Published: (2023)
Reward Sharpness-Aware Fine-Tuning for Diffusion Models
by: Kim, Kwanyoung, et al.
Published: (2026)
by: Kim, Kwanyoung, et al.
Published: (2026)
MobileRL: Online Agentic Reinforcement Learning for Mobile GUI Agents
by: Xu, Yifan, et al.
Published: (2025)
by: Xu, Yifan, et al.
Published: (2025)
Which Experiences Are Influential for RL Agents? Efficiently Estimating The Influence of Experiences
by: Hiraoka, Takuya, et al.
Published: (2024)
by: Hiraoka, Takuya, et al.
Published: (2024)
Process Reward Models for LLM Agents: Practical Framework and Directions
by: Choudhury, Sanjiban
Published: (2025)
by: Choudhury, Sanjiban
Published: (2025)
Online Continual Learning For Interactive Instruction Following Agents
by: Kim, Byeonghwi, et al.
Published: (2024)
by: Kim, Byeonghwi, et al.
Published: (2024)
Adaptive Milestone Reward for GUI Agents
by: Zheng, Congmin, et al.
Published: (2026)
by: Zheng, Congmin, et al.
Published: (2026)
LeakAgent: RL-based Red-teaming Agent for LLM Privacy Leakage
by: Nie, Yuzhou, et al.
Published: (2024)
by: Nie, Yuzhou, et al.
Published: (2024)
ProgVLA: Progress-Aware Robot Manipulation Skill Learning
by: Kim, Seungsu, et al.
Published: (2026)
by: Kim, Seungsu, et al.
Published: (2026)
AgentRewardBench: Evaluating Automatic Evaluations of Web Agent Trajectories
by: Lù, Xing Han, et al.
Published: (2025)
by: Lù, Xing Han, et al.
Published: (2025)
Your Reward Function for RL is Your Best PRM for Search: Unifying RL and Search-Based TTS
by: Jin, Can, et al.
Published: (2025)
by: Jin, Can, et al.
Published: (2025)
EnterpriseBench Corecraft: Training Generalizable Agents on High-Fidelity RL Environments
by: Mehta, Sushant, et al.
Published: (2026)
by: Mehta, Sushant, et al.
Published: (2026)
Plan Before You Trade: Inference-Time Optimization for RL Trading Agents
by: Go, Eun, et al.
Published: (2026)
by: Go, Eun, et al.
Published: (2026)
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities
by: Schmied, Thomas, et al.
Published: (2025)
by: Schmied, Thomas, et al.
Published: (2025)
Reward Hacking Benchmark: Measuring Exploits in LLM Agents with Tool Use
by: Thaman, Kunvar
Published: (2026)
by: Thaman, Kunvar
Published: (2026)
MAESTRO: Multi-Agent Environment Shaping through Task and Reward Optimization
by: Wu, Boyuan
Published: (2025)
by: Wu, Boyuan
Published: (2025)
Verifying Meta-Awareness via Predictive Rewards in Reasoning Models
by: Kim, Yoonjeon, et al.
Published: (2025)
by: Kim, Yoonjeon, et al.
Published: (2025)
AgentGA: Evolving Code Solutions in Agent-Seed Space
by: Tan, David Y. Y., et al.
Published: (2026)
by: Tan, David Y. Y., et al.
Published: (2026)
Agent-Omit: Adaptive Context Omission for Efficient LLM Agents
by: Ning, Yansong, et al.
Published: (2026)
by: Ning, Yansong, et al.
Published: (2026)
Spatially-Aware Transformer for Embodied Agents
by: Cho, Junmo, et al.
Published: (2024)
by: Cho, Junmo, et al.
Published: (2024)
FairAgent: Democratizing Fairness-Aware Machine Learning with LLM-Powered Agents
by: Dai, Yucong, et al.
Published: (2025)
by: Dai, Yucong, et al.
Published: (2025)
Multi-Objective Instruction-Aware Representation Learning in Procedural Content Generation RL
by: Kim, Sung-Hyun, et al.
Published: (2025)
by: Kim, Sung-Hyun, et al.
Published: (2025)
Similar Items
-
ResWM: Residual-Action World Model for Visual RL
by: Zhang, Jseen, et al.
Published: (2026) -
ProgRM: Build Better GUI Agents with Progress Rewards
by: Zhang, Danyang, et al.
Published: (2025) -
ProgFed: Effective, Communication, and Computation Efficient Federated Learning by Progressive Training
by: Wang, Hui-Po, et al.
Published: (2021) -
Debate as Reward: A Multi-Agent Reward System for Scientific Ideation via RL Post-Training
by: Salimi, Moein, et al.
Published: (2026) -
Towards Improving Reward Design in RL: A Reward Alignment Metric for RL Practitioners
by: Muslimani, Calarina, et al.
Published: (2025)