Agentic Reward Modeling: Verifying GUI Agent via Online Proactive Interaction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cui, Chaoqun, Huang, Jing, Wang, Shijing, Zheng, Liming, Kong, Qingchao, Zeng, Zhixiong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TreeCUA: Efficiently Scaling GUI Automation with Tree-Structured Verifiable Evolution
von: Jiang, Deyang, et al.
Veröffentlicht: (2026)
von: Jiang, Deyang, et al.
Veröffentlicht: (2026)
MobileDreamer: Generative Sketch World Model for GUI Agent
von: Cao, Yilin, et al.
Veröffentlicht: (2026)
von: Cao, Yilin, et al.
Veröffentlicht: (2026)
ScaleTrack: Scaling and back-tracking Automated GUI Agents
von: Huang, Jing, et al.
Veröffentlicht: (2025)
von: Huang, Jing, et al.
Veröffentlicht: (2025)
RoVer: Robot Reward Model as Test-Time Verifier for Vision-Language-Action Model
von: Dai, Mingtong, et al.
Veröffentlicht: (2025)
von: Dai, Mingtong, et al.
Veröffentlicht: (2025)
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
von: Wu, Yanru, et al.
Veröffentlicht: (2026)
von: Wu, Yanru, et al.
Veröffentlicht: (2026)
Online Intrinsic Rewards for Decision Making Agents from Large Language Model Feedback
von: Zheng, Qinqing, et al.
Veröffentlicht: (2024)
von: Zheng, Qinqing, et al.
Veröffentlicht: (2024)
Imagine, Verify, Execute: Memory-guided Agentic Exploration with Vision-Language Models
von: Lee, Seungjae, et al.
Veröffentlicht: (2025)
von: Lee, Seungjae, et al.
Veröffentlicht: (2025)
Spatiotemporal Receding Horizon Control with Proactive Interaction Towards Autonomous Driving in Dense Traffic
von: Zheng, Lei, et al.
Veröffentlicht: (2023)
von: Zheng, Lei, et al.
Veröffentlicht: (2023)
AgentVLN: Towards Agentic Vision-and-Language Navigation
von: Xin, Zihao, et al.
Veröffentlicht: (2026)
von: Xin, Zihao, et al.
Veröffentlicht: (2026)
Reinforced Embodied Planning with Verifiable Reward for Real-World Robotic Manipulation
von: Bo, Zitong, et al.
Veröffentlicht: (2025)
von: Bo, Zitong, et al.
Veröffentlicht: (2025)
UItron: Foundational GUI Agent with Advanced Perception and Planning
von: Zeng, Zhixiong, et al.
Veröffentlicht: (2025)
von: Zeng, Zhixiong, et al.
Veröffentlicht: (2025)
From Reaction to Anticipation: Proactive Failure Recovery through Agentic Task Graph for Robotic Manipulation
von: Xu, Sheng, et al.
Veröffentlicht: (2026)
von: Xu, Sheng, et al.
Veröffentlicht: (2026)
Reward-Driven Automated Curriculum Learning for Interaction-Aware Self-Driving at Unsignalized Intersections
von: Peng, Zengqi, et al.
Veröffentlicht: (2024)
von: Peng, Zengqi, et al.
Veröffentlicht: (2024)
SpiritSight Agent: Advanced GUI Agent with One Look
von: Huang, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Huang, Zhiyuan, et al.
Veröffentlicht: (2025)
Cooperative Reward Shaping for Multi-Agent Pathfinding
von: Song, Zhenyu, et al.
Veröffentlicht: (2024)
von: Song, Zhenyu, et al.
Veröffentlicht: (2024)
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
von: Zhang, Chen Bo Calvin, et al.
Veröffentlicht: (2024)
von: Zhang, Chen Bo Calvin, et al.
Veröffentlicht: (2024)
UITron-Speech: Towards Automated GUI Agents Based on Speech Instructions
von: Han, Wenkang, et al.
Veröffentlicht: (2025)
von: Han, Wenkang, et al.
Veröffentlicht: (2025)
MobileUse: A GUI Agent with Hierarchical Reflection for Autonomous Mobile Operation
von: Li, Ning, et al.
Veröffentlicht: (2025)
von: Li, Ning, et al.
Veröffentlicht: (2025)
Learning Reward for Robot Skills Using Large Language Models via Self-Alignment
von: Zeng, Yuwei, et al.
Veröffentlicht: (2024)
von: Zeng, Yuwei, et al.
Veröffentlicht: (2024)
TOLEBI: Learning Fault-Tolerant Bipedal Locomotion via Online Status Estimation and Fallibility Rewards
von: Lee, Hokyun, et al.
Veröffentlicht: (2026)
von: Lee, Hokyun, et al.
Veröffentlicht: (2026)
GRAPPA: Generalizing and Adapting Robot Policies via Online Agentic Guidance
von: Bucker, Arthur, et al.
Veröffentlicht: (2024)
von: Bucker, Arthur, et al.
Veröffentlicht: (2024)
SInViG: A Self-Evolving Interactive Visual Agent for Human-Robot Interaction
von: Xu, Jie, et al.
Veröffentlicht: (2024)
von: Xu, Jie, et al.
Veröffentlicht: (2024)
Legible and Proactive Robot Planning for Prosocial Human-Robot Interactions
von: Geldenbott, Jasper, et al.
Veröffentlicht: (2024)
von: Geldenbott, Jasper, et al.
Veröffentlicht: (2024)
HPTune: Hierarchical Proactive Tuning for Collision-Free Model Predictive Control
von: Zuo, Wei, et al.
Veröffentlicht: (2026)
von: Zuo, Wei, et al.
Veröffentlicht: (2026)
Efficient Language-instructed Skill Acquisition via Reward-Policy Co-Evolution
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
Action Draft and Verify: A Self-Verifying Framework for Vision-Language-Action Model
von: Zhao, Chen, et al.
Veröffentlicht: (2026)
von: Zhao, Chen, et al.
Veröffentlicht: (2026)
AgentV-RL: Scaling Reward Modeling with Agentic Verifier
von: Zhang, Jiazheng, et al.
Veröffentlicht: (2026)
von: Zhang, Jiazheng, et al.
Veröffentlicht: (2026)
Observability-Aware Active Calibration of Multi-Sensor Extrinsics for Ground Robots via Online Trajectory Optimization
von: Wang, Jiang, et al.
Veröffentlicht: (2025)
von: Wang, Jiang, et al.
Veröffentlicht: (2025)
Good Weights: Proactive, Adaptive Dead Reckoning Fusion for Continuous and Robust Visual SLAM
von: Du, Yanwei, et al.
Veröffentlicht: (2025)
von: Du, Yanwei, et al.
Veröffentlicht: (2025)
OmniActor: A Generalist GUI and Embodied Agent for 2D&3D Worlds
von: Yang, Longrong, et al.
Veröffentlicht: (2025)
von: Yang, Longrong, et al.
Veröffentlicht: (2025)
ProDrive: Proactive Planning for Autonomous Driving via Ego-Environment Co-Evolution
von: Fu, Chuyao, et al.
Veröffentlicht: (2026)
von: Fu, Chuyao, et al.
Veröffentlicht: (2026)
Towards Proactive Safe Human-Robot Collaborations via Data-Efficient Conditional Behavior Prediction
von: Pandya, Ravi, et al.
Veröffentlicht: (2023)
von: Pandya, Ravi, et al.
Veröffentlicht: (2023)
Genetic Informed Trees (GIT*): Path Planning via Reinforced Genetic Programming Heuristics
von: Zhang, Liding, et al.
Veröffentlicht: (2025)
von: Zhang, Liding, et al.
Veröffentlicht: (2025)
World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry
von: Liu, Yuejiang, et al.
Veröffentlicht: (2026)
von: Liu, Yuejiang, et al.
Veröffentlicht: (2026)
Online Continual Learning For Interactive Instruction Following Agents
von: Kim, Byeonghwi, et al.
Veröffentlicht: (2024)
von: Kim, Byeonghwi, et al.
Veröffentlicht: (2024)
TacMan-Turbo: Proactive Tactile Control for Robust and Efficient Articulated Object Manipulation
von: Zhao, Zihang, et al.
Veröffentlicht: (2025)
von: Zhao, Zihang, et al.
Veröffentlicht: (2025)
Update-Free On-Policy Steering via Verifiers
von: Attarian, Maria, et al.
Veröffentlicht: (2026)
von: Attarian, Maria, et al.
Veröffentlicht: (2026)
Learning Transparent Reward Models via Unsupervised Feature Selection
von: Baimukashev, Daulet, et al.
Veröffentlicht: (2024)
von: Baimukashev, Daulet, et al.
Veröffentlicht: (2024)
VLA-RFT: Vision-Language-Action Reinforcement Fine-tuning with Verified Rewards in World Simulators
von: Li, Hengtao, et al.
Veröffentlicht: (2025)
von: Li, Hengtao, et al.
Veröffentlicht: (2025)
RoboMemory: A Brain-inspired Multi-memory Agentic Framework for Interactive Environmental Learning in Physical Embodied Systems
von: Lei, Mingcong, et al.
Veröffentlicht: (2025)
von: Lei, Mingcong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
TreeCUA: Efficiently Scaling GUI Automation with Tree-Structured Verifiable Evolution
von: Jiang, Deyang, et al.
Veröffentlicht: (2026) -
MobileDreamer: Generative Sketch World Model for GUI Agent
von: Cao, Yilin, et al.
Veröffentlicht: (2026) -
ScaleTrack: Scaling and back-tracking Automated GUI Agents
von: Huang, Jing, et al.
Veröffentlicht: (2025) -
RoVer: Robot Reward Model as Test-Time Verifier for Vision-Language-Action Model
von: Dai, Mingtong, et al.
Veröffentlicht: (2025) -
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
von: Wu, Yanru, et al.
Veröffentlicht: (2026)