OS-Themis: A Scalable Critic Framework for Generalist GUI Rewards
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Zehao, Wu, Zhenyu, Zhao, Yibo, Yang, Bowen, Xie, Jingjing, Liu, Zhaoyang, Liu, Zhoumianze, Jin, Kaiming, Liang, Jianze, Li, Zonglin, Wu, Feng, Zhou, Bowen, Wang, Zun, Ding, Zichen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OS-Oracle: A Comprehensive Framework for Cross-Platform GUI Critic Models
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025)
OS-Symphony: A Holistic Framework for Robust and Generalist Computer-Using Agent
von: Yang, Bowen, et al.
Veröffentlicht: (2026)
von: Yang, Bowen, et al.
Veröffentlicht: (2026)
OS-Copilot: Towards Generalist Computer Agents with Self-Improvement
von: Wu, Zhiyong, et al.
Veröffentlicht: (2024)
von: Wu, Zhiyong, et al.
Veröffentlicht: (2024)
Retrieval, Reward, and Training Protocols: What Matters in Training Search Agents?
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows
von: Sun, Qiushi, et al.
Veröffentlicht: (2025)
von: Sun, Qiushi, et al.
Veröffentlicht: (2025)
OS-ATLAS: A Foundation Action Model for Generalist GUI Agents
von: Wu, Zhiyong, et al.
Veröffentlicht: (2024)
von: Wu, Zhiyong, et al.
Veröffentlicht: (2024)
OS-Genesis: Automating GUI Agent Trajectory Construction via Reverse Task Synthesis
von: Sun, Qiushi, et al.
Veröffentlicht: (2024)
von: Sun, Qiushi, et al.
Veröffentlicht: (2024)
MMBench-GUI: Hierarchical Multi-Platform Evaluation Framework for GUI Agents
von: Wang, Xuehui, et al.
Veröffentlicht: (2025)
von: Wang, Xuehui, et al.
Veröffentlicht: (2025)
Inference-Time Scaling for Generalist Reward Modeling
von: Liu, Zijun, et al.
Veröffentlicht: (2025)
von: Liu, Zijun, et al.
Veröffentlicht: (2025)
InfiGUIAgent: A Multimodal Generalist GUI Agent with Native Reasoning and Reflection
von: Liu, Yuhang, et al.
Veröffentlicht: (2025)
von: Liu, Yuhang, et al.
Veröffentlicht: (2025)
GUI-PRA: Process Reward Agent for GUI Tasks
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
Thēmis
Veröffentlicht: (2018)
Veröffentlicht: (2018)
DocOS: Towards Proactive Document-Guided Actions in GUI Agents
von: Liu, Jingjing, et al.
Veröffentlicht: (2026)
von: Liu, Jingjing, et al.
Veröffentlicht: (2026)
GUI-GENESIS: Automated Synthesis of Efficient Environments with Verifiable Rewards for GUI Agent Post-Training
von: Cao, Yuan, et al.
Veröffentlicht: (2026)
von: Cao, Yuan, et al.
Veröffentlicht: (2026)
Omni-Reward: Towards Generalist Omni-Modal Reward Modeling with Free-Form Preferences
von: Jin, Zhuoran, et al.
Veröffentlicht: (2025)
von: Jin, Zhuoran, et al.
Veröffentlicht: (2025)
WavReward: Spoken Dialogue Models With Generalist Reward Evaluators
von: Ji, Shengpeng, et al.
Veröffentlicht: (2025)
von: Ji, Shengpeng, et al.
Veröffentlicht: (2025)
Clio@Themis
Veröffentlicht: (2022)
Veröffentlicht: (2022)
UMG-CLIP: A Unified Multi-Granularity Vision Generalist for Open-World Understanding
von: Shi, Bowen, et al.
Veröffentlicht: (2024)
von: Shi, Bowen, et al.
Veröffentlicht: (2024)
OS-Kairos: Adaptive Interaction for MLLM-Powered GUI Agents
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2025)
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2025)
Relative Stability Conditions on Triangulated Categories
von: Liu, Bowen, et al.
Veröffentlicht: (2024)
von: Liu, Bowen, et al.
Veröffentlicht: (2024)
Confidence as a Reward: Transforming LLMs into Reward Models
von: Du, He, et al.
Veröffentlicht: (2025)
von: Du, He, et al.
Veröffentlicht: (2025)
Training High-Level Schedulers with Execution-Feedback Reinforcement Learning for Long-Horizon GUI Automation
von: Deng, Zehao, et al.
Veröffentlicht: (2025)
von: Deng, Zehao, et al.
Veröffentlicht: (2025)
Convergence analysis of the transformed gradient projection algorithms on compact matrix manifolds
von: Ding, Wentao, et al.
Veröffentlicht: (2024)
von: Ding, Wentao, et al.
Veröffentlicht: (2024)
Structural Reward Model: Enhancing Interpretability, Efficiency, and Scalability in Reward Modeling
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2025)
ProgRM: Build Better GUI Agents with Progress Rewards
von: Zhang, Danyang, et al.
Veröffentlicht: (2025)
von: Zhang, Danyang, et al.
Veröffentlicht: (2025)
Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining
von: Xiong, Weimin, et al.
Veröffentlicht: (2026)
von: Xiong, Weimin, et al.
Veröffentlicht: (2026)
Scalable Binary CUR Low-Rank Approximation Algorithm
von: Su, Bowen
Veröffentlicht: (2025)
von: Su, Bowen
Veröffentlicht: (2025)
Towards Building Specialized Generalist AI with System 1 and System 2 Fusion
von: Zhang, Kaiyan, et al.
Veröffentlicht: (2024)
von: Zhang, Kaiyan, et al.
Veröffentlicht: (2024)
Advancing LLM Reasoning Generalists with Preference Trees
von: Yuan, Lifan, et al.
Veröffentlicht: (2024)
von: Yuan, Lifan, et al.
Veröffentlicht: (2024)
GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents
von: Luo, Run, et al.
Veröffentlicht: (2025)
von: Luo, Run, et al.
Veröffentlicht: (2025)
Themis: Training Robust Multilingual Code Reward Models for Flexible Multi-Criteria Scoring
von: Paul, Indraneil, et al.
Veröffentlicht: (2026)
von: Paul, Indraneil, et al.
Veröffentlicht: (2026)
ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows
von: Sun, Qiushi, et al.
Veröffentlicht: (2025)
von: Sun, Qiushi, et al.
Veröffentlicht: (2025)
Semi-Unsupervised Microscopy Segmentation with Fuzzy Logic and Spatial Statistics for Cross-Domain Analysis Using a GUI
von: Das, Surajit, et al.
Veröffentlicht: (2025)
von: Das, Surajit, et al.
Veröffentlicht: (2025)
GUI-Shepherd: Reliable Process Reward and Verification for Long-Sequence GUI Tasks
von: Chen, Cong, et al.
Veröffentlicht: (2025)
von: Chen, Cong, et al.
Veröffentlicht: (2025)
Themis: Automatic and Efficient Deep Learning System Testing with Strong Fault Detection Capability
von: Huang, Dong, et al.
Veröffentlicht: (2024)
von: Huang, Dong, et al.
Veröffentlicht: (2024)
Scaling the Scaling Logic: Agentic Meta-Synthesis of Logic Reasoning
von: Liu, Bowen, et al.
Veröffentlicht: (2026)
von: Liu, Bowen, et al.
Veröffentlicht: (2026)
GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding
von: Tang, Fei, et al.
Veröffentlicht: (2025)
von: Tang, Fei, et al.
Veröffentlicht: (2025)
EchoTrail-GUI: Building Actionable Memory for GUI Agents via Critic-Guided Self-Exploration
von: Li, Runze, et al.
Veröffentlicht: (2025)
von: Li, Runze, et al.
Veröffentlicht: (2025)
Los palacios de Themis
von: Hugo Arciniega
Veröffentlicht: (2000)
von: Hugo Arciniega
Veröffentlicht: (2000)
GRACE: Reinforcement Learning for Grounded Response and Abstention under Contextual Evidence
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
OS-Oracle: A Comprehensive Framework for Cross-Platform GUI Critic Models
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025) -
OS-Symphony: A Holistic Framework for Robust and Generalist Computer-Using Agent
von: Yang, Bowen, et al.
Veröffentlicht: (2026) -
OS-Copilot: Towards Generalist Computer Agents with Self-Improvement
von: Wu, Zhiyong, et al.
Veröffentlicht: (2024) -
Retrieval, Reward, and Training Protocols: What Matters in Training Search Agents?
von: Zhao, Yibo, et al.
Veröffentlicht: (2026) -
OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows
von: Sun, Qiushi, et al.
Veröffentlicht: (2025)