OmegaUse: Building a General-Purpose GUI Agent for Autonomous Task Execution
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Le, Xiao, Yixiong, Lu, Xinjiang, Cao, Jingjia, Zhao, Yusai, Zhou, Jingbo, An, Lang, Feng, Zikan, Sha, Wanxiang, Shi, Yu, Xiao, Congxi, Xiong, Jian, Zhang, Yankai, Wu, Hua, Wang, Haifeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TimeFound: A Foundation Model for Time Series Forecasting
von: Xiao, Congxi, et al.
Veröffentlicht: (2025)
von: Xiao, Congxi, et al.
Veröffentlicht: (2025)
SDWPF: A Dataset for Spatial Dynamic Wind Power Forecasting Challenge at KDD Cup 2022
von: Zhou, Jingbo, et al.
Veröffentlicht: (2022)
von: Zhou, Jingbo, et al.
Veröffentlicht: (2022)
Identifiable Representation and Model Learning for Latent Dynamic Systems
von: Zhang, Congxi, et al.
Veröffentlicht: (2024)
von: Zhang, Congxi, et al.
Veröffentlicht: (2024)
Tracking control of latent dynamic systems with application to spacecraft attitude control
von: Zhang, Congxi, et al.
Veröffentlicht: (2024)
von: Zhang, Congxi, et al.
Veröffentlicht: (2024)
AgentCPM-GUI: Building Mobile-Use Agents with Reinforcement Fine-Tuning
von: Zhang, Zhong, et al.
Veröffentlicht: (2025)
von: Zhang, Zhong, et al.
Veröffentlicht: (2025)
Executable Agentic Memory for GUI Agent
von: Qin, Zerui, et al.
Veröffentlicht: (2026)
von: Qin, Zerui, et al.
Veröffentlicht: (2026)
"It Felt Real" Victim Perspectives on Platform Design and Longer-Running Scams
von: Xiao, Jingjia, et al.
Veröffentlicht: (2025)
von: Xiao, Jingjia, et al.
Veröffentlicht: (2025)
GUI-PRA: Process Reward Agent for GUI Tasks
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
Can GenAI Move from Individual Use to Collaborative Work? Experiences, Challenges, and Opportunities of Coordinating GenAI into Collaborative Newswork
von: Xiao, Qing, et al.
Veröffentlicht: (2025)
von: Xiao, Qing, et al.
Veröffentlicht: (2025)
“Exu, a Sapucaí é vossa”: as múltiplas presenças de exu na performance do carnaval
von: Vânia Zikán Cardoso
Veröffentlicht: (2023)
von: Vânia Zikán Cardoso
Veröffentlicht: (2023)
Danger of words: risk and (mis)comprehension in consultations with the spirits of the povo da rua
von: Vânia Zikán Cardoso
Veröffentlicht: (2017)
von: Vânia Zikán Cardoso
Veröffentlicht: (2017)
Breaking the Data Barrier -- Building GUI Agents Through Task Generalization
von: Zhang, Junlei, et al.
Veröffentlicht: (2025)
von: Zhang, Junlei, et al.
Veröffentlicht: (2025)
ChartAnchor: Chart Grounding with Structural-Semantic Fidelity
von: Li, Xinhang, et al.
Veröffentlicht: (2025)
von: Li, Xinhang, et al.
Veröffentlicht: (2025)
GUI Testing Arena: A Unified Benchmark for Advancing Autonomous GUI Testing Agent
von: Zhao, Kangjia, et al.
Veröffentlicht: (2024)
von: Zhao, Kangjia, et al.
Veröffentlicht: (2024)
Building Autonomous GUI Navigation via Agentic-Q Estimation and Step-Wise Policy Optimization
von: Wang, Yibo, et al.
Veröffentlicht: (2026)
von: Wang, Yibo, et al.
Veröffentlicht: (2026)
Some ergodic theorems involving Omega function and their applications
von: Xiao, Rongzhong
Veröffentlicht: (2024)
von: Xiao, Rongzhong
Veröffentlicht: (2024)
MobileUse: A GUI Agent with Hierarchical Reflection for Autonomous Mobile Operation
von: Li, Ning, et al.
Veröffentlicht: (2025)
von: Li, Ning, et al.
Veröffentlicht: (2025)
Unrewarded Exploration in Large Language Models Reveals Latent Learning from Psychology
von: Xiong, Jian, et al.
Veröffentlicht: (2026)
von: Xiong, Jian, et al.
Veröffentlicht: (2026)
Environment Modeling for Service Robots From a Task Execution Perspective
von: Zhang, Ying, et al.
Veröffentlicht: (2025)
von: Zhang, Ying, et al.
Veröffentlicht: (2025)
GUI-explorer: Autonomous Exploration and Mining of Transition-aware Knowledge for GUI Agent
von: Xie, Bin, et al.
Veröffentlicht: (2025)
von: Xie, Bin, et al.
Veröffentlicht: (2025)
GUI-Bee: Align GUI Action Grounding to Novel Environments via Autonomous Exploration
von: Fan, Yue, et al.
Veröffentlicht: (2025)
von: Fan, Yue, et al.
Veröffentlicht: (2025)
BEVWorld: A Multimodal World Simulator for Autonomous Driving via Scene-Level BEV Latents
von: Zhang, Yumeng, et al.
Veröffentlicht: (2024)
von: Zhang, Yumeng, et al.
Veröffentlicht: (2024)
BridgeTower: Building Bridges Between Encoders in Vision-Language Representation Learning
von: Xu, Xiao, et al.
Veröffentlicht: (2022)
von: Xu, Xiao, et al.
Veröffentlicht: (2022)
ClawGUI: A Unified Framework for Training, Evaluating, and Deploying GUI Agents
von: Tang, Fei, et al.
Veröffentlicht: (2026)
von: Tang, Fei, et al.
Veröffentlicht: (2026)
ProgRM: Build Better GUI Agents with Progress Rewards
von: Zhang, Danyang, et al.
Veröffentlicht: (2025)
von: Zhang, Danyang, et al.
Veröffentlicht: (2025)
CRAFT-GUI: Curriculum-Reinforced Agent For GUI Tasks
von: Nong, Songqin, et al.
Veröffentlicht: (2025)
von: Nong, Songqin, et al.
Veröffentlicht: (2025)
The Digital Landscape of God: Narrative, Visuals and Viewer Engagement of Religious Videos on YouTube
von: Chen, Rongyi, et al.
Veröffentlicht: (2025)
von: Chen, Rongyi, et al.
Veröffentlicht: (2025)
Android in the Zoo: Chain-of-Action-Thought for GUI Agents
von: Zhang, Jiwen, et al.
Veröffentlicht: (2024)
von: Zhang, Jiwen, et al.
Veröffentlicht: (2024)
EchoTrail-GUI: Building Actionable Memory for GUI Agents via Critic-Guided Self-Exploration
von: Li, Runze, et al.
Veröffentlicht: (2025)
von: Li, Runze, et al.
Veröffentlicht: (2025)
Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction
von: Xu, Yiheng, et al.
Veröffentlicht: (2024)
von: Xu, Yiheng, et al.
Veröffentlicht: (2024)
Purposive Selection of Building Typologies
von: Onwukwe, Chukwuemeka Ozioma Stanislaus
Veröffentlicht: (2026)
von: Onwukwe, Chukwuemeka Ozioma Stanislaus
Veröffentlicht: (2026)
GoClick: Lightweight Element Grounding Model for Autonomous GUI Interaction
von: Li, Hongxin, et al.
Veröffentlicht: (2026)
von: Li, Hongxin, et al.
Veröffentlicht: (2026)
Advances in the Development of Dual‐Atom Catalysts: Structural Characterization and Diversified Catalytic Applications
von: Liping Zhang, et al.
Veröffentlicht: (2024)
von: Liping Zhang, et al.
Veröffentlicht: (2024)
Finite‐Time Control of Time‐Varying Delay Nonlinear Systems With Multiple Disturbances
von: Xinjiang Wei, et al.
Veröffentlicht: (2025)
von: Xinjiang Wei, et al.
Veröffentlicht: (2025)
Composite observer based finite time control for nonlinear systems subjecting to multiple disturbances
von: Huifeng Zhang, et al.
Veröffentlicht: (2024)
von: Huifeng Zhang, et al.
Veröffentlicht: (2024)
Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction
von: Zhang, Danyang, et al.
Veröffentlicht: (2023)
von: Zhang, Danyang, et al.
Veröffentlicht: (2023)
GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding
von: Tang, Fei, et al.
Veröffentlicht: (2025)
von: Tang, Fei, et al.
Veröffentlicht: (2025)
GUI Knowledge Bench: Revealing the Knowledge Gap of VLMs in GUI Tasks
von: Shi, Chenrui, et al.
Veröffentlicht: (2025)
von: Shi, Chenrui, et al.
Veröffentlicht: (2025)
BEAP-Agent: Backtrackable Execution and Adaptive Planning for GUI Agents
von: Lu, Ziyu, et al.
Veröffentlicht: (2026)
von: Lu, Ziyu, et al.
Veröffentlicht: (2026)
Training High-Level Schedulers with Execution-Feedback Reinforcement Learning for Long-Horizon GUI Automation
von: Deng, Zehao, et al.
Veröffentlicht: (2025)
von: Deng, Zehao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
TimeFound: A Foundation Model for Time Series Forecasting
von: Xiao, Congxi, et al.
Veröffentlicht: (2025) -
SDWPF: A Dataset for Spatial Dynamic Wind Power Forecasting Challenge at KDD Cup 2022
von: Zhou, Jingbo, et al.
Veröffentlicht: (2022) -
Identifiable Representation and Model Learning for Latent Dynamic Systems
von: Zhang, Congxi, et al.
Veröffentlicht: (2024) -
Tracking control of latent dynamic systems with application to spacecraft attitude control
von: Zhang, Congxi, et al.
Veröffentlicht: (2024) -
AgentCPM-GUI: Building Mobile-Use Agents with Reinforcement Fine-Tuning
von: Zhang, Zhong, et al.
Veröffentlicht: (2025)