UI-Venus Technical Report: Building High-performance UI Agents with RFT
Fuente:
arXiv
Saved in:
| Main Authors: | Gu, Zhangxuan, Zeng, Zhengwen, Xu, Zhenyu, Zhou, Xingran, Shen, Shuheng, Liu, Yunfei, Zhou, Beitong, Meng, Changhua, Xia, Tianyu, Chen, Weizhi, Wen, Yue, Dou, Jingya, Tang, Fei, Lin, Jinzhen, Liu, Yulin, Guo, Zhenlin, Gong, Yichen, Jia, Heng, Gao, Changlong, Guo, Yuan, Deng, Yong, Guo, Zhenyu, Chen, Liang, Wang, Weiqiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UI-Venus-1.5 Technical Report
by: Venus Team, et al.
Published: (2026)
by: Venus Team, et al.
Published: (2026)
VenusBench-GD: A Comprehensive Multi-Platform GUI Benchmark for Diverse Grounding Tasks
by: Zhou, Beitong, et al.
Published: (2025)
by: Zhou, Beitong, et al.
Published: (2025)
VenusBench-Mobile: A Challenging and User-Centric Benchmark for Mobile GUI Agents with Capability Diagnostics
by: Gong, Yichen, et al.
Published: (2026)
by: Gong, Yichen, et al.
Published: (2026)
UI-UG: A Unified MLLM for UI Understanding and Generation
by: Yang, Hao, et al.
Published: (2025)
by: Yang, Hao, et al.
Published: (2025)
E-ANT: A Large-Scale Dataset for Efficient Automatic GUI NavigaTion
by: Wang, Ke, et al.
Published: (2024)
by: Wang, Ke, et al.
Published: (2024)
MAIC-UI: Making Interactive Courseware with Generative UI
by: Tu, Shangqing, et al.
Published: (2026)
by: Tu, Shangqing, et al.
Published: (2026)
Beyond Screenshots: Evaluating VLMs' Understanding of UI Animations
by: Liang, Chen, et al.
Published: (2026)
by: Liang, Chen, et al.
Published: (2026)
ConsentDiff at Scale: Longitudinal Audits of Web Privacy Policy Changes and UI Frictions
by: Guo, Haoze
Published: (2025)
by: Guo, Haoze
Published: (2025)
Generative UI: LLMs are Effective UI Generators
by: Leviathan, Yaniv, et al.
Published: (2026)
by: Leviathan, Yaniv, et al.
Published: (2026)
Macaron-A2UI: A Model for Generative UI in Personal Agents
by: Kong, Fancy, et al.
Published: (2026)
by: Kong, Fancy, et al.
Published: (2026)
CrowdGenUI: Aligning LLM-Based UI Generation with Crowdsourced User Preferences
by: Liu, Yimeng, et al.
Published: (2024)
by: Liu, Yimeng, et al.
Published: (2024)
ShowUI-Aloha: Human-Taught GUI Agent
by: Zhang, Yichun, et al.
Published: (2026)
by: Zhang, Yichun, et al.
Published: (2026)
SpecifyUI: Supporting Iterative UI Design Intent Expression through Structured Specifications and Generative AI
by: Chen, Yunnong, et al.
Published: (2025)
by: Chen, Yunnong, et al.
Published: (2025)
GhostUI: Unveiling Hidden Interactions in Mobile UI
by: Kweon, Minkyu, et al.
Published: (2026)
by: Kweon, Minkyu, et al.
Published: (2026)
RWKV-UI: UI Understanding with Enhanced Perception and Reasoning
by: Yang, Jiaxi, et al.
Published: (2025)
by: Yang, Jiaxi, et al.
Published: (2025)
The GenUI Study: Exploring the Design of Generative UI Tools to Support UX Practitioners and Beyond
by: Chen, Xiang 'Anthony', et al.
Published: (2025)
by: Chen, Xiang 'Anthony', et al.
Published: (2025)
ReDemon UI: Reactive Synthesis by Demonstration for Web UI
by: Lee, Jay, et al.
Published: (2025)
by: Lee, Jay, et al.
Published: (2025)
Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMs
by: You, Keen, et al.
Published: (2024)
by: You, Keen, et al.
Published: (2024)
UI Semantic Group Detection: Grouping UI Elements with Similar Semantics in Mobile Graphical User Interface
by: Xiao, Shuhong, et al.
Published: (2024)
by: Xiao, Shuhong, et al.
Published: (2024)
MAI-UI Technical Report: Real-World Centric Foundation GUI Agents
by: Zhou, Hanzhang, et al.
Published: (2025)
by: Zhou, Hanzhang, et al.
Published: (2025)
wcraytor/glmnetUI: glmnetUI: User Interface for R/glmnet()
by: William Bert Craytor
Published: (2026)
by: William Bert Craytor
Published: (2026)
UI2V-Bench: An Understanding-based Image-to-video Generation Benchmark
by: Zhang, Ailing, et al.
Published: (2025)
by: Zhang, Ailing, et al.
Published: (2025)
MUD: Towards a Large-Scale and Noise-Filtered UI Dataset for Modern Style UI Modeling
by: Feng, Sidong, et al.
Published: (2024)
by: Feng, Sidong, et al.
Published: (2024)
UI-Ins: Enhancing GUI Grounding with Multi-Perspective Instruction-as-Reasoning
by: Chen, Liangyu, et al.
Published: (2025)
by: Chen, Liangyu, et al.
Published: (2025)
Decoupling Knowledge and Reasoning in Transformers: A Modular Architecture with Generalized Cross-Attention
by: Guo, Zhenyu, et al.
Published: (2025)
by: Guo, Zhenyu, et al.
Published: (2025)
Bridging Design and Development with Automated Declarative UI Code Generation
by: Zhou, Ting, et al.
Published: (2024)
by: Zhou, Ting, et al.
Published: (2024)
ReFinE: Streamlining UI Mockup Iteration with Research Findings
by: Shin, Donghoon, et al.
Published: (2026)
by: Shin, Donghoon, et al.
Published: (2026)
MLLM-Based UI2Code Automation Guided by UI Layout Information
by: Wu, Fan, et al.
Published: (2025)
by: Wu, Fan, et al.
Published: (2025)
Privacy Starts with UI: Privacy Patterns and Designer Perspectives in UI/UX Practice
by: Maloku, Anxhela, et al.
Published: (2026)
by: Maloku, Anxhela, et al.
Published: (2026)
UI Remix: Supporting UI Design Through Interactive Example Retrieval and Remixing
by: Wang, Junling, et al.
Published: (2026)
by: Wang, Junling, et al.
Published: (2026)
UI2Code^N: UI-to-Code Generation as Interactive Visual Optimization
by: Yang, Zhen, et al.
Published: (2025)
by: Yang, Zhen, et al.
Published: (2025)
DR-Venus: Towards Frontier Edge-Scale Deep Research Agents with Only 10K Open Data
by: Venus Team, et al.
Published: (2026)
by: Venus Team, et al.
Published: (2026)
Falcon-UI: Understanding GUI Before Following User Instructions
by: Shen, Huawen, et al.
Published: (2024)
by: Shen, Huawen, et al.
Published: (2024)
Magentic-UI: Towards Human-in-the-loop Agentic Systems
by: Mozannar, Hussein, et al.
Published: (2025)
by: Mozannar, Hussein, et al.
Published: (2025)
MVP: Multiple View Prediction Improves GUI Grounding
by: Zhang, Yunzhu, et al.
Published: (2025)
by: Zhang, Yunzhu, et al.
Published: (2025)
VSA:Visual-Structural Alignment for UI-to-Code
by: Wu, Xian, et al.
Published: (2025)
by: Wu, Xian, et al.
Published: (2025)
Predicting and Explaining Mobile UI Tappability with Vision Modeling and Saliency Analysis
by: Schoop, Eldon, et al.
Published: (2022)
by: Schoop, Eldon, et al.
Published: (2022)
FocusUI: Efficient UI Grounding via Position-Preserving Visual Token Selection
by: Ouyang, Mingyu, et al.
Published: (2026)
by: Ouyang, Mingyu, et al.
Published: (2026)
UI-Evol: Automatic Knowledge Evolving for Computer Use Agents
by: Zhang, Ziyun, et al.
Published: (2025)
by: Zhang, Ziyun, et al.
Published: (2025)
UI as Observer–Platform Physics
by: Ainstein
Published: (2026)
by: Ainstein
Published: (2026)
Similar Items
-
UI-Venus-1.5 Technical Report
by: Venus Team, et al.
Published: (2026) -
VenusBench-GD: A Comprehensive Multi-Platform GUI Benchmark for Diverse Grounding Tasks
by: Zhou, Beitong, et al.
Published: (2025) -
VenusBench-Mobile: A Challenging and User-Centric Benchmark for Mobile GUI Agents with Capability Diagnostics
by: Gong, Yichen, et al.
Published: (2026) -
UI-UG: A Unified MLLM for UI Understanding and Generation
by: Yang, Hao, et al.
Published: (2025) -
E-ANT: A Large-Scale Dataset for Efficient Automatic GUI NavigaTion
by: Wang, Ke, et al.
Published: (2024)