VenusBench-GD: A Comprehensive Multi-Platform GUI Benchmark for Diverse Grounding Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Beitong, Huang, Zhexiao, Guo, Yuan, Gu, Zhangxuan, Xia, Tianyu, Luo, Zichen, Tang, Fei, Kong, Dehan, Shang, Yanyi, Ou, Suling, Guo, Zhenlin, Meng, Changhua, Shen, Shuheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VenusBench-Mobile: A Challenging and User-Centric Benchmark for Mobile GUI Agents with Capability Diagnostics
by: Gong, Yichen, et al.
Published: (2026)
by: Gong, Yichen, et al.
Published: (2026)
E-ANT: A Large-Scale Dataset for Efficient Automatic GUI NavigaTion
by: Wang, Ke, et al.
Published: (2024)
by: Wang, Ke, et al.
Published: (2024)
GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding
by: Tang, Fei, et al.
Published: (2025)
by: Tang, Fei, et al.
Published: (2025)
UI-Venus Technical Report: Building High-performance UI Agents with RFT
by: Gu, Zhangxuan, et al.
Published: (2025)
by: Gu, Zhangxuan, et al.
Published: (2025)
UI-Venus-1.5 Technical Report
by: Venus Team, et al.
Published: (2026)
by: Venus Team, et al.
Published: (2026)
MVP: Multiple View Prediction Improves GUI Grounding
by: Zhang, Yunzhu, et al.
Published: (2025)
by: Zhang, Yunzhu, et al.
Published: (2025)
WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents
by: Fan, Sicheng, et al.
Published: (2026)
by: Fan, Sicheng, et al.
Published: (2026)
FDM-Bench: A Comprehensive Benchmark for Evaluating Large Language Models in Additive Manufacturing Tasks
by: Eslaminia, Ahmadreza, et al.
Published: (2024)
by: Eslaminia, Ahmadreza, et al.
Published: (2024)
WebChain: A Large-Scale Human-Annotated Dataset of Real-World Web Interaction Traces
by: Fan, Sicheng, et al.
Published: (2026)
by: Fan, Sicheng, et al.
Published: (2026)
OS-Oracle: A Comprehensive Framework for Cross-Platform GUI Critic Models
by: Wu, Zhenyu, et al.
Published: (2025)
by: Wu, Zhenyu, et al.
Published: (2025)
GUI Knowledge Bench: Revealing the Knowledge Gap of VLMs in GUI Tasks
by: Shi, Chenrui, et al.
Published: (2025)
by: Shi, Chenrui, et al.
Published: (2025)
Universal Retrieval for Multimodal Trajectory Modeling
by: Zhang, Xuan, et al.
Published: (2025)
by: Zhang, Xuan, et al.
Published: (2025)
PC$^2$: Pseudo-Classification Based Pseudo-Captioning for Noisy Correspondence Learning in Cross-Modal Retrieval
by: Duan, Yue, et al.
Published: (2024)
by: Duan, Yue, et al.
Published: (2024)
X-IONet: Cross-Platform Inertial Odometry Network for Pedestrian and Legged Robot
by: Shen, Dehan, et al.
Published: (2025)
by: Shen, Dehan, et al.
Published: (2025)
MMBench-GUI: Hierarchical Multi-Platform Evaluation Framework for GUI Agents
by: Wang, Xuehui, et al.
Published: (2025)
by: Wang, Xuehui, et al.
Published: (2025)
Improving GUI Grounding with Explicit Position-to-Coordinate Mapping
by: Wang, Suyuchen, et al.
Published: (2025)
by: Wang, Suyuchen, et al.
Published: (2025)
Phi-Ground Tech Report: Advancing Perception in GUI Grounding
by: Zhang, Miaosen, et al.
Published: (2025)
by: Zhang, Miaosen, et al.
Published: (2025)
Breaking the Data Barrier -- Building GUI Agents Through Task Generalization
by: Zhang, Junlei, et al.
Published: (2025)
by: Zhang, Junlei, et al.
Published: (2025)
AutoFocus: Uncertainty-Aware Active Visual Search for GUI Grounding
by: Yao, Ruilin, et al.
Published: (2026)
by: Yao, Ruilin, et al.
Published: (2026)
POINTS-GUI-G: GUI-Grounding Journey
by: Zhao, Zhongyin, et al.
Published: (2026)
by: Zhao, Zhongyin, et al.
Published: (2026)
GUI-Shepherd: Reliable Process Reward and Verification for Long-Sequence GUI Tasks
by: Chen, Cong, et al.
Published: (2025)
by: Chen, Cong, et al.
Published: (2025)
GroundingBooth: Grounding Text-to-Image Customization
by: Xiong, Zhexiao, et al.
Published: (2024)
by: Xiong, Zhexiao, et al.
Published: (2024)
A Unified, Cross-Platform Framework for Automatic GUI and Plugin Generation in Structural Bioinformatics and Beyond
by: Guo, Sikao, et al.
Published: (2026)
by: Guo, Sikao, et al.
Published: (2026)
META-GUI: Towards Multi-modal Conversational Agents on Mobile GUI
by: Sun, Liangtai, et al.
Published: (2022)
by: Sun, Liangtai, et al.
Published: (2022)
FineState-Bench: Benchmarking State-Conditioned Grounding for Fine-grained GUI State Setting
by: Ji, Fengxian, et al.
Published: (2026)
by: Ji, Fengxian, et al.
Published: (2026)
Convergence and Optimality of the EM Algorithm Under Multi-Component Gaussian Mixture Models
by: Bing, Xin, et al.
Published: (2025)
by: Bing, Xin, et al.
Published: (2025)
The synthetic instrument: From sparse association to sparse causation
by: Tang, Dingke, et al.
Published: (2023)
by: Tang, Dingke, et al.
Published: (2023)
Simultaneous hypothesis testing for comparing many functional means
by: Decker, Colin, et al.
Published: (2025)
by: Decker, Colin, et al.
Published: (2025)
Successive classification learning for estimating quantile optimal treatment regimes
by: Xia, Junwen, et al.
Published: (2025)
by: Xia, Junwen, et al.
Published: (2025)
Towards R-learner with Continuous Treatments
by: Zhang, Yichi, et al.
Published: (2022)
by: Zhang, Yichi, et al.
Published: (2022)
Simultaneous Estimation of Multiple Treatment Effects from Observational Studies
by: Shi, Xiaochuan, et al.
Published: (2025)
by: Shi, Xiaochuan, et al.
Published: (2025)
Functional principal component analysis with informative observation times
by: Sang, Peijun, et al.
Published: (2022)
by: Sang, Peijun, et al.
Published: (2022)
WebCanvas: Benchmarking Web Agents in Online Environments
by: Pan, Yichen, et al.
Published: (2024)
by: Pan, Yichen, et al.
Published: (2024)
FineState-Bench: A Comprehensive Benchmark for Fine-Grained State Control in GUI Agents
by: Ji, Fengxian, et al.
Published: (2025)
by: Ji, Fengxian, et al.
Published: (2025)
From Perception to Symbolic Task Planning: Vision-Language Guided Human-Robot Collaborative Structured Assembly
by: Chen, Yanyi, et al.
Published: (2026)
by: Chen, Yanyi, et al.
Published: (2026)
GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents
by: Wu, Qianhui, et al.
Published: (2025)
by: Wu, Qianhui, et al.
Published: (2025)
MemGUI-Bench: Benchmarking Memory of Mobile GUI Agents in Dynamic Environments
by: Liu, Guangyi, et al.
Published: (2026)
by: Liu, Guangyi, et al.
Published: (2026)
Semantics-Guided Moving Object Segmentation with 3D LiDAR
by: Gu, Shuo, et al.
Published: (2022)
by: Gu, Shuo, et al.
Published: (2022)
CRAFT-GUI: Curriculum-Reinforced Agent For GUI Tasks
by: Nong, Songqin, et al.
Published: (2025)
by: Nong, Songqin, et al.
Published: (2025)
GUI-PRA: Process Reward Agent for GUI Tasks
by: Xiong, Tao, et al.
Published: (2025)
by: Xiong, Tao, et al.
Published: (2025)
Similar Items
-
VenusBench-Mobile: A Challenging and User-Centric Benchmark for Mobile GUI Agents with Capability Diagnostics
by: Gong, Yichen, et al.
Published: (2026) -
E-ANT: A Large-Scale Dataset for Efficient Automatic GUI NavigaTion
by: Wang, Ke, et al.
Published: (2024) -
GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding
by: Tang, Fei, et al.
Published: (2025) -
UI-Venus Technical Report: Building High-performance UI Agents with RFT
by: Gu, Zhangxuan, et al.
Published: (2025) -
UI-Venus-1.5 Technical Report
by: Venus Team, et al.
Published: (2026)