WebRPG: Automatic Web Rendering Parameters Generation for Visual Presentation
Fuente:
arXiv
Saved in:
| Main Authors: | Shao, Zirui, Gao, Feiyu, Xing, Hangdi, Zhu, Zepeng, Yu, Zhi, Bu, Jiajun, Zheng, Qi, Yao, Cong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Is Cognition Consistent with Perception? Assessing and Mitigating Multimodal Knowledge Conflicts in Document Understanding
by: Shao, Zirui, et al.
Published: (2024)
by: Shao, Zirui, et al.
Published: (2024)
BrailleLLM: Braille Instruction Tuning with Large Language Models for Braille Domain Tasks
by: Huang, Tianyuan, et al.
Published: (2025)
by: Huang, Tianyuan, et al.
Published: (2025)
A Simple yet Effective Layout Token in Large Language Models for Document Understanding
by: Zhu, Zhaoqing, et al.
Published: (2025)
by: Zhu, Zhaoqing, et al.
Published: (2025)
Towards Scalable Web Accessibility Audit with MLLMs as Copilots
by: Gu, Ming, et al.
Published: (2025)
by: Gu, Ming, et al.
Published: (2025)
LORE++: Logical Location Regression Network for Table Structure Recognition with Pre-training
by: Long, Rujiao, et al.
Published: (2024)
by: Long, Rujiao, et al.
Published: (2024)
Doc-CoB: Enhancing Document Understanding with Visual Chain-of-Boxes Reasoning
by: Mo, Ye, et al.
Published: (2025)
by: Mo, Ye, et al.
Published: (2025)
Learning Only with Images: Visual Reinforcement Learning with Reasoning, Rendering, and Visual Feedback
by: Chen, Yang, et al.
Published: (2025)
by: Chen, Yang, et al.
Published: (2025)
ProcTag: Process Tagging for Assessing the Efficacy of Document Instruction Data
by: Shen, Yufan, et al.
Published: (2024)
by: Shen, Yufan, et al.
Published: (2024)
WebRenderBench: Enhancing Web Interface Generation through Layout-Style Consistency and Reinforcement Learning
by: Lai, Peichao, et al.
Published: (2025)
by: Lai, Peichao, et al.
Published: (2025)
City-on-Web: Real-time Neural Rendering of Large-scale Scenes on the Web
by: Song, Kaiwen, et al.
Published: (2023)
by: Song, Kaiwen, et al.
Published: (2023)
Automatic Generation of Web Censorship Probe Lists
by: Tang, Jenny, et al.
Published: (2024)
by: Tang, Jenny, et al.
Published: (2024)
Automatic Welding of Corrugated Steel Webs on Composite Box Girder with Corrugated Steel Webs
by: Yunfei Wu, et al.
Published: (2024)
by: Yunfei Wu, et al.
Published: (2024)
TongUI: Internet-Scale Trajectories from Multimodal Web Tutorials for Generalized GUI Agents
by: Zhang, Bofei, et al.
Published: (2025)
by: Zhang, Bofei, et al.
Published: (2025)
Visual Text Generation in the Wild
by: Zhu, Yuanzhi, et al.
Published: (2024)
by: Zhu, Yuanzhi, et al.
Published: (2024)
Web-Based Slide Presentations.
by: Just, Melissa L.
Published: (1997)
by: Just, Melissa L.
Published: (1997)
MolmoWeb: Open Visual Web Agent and Open Data for the Open Web
by: Gupta, Tanmay, et al.
Published: (2026)
by: Gupta, Tanmay, et al.
Published: (2026)
VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks
by: Koh, Jing Yu, et al.
Published: (2024)
by: Koh, Jing Yu, et al.
Published: (2024)
GuideWeb: A Benchmark for Automatic In-App Guide Generation on Real-World Web UIs
by: Gan, Chengguang, et al.
Published: (2026)
by: Gan, Chengguang, et al.
Published: (2026)
WebGen-V Bench: Structured Representation for Enhancing Visual Design in LLM-based Web Generation and Evaluation
by: Wang, Kuang-Da, et al.
Published: (2025)
by: Wang, Kuang-Da, et al.
Published: (2025)
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents
by: Yang, Rui, et al.
Published: (2026)
by: Yang, Rui, et al.
Published: (2026)
MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing
by: Zhou, Bangbang, et al.
Published: (2026)
by: Zhou, Bangbang, et al.
Published: (2026)
GUIDE: Resolving Domain Bias in GUI Agents through Real-Time Web Video Retrieval and Plug-and-Play Annotation
by: Xie, Rui, et al.
Published: (2026)
by: Xie, Rui, et al.
Published: (2026)
The Turán number of the triangular pyramid of 4-layers
by: Chen, Hangdi, et al.
Published: (2026)
by: Chen, Hangdi, et al.
Published: (2026)
Dual-View Visual Contextualization for Web Navigation
by: Kil, Jihyung, et al.
Published: (2024)
by: Kil, Jihyung, et al.
Published: (2024)
Coverage-Aware Web Crawling for Domain-Specific Supplier Discovery via a Web--Knowledge--Web Pipeline
by: Qi, Yijiashun, et al.
Published: (2026)
by: Qi, Yijiashun, et al.
Published: (2026)
Enhancing Web Agents with a Hierarchical Memory Tree
by: Tan, Yunteng, et al.
Published: (2026)
by: Tan, Yunteng, et al.
Published: (2026)
Structured Distillation of Web Agent Capabilities Enables Generalization
by: Lù, Xing Han, et al.
Published: (2026)
by: Lù, Xing Han, et al.
Published: (2026)
WebXSkill: Skill Learning for Autonomous Web Agents
by: Wang, Zhaoyang, et al.
Published: (2026)
by: Wang, Zhaoyang, et al.
Published: (2026)
AgentRewardBench: Evaluating Automatic Evaluations of Web Agent Trajectories
by: Lù, Xing Han, et al.
Published: (2025)
by: Lù, Xing Han, et al.
Published: (2025)
Web Diagrams of Cluster Variables for Grassmannian Gr(4,8)
by: Zhang, Wen Ting, et al.
Published: (2025)
by: Zhang, Wen Ting, et al.
Published: (2025)
WebInject: Prompt Injection Attack to Web Agents
by: Wang, Xilong, et al.
Published: (2025)
by: Wang, Xilong, et al.
Published: (2025)
REAL: Resolving Knowledge Conflicts in Knowledge-Intensive Visual Question Answering via Reasoning-Pivot Alignment
by: Ye, Kai, et al.
Published: (2026)
by: Ye, Kai, et al.
Published: (2026)
OphIn-500K: Curating Web-Scale Visual Instructions for Scaling Ophthalmic Multimodal Large Language Models
by: Dong, Xuanzhao, et al.
Published: (2026)
by: Dong, Xuanzhao, et al.
Published: (2026)
Constructing a Spider‐Web Polymer Blocking Layer on Separator for the High‐Loading Li‐S Battery
by: Qian Zhang, et al.
Published: (2024)
by: Qian Zhang, et al.
Published: (2024)
WebGuard: Building a Generalizable Guardrail for Web Agents
by: Zheng, Boyuan, et al.
Published: (2025)
by: Zheng, Boyuan, et al.
Published: (2025)
WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
by: He, Hongliang, et al.
Published: (2024)
by: He, Hongliang, et al.
Published: (2024)
Selecting a Web 2.0 Presentation Tool
by: Hodges, Charles B., et al.
Published: (2011)
by: Hodges, Charles B., et al.
Published: (2011)
RPG: A Repository Planning Graph for Unified and Scalable Codebase Generation
by: Luo, Jane, et al.
Published: (2025)
by: Luo, Jane, et al.
Published: (2025)
AI-Assisted Adaptive Rendering for High-Frequency Security Telemetry in Web Interfaces
by: Rajhans, Mona
Published: (2026)
by: Rajhans, Mona
Published: (2026)
WebGym: Scaling Training Environments for Visual Web Agents with Realistic Tasks
by: Bai, Hao, et al.
Published: (2026)
by: Bai, Hao, et al.
Published: (2026)
Similar Items
-
Is Cognition Consistent with Perception? Assessing and Mitigating Multimodal Knowledge Conflicts in Document Understanding
by: Shao, Zirui, et al.
Published: (2024) -
BrailleLLM: Braille Instruction Tuning with Large Language Models for Braille Domain Tasks
by: Huang, Tianyuan, et al.
Published: (2025) -
A Simple yet Effective Layout Token in Large Language Models for Document Understanding
by: Zhu, Zhaoqing, et al.
Published: (2025) -
Towards Scalable Web Accessibility Audit with MLLMs as Copilots
by: Gu, Ming, et al.
Published: (2025) -
LORE++: Logical Location Regression Network for Table Structure Recognition with Pre-training
by: Long, Rujiao, et al.
Published: (2024)