Code2World: A GUI World Model via Renderable Code Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zheng, Yuhao, Zhong, Li'an, Wang, Yi, Dai, Rui, Liu, Kaikui, Chu, Xiangxiang, Lv, Linyuan, Torr, Philip, Lin, Kevin Qinghong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Code2Video: A Code-centric Paradigm for Educational Video Generation
di: Chen, Yanzhe, et al.
Pubblicazione: (2025)
di: Chen, Yanzhe, et al.
Pubblicazione: (2025)
ShowUI-$π$: Flow-based Generative Models as GUI Dexterous Hands
di: Hu, Siyuan, et al.
Pubblicazione: (2025)
di: Hu, Siyuan, et al.
Pubblicazione: (2025)
Computer-Use Agents as Judges for Generative User Interface
di: Lin, Kevin Qinghong, et al.
Pubblicazione: (2025)
di: Lin, Kevin Qinghong, et al.
Pubblicazione: (2025)
GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents
di: Ouyang, Mingyu, et al.
Pubblicazione: (2026)
di: Ouyang, Mingyu, et al.
Pubblicazione: (2026)
GUIDE: A Benchmark for Understanding and Assisting Users in Open-Ended GUI Tasks
di: Yang, Saelyne, et al.
Pubblicazione: (2026)
di: Yang, Saelyne, et al.
Pubblicazione: (2026)
Coding with Eyes: Visual Feedback Unlocks Reliable GUI Code Generating and Debugging
di: Liu, Zhilin, et al.
Pubblicazione: (2026)
di: Liu, Zhilin, et al.
Pubblicazione: (2026)
"Newspaper Eat" Means "Not Tasty": A Taxonomy and Benchmark for Coded Language in Real-World Chinese Online Reviews
di: Wan, Ruyuan, et al.
Pubblicazione: (2026)
di: Wan, Ruyuan, et al.
Pubblicazione: (2026)
ShowUI: One Vision-Language-Action Model for GUI Visual Agent
di: Lin, Kevin Qinghong, et al.
Pubblicazione: (2024)
di: Lin, Kevin Qinghong, et al.
Pubblicazione: (2024)
World Craft: Agentic Framework to Create Visualizable Worlds via Text
di: Sun, Jianwen, et al.
Pubblicazione: (2026)
di: Sun, Jianwen, et al.
Pubblicazione: (2026)
Sketch Then Generate: Providing Incremental User Feedback and Guiding LLM Code Generation through Language-Oriented Code Sketches
di: Zhu-Tian, Chen, et al.
Pubblicazione: (2024)
di: Zhu-Tian, Chen, et al.
Pubblicazione: (2024)
Assessing Photorealism of Rendered Objects in Real-World Images: A Transparent and Reproducible User Study
di: Kluge, Sven, et al.
Pubblicazione: (2024)
di: Kluge, Sven, et al.
Pubblicazione: (2024)
ViMo: A Generative Visual GUI World Model for App Agents
di: Luo, Dezhao, et al.
Pubblicazione: (2025)
di: Luo, Dezhao, et al.
Pubblicazione: (2025)
Ivie: Lightweight Anchored Explanations of Just-Generated Code
di: Yan, Litao, et al.
Pubblicazione: (2024)
di: Yan, Litao, et al.
Pubblicazione: (2024)
Hedwig: Dynamic Autonomy for Coding Agents Under Local Oversight
di: Shukla, Tanjal, et al.
Pubblicazione: (2026)
di: Shukla, Tanjal, et al.
Pubblicazione: (2026)
SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents
di: Cheng, Kanzhi, et al.
Pubblicazione: (2024)
di: Cheng, Kanzhi, et al.
Pubblicazione: (2024)
WinClick: GUI Grounding with Multimodal Large Language Models
di: Hui, Zheng, et al.
Pubblicazione: (2025)
di: Hui, Zheng, et al.
Pubblicazione: (2025)
Code Semantic Zooming
di: Ba, Jinsheng, et al.
Pubblicazione: (2025)
di: Ba, Jinsheng, et al.
Pubblicazione: (2025)
From Standard English to Singlish: A Retrieval-Augmented Approach for Code-Switched Creole Generation in Large Language Models
di: Lai, Foong Ming, et al.
Pubblicazione: (2026)
di: Lai, Foong Ming, et al.
Pubblicazione: (2026)
HICode: Hierarchical Inductive Coding with LLMs
di: Zhong, Mian, et al.
Pubblicazione: (2025)
di: Zhong, Mian, et al.
Pubblicazione: (2025)
Code Shaping: Iterative Code Editing with Free-form AI-Interpreted Sketching
di: Yen, Ryan, et al.
Pubblicazione: (2025)
di: Yen, Ryan, et al.
Pubblicazione: (2025)
Characterizing Unintended Consequences in Human-GUI Agent Collaboration for Web Browsing
di: Zhang, Shuning, et al.
Pubblicazione: (2025)
di: Zhang, Shuning, et al.
Pubblicazione: (2025)
PlotGen-Bench: Evaluating VLMs on Generating Visualization Code from Diverse Plots across Multiple Libraries
di: Zhao, Yi, et al.
Pubblicazione: (2025)
di: Zhao, Yi, et al.
Pubblicazione: (2025)
Aria-UI: Visual Grounding for GUI Instructions
di: Yang, Yuhao, et al.
Pubblicazione: (2024)
di: Yang, Yuhao, et al.
Pubblicazione: (2024)
Understanding and Mitigating Harmful Design in User-Generated Virtual Worlds
di: Zhang, Zinan, et al.
Pubblicazione: (2024)
di: Zhang, Zinan, et al.
Pubblicazione: (2024)
Blending the Worlds: World-Fixed Visual Appearances in Automotive Augmented Reality
di: Schramm, Robin Connor, et al.
Pubblicazione: (2025)
di: Schramm, Robin Connor, et al.
Pubblicazione: (2025)
InterLink: Linking Text with Code and Output in Computational Notebooks
di: Lin, Yanna, et al.
Pubblicazione: (2025)
di: Lin, Yanna, et al.
Pubblicazione: (2025)
Beyond Code Generation: LLM-supported Exploration of the Program Design Space
di: Zamfirescu-Pereira, J. D., et al.
Pubblicazione: (2025)
di: Zamfirescu-Pereira, J. D., et al.
Pubblicazione: (2025)
IntelliExplain: Enhancing Conversational Code Generation for Non-Professional Programmers
di: Yan, Hao, et al.
Pubblicazione: (2024)
di: Yan, Hao, et al.
Pubblicazione: (2024)
SparkUI-Parser: Enhancing GUI Perception with Robust Grounding and Parsing
di: Jing, Hongyi, et al.
Pubblicazione: (2025)
di: Jing, Hongyi, et al.
Pubblicazione: (2025)
Comparing Visual Metaphors with Textual Code For Learning Basic Computer Science Concepts in Virtual Reality
di: Baron, Kevin William
Pubblicazione: (2024)
di: Baron, Kevin William
Pubblicazione: (2024)
Vibe Coding, Interface Flattening
di: Jin, Hongrui
Pubblicazione: (2025)
di: Jin, Hongrui
Pubblicazione: (2025)
Case Study of GAI for Generating Novel Images for Real-World Embroidery
di: Glazko, Kate, et al.
Pubblicazione: (2025)
di: Glazko, Kate, et al.
Pubblicazione: (2025)
No General Code of Ethics for All: Ethical Considerations in Human-bot Psycho-counseling
di: Ma, Lizhi, et al.
Pubblicazione: (2024)
di: Ma, Lizhi, et al.
Pubblicazione: (2024)
Think, Act, and Ask: Open-World Interactive Personalized Robot Navigation
di: Dai, Yinpei, et al.
Pubblicazione: (2023)
di: Dai, Yinpei, et al.
Pubblicazione: (2023)
LogoMotion: Visually-Grounded Code Synthesis for Creating and Editing Animation
di: Liu, Vivian, et al.
Pubblicazione: (2024)
di: Liu, Vivian, et al.
Pubblicazione: (2024)
Envisioning Generative Artificial Intelligence in Cartography and Mapmaking
di: Kang, Yuhao, et al.
Pubblicazione: (2025)
di: Kang, Yuhao, et al.
Pubblicazione: (2025)
Layout2Rendering: AI-aided Greenspace design
di: Chen, Ran, et al.
Pubblicazione: (2024)
di: Chen, Ran, et al.
Pubblicazione: (2024)
CinemaWorld: Generative Augmented Reality with LLMs and 3D Scene Generation for Movie Augmentation
di: Ihara, Keiichi, et al.
Pubblicazione: (2026)
di: Ihara, Keiichi, et al.
Pubblicazione: (2026)
OmniGUI: Benchmarking GUI Agents in Omni-Modal Smartphone Environments
di: Henry, Felix, et al.
Pubblicazione: (2026)
di: Henry, Felix, et al.
Pubblicazione: (2026)
How do Observable Users Decompose D3 Code? A Qualitative Study
di: Lin, Melissa, et al.
Pubblicazione: (2024)
di: Lin, Melissa, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Code2Video: A Code-centric Paradigm for Educational Video Generation
di: Chen, Yanzhe, et al.
Pubblicazione: (2025) -
ShowUI-$π$: Flow-based Generative Models as GUI Dexterous Hands
di: Hu, Siyuan, et al.
Pubblicazione: (2025) -
Computer-Use Agents as Judges for Generative User Interface
di: Lin, Kevin Qinghong, et al.
Pubblicazione: (2025) -
GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents
di: Ouyang, Mingyu, et al.
Pubblicazione: (2026) -
GUIDE: A Benchmark for Understanding and Assisting Users in Open-Ended GUI Tasks
di: Yang, Saelyne, et al.
Pubblicazione: (2026)