AutoDroid-V2: Boosting SLM-based GUI Agents via Code Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Wen, Hao, Tian, Shizuo, Pavlov, Borislav, Du, Wenjie, Li, Yixuan, Chang, Ge, Zhao, Shanhui, Liu, Jiacheng, Liu, Yunxin, Zhang, Ya-Qin, Li, Yuanchun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AutoDroid: LLM-powered Task Automation in Android
por: Wen, Hao, et al.
Publicado: (2023)
por: Wen, Hao, et al.
Publicado: (2023)
AgentProg: Empowering Long-Horizon GUI Agents with Program-Guided Context Management
por: Tian, Shizuo, et al.
Publicado: (2025)
por: Tian, Shizuo, et al.
Publicado: (2025)
Mobile GUI Agents under Real-world Threats: Are We There Yet?
por: Liu, Guohong, et al.
Publicado: (2025)
por: Liu, Guohong, et al.
Publicado: (2025)
GUI-Xplore: Empowering Generalizable GUI Agents with One Exploration
por: Sun, Yuchen, et al.
Publicado: (2025)
por: Sun, Yuchen, et al.
Publicado: (2025)
Joint Agent Memory and Exploration Learning via Novelty Signals
por: Tian, Shizuo, et al.
Publicado: (2026)
por: Tian, Shizuo, et al.
Publicado: (2026)
SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking
por: Liu, Guohong, et al.
Publicado: (2026)
por: Liu, Guohong, et al.
Publicado: (2026)
LLM-Explorer: Towards Efficient and Affordable LLM-based Exploration for Mobile Apps
por: Zhao, Shanhui, et al.
Publicado: (2025)
por: Zhao, Shanhui, et al.
Publicado: (2025)
DroidBot-GPT: GPT-powered UI Automation for Android
por: Wen, Hao, et al.
Publicado: (2023)
por: Wen, Hao, et al.
Publicado: (2023)
Routine Computing: A Systematic Review of Sensing Daily Life Dimensions Towards Human-Centered Goals
por: Pavlov, Borislav, et al.
Publicado: (2026)
por: Pavlov, Borislav, et al.
Publicado: (2026)
ParaThinker: Native Parallel Thinking as a New Paradigm to Scale LLM Test-time Compute
por: Wen, Hao, et al.
Publicado: (2025)
por: Wen, Hao, et al.
Publicado: (2025)
UI-TARS: Pioneering Automated GUI Interaction with Native Agents
por: Qin, Yujia, et al.
Publicado: (2025)
por: Qin, Yujia, et al.
Publicado: (2025)
ChainStream: An LLM-based Framework for Unified Synthetic Sensing
por: Liu, Jiacheng, et al.
Publicado: (2024)
por: Liu, Jiacheng, et al.
Publicado: (2024)
ReuseDroid: A VLM-empowered Android UI Test Migrator Boosted by Active Feedback
por: Li, Xiaolei, et al.
Publicado: (2025)
por: Li, Xiaolei, et al.
Publicado: (2025)
FuncDroid: Towards Inter-Functional Flows for Comprehensive Mobile App GUI Testing
por: He, Jinlong, et al.
Publicado: (2026)
por: He, Jinlong, et al.
Publicado: (2026)
EpiDroid: Dependency-Guided Recomposition for Deep State Discovery in Mobile GUI Testing
por: Song, Jiahui, et al.
Publicado: (2026)
por: Song, Jiahui, et al.
Publicado: (2026)
GRAIL:Learning to Interact with Large Knowledge Graphs for Retrieval Augmented Reasoning
por: Chang, Ge, et al.
Publicado: (2025)
por: Chang, Ge, et al.
Publicado: (2025)
Enhancing Agentic Textual Graph Retrieval with Synthetic Stepwise Supervision
por: Chang, Ge, et al.
Publicado: (2025)
por: Chang, Ge, et al.
Publicado: (2025)
A First Look At Efficient And Secure On-Device LLM Inference Against KV Leakage
por: Yang, Huan, et al.
Publicado: (2024)
por: Yang, Huan, et al.
Publicado: (2024)
WindowsWorld: A Process-Centric Benchmark of Autonomous GUI Agents in Professional Cross-Application Environments
por: Li, Jinchao, et al.
Publicado: (2026)
por: Li, Jinchao, et al.
Publicado: (2026)
Auto-scaling Continuous Memory for GUI Agent
por: Wu, Wenyi, et al.
Publicado: (2025)
por: Wu, Wenyi, et al.
Publicado: (2025)
Next-Gen CAPTCHAs: Leveraging the Cognitive Gap for Scalable and Diverse GUI-Agent Defense
por: Liu, Jiacheng, et al.
Publicado: (2026)
por: Liu, Jiacheng, et al.
Publicado: (2026)
Advancing Mobile GUI Agents: A Verifier-Driven Approach to Practical Deployment
por: Dai, Gaole, et al.
Publicado: (2025)
por: Dai, Gaole, et al.
Publicado: (2025)
AutoGUI: Scaling GUI Grounding with Automatic Functionality Annotations from LLMs
por: Li, Hongxin, et al.
Publicado: (2025)
por: Li, Hongxin, et al.
Publicado: (2025)
SHARE: An SLM-based Hierarchical Action CorREction Assistant for Text-to-SQL
por: Qu, Ge, et al.
Publicado: (2025)
por: Qu, Ge, et al.
Publicado: (2025)
LoRA-Switch: Boosting the Efficiency of Dynamic LLM Adapters via System-Algorithm Co-design
por: Kong, Rui, et al.
Publicado: (2024)
por: Kong, Rui, et al.
Publicado: (2024)
Rational Decision-Making Agent with Internalized Utility Judgment
por: Ye, Yining, et al.
Publicado: (2023)
por: Ye, Yining, et al.
Publicado: (2023)
ProRe: A Proactive Reward System for GUI Agents via Reasoner-Actor Collaboration
por: Dai, Gaole, et al.
Publicado: (2025)
por: Dai, Gaole, et al.
Publicado: (2025)
Continual GUI Agents
por: Liu, Ziwei, et al.
Publicado: (2026)
por: Liu, Ziwei, et al.
Publicado: (2026)
BudgetThinker: Empowering Budget-aware LLM Reasoning with Control Tokens
por: Wen, Hao, et al.
Publicado: (2025)
por: Wen, Hao, et al.
Publicado: (2025)
Threshold Neuron: A Brain-inspired Artificial Neuron for Efficient On-device Inference
por: Zheng, Zihao, et al.
Publicado: (2024)
por: Zheng, Zihao, et al.
Publicado: (2024)
Mobile-Bench-v2: A More Realistic and Comprehensive Benchmark for VLM-based Mobile Agents
por: Xu, Weikai, et al.
Publicado: (2025)
por: Xu, Weikai, et al.
Publicado: (2025)
From Automated to Autonomous: Hierarchical Agent-native Network Architecture (HANA)
por: Wu, Binghan, et al.
Publicado: (2026)
por: Wu, Binghan, et al.
Publicado: (2026)
An Empirical Study of LLM Reasoning Ability Under Strict Output Length Constraint
por: Sun, Yi, et al.
Publicado: (2025)
por: Sun, Yi, et al.
Publicado: (2025)
Leveraging AI Agents for Autonomous Networks: A Reference Architecture and Empirical Studies
por: Wu, Binghan, et al.
Publicado: (2025)
por: Wu, Binghan, et al.
Publicado: (2025)
AutoGUI-v2: A Comprehensive Multi-Modal GUI Functionality Understanding Benchmark
por: Li, Hongxin, et al.
Publicado: (2026)
por: Li, Hongxin, et al.
Publicado: (2026)
PIRA-Bench: A Transition from Reactive GUI Agents to GUI-based Proactive Intent Recommendation Agents
por: Chai, Yuxiang, et al.
Publicado: (2026)
por: Chai, Yuxiang, et al.
Publicado: (2026)
GUI Agents for Continual Game Generation
por: Huang, Yixu, et al.
Publicado: (2026)
por: Huang, Yixu, et al.
Publicado: (2026)
MobileViews: A Million-scale and Diverse Mobile GUI Dataset
por: Gao, Longxi, et al.
Publicado: (2024)
por: Gao, Longxi, et al.
Publicado: (2024)
GUI-PRA: Process Reward Agent for GUI Tasks
por: Xiong, Tao, et al.
Publicado: (2025)
por: Xiong, Tao, et al.
Publicado: (2025)
$\texttt{Droid}$: A Resource Suite for AI-Generated Code Detection
por: Orel, Daniil, et al.
Publicado: (2025)
por: Orel, Daniil, et al.
Publicado: (2025)
Ejemplares similares
-
AutoDroid: LLM-powered Task Automation in Android
por: Wen, Hao, et al.
Publicado: (2023) -
AgentProg: Empowering Long-Horizon GUI Agents with Program-Guided Context Management
por: Tian, Shizuo, et al.
Publicado: (2025) -
Mobile GUI Agents under Real-world Threats: Are We There Yet?
por: Liu, Guohong, et al.
Publicado: (2025) -
GUI-Xplore: Empowering Generalizable GUI Agents with One Exploration
por: Sun, Yuchen, et al.
Publicado: (2025) -
Joint Agent Memory and Exploration Learning via Novelty Signals
por: Tian, Shizuo, et al.
Publicado: (2026)