GUI-CIDER: Mid-training GUI Agents via Causal Internalization and Density-aware Exemplar Reselection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Zheng, Han, Chengcheng, Lu, Zhengxi, Ju, Tianjie, Chen, Yanyu, Gu, Qi, Cai, Xunliang, Zhang, Zhuosheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Smoothing Grounding and Reasoning for MLLM-Powered GUI Agents with Query-Oriented Pivot Tasks
von: Wu, Zongru, et al.
Veröffentlicht: (2025)
von: Wu, Zongru, et al.
Veröffentlicht: (2025)
Training High-Level Schedulers with Execution-Feedback Reinforcement Learning for Long-Horizon GUI Automation
von: Deng, Zehao, et al.
Veröffentlicht: (2025)
von: Deng, Zehao, et al.
Veröffentlicht: (2025)
Hidden Ghost Hand: Unveiling Backdoor Vulnerabilities in MLLM-Powered Mobile GUI Agents
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2025)
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2025)
See, Think, Act: Teaching Multimodal Agents to Effectively Interact with GUI by Identifying Toggles
von: Wu, Zongru, et al.
Veröffentlicht: (2025)
von: Wu, Zongru, et al.
Veröffentlicht: (2025)
EVA: Red-Teaming GUI Agents via Evolving Indirect Prompt Injection
von: Lu, Yijie, et al.
Veröffentlicht: (2025)
von: Lu, Yijie, et al.
Veröffentlicht: (2025)
GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding
von: Tang, Fei, et al.
Veröffentlicht: (2025)
von: Tang, Fei, et al.
Veröffentlicht: (2025)
SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization
von: Lu, Zhengxi, et al.
Veröffentlicht: (2026)
von: Lu, Zhengxi, et al.
Veröffentlicht: (2026)
MemGUI-Bench: Benchmarking Memory of Mobile GUI Agents in Dynamic Environments
von: Liu, Guangyi, et al.
Veröffentlicht: (2026)
von: Liu, Guangyi, et al.
Veröffentlicht: (2026)
SE-GA: Memory-Augmented Self-Evolution for GUI Agents
von: Jin, Shilong, et al.
Veröffentlicht: (2026)
von: Jin, Shilong, et al.
Veröffentlicht: (2026)
Faithful Mobile GUI Agents with Guided Advantage Estimator
von: Hu, Haowen, et al.
Veröffentlicht: (2026)
von: Hu, Haowen, et al.
Veröffentlicht: (2026)
Mobile-Agent-v3: Fundamental Agents for GUI Automation
von: Ye, Jiabo, et al.
Veröffentlicht: (2025)
von: Ye, Jiabo, et al.
Veröffentlicht: (2025)
GEM: Gaussian Embedding Modeling for Out-of-Distribution Detection in GUI Agents
von: Wu, Zheng, et al.
Veröffentlicht: (2025)
von: Wu, Zheng, et al.
Veröffentlicht: (2025)
OS-Kairos: Adaptive Interaction for MLLM-Powered GUI Agents
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2025)
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2025)
CoCo-Agent: A Comprehensive Cognitive MLLM Agent for Smartphone GUI Automation
von: Ma, Xinbei, et al.
Veröffentlicht: (2024)
von: Ma, Xinbei, et al.
Veröffentlicht: (2024)
Adaptive Milestone Reward for GUI Agents
von: Zheng, Congmin, et al.
Veröffentlicht: (2026)
von: Zheng, Congmin, et al.
Veröffentlicht: (2026)
GUI-explorer: Autonomous Exploration and Mining of Transition-aware Knowledge for GUI Agent
von: Xie, Bin, et al.
Veröffentlicht: (2025)
von: Xie, Bin, et al.
Veröffentlicht: (2025)
Causal Probing for Internal Visual Representations in Multimodal Large Language Models
von: Deng, Zehao, et al.
Veröffentlicht: (2026)
von: Deng, Zehao, et al.
Veröffentlicht: (2026)
Test-Time Reinforcement Learning for GUI Grounding via Region Consistency
von: Du, Yong, et al.
Veröffentlicht: (2025)
von: Du, Yong, et al.
Veröffentlicht: (2025)
MPR-GUI: Benchmarking and Enhancing Multilingual Perception and Reasoning in GUI Agents
von: Chen, Ruihan, et al.
Veröffentlicht: (2025)
von: Chen, Ruihan, et al.
Veröffentlicht: (2025)
LiteGUI: Distilling Compact GUI Agents with Reinforcement Learning
von: Wu, Yubin, et al.
Veröffentlicht: (2026)
von: Wu, Yubin, et al.
Veröffentlicht: (2026)
UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning
von: Lu, Zhengxi, et al.
Veröffentlicht: (2025)
von: Lu, Zhengxi, et al.
Veröffentlicht: (2025)
UI-S1: Advancing GUI Automation via Semi-online Reinforcement Learning
von: Lu, Zhengxi, et al.
Veröffentlicht: (2025)
von: Lu, Zhengxi, et al.
Veröffentlicht: (2025)
LaSM: Layer-wise Scaling Mechanism for Defending Pop-up Attack on GUI Agents
von: Yan, Zihe, et al.
Veröffentlicht: (2025)
von: Yan, Zihe, et al.
Veröffentlicht: (2025)
Executable Agentic Memory for GUI Agent
von: Qin, Zerui, et al.
Veröffentlicht: (2026)
von: Qin, Zerui, et al.
Veröffentlicht: (2026)
GUI-CEval: A Hierarchical and Comprehensive Chinese Benchmark for Mobile GUI Agents
von: Li, Yang, et al.
Veröffentlicht: (2026)
von: Li, Yang, et al.
Veröffentlicht: (2026)
ClawGUI: A Unified Framework for Training, Evaluating, and Deploying GUI Agents
von: Tang, Fei, et al.
Veröffentlicht: (2026)
von: Tang, Fei, et al.
Veröffentlicht: (2026)
META-GUI: Towards Multi-modal Conversational Agents on Mobile GUI
von: Sun, Liangtai, et al.
Veröffentlicht: (2022)
von: Sun, Liangtai, et al.
Veröffentlicht: (2022)
GUI-Eyes: Tool-Augmented Perception for Visual Grounding in GUI Agents
von: Chen, Chen, et al.
Veröffentlicht: (2026)
von: Chen, Chen, et al.
Veröffentlicht: (2026)
CRAFT-GUI: Curriculum-Reinforced Agent For GUI Tasks
von: Nong, Songqin, et al.
Veröffentlicht: (2025)
von: Nong, Songqin, et al.
Veröffentlicht: (2025)
GUI-PRA: Process Reward Agent for GUI Tasks
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
GUI-ARP: Enhancing Grounding with Adaptive Region Perception for GUI Agents
von: Ye, Xianhang, et al.
Veröffentlicht: (2025)
von: Ye, Xianhang, et al.
Veröffentlicht: (2025)
GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL
von: Yang, Rui, et al.
Veröffentlicht: (2026)
von: Yang, Rui, et al.
Veröffentlicht: (2026)
UIPro: Unleashing Superior Interaction Capability For GUI Agents
von: Li, Hongxin, et al.
Veröffentlicht: (2025)
von: Li, Hongxin, et al.
Veröffentlicht: (2025)
MementoGUI: Learning Agentic Multimodal Memory Control for Long-Horizon GUI Agents
von: Zeng, Ziyun, et al.
Veröffentlicht: (2026)
von: Zeng, Ziyun, et al.
Veröffentlicht: (2026)
HiconAgent: History Context-aware Policy Optimization for GUI Agents
von: Zhou, Xurui, et al.
Veröffentlicht: (2025)
von: Zhou, Xurui, et al.
Veröffentlicht: (2025)
LLM-Powered GUI Agents in Phone Automation: Surveying Progress and Prospects
von: Liu, Guangyi, et al.
Veröffentlicht: (2025)
von: Liu, Guangyi, et al.
Veröffentlicht: (2025)
GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents
von: Wu, Qianhui, et al.
Veröffentlicht: (2025)
von: Wu, Qianhui, et al.
Veröffentlicht: (2025)
GUI-Xplore: Empowering Generalizable GUI Agents with One Exploration
von: Sun, Yuchen, et al.
Veröffentlicht: (2025)
von: Sun, Yuchen, et al.
Veröffentlicht: (2025)
Continual GUI Agents
von: Liu, Ziwei, et al.
Veröffentlicht: (2026)
von: Liu, Ziwei, et al.
Veröffentlicht: (2026)
GUI-KV: Efficient GUI Agents via KV Cache with Spatio-Temporal Awareness
von: Huang, Kung-Hsiang, et al.
Veröffentlicht: (2025)
von: Huang, Kung-Hsiang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Smoothing Grounding and Reasoning for MLLM-Powered GUI Agents with Query-Oriented Pivot Tasks
von: Wu, Zongru, et al.
Veröffentlicht: (2025) -
Training High-Level Schedulers with Execution-Feedback Reinforcement Learning for Long-Horizon GUI Automation
von: Deng, Zehao, et al.
Veröffentlicht: (2025) -
Hidden Ghost Hand: Unveiling Backdoor Vulnerabilities in MLLM-Powered Mobile GUI Agents
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2025) -
See, Think, Act: Teaching Multimodal Agents to Effectively Interact with GUI by Identifying Toggles
von: Wu, Zongru, et al.
Veröffentlicht: (2025) -
EVA: Red-Teaming GUI Agents via Evolving Indirect Prompt Injection
von: Lu, Yijie, et al.
Veröffentlicht: (2025)