Saved in:
| Main Authors: | Kong, Yi, Shi, Dianxi, Yang, Guoli, ke-di, Zhang, Huang, Chenlin, Li, Xiaopeng, Jin, Songchang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2507.21953 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VisionTasker: Mobile Task Automation Using Vision Based UI Understanding and LLM Task Planning
by: Song, Yunpeng, et al.
Published: (2023)
by: Song, Yunpeng, et al.
Published: (2023)
Explore, Select, Derive, and Recall: Augmenting LLM with Human-like Memory for Mobile Task Automation
by: Lee, Sunjae, et al.
Published: (2023)
by: Lee, Sunjae, et al.
Published: (2023)
MobA: Multifaceted Memory-Enhanced Adaptive Planning for Efficient Mobile Task Automation
by: Zhu, Zichen, et al.
Published: (2024)
by: Zhu, Zichen, et al.
Published: (2024)
OS-Genesis: Automating GUI Agent Trajectory Construction via Reverse Task Synthesis
by: Sun, Qiushi, et al.
Published: (2024)
by: Sun, Qiushi, et al.
Published: (2024)
Prompt2Task: Automating UI Tasks on Smartphones from Textual Prompts
by: Huang, Tian, et al.
Published: (2024)
by: Huang, Tian, et al.
Published: (2024)
Multi-User Mobile Augmented Reality for Cardiovascular Surgical Planning
by: Mehta, Pratham, et al.
Published: (2024)
by: Mehta, Pratham, et al.
Published: (2024)
From Human Negotiation to Agent Negotiation: Personal Mobility Agents in Automated Traffic
by: Jansen, Pascal
Published: (2026)
by: Jansen, Pascal
Published: (2026)
OpenMobile: Building Open Mobile Agents with Task and Trajectory Synthesis
by: Cheng, Kanzhi, et al.
Published: (2026)
by: Cheng, Kanzhi, et al.
Published: (2026)
ARCollab: Towards Multi-User Interactive Cardiovascular Surgical Planning in Mobile Augmented Reality
by: Mehta, Pratham, et al.
Published: (2024)
by: Mehta, Pratham, et al.
Published: (2024)
Human-Aided Trajectory Planning for Automated Vehicles through Teleoperation and Arbitration Graphs
by: Large, Nick Le, et al.
Published: (2025)
by: Large, Nick Le, et al.
Published: (2025)
Augmenting Minds or Automating Skills: The Differential Role of Human Capital in Generative AI's Impact on Creative Tasks
by: Huang, Meiling, et al.
Published: (2024)
by: Huang, Meiling, et al.
Published: (2024)
HybridCollab: Unifying In-Person and Remote Collaboration for Cardiovascular Surgical Planning in Mobile Augmented Reality
by: Mehta, Pratham Darrpan, et al.
Published: (2025)
by: Mehta, Pratham Darrpan, et al.
Published: (2025)
AR Secretary Agent: Real-time Memory Augmentation via LLM-powered Augmented Reality Glasses
by: Haddad, Raphaël A. El, et al.
Published: (2025)
by: Haddad, Raphaël A. El, et al.
Published: (2025)
Studying Mobile Spatial Collaboration across Video Calls and Augmented Reality
by: Vanukuru, Rishi, et al.
Published: (2026)
by: Vanukuru, Rishi, et al.
Published: (2026)
Beyond Levels of Driving Automation: A Triadic Framework of Human-AI Collaboration in On-Road Mobility
by: Huang, Gaojian, et al.
Published: (2025)
by: Huang, Gaojian, et al.
Published: (2025)
LlamaTouch: A Faithful and Scalable Testbed for Mobile UI Task Automation
by: Zhang, Li, et al.
Published: (2024)
by: Zhang, Li, et al.
Published: (2024)
Effect factors of motion aftereffect in depth: Adaptation direction and induced eyes vergence
by: di, Zhang
Published: (2025)
by: di, Zhang
Published: (2025)
LightVA: Lightweight Visual Analytics with LLM Agent-Based Task Planning and Execution
by: Zhao, Yuheng, et al.
Published: (2024)
by: Zhao, Yuheng, et al.
Published: (2024)
Metabook: A Mobile-to-Headset Pipeline for 3D Story Book Creation in Augmented Reality
by: Wang, Yibo, et al.
Published: (2024)
by: Wang, Yibo, et al.
Published: (2024)
Unveiling the Tricks: Automated Detection of Dark Patterns in Mobile Applications
by: Chen, Jieshan, et al.
Published: (2023)
by: Chen, Jieshan, et al.
Published: (2023)
Persode: Personalized Visual Journaling with Episodic Memory-Aware AI Agent
by: Jin, Seokho, et al.
Published: (2025)
by: Jin, Seokho, et al.
Published: (2025)
Holistic Construction Automation with Modular Robots: From High-Level Task Specification to Execution
by: Külz, Jonathan, et al.
Published: (2024)
by: Külz, Jonathan, et al.
Published: (2024)
User Understanding of Privacy Permissions in Mobile Augmented Reality: Perceptions and Misconceptions
by: Paneva, Viktorija, et al.
Published: (2025)
by: Paneva, Viktorija, et al.
Published: (2025)
Examining Augmented Virtuality Impairment Simulation for Mobile App Accessibility Design
by: Choo, Kenny Tsu Wei, et al.
Published: (2025)
by: Choo, Kenny Tsu Wei, et al.
Published: (2025)
MapIO: Embodied Interaction for the Accessibility of Tactile Maps Through Augmented Touch Exploration and Conversation
by: Manzoni, Matteo, et al.
Published: (2024)
by: Manzoni, Matteo, et al.
Published: (2024)
Unlocking Memories with AI: Exploring the Role of AI-Generated Cues in Personal Reminiscing
by: Jeung, Jun Li, et al.
Published: (2024)
by: Jeung, Jun Li, et al.
Published: (2024)
MobileAgentBench: An Efficient and User-Friendly Benchmark for Mobile LLM Agents
by: Wang, Luyuan, et al.
Published: (2024)
by: Wang, Luyuan, et al.
Published: (2024)
MagicAgent: Towards Generalized Agent Planning
by: Ren, Xuhui, et al.
Published: (2026)
by: Ren, Xuhui, et al.
Published: (2026)
Memory Printer: Exploring Everyday Reminiscing by Combining Slow Design with Generative AI-based Image Creation
by: Fang, Zhou, et al.
Published: (2026)
by: Fang, Zhou, et al.
Published: (2026)
Agent-Initiated Interaction in Phone UI Automation
by: Kahlon, Noam, et al.
Published: (2025)
by: Kahlon, Noam, et al.
Published: (2025)
Multimodal Feedback for Task Guidance in Augmented Reality
by: Guo, Hu, et al.
Published: (2025)
by: Guo, Hu, et al.
Published: (2025)
EZBlender: Efficient 3D Editing with Plan-and-ReAct Agent
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
MARVisT: Authoring Glyph-based Visualization in Mobile Augmented Reality
by: Zhu-Tian, Chen, et al.
Published: (2023)
by: Zhu-Tian, Chen, et al.
Published: (2023)
AppGen: Mobility-aware App Usage Behavior Generation for Mobile Users
by: Huang, Zihan, et al.
Published: (2024)
by: Huang, Zihan, et al.
Published: (2024)
Task-Aware Delegation Cues for LLM Agents
by: Gu, Xingrui
Published: (2026)
by: Gu, Xingrui
Published: (2026)
Anytime Trust Rating Dynamics in a Human-Robot Interaction Task
by: Dekarske, Jason, et al.
Published: (2024)
by: Dekarske, Jason, et al.
Published: (2024)
Map as a By-product: Collective Landmark Mapping from IMU Data and User-provided Texts in Situated Tasks
by: Yonetani, Ryo, et al.
Published: (2025)
by: Yonetani, Ryo, et al.
Published: (2025)
Analysis of Locally Coupled 3D Manipulation Mappings Based on Mobile Device Motion
by: Issartel, Paul, et al.
Published: (2016)
by: Issartel, Paul, et al.
Published: (2016)
Vocabuild: An Accessible Augmented Tangible Interface for Gamified Vocabulary Learning of Constructing Meaning
by: Hu, Siying, et al.
Published: (2025)
by: Hu, Siying, et al.
Published: (2025)
Exploring the Effect of Viewing Attributes of Mobile AR Interfaces on Remote Collaborative and Competitive Tasks
by: Nugegoda, Nelusha, et al.
Published: (2025)
by: Nugegoda, Nelusha, et al.
Published: (2025)
Similar Items
-
VisionTasker: Mobile Task Automation Using Vision Based UI Understanding and LLM Task Planning
by: Song, Yunpeng, et al.
Published: (2023) -
Explore, Select, Derive, and Recall: Augmenting LLM with Human-like Memory for Mobile Task Automation
by: Lee, Sunjae, et al.
Published: (2023) -
MobA: Multifaceted Memory-Enhanced Adaptive Planning for Efficient Mobile Task Automation
by: Zhu, Zichen, et al.
Published: (2024) -
OS-Genesis: Automating GUI Agent Trajectory Construction via Reverse Task Synthesis
by: Sun, Qiushi, et al.
Published: (2024) -
Prompt2Task: Automating UI Tasks on Smartphones from Textual Prompts
by: Huang, Tian, et al.
Published: (2024)