Guardado en:
| Autores principales: | Kong, Yi, Shi, Dianxi, Yang, Guoli, ke-di, Zhang, Huang, Chenlin, Li, Xiaopeng, Jin, Songchang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2507.21953 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
VisionTasker: Mobile Task Automation Using Vision Based UI Understanding and LLM Task Planning
por: Song, Yunpeng, et al.
Publicado: (2023)
por: Song, Yunpeng, et al.
Publicado: (2023)
Explore, Select, Derive, and Recall: Augmenting LLM with Human-like Memory for Mobile Task Automation
por: Lee, Sunjae, et al.
Publicado: (2023)
por: Lee, Sunjae, et al.
Publicado: (2023)
MobA: Multifaceted Memory-Enhanced Adaptive Planning for Efficient Mobile Task Automation
por: Zhu, Zichen, et al.
Publicado: (2024)
por: Zhu, Zichen, et al.
Publicado: (2024)
OS-Genesis: Automating GUI Agent Trajectory Construction via Reverse Task Synthesis
por: Sun, Qiushi, et al.
Publicado: (2024)
por: Sun, Qiushi, et al.
Publicado: (2024)
Prompt2Task: Automating UI Tasks on Smartphones from Textual Prompts
por: Huang, Tian, et al.
Publicado: (2024)
por: Huang, Tian, et al.
Publicado: (2024)
Multi-User Mobile Augmented Reality for Cardiovascular Surgical Planning
por: Mehta, Pratham, et al.
Publicado: (2024)
por: Mehta, Pratham, et al.
Publicado: (2024)
From Human Negotiation to Agent Negotiation: Personal Mobility Agents in Automated Traffic
por: Jansen, Pascal
Publicado: (2026)
por: Jansen, Pascal
Publicado: (2026)
OpenMobile: Building Open Mobile Agents with Task and Trajectory Synthesis
por: Cheng, Kanzhi, et al.
Publicado: (2026)
por: Cheng, Kanzhi, et al.
Publicado: (2026)
ARCollab: Towards Multi-User Interactive Cardiovascular Surgical Planning in Mobile Augmented Reality
por: Mehta, Pratham, et al.
Publicado: (2024)
por: Mehta, Pratham, et al.
Publicado: (2024)
Human-Aided Trajectory Planning for Automated Vehicles through Teleoperation and Arbitration Graphs
por: Large, Nick Le, et al.
Publicado: (2025)
por: Large, Nick Le, et al.
Publicado: (2025)
Augmenting Minds or Automating Skills: The Differential Role of Human Capital in Generative AI's Impact on Creative Tasks
por: Huang, Meiling, et al.
Publicado: (2024)
por: Huang, Meiling, et al.
Publicado: (2024)
HybridCollab: Unifying In-Person and Remote Collaboration for Cardiovascular Surgical Planning in Mobile Augmented Reality
por: Mehta, Pratham Darrpan, et al.
Publicado: (2025)
por: Mehta, Pratham Darrpan, et al.
Publicado: (2025)
AR Secretary Agent: Real-time Memory Augmentation via LLM-powered Augmented Reality Glasses
por: Haddad, Raphaël A. El, et al.
Publicado: (2025)
por: Haddad, Raphaël A. El, et al.
Publicado: (2025)
Studying Mobile Spatial Collaboration across Video Calls and Augmented Reality
por: Vanukuru, Rishi, et al.
Publicado: (2026)
por: Vanukuru, Rishi, et al.
Publicado: (2026)
Beyond Levels of Driving Automation: A Triadic Framework of Human-AI Collaboration in On-Road Mobility
por: Huang, Gaojian, et al.
Publicado: (2025)
por: Huang, Gaojian, et al.
Publicado: (2025)
LlamaTouch: A Faithful and Scalable Testbed for Mobile UI Task Automation
por: Zhang, Li, et al.
Publicado: (2024)
por: Zhang, Li, et al.
Publicado: (2024)
Effect factors of motion aftereffect in depth: Adaptation direction and induced eyes vergence
por: di, Zhang
Publicado: (2025)
por: di, Zhang
Publicado: (2025)
LightVA: Lightweight Visual Analytics with LLM Agent-Based Task Planning and Execution
por: Zhao, Yuheng, et al.
Publicado: (2024)
por: Zhao, Yuheng, et al.
Publicado: (2024)
Metabook: A Mobile-to-Headset Pipeline for 3D Story Book Creation in Augmented Reality
por: Wang, Yibo, et al.
Publicado: (2024)
por: Wang, Yibo, et al.
Publicado: (2024)
Unveiling the Tricks: Automated Detection of Dark Patterns in Mobile Applications
por: Chen, Jieshan, et al.
Publicado: (2023)
por: Chen, Jieshan, et al.
Publicado: (2023)
Persode: Personalized Visual Journaling with Episodic Memory-Aware AI Agent
por: Jin, Seokho, et al.
Publicado: (2025)
por: Jin, Seokho, et al.
Publicado: (2025)
Holistic Construction Automation with Modular Robots: From High-Level Task Specification to Execution
por: Külz, Jonathan, et al.
Publicado: (2024)
por: Külz, Jonathan, et al.
Publicado: (2024)
User Understanding of Privacy Permissions in Mobile Augmented Reality: Perceptions and Misconceptions
por: Paneva, Viktorija, et al.
Publicado: (2025)
por: Paneva, Viktorija, et al.
Publicado: (2025)
Examining Augmented Virtuality Impairment Simulation for Mobile App Accessibility Design
por: Choo, Kenny Tsu Wei, et al.
Publicado: (2025)
por: Choo, Kenny Tsu Wei, et al.
Publicado: (2025)
MapIO: Embodied Interaction for the Accessibility of Tactile Maps Through Augmented Touch Exploration and Conversation
por: Manzoni, Matteo, et al.
Publicado: (2024)
por: Manzoni, Matteo, et al.
Publicado: (2024)
Unlocking Memories with AI: Exploring the Role of AI-Generated Cues in Personal Reminiscing
por: Jeung, Jun Li, et al.
Publicado: (2024)
por: Jeung, Jun Li, et al.
Publicado: (2024)
MobileAgentBench: An Efficient and User-Friendly Benchmark for Mobile LLM Agents
por: Wang, Luyuan, et al.
Publicado: (2024)
por: Wang, Luyuan, et al.
Publicado: (2024)
MagicAgent: Towards Generalized Agent Planning
por: Ren, Xuhui, et al.
Publicado: (2026)
por: Ren, Xuhui, et al.
Publicado: (2026)
Memory Printer: Exploring Everyday Reminiscing by Combining Slow Design with Generative AI-based Image Creation
por: Fang, Zhou, et al.
Publicado: (2026)
por: Fang, Zhou, et al.
Publicado: (2026)
Agent-Initiated Interaction in Phone UI Automation
por: Kahlon, Noam, et al.
Publicado: (2025)
por: Kahlon, Noam, et al.
Publicado: (2025)
Multimodal Feedback for Task Guidance in Augmented Reality
por: Guo, Hu, et al.
Publicado: (2025)
por: Guo, Hu, et al.
Publicado: (2025)
EZBlender: Efficient 3D Editing with Plan-and-ReAct Agent
por: Wang, Hao, et al.
Publicado: (2026)
por: Wang, Hao, et al.
Publicado: (2026)
MARVisT: Authoring Glyph-based Visualization in Mobile Augmented Reality
por: Zhu-Tian, Chen, et al.
Publicado: (2023)
por: Zhu-Tian, Chen, et al.
Publicado: (2023)
AppGen: Mobility-aware App Usage Behavior Generation for Mobile Users
por: Huang, Zihan, et al.
Publicado: (2024)
por: Huang, Zihan, et al.
Publicado: (2024)
Task-Aware Delegation Cues for LLM Agents
por: Gu, Xingrui
Publicado: (2026)
por: Gu, Xingrui
Publicado: (2026)
Anytime Trust Rating Dynamics in a Human-Robot Interaction Task
por: Dekarske, Jason, et al.
Publicado: (2024)
por: Dekarske, Jason, et al.
Publicado: (2024)
Map as a By-product: Collective Landmark Mapping from IMU Data and User-provided Texts in Situated Tasks
por: Yonetani, Ryo, et al.
Publicado: (2025)
por: Yonetani, Ryo, et al.
Publicado: (2025)
Analysis of Locally Coupled 3D Manipulation Mappings Based on Mobile Device Motion
por: Issartel, Paul, et al.
Publicado: (2016)
por: Issartel, Paul, et al.
Publicado: (2016)
Vocabuild: An Accessible Augmented Tangible Interface for Gamified Vocabulary Learning of Constructing Meaning
por: Hu, Siying, et al.
Publicado: (2025)
por: Hu, Siying, et al.
Publicado: (2025)
Exploring the Effect of Viewing Attributes of Mobile AR Interfaces on Remote Collaborative and Competitive Tasks
por: Nugegoda, Nelusha, et al.
Publicado: (2025)
por: Nugegoda, Nelusha, et al.
Publicado: (2025)
Ejemplares similares
-
VisionTasker: Mobile Task Automation Using Vision Based UI Understanding and LLM Task Planning
por: Song, Yunpeng, et al.
Publicado: (2023) -
Explore, Select, Derive, and Recall: Augmenting LLM with Human-like Memory for Mobile Task Automation
por: Lee, Sunjae, et al.
Publicado: (2023) -
MobA: Multifaceted Memory-Enhanced Adaptive Planning for Efficient Mobile Task Automation
por: Zhu, Zichen, et al.
Publicado: (2024) -
OS-Genesis: Automating GUI Agent Trajectory Construction via Reverse Task Synthesis
por: Sun, Qiushi, et al.
Publicado: (2024) -
Prompt2Task: Automating UI Tasks on Smartphones from Textual Prompts
por: Huang, Tian, et al.
Publicado: (2024)