PC-Agent: A Hierarchical Multi-Agent Collaboration Framework for Complex Task Automation on PC
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Haowei, Zhang, Xi, Xu, Haiyang, Wanyan, Yuyang, Wang, Junyang, Yan, Ming, Zhang, Ji, Yuan, Chunfeng, Xu, Changsheng, Hu, Weiming, Huang, Fei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Look Before You Leap: A GUI-Critic-R1 Model for Pre-Operative Error Diagnosis in GUI Automation
by: Wanyan, Yuyang, et al.
Published: (2025)
by: Wanyan, Yuyang, et al.
Published: (2025)
Mobile-Agent-E: Self-Evolving Mobile Assistant for Complex Tasks
by: Wang, Zhenhailong, et al.
Published: (2025)
by: Wang, Zhenhailong, et al.
Published: (2025)
Modality-Collaborative Low-Rank Decomposers for Few-Shot Video Domain Adaptation
by: Wanyan, Yuyang, et al.
Published: (2025)
by: Wanyan, Yuyang, et al.
Published: (2025)
Mobile-Agent-V: A Video-Guided Approach for Effortless and Efficient Operational Knowledge Injection in Mobile Automation
by: Wang, Junyang, et al.
Published: (2025)
by: Wang, Junyang, et al.
Published: (2025)
Mobile-Agent-V: A Video-Guided Approach for Effortless and Efficient Operational Knowledge Injection in Mobile Automation
by: Wang, Junyang, et al.
Published: (2025)
by: Wang, Junyang, et al.
Published: (2025)
A Comprehensive Review of Few-shot Action Recognition
by: Wanyan, Yuyang, et al.
Published: (2024)
by: Wanyan, Yuyang, et al.
Published: (2024)
Mobile-Agent-v2: Mobile Device Operation Assistant with Effective Navigation via Multi-Agent Collaboration
by: Wang, Junyang, et al.
Published: (2024)
by: Wang, Junyang, et al.
Published: (2024)
MIBench: Evaluating Multimodal Large Language Models over Multiple Images
by: Liu, Haowei, et al.
Published: (2024)
by: Liu, Haowei, et al.
Published: (2024)
Mobile-Agent-v3: Fundamental Agents for GUI Automation
by: Ye, Jiabo, et al.
Published: (2025)
by: Ye, Jiabo, et al.
Published: (2025)
Mobile-Agent: Autonomous Multi-Modal Mobile Device Agent with Visual Perception
by: Wang, Junyang, et al.
Published: (2024)
by: Wang, Junyang, et al.
Published: (2024)
Unifying Latent and Lexicon Representations for Effective Video-Text Retrieval
by: Liu, Haowei, et al.
Published: (2024)
by: Liu, Haowei, et al.
Published: (2024)
Semantics-enhanced Cross-modal Masked Image Modeling for Vision-Language Pre-training
by: Liu, Haowei, et al.
Published: (2024)
by: Liu, Haowei, et al.
Published: (2024)
HAWK: A Hierarchical Workflow Framework for Multi-Agent Collaboration
by: Cheng, Yuyang, et al.
Published: (2025)
by: Cheng, Yuyang, et al.
Published: (2025)
STAMP: Training Explicit Memory for Mobile GUI Agents in Controllable and Scalable Virtual Environments
by: Wang, Junyang, et al.
Published: (2026)
by: Wang, Junyang, et al.
Published: (2026)
Mobile-Agent-v3.5: Multi-platform Fundamental GUI Agents
by: Xu, Haiyang, et al.
Published: (2026)
by: Xu, Haiyang, et al.
Published: (2026)
Large Language Model-based Human-Agent Collaboration for Complex Task Solving
by: Feng, Xueyang, et al.
Published: (2024)
by: Feng, Xueyang, et al.
Published: (2024)
BMW Agents -- A Framework For Task Automation Through Multi-Agent Collaboration
by: Crawford, Noel, et al.
Published: (2024)
by: Crawford, Noel, et al.
Published: (2024)
EcoAgent: An Efficient Device-Cloud Collaborative Multi-Agent Framework for Mobile Automation
by: Yi, Biao, et al.
Published: (2025)
by: Yi, Biao, et al.
Published: (2025)
MiTa: A Hierarchical Multi-Agent Collaboration Framework with Memory-integrated and Task Allocation
by: Zhang, XiaoJie, et al.
Published: (2026)
by: Zhang, XiaoJie, et al.
Published: (2026)
Spectrum of $J^{PC} = 0^{\pm\pm}$ Gluonic Hidden-Charm Tetraquark States
by: Wan, Bing-Dong, et al.
Published: (2025)
by: Wan, Bing-Dong, et al.
Published: (2025)
Task-Aware Automated User Profile Generation for Recommendation Simulation Using Large Language Models
by: Wanyan, Xinye, et al.
Published: (2026)
by: Wanyan, Xinye, et al.
Published: (2026)
Hidden-charm and -bottom tetraquark states with $J^{PC}=1^{-+}$ via QCD sum rules
by: Wan, Bing-Dong, et al.
Published: (2025)
by: Wan, Bing-Dong, et al.
Published: (2025)
Spectroscopy of hidden-heavy tetraquark states with $J^{PC}=0^{--}$ in a color-octet configuration
by: Wan, Bing-Dong, et al.
Published: (2026)
by: Wan, Bing-Dong, et al.
Published: (2026)
Complementary Text-Guided Attention for Zero-Shot Adversarial Robustness
by: Yu, Lu, et al.
Published: (2026)
by: Yu, Lu, et al.
Published: (2026)
Text-Guided Attention is All You Need for Zero-Shot Robustness in Vision-Language Models
by: Yu, Lu, et al.
Published: (2024)
by: Yu, Lu, et al.
Published: (2024)
mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models
by: Ye, Jiabo, et al.
Published: (2024)
by: Ye, Jiabo, et al.
Published: (2024)
Subgoal-based Hierarchical Reinforcement Learning for Multi-Agent Collaboration
by: Xu, Cheng, et al.
Published: (2024)
by: Xu, Cheng, et al.
Published: (2024)
OSWorld-MCP: Benchmarking MCP Tool Invocation In Computer-Use Agents
by: Jia, Hongrui, et al.
Published: (2025)
by: Jia, Hongrui, et al.
Published: (2025)
V2X-PC: Vehicle-to-everything Collaborative Perception via Point Cluster
by: Liu, Si, et al.
Published: (2024)
by: Liu, Si, et al.
Published: (2024)
AI Urban Scientist: Multi-Agent Collaborative Automation for Urban Research
by: Xia, Tong, et al.
Published: (2025)
by: Xia, Tong, et al.
Published: (2025)
RS-Agent: Automating Remote Sensing Tasks through Intelligent Agent
by: Xu, Wenjia, et al.
Published: (2024)
by: Xu, Wenjia, et al.
Published: (2024)
ToolCUA: Towards Optimal GUI-Tool Path Orchestration for Computer Use Agents
by: Hu, Xuhao, et al.
Published: (2026)
by: Hu, Xuhao, et al.
Published: (2026)
Inclination of the South China Sea from sediment core PC27, PC83 and PC111
by: Yang, Xiaoqiang, et al.
Published: (2024)
by: Yang, Xiaoqiang, et al.
Published: (2024)
Visual Document Understanding and Reasoning: A Multi-Agent Collaboration Framework with Agent-Wise Adaptive Test-Time Scaling
by: Yu, Xinlei, et al.
Published: (2025)
by: Yu, Xinlei, et al.
Published: (2025)
CompileAgent: Automated Real-World Repo-Level Compilation with Tool-Integrated LLM-based Agent System
by: Hu, Li, et al.
Published: (2025)
by: Hu, Li, et al.
Published: (2025)
PC2P: Multi-Agent Path Finding via Personalized-Enhanced Communication and Crowd Perception
by: Li, Guotao, et al.
Published: (2026)
by: Li, Guotao, et al.
Published: (2026)
Nexus: A Lightweight and Scalable Multi-Agent Framework for Complex Tasks Automation
by: Sami, Humza, et al.
Published: (2025)
by: Sami, Humza, et al.
Published: (2025)
DatawiseAgent: A Notebook-Centric LLM Agent Framework for Adaptive and Robust Data Science Automation
by: You, Ziming, et al.
Published: (2025)
by: You, Ziming, et al.
Published: (2025)
VulnResolver: A Hybrid Agent Framework for LLM-Based Automated Vulnerability Issue Resolution
by: Zhang, Mingming, et al.
Published: (2026)
by: Zhang, Mingming, et al.
Published: (2026)
TinyChart: Efficient Chart Understanding with Visual Token Merging and Program-of-Thoughts Learning
by: Zhang, Liang, et al.
Published: (2024)
by: Zhang, Liang, et al.
Published: (2024)
Similar Items
-
Look Before You Leap: A GUI-Critic-R1 Model for Pre-Operative Error Diagnosis in GUI Automation
by: Wanyan, Yuyang, et al.
Published: (2025) -
Mobile-Agent-E: Self-Evolving Mobile Assistant for Complex Tasks
by: Wang, Zhenhailong, et al.
Published: (2025) -
Modality-Collaborative Low-Rank Decomposers for Few-Shot Video Domain Adaptation
by: Wanyan, Yuyang, et al.
Published: (2025) -
Mobile-Agent-V: A Video-Guided Approach for Effortless and Efficient Operational Knowledge Injection in Mobile Automation
by: Wang, Junyang, et al.
Published: (2025) -
Mobile-Agent-V: A Video-Guided Approach for Effortless and Efficient Operational Knowledge Injection in Mobile Automation
by: Wang, Junyang, et al.
Published: (2025)