Foundations and Recent Trends in Multimodal Mobile Agents: A Survey
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Biao, Li, Yanda, Zhang, Zhiwei, Wei, Yunchao, Fang, Meng, Chen, Ling |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SpeechR: A Benchmark for Speech Reasoning in Large Audio-Language Models
von: Yang, Wanqi, et al.
Veröffentlicht: (2025)
von: Yang, Wanqi, et al.
Veröffentlicht: (2025)
MTPChat: A Multimodal Time-Aware Persona Dataset for Conversational Agents
von: Yang, Wanqi, et al.
Veröffentlicht: (2025)
von: Yang, Wanqi, et al.
Veröffentlicht: (2025)
Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models
von: Yang, Wanqi, et al.
Veröffentlicht: (2024)
von: Yang, Wanqi, et al.
Veröffentlicht: (2024)
AppAgent v2: Advanced Agent for Flexible Mobile Interactions
von: Li, Yanda, et al.
Veröffentlicht: (2024)
von: Li, Yanda, et al.
Veröffentlicht: (2024)
Enhancing Temporal Sensitivity and Reasoning for Time-Sensitive Question Answering
von: Yang, Wanqi, et al.
Veröffentlicht: (2024)
von: Yang, Wanqi, et al.
Veröffentlicht: (2024)
Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models
von: Li, Yanda, et al.
Veröffentlicht: (2026)
von: Li, Yanda, et al.
Veröffentlicht: (2026)
MMAC-Copilot: Multi-modal Agent Collaboration Operating Copilot
von: Song, Zirui, et al.
Veröffentlicht: (2024)
von: Song, Zirui, et al.
Veröffentlicht: (2024)
EcoAgent: An Efficient Device-Cloud Collaborative Multi-Agent Framework for Mobile Automation
von: Yi, Biao, et al.
Veröffentlicht: (2025)
von: Yi, Biao, et al.
Veröffentlicht: (2025)
Multi-Agent Autonomous Driving Systems with Large Language Models: A Survey of Recent Advances
von: Wu, Yaozu, et al.
Veröffentlicht: (2025)
von: Wu, Yaozu, et al.
Veröffentlicht: (2025)
Recent Advances of Multimodal Continual Learning: A Comprehensive Survey
von: Yu, Dianzhi, et al.
Veröffentlicht: (2024)
von: Yu, Dianzhi, et al.
Veröffentlicht: (2024)
A Survey of Personalized Federated Foundation Models for Privacy-Preserving Recommendation
von: Li, Zhiwei, et al.
Veröffentlicht: (2025)
von: Li, Zhiwei, et al.
Veröffentlicht: (2025)
A Survey on Current Trends and Recent Advances in Text Anonymization
von: Deußer, Tobias, et al.
Veröffentlicht: (2025)
von: Deußer, Tobias, et al.
Veröffentlicht: (2025)
Conditional Diffusion Model for Multi-Agent Dynamic Task Decomposition
von: Zhu, Yanda, et al.
Veröffentlicht: (2025)
von: Zhu, Yanda, et al.
Veröffentlicht: (2025)
Thinking on Maps: How Foundation Model Agents Explore, Remember, and Reason Map Environments
von: Wei, Zhiwei, et al.
Veröffentlicht: (2025)
von: Wei, Zhiwei, et al.
Veröffentlicht: (2025)
Large Multimodal Agents: A Survey
von: Xie, Junlin, et al.
Veröffentlicht: (2024)
von: Xie, Junlin, et al.
Veröffentlicht: (2024)
Rethinking Memory Mechanisms of Foundation Agents in the Second Half: A Survey
von: Huang, Wei-Chieh, et al.
Veröffentlicht: (2026)
von: Huang, Wei-Chieh, et al.
Veröffentlicht: (2026)
Game-TARS: Pretrained Foundation Models for Scalable Generalist Multimodal Game Agents
von: Wang, Zihao, et al.
Veröffentlicht: (2025)
von: Wang, Zihao, et al.
Veröffentlicht: (2025)
A Survey on GUI Agents with Foundation Models Enhanced by Reinforcement Learning
von: Li, Jiahao, et al.
Veröffentlicht: (2025)
von: Li, Jiahao, et al.
Veröffentlicht: (2025)
Mobile Foundation Model as Firmware
von: Yuan, Jinliang, et al.
Veröffentlicht: (2023)
von: Yuan, Jinliang, et al.
Veröffentlicht: (2023)
Benchmarking Foundation Models with Retrieval-Augmented Generation in Olympic-Level Physics Problem Solving
von: Zheng, Shunfeng, et al.
Veröffentlicht: (2025)
von: Zheng, Shunfeng, et al.
Veröffentlicht: (2025)
Brain-Conditional Multimodal Synthesis: A Survey and Taxonomy
von: Mai, Weijian, et al.
Veröffentlicht: (2023)
von: Mai, Weijian, et al.
Veröffentlicht: (2023)
Curriculum Learning with Quality-Driven Data Selection
von: Wu, Biao, et al.
Veröffentlicht: (2024)
von: Wu, Biao, et al.
Veröffentlicht: (2024)
Synergizing Foundation Models and Federated Learning: A Survey
von: Li, Shenghui, et al.
Veröffentlicht: (2024)
von: Li, Shenghui, et al.
Veröffentlicht: (2024)
Representation Potentials of Foundation Models for Multimodal Alignment: A Survey
von: Lu, Jianglin, et al.
Veröffentlicht: (2025)
von: Lu, Jianglin, et al.
Veröffentlicht: (2025)
SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents
von: Liang, Siyuan, et al.
Veröffentlicht: (2025)
von: Liang, Siyuan, et al.
Veröffentlicht: (2025)
A Survey of Scientific Large Language Models: From Data Foundations to Agent Frontiers
von: Hu, Ming, et al.
Veröffentlicht: (2025)
von: Hu, Ming, et al.
Veröffentlicht: (2025)
A Survey of WebAgents: Towards Next-Generation AI Agents for Web Automation with Large Foundation Models
von: Ning, Liangbo, et al.
Veröffentlicht: (2025)
von: Ning, Liangbo, et al.
Veröffentlicht: (2025)
GUI Agents with Foundation Models: A Comprehensive Survey
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
A Survey of Resource-efficient LLM and Multimodal Foundation Models
von: Xu, Mengwei, et al.
Veröffentlicht: (2024)
von: Xu, Mengwei, et al.
Veröffentlicht: (2024)
Recent Trends in Modelling the Continuous Time Series using Deep Learning: A Survey
von: Habiba, Mansura, et al.
Veröffentlicht: (2024)
von: Habiba, Mansura, et al.
Veröffentlicht: (2024)
A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems
von: Fang, Jinyuan, et al.
Veröffentlicht: (2025)
von: Fang, Jinyuan, et al.
Veröffentlicht: (2025)
Impact of Stickers on Multimodal Sentiment and Intent in Social Media: A New Task, Dataset and Baseline
von: Shi, Yuanchen, et al.
Veröffentlicht: (2024)
von: Shi, Yuanchen, et al.
Veröffentlicht: (2024)
Residual RL--MPC for Robust Microrobotic Cell Pushing Under Time-Varying Flow
von: Yang, Yanda, et al.
Veröffentlicht: (2026)
von: Yang, Yanda, et al.
Veröffentlicht: (2026)
ReasonLight: A Multimodal Foundation Model-Enhanced Reinforcement Learning Framework for Zero-Shot Traffic Signal Control
von: Pang, Aoyu, et al.
Veröffentlicht: (2026)
von: Pang, Aoyu, et al.
Veröffentlicht: (2026)
DeepSurvey-Bench: Evaluating Academic Value of Automatically Generated Scientific Survey
von: Zhang, Guo-Biao, et al.
Veröffentlicht: (2026)
von: Zhang, Guo-Biao, et al.
Veröffentlicht: (2026)
MoveGPT: Scaling Mobility Foundation Models with Spatially-Aware Mixture of Experts
von: Han, Chonghua, et al.
Veröffentlicht: (2025)
von: Han, Chonghua, et al.
Veröffentlicht: (2025)
A Survey: Towards Privacy and Security in Mobile Large Language Models
von: Xu, Honghui, et al.
Veröffentlicht: (2025)
von: Xu, Honghui, et al.
Veröffentlicht: (2025)
Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning
von: Wang, Yiqi, et al.
Veröffentlicht: (2024)
von: Wang, Yiqi, et al.
Veröffentlicht: (2024)
Training and Serving System of Foundation Models: A Comprehensive Survey
von: Zhou, Jiahang, et al.
Veröffentlicht: (2024)
von: Zhou, Jiahang, et al.
Veröffentlicht: (2024)
SLCA++: Unleash the Power of Sequential Fine-tuning for Continual Learning with Pre-training
von: Zhang, Gengwei, et al.
Veröffentlicht: (2024)
von: Zhang, Gengwei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SpeechR: A Benchmark for Speech Reasoning in Large Audio-Language Models
von: Yang, Wanqi, et al.
Veröffentlicht: (2025) -
MTPChat: A Multimodal Time-Aware Persona Dataset for Conversational Agents
von: Yang, Wanqi, et al.
Veröffentlicht: (2025) -
Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models
von: Yang, Wanqi, et al.
Veröffentlicht: (2024) -
AppAgent v2: Advanced Agent for Flexible Mobile Interactions
von: Li, Yanda, et al.
Veröffentlicht: (2024) -
Enhancing Temporal Sensitivity and Reasoning for Time-Sensitive Question Answering
von: Yang, Wanqi, et al.
Veröffentlicht: (2024)