AndroTMem: From Interaction Trajectories to Anchored Memory in Long-Horizon GUI Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Shi, Yibo, Li, Jungang, Zhang, Linghao, Dongfang, Zihao, Wu, Biao, Tao, Sicheng, Yan, Yibo, Qin, Chenxi, Liu, Weiting, Lin, Zhixin, Li, Hanqian, Huang, Yu, Dai, Song, Hei, Yonghua, Ding, Yue, Li, Xiang, Wang, Shikang, Xu, Chengdong, Liu, Jingqi, Ma, Xueying, Zheng, Zhiwen, Zhang, Xiaofei, Wang, Bincheng, Yang, Nichen, Wu, Jie, Tian, Lihua, Li, Chen, Hu, Xuming |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Temporal Gains, Spatial Costs: Revisiting Video Fine-Tuning in Multimodal Large Language Models
by: Zhang, Linghao, et al.
Published: (2026)
by: Zhang, Linghao, et al.
Published: (2026)
RTV-Bench: Benchmarking MLLM Continuous Perception, Understanding and Reasoning through Real-Time Video
by: Xun, Shuhang, et al.
Published: (2025)
by: Xun, Shuhang, et al.
Published: (2025)
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
by: Dai, Song, et al.
Published: (2025)
by: Dai, Song, et al.
Published: (2025)
Unveiling Instruction-Specific Neurons & Experts: An Analytical Framework for LLM's Instruction-Following Capabilities
by: Zhang, Junyan, et al.
Published: (2025)
by: Zhang, Junyan, et al.
Published: (2025)
Mobile GUI Agent Privacy Personalization with Trajectory Induced Preference Optimization
by: Lin, Zhixin, et al.
Published: (2026)
by: Lin, Zhixin, et al.
Published: (2026)
MOSS-ChatV: Reinforcement Learning with Process Reasoning Reward for Video Temporal Reasoning
by: Tao, Sicheng, et al.
Published: (2025)
by: Tao, Sicheng, et al.
Published: (2025)
Unlocking Speech Instruction Data Potential with Query Rewriting
by: Hei, Yonghua, et al.
Published: (2025)
by: Hei, Yonghua, et al.
Published: (2025)
Thinking Economically: A Hierarchical Framework for Adaptive-Complexity Reasoning in LLMs
by: Gao, Yubo, et al.
Published: (2026)
by: Gao, Yubo, et al.
Published: (2026)
A Visual Semantic Adaptive Watermark grounded by Prefix-Tuning for Large Vision-Language Model
by: Zheng, Qi, et al.
Published: (2026)
by: Zheng, Qi, et al.
Published: (2026)
Mind the Third Eye! Benchmarking Privacy Awareness in MLLM-powered Smartphone Agents
by: Lin, Zhixin, et al.
Published: (2025)
by: Lin, Zhixin, et al.
Published: (2025)
SAVEn-Vid: Synergistic Audio-Visual Integration for Enhanced Understanding in Long Video Context
by: Li, Jungang, et al.
Published: (2024)
by: Li, Jungang, et al.
Published: (2024)
VideoMark: A Distortion-Free Robust Watermarking Framework for Video Diffusion Models
by: Hu, Xuming, et al.
Published: (2025)
by: Hu, Xuming, et al.
Published: (2025)
EffiReason-Bench: A Unified Benchmark for Evaluating and Advancing Efficient Reasoning in Large Language Models
by: Huang, Junquan, et al.
Published: (2025)
by: Huang, Junquan, et al.
Published: (2025)
KnowMT-Bench: Benchmarking Knowledge-Intensive Long-Form Question Answering in Multi-Turn Dialogues
by: Chen, Junhao, et al.
Published: (2025)
by: Chen, Junhao, et al.
Published: (2025)
Video Signature: Implicit Watermarking for Video Diffusion Models
by: Huang, Yu, et al.
Published: (2025)
by: Huang, Yu, et al.
Published: (2025)
SOMP: Scalable Gradient Inversion for Large Language Models via Subspace-Guided Orthogonal Matching Pursuit
by: Li, Yibo, et al.
Published: (2026)
by: Li, Yibo, et al.
Published: (2026)
HyperG: Hypergraph-Enhanced LLMs for Structured Knowledge
by: Huang, Sirui, et al.
Published: (2025)
by: Huang, Sirui, et al.
Published: (2025)
GUI-Rise: Structured Reasoning and History Summarization for GUI Navigation
by: Liu, Tao, et al.
Published: (2025)
by: Liu, Tao, et al.
Published: (2025)
The $χ_y$-genus, Chern number inequalities and signature
by: Li, Ping, et al.
Published: (2026)
by: Li, Ping, et al.
Published: (2026)
SAMark: A Self-Anchored Text Watermarking with Paragraph-Level Paraphrase Robustness
by: Huo, Jiahao, et al.
Published: (2026)
by: Huo, Jiahao, et al.
Published: (2026)
OS-Themis: A Scalable Critic Framework for Generalist GUI Rewards
by: Li, Zehao, et al.
Published: (2026)
by: Li, Zehao, et al.
Published: (2026)
Variance-Adaptive Muon: Accelerating LLM Pretraining with NSR-Modulated and Variance-Scaled Momentum
by: Li, Jingru, et al.
Published: (2026)
by: Li, Jingru, et al.
Published: (2026)
Less is More: Empowering GUI Agent with Context-Aware Simplification
by: Chen, Gongwei, et al.
Published: (2025)
by: Chen, Gongwei, et al.
Published: (2025)
ConfTuner: Training Large Language Models to Express Their Confidence Verbally
by: Li, Yibo, et al.
Published: (2025)
by: Li, Yibo, et al.
Published: (2025)
PhyRPR: Training-Free Physics-Constrained Video Generation
by: Zhao, Yibo, et al.
Published: (2026)
by: Zhao, Yibo, et al.
Published: (2026)
How Customers' Priority Is Shifting From Welcoming to Boycotting “ TikTok Refugees”? Influences of Algorithmic Recommendations and AI ‐Generated Content ( AIGC )
by: Zhuorong Wu, et al.
Published: (2026)
by: Zhuorong Wu, et al.
Published: (2026)
Unveiling Language Routing Isolation in Multilingual MoE Models for Interpretable Subnetwork Adaptation
by: Zheng, Kening, et al.
Published: (2026)
by: Zheng, Kening, et al.
Published: (2026)
Exploring Response Uncertainty in MLLMs: An Empirical Evaluation under Misleading Scenarios
by: Dang, Yunkai, et al.
Published: (2024)
by: Dang, Yunkai, et al.
Published: (2024)
A Dental Fluorosis Segmentation Model Combining Dynamic Snake Convolution and Learnable Shape Prior
by: Zhihao Li, et al.
Published: (2026)
by: Zhihao Li, et al.
Published: (2026)
ThreatIntel-Andro: Expert-Verified Benchmarking for Robust Android Malware Research
by: Bai, Hongpeng, et al.
Published: (2025)
by: Bai, Hongpeng, et al.
Published: (2025)
AndroCon: Conning Location Services in Android
by: Nag, Soham, et al.
Published: (2024)
by: Nag, Soham, et al.
Published: (2024)
A neighborhood union condition for the existence of a spanning tree without samll degree vertices
by: Li, Yibo, et al.
Published: (2025)
by: Li, Yibo, et al.
Published: (2025)
Influence of Cooling Rate on Residual Strain Evolution and Interlaminar Properties in Thermoplastic Composites With Embedded FBG Monitoring
by: Zijing Qu, et al.
Published: (2025)
by: Zijing Qu, et al.
Published: (2025)
Pierce the Mists, Greet the Sky: Decipher Knowledge Overshadowing via Knowledge Circuit Analysis
by: Huang, Haoming, et al.
Published: (2025)
by: Huang, Haoming, et al.
Published: (2025)
Ancestral Sequence Reconstruction and Comprehensive Computational Simulations Unmask an Efficient PET Hydrolase with the Wobbled Catalytic Triad
by: Yibo Song, et al.
Published: (2025)
by: Yibo Song, et al.
Published: (2025)
Unlocking a Sustainable Future for Plastics: A Chemical‐Enzymatic Pathway for Efficient Conversion of Mixed Waste to MHET and Energy‐Saving PET Recycling
by: Anni Li, et al.
Published: (2024)
by: Anni Li, et al.
Published: (2024)
Retrieval, Reward, and Training Protocols: What Matters in Training Search Agents?
by: Zhao, Yibo, et al.
Published: (2026)
by: Zhao, Yibo, et al.
Published: (2026)
VSI: Visual Subtitle Integration for Keyframe Selection to enhance Long Video Understanding
by: He, Jianxiang, et al.
Published: (2025)
by: He, Jianxiang, et al.
Published: (2025)
Optimal Trudinger-Moser inequalities on complete noncompact Riemannian manifolds: Revisit of the argument from the local inequalities to global ones
by: Li, Jungang, et al.
Published: (2026)
by: Li, Jungang, et al.
Published: (2026)
Global Compactness and Existence for Higher Order Critical Equations on Hyperbolic Spaces
by: Li, Jungang, et al.
Published: (2024)
by: Li, Jungang, et al.
Published: (2024)
Similar Items
-
Temporal Gains, Spatial Costs: Revisiting Video Fine-Tuning in Multimodal Large Language Models
by: Zhang, Linghao, et al.
Published: (2026) -
RTV-Bench: Benchmarking MLLM Continuous Perception, Understanding and Reasoning through Real-Time Video
by: Xun, Shuhang, et al.
Published: (2025) -
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
by: Dai, Song, et al.
Published: (2025) -
Unveiling Instruction-Specific Neurons & Experts: An Analytical Framework for LLM's Instruction-Following Capabilities
by: Zhang, Junyan, et al.
Published: (2025) -
Mobile GUI Agent Privacy Personalization with Trajectory Induced Preference Optimization
by: Lin, Zhixin, et al.
Published: (2026)