Beyond Syntax: Action Semantics Learning for App Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tang, Bohan, Luo, Dezhao, Liu, Jianheng, Chen, Jingxuan, Gong, Shaogang, Hao, Jianye, Wang, Jun, Shao, Kun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning
von: Wu, Qingyuan, et al.
Veröffentlicht: (2025)
von: Wu, Qingyuan, et al.
Veröffentlicht: (2025)
ViMo: A Generative Visual GUI World Model for App Agents
von: Luo, Dezhao, et al.
Veröffentlicht: (2025)
von: Luo, Dezhao, et al.
Veröffentlicht: (2025)
DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
von: Wang, Taiyi, et al.
Veröffentlicht: (2024)
von: Wang, Taiyi, et al.
Veröffentlicht: (2024)
AppVLM: A Lightweight Vision Language Model for Online App Control
von: Papoudakis, Georgios, et al.
Veröffentlicht: (2025)
von: Papoudakis, Georgios, et al.
Veröffentlicht: (2025)
Lightweight Neural App Control
von: Christianos, Filippos, et al.
Veröffentlicht: (2024)
von: Christianos, Filippos, et al.
Veröffentlicht: (2024)
K^2-Agent: Co-Evolving Know-What and Know-How for Hierarchical Mobile Device Control
von: Wu, Zhe, et al.
Veröffentlicht: (2026)
von: Wu, Zhe, et al.
Veröffentlicht: (2026)
SPA-Bench: A Comprehensive Benchmark for SmartPhone Agent Evaluation
von: Chen, Jingxuan, et al.
Veröffentlicht: (2024)
von: Chen, Jingxuan, et al.
Veröffentlicht: (2024)
GUI Agents with Foundation Models: A Comprehensive Survey
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
Hi-Agent: Hierarchical Vision-Language Agents for Mobile Device Control
von: Wu, Zhe, et al.
Veröffentlicht: (2025)
von: Wu, Zhe, et al.
Veröffentlicht: (2025)
Deep Research Agents: A Systematic Examination And Roadmap
von: Huang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Huang, Yuxuan, et al.
Veröffentlicht: (2025)
ActionCodec: What Makes for Good Action Tokenizers
von: Dong, Zibin, et al.
Veröffentlicht: (2026)
von: Dong, Zibin, et al.
Veröffentlicht: (2026)
Unveiling Code Pre-Trained Models: Investigating Syntax and Semantics Capacities
von: Ma, Wei, et al.
Veröffentlicht: (2022)
von: Ma, Wei, et al.
Veröffentlicht: (2022)
Generative Video Diffusion for Unseen Novel Semantic Video Moment Retrieval
von: Luo, Dezhao, et al.
Veröffentlicht: (2024)
von: Luo, Dezhao, et al.
Veröffentlicht: (2024)
Improving Code Translation with Syntax-Guided and Semantic-aware Preference Optimization
von: Wu, Yuhan, et al.
Veröffentlicht: (2026)
von: Wu, Yuhan, et al.
Veröffentlicht: (2026)
Visual Self-paced Iterative Learning for Unsupervised Temporal Action Localization
von: Hu, Yupeng, et al.
Veröffentlicht: (2023)
von: Hu, Yupeng, et al.
Veröffentlicht: (2023)
PMAT: Optimizing Action Generation Order in Multi-Agent Reinforcement Learning
von: Hu, Kun, et al.
Veröffentlicht: (2025)
von: Hu, Kun, et al.
Veröffentlicht: (2025)
Enhancing Coreference Resolution with Pretrained Language Models: Bridging the Gap Between Syntax and Semantics
von: Liu, Xingzu, et al.
Veröffentlicht: (2025)
von: Liu, Xingzu, et al.
Veröffentlicht: (2025)
FilDeep: Learning Large Deformations of Elastic-Plastic Solids with Multi-Fidelity Data
von: Tang, Jianheng, et al.
Veröffentlicht: (2026)
von: Tang, Jianheng, et al.
Veröffentlicht: (2026)
BinCtx: Multi-Modal Representation Learning for Robust Android App Behavior Detection
von: Liu, Zichen, et al.
Veröffentlicht: (2025)
von: Liu, Zichen, et al.
Veröffentlicht: (2025)
Whispering Context: Distilling Syntax and Semantics for Long Speech Transcripts
von: Altinok, Duygu
Veröffentlicht: (2025)
von: Altinok, Duygu
Veröffentlicht: (2025)
$\textbf{Re}^{2}$: Unlocking LLM Reasoning via Reinforcement Learning with Re-solving
von: Wang, Pinzheng, et al.
Veröffentlicht: (2026)
von: Wang, Pinzheng, et al.
Veröffentlicht: (2026)
How does Misinformation Affect Large Language Model Behaviors and Preferences?
von: Peng, Miao, et al.
Veröffentlicht: (2025)
von: Peng, Miao, et al.
Veröffentlicht: (2025)
Succeed or Learn Slowly: Sample Efficient Off-Policy Reinforcement Learning for Mobile App Control
von: Papoudakis, Georgios, et al.
Veröffentlicht: (2025)
von: Papoudakis, Georgios, et al.
Veröffentlicht: (2025)
LLM Based Bayesian Optimization for Prompt Search
von: Ballew, Adam, et al.
Veröffentlicht: (2025)
von: Ballew, Adam, et al.
Veröffentlicht: (2025)
Causal Interventions on Causal Paths: Mapping GPT-2's Reasoning From Syntax to Semantics
von: Lee, Isabelle, et al.
Veröffentlicht: (2024)
von: Lee, Isabelle, et al.
Veröffentlicht: (2024)
Beyond One-shot: AI Agents for Learning in Field Experiments
von: Luo, Junjie, et al.
Veröffentlicht: (2026)
von: Luo, Junjie, et al.
Veröffentlicht: (2026)
Syntax Is Easy, Semantics Is Hard: Evaluating LLMs for LTL Translation
von: Danso, Priscilla Kyei, et al.
Veröffentlicht: (2026)
von: Danso, Priscilla Kyei, et al.
Veröffentlicht: (2026)
Reinforcement Learning and Data-Generation for Syntax-Guided Synthesis
von: Parsert, Julian, et al.
Veröffentlicht: (2023)
von: Parsert, Julian, et al.
Veröffentlicht: (2023)
Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM Understanding of Paired-Focus Constructions
von: Scivetti, Wesley, et al.
Veröffentlicht: (2026)
von: Scivetti, Wesley, et al.
Veröffentlicht: (2026)
InfoSeeker: A Scalable Hierarchical Parallel Agent Framework for Web Information Seeking
von: Lee, Ka Yiu, et al.
Veröffentlicht: (2026)
von: Lee, Ka Yiu, et al.
Veröffentlicht: (2026)
Kolb-Based Experiential Learning for Generalist Agents with Human-Level Kaggle Data Science Performance
von: Grosnit, Antoine, et al.
Veröffentlicht: (2024)
von: Grosnit, Antoine, et al.
Veröffentlicht: (2024)
Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation
von: Li, Zaijing, et al.
Veröffentlicht: (2026)
von: Li, Zaijing, et al.
Veröffentlicht: (2026)
Beyond False Discovery Rate: A Stepdown Group SLOPE Approach for Grouped Variable Selection
von: Zhang, Xuelin, et al.
Veröffentlicht: (2026)
von: Zhang, Xuelin, et al.
Veröffentlicht: (2026)
A Framework for Formalizing LLM Agent Security
von: Siu, Vincent, et al.
Veröffentlicht: (2026)
von: Siu, Vincent, et al.
Veröffentlicht: (2026)
Enhancing the Medical Context-Awareness Ability of LLMs via Multifaceted Self-Refinement Learning
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2025)
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2025)
AppAgentX: Evolving GUI Agents as Proficient Smartphone Users
von: Jiang, Wenjia, et al.
Veröffentlicht: (2025)
von: Jiang, Wenjia, et al.
Veröffentlicht: (2025)
Beyond World-Frame Action Heads: Motion-Centric Action Frames for Vision-Language-Action Models
von: Yang, Huoren, et al.
Veröffentlicht: (2026)
von: Yang, Huoren, et al.
Veröffentlicht: (2026)
UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning
von: Lu, Zhengxi, et al.
Veröffentlicht: (2025)
von: Lu, Zhengxi, et al.
Veröffentlicht: (2025)
SyntaxShap: Syntax-aware Explainability Method for Text Generation
von: Amara, Kenza, et al.
Veröffentlicht: (2024)
von: Amara, Kenza, et al.
Veröffentlicht: (2024)
Latent Action Reparameterization for Efficient Agent Inference
von: Huang, Wenhao, et al.
Veröffentlicht: (2026)
von: Huang, Wenhao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning
von: Wu, Qingyuan, et al.
Veröffentlicht: (2025) -
ViMo: A Generative Visual GUI World Model for App Agents
von: Luo, Dezhao, et al.
Veröffentlicht: (2025) -
DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
von: Wang, Taiyi, et al.
Veröffentlicht: (2024) -
AppVLM: A Lightweight Vision Language Model for Online App Control
von: Papoudakis, Georgios, et al.
Veröffentlicht: (2025) -
Lightweight Neural App Control
von: Christianos, Filippos, et al.
Veröffentlicht: (2024)