OSExpert: Computer-Use Agents Learning Professional Skills via Exploration
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Jiateng, Wang, Zhenhailong, Wang, Rushi, Li, Bingxuan, Kim, Jeonghwan, Tiwari, Aditi, Yu, Pengfei, Zhang, Denghui, Ji, Heng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing
von: Qian, Cheng, et al.
Veröffentlicht: (2026)
von: Qian, Cheng, et al.
Veröffentlicht: (2026)
Augmenting Interface Usability Heuristics for Reliable Computer-Use Agents
von: Liu, Jiateng, et al.
Veröffentlicht: (2026)
von: Liu, Jiateng, et al.
Veröffentlicht: (2026)
Infogent: An Agent-Based Framework for Web Information Aggregation
von: Reddy, Revanth Gangi, et al.
Veröffentlicht: (2024)
von: Reddy, Revanth Gangi, et al.
Veröffentlicht: (2024)
Open Vocabulary Electroencephalography-To-Text Decoding and Zero-shot Sentiment Classification
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2021)
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2021)
Context Engineering for Trustworthiness: Rescorla Wagner Steering Under Mixed and Inappropriate Contexts
von: Wang, Rushi, et al.
Veröffentlicht: (2025)
von: Wang, Rushi, et al.
Veröffentlicht: (2025)
Predicting Camera Pose from Perspective Descriptions for Spatial Reasoning
von: Zhang, Xuejun, et al.
Veröffentlicht: (2026)
von: Zhang, Xuejun, et al.
Veröffentlicht: (2026)
Analyzing and Internalizing Complex Policy Documents for LLM Agents
von: Liu, Jiateng, et al.
Veröffentlicht: (2025)
von: Liu, Jiateng, et al.
Veröffentlicht: (2025)
SYNTHIA: Novel Concept Design with Affordance Composition
von: Ha, Hyeonjeong, et al.
Veröffentlicht: (2025)
von: Ha, Hyeonjeong, et al.
Veröffentlicht: (2025)
Multimodal Policy Internalization for Conversational Agents
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2025)
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2025)
Automating Financial Statement Audits with Large Language Models
von: Wang, Rushi, et al.
Veröffentlicht: (2025)
von: Wang, Rushi, et al.
Veröffentlicht: (2025)
Democratizing LLMs: An Exploration of Cost-Performance Trade-offs in Self-Refined Open-Source Models
von: Shashidhar, Sumuk, et al.
Veröffentlicht: (2023)
von: Shashidhar, Sumuk, et al.
Veröffentlicht: (2023)
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2023)
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2023)
From Context to Skills: Can Language Models Learn from Context Skillfully?
von: Si, Shuzheng, et al.
Veröffentlicht: (2026)
von: Si, Shuzheng, et al.
Veröffentlicht: (2026)
Efficient Agent Training for Computer Use
von: He, Yanheng, et al.
Veröffentlicht: (2025)
von: He, Yanheng, et al.
Veröffentlicht: (2025)
Learning to Transfer Human Hand Skills for Robot Manipulations
von: Park, Sungjae, et al.
Veröffentlicht: (2025)
von: Park, Sungjae, et al.
Veröffentlicht: (2025)
Beyond Reactive Safety: Risk-Aware LLM Alignment via Long-Horizon Simulation
von: Sun, Chenkai, et al.
Veröffentlicht: (2025)
von: Sun, Chenkai, et al.
Veröffentlicht: (2025)
DyMU: Dynamic Merging and Virtual Unmerging for Efficient VLMs
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2025)
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2025)
Finer: Investigating and Enhancing Fine-Grained Visual Concept Recognition in Large Vision Language Models
von: Kim, Jeonghwan, et al.
Veröffentlicht: (2024)
von: Kim, Jeonghwan, et al.
Veröffentlicht: (2024)
RS-Claw: Progressive Active Tool Exploration via Hierarchical Skill Trees for Remote Sensing Agents
von: Liu, Liangtian, et al.
Veröffentlicht: (2026)
von: Liu, Liangtian, et al.
Veröffentlicht: (2026)
Aligning LLMs with Individual Preferences via Interaction
von: Wu, Shujin, et al.
Veröffentlicht: (2024)
von: Wu, Shujin, et al.
Veröffentlicht: (2024)
IntentCUA: Learning Intent-level Representations for Skill Abstraction and Multi-Agent Planning in Computer-Use Agents
von: Lee, Seoyoung, et al.
Veröffentlicht: (2026)
von: Lee, Seoyoung, et al.
Veröffentlicht: (2026)
Protein Language Models Diverge from Natural Language: Comparative Analysis and Improved Inference
von: Hart, Anna, et al.
Veröffentlicht: (2026)
von: Hart, Anna, et al.
Veröffentlicht: (2026)
PEARL: Self-Evolving Assistant for Time Management with Reinforcement Learning
von: Li, Bingxuan, et al.
Veröffentlicht: (2026)
von: Li, Bingxuan, et al.
Veröffentlicht: (2026)
ERA: Transforming VLMs into Embodied Agents via Embodied Prior Learning and Online Reinforcement Learning
von: Chen, Hanyang, et al.
Veröffentlicht: (2025)
von: Chen, Hanyang, et al.
Veröffentlicht: (2025)
MINT: Evaluating LLMs in Multi-turn Interaction with Tools and Language Feedback
von: Wang, Xingyao, et al.
Veröffentlicht: (2023)
von: Wang, Xingyao, et al.
Veröffentlicht: (2023)
Intelligent Co-Design: An Interactive LLM Framework for Interior Spatial Design via Multi-Modal Agents
von: Lim, Ren Jian, et al.
Veröffentlicht: (2026)
von: Lim, Ren Jian, et al.
Veröffentlicht: (2026)
Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning
von: Shi, Yaorui, et al.
Veröffentlicht: (2026)
von: Shi, Yaorui, et al.
Veröffentlicht: (2026)
SkillOrchestra: Learning to Route Agents via Skill Transfer
von: Wang, Jiayu, et al.
Veröffentlicht: (2026)
von: Wang, Jiayu, et al.
Veröffentlicht: (2026)
Agent Alpha: Tree Search Unifying Generation, Exploration and Evaluation for Computer-Use Agents
von: Tang, Sizhe, et al.
Veröffentlicht: (2026)
von: Tang, Sizhe, et al.
Veröffentlicht: (2026)
CUA-Skill: Develop Skills for Computer Using Agent
von: Chen, Tianyi, et al.
Veröffentlicht: (2026)
von: Chen, Tianyi, et al.
Veröffentlicht: (2026)
Visually Descriptive Language Model for Vector Graphics Reasoning
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2024)
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2024)
MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
von: Zhu, Kunlun, et al.
Veröffentlicht: (2025)
von: Zhu, Kunlun, et al.
Veröffentlicht: (2025)
Learning from Online Videos at Inference Time for Computer-Use Agents
von: Liu, Yujian, et al.
Veröffentlicht: (2025)
von: Liu, Yujian, et al.
Veröffentlicht: (2025)
Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents
von: Mi, Qirui, et al.
Veröffentlicht: (2026)
von: Mi, Qirui, et al.
Veröffentlicht: (2026)
ARMADA: Attribute-Based Multimodal Data Augmentation
von: Jin, Xiaomeng, et al.
Veröffentlicht: (2024)
von: Jin, Xiaomeng, et al.
Veröffentlicht: (2024)
On the Reliability of Computer Use Agents
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2026)
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2026)
Terminal-World: Scaling Terminal-Agent Environments via Agent Skills
von: Cheng, Zihao, et al.
Veröffentlicht: (2026)
von: Cheng, Zihao, et al.
Veröffentlicht: (2026)
Skill-R1: Agent Skill Evolution via Reinforcement Learning
von: Vishe, Yash, et al.
Veröffentlicht: (2026)
von: Vishe, Yash, et al.
Veröffentlicht: (2026)
Fire360: A Benchmark for Robust Perception and Episodic Memory in Degraded 360-Degree Firefighting Videos
von: Tiwari, Aditi, et al.
Veröffentlicht: (2025)
von: Tiwari, Aditi, et al.
Veröffentlicht: (2025)
SoK: Agentic Skills -- Beyond Tool Use in LLM Agents
von: Jiang, Yanna, et al.
Veröffentlicht: (2026)
von: Jiang, Yanna, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing
von: Qian, Cheng, et al.
Veröffentlicht: (2026) -
Augmenting Interface Usability Heuristics for Reliable Computer-Use Agents
von: Liu, Jiateng, et al.
Veröffentlicht: (2026) -
Infogent: An Agent-Based Framework for Web Information Aggregation
von: Reddy, Revanth Gangi, et al.
Veröffentlicht: (2024) -
Open Vocabulary Electroencephalography-To-Text Decoding and Zero-shot Sentiment Classification
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2021) -
Context Engineering for Trustworthiness: Rescorla Wagner Steering Under Mixed and Inappropriate Contexts
von: Wang, Rushi, et al.
Veröffentlicht: (2025)