Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Chaoran, Zhang, Zhiping, Khalilov, Ibrahim, Guo, Bingcan, Gebreegziabher, Simret A, Ye, Yanfang, Xiao, Ziang, Yao, Yaxing, Li, Tianshi, Li, Toby Jia-Jun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Obvious Invisible Threat: LLM-Powered GUI Agents' Vulnerability to Fine-Print Injections
by: Chen, Chaoran, et al.
Published: (2025)
by: Chen, Chaoran, et al.
Published: (2025)
Dark Patterns Meet GUI Agents: LLM Agent Susceptibility to Manipulative Interfaces and the Role of Human Oversight
by: Tang, Jingyu, et al.
Published: (2025)
by: Tang, Jingyu, et al.
Published: (2025)
Comparing Human Oversight Strategies for Computer-Use Agents
by: Chen, Chaoran, et al.
Published: (2026)
by: Chen, Chaoran, et al.
Published: (2026)
Beyond Permissions: Investigating Mobile Personalization with Simulated Personas
by: Khalilov, Ibrahim, et al.
Published: (2025)
by: Khalilov, Ibrahim, et al.
Published: (2025)
PriviSense: A Frida-Based Framework for Multi-Sensor Spoofing on Android
by: Khalilov, Ibrahim, et al.
Published: (2026)
by: Khalilov, Ibrahim, et al.
Published: (2026)
The Behavioral Fabric of LLM-Powered GUI Agents: Human Values and Interaction Outcomes
by: Gebreegziabher, Simret Araya, et al.
Published: (2026)
by: Gebreegziabher, Simret Araya, et al.
Published: (2026)
CLEAR: Towards Contextual LLM-Empowered Privacy Policy Analysis and Risk Generation for Large Language Model Applications
by: Chen, Chaoran, et al.
Published: (2024)
by: Chen, Chaoran, et al.
Published: (2024)
Privacy Leakage Overshadowed by Views of AI: A Study on Human Oversight of Privacy in Language Model Agent
by: Zhang, Zhiping, et al.
Published: (2024)
by: Zhang, Zhiping, et al.
Published: (2024)
Stayin' Aligned Over Time: Towards Longitudinal Human-LLM Alignment via Contextual Reflection and Privacy-Preserving Behavioral Data
by: Gebreegziabher, Simret Araya, et al.
Published: (2026)
by: Gebreegziabher, Simret Araya, et al.
Published: (2026)
Why am I seeing this: Democratizing End User Auditing for Online Content Recommendations
by: Chen, Chaoran, et al.
Published: (2024)
by: Chen, Chaoran, et al.
Published: (2024)
Supporting Co-Adaptive Machine Teaching through Human Concept Learning and Cognitive Theories
by: Gebreegziabher, Simret Araya, et al.
Published: (2024)
by: Gebreegziabher, Simret Araya, et al.
Published: (2024)
An Empathy-Based Sandbox Approach to Bridge the Privacy Gap among Attitudes, Goals, Knowledge, and Behaviors
by: Chen, Chaoran, et al.
Published: (2023)
by: Chen, Chaoran, et al.
Published: (2023)
A Taxonomy for Human-LLM Interaction Modes: An Initial Exploration
by: Gao, Jie, et al.
Published: (2024)
by: Gao, Jie, et al.
Published: (2024)
Leveraging Variation Theory in Counterfactual Data Augmentation for Optimized Active Learning
by: Gebreegziabher, Simret Araya, et al.
Published: (2024)
by: Gebreegziabher, Simret Araya, et al.
Published: (2024)
Designing Staged Evaluation Workflows for LLMs: Integrating Domain Experts, Lay Users, and Model-Generated Evaluation Criteria
by: Szymanski, Annalisa, et al.
Published: (2024)
by: Szymanski, Annalisa, et al.
Published: (2024)
Towards a Design Guideline for RPA Evaluation: A Survey of Large Language Model-Based Role-Playing Agents
by: Chen, Chaoran, et al.
Published: (2025)
by: Chen, Chaoran, et al.
Published: (2025)
My Favorite Streamer is an LLM: Discovering, Bonding, and Co-Creating in AI VTuber Fandom
by: Ye, Jiayi, et al.
Published: (2025)
by: Ye, Jiayi, et al.
Published: (2025)
From Human-Human Collaboration to Human-Agent Collaboration: A Vision, Design Philosophy, and an Empirical Framework for Achieving Successful Partnerships Between Humans and LLM Agents
by: Yao, Bingsheng, et al.
Published: (2026)
by: Yao, Bingsheng, et al.
Published: (2026)
CoCo Matrix: Taxonomy of Cognitive Contributions in Co-writing with Intelligent Agents
by: Wan, Ruyuan, et al.
Published: (2024)
by: Wan, Ruyuan, et al.
Published: (2024)
Through the Lens of Human-Human Collaboration: A Configurable Research Platform for Exploring Human-Agent Collaboration
by: Yao, Bingsheng, et al.
Published: (2025)
by: Yao, Bingsheng, et al.
Published: (2025)
MultEval: Supporting Collaborative Alignment for LLM-as-a-Judge Evaluation Criteria
by: Chiang, Charles, et al.
Published: (2026)
by: Chiang, Charles, et al.
Published: (2026)
Not My Agent, Not My Boundary? Elicitation of Personal Privacy Boundaries in AI-Delegated Information Sharing
by: Guo, Bingcan, et al.
Published: (2025)
by: Guo, Bingcan, et al.
Published: (2025)
Human-Centered Privacy Research in the Age of Large Language Models
by: Li, Tianshi, et al.
Published: (2024)
by: Li, Tianshi, et al.
Published: (2024)
Towards Human-Centered RegTech: Unpacking Professionals' Strategies and Needs for Using LLMs Safely
by: Hu, Siying, et al.
Published: (2025)
by: Hu, Siying, et al.
Published: (2025)
Autonomy Reshapes How Personalization Affects Privacy Concerns and Trust in LLM Agents
by: Zhang, Zhiping, et al.
Published: (2025)
by: Zhang, Zhiping, et al.
Published: (2025)
Secret Use of Large Language Model (LLM)
by: Zhang, Zhiping, et al.
Published: (2024)
by: Zhang, Zhiping, et al.
Published: (2024)
"I'm categorizing LLM as a productivity tool": Examining ethics of LLM use in HCI research practices
by: Kapania, Shivani, et al.
Published: (2024)
by: Kapania, Shivani, et al.
Published: (2024)
Disclose with Care: Designing Privacy Controls in Interview Chatbots
by: Li, Ziwen, et al.
Published: (2026)
by: Li, Ziwen, et al.
Published: (2026)
VeriOS: Query-Driven Proactive Human-Agent-GUI Interaction for Trustworthy OS Agents
by: Wu, Zheng, et al.
Published: (2025)
by: Wu, Zheng, et al.
Published: (2025)
From Fragmentation to Integration: Exploring the Design Space of AI Agents for Human-as-the-Unit Privacy Management
by: Xu, Eryue, et al.
Published: (2026)
by: Xu, Eryue, et al.
Published: (2026)
LLM-Powered GUI Agents in Phone Automation: Surveying Progress and Prospects
by: Liu, Guangyi, et al.
Published: (2025)
by: Liu, Guangyi, et al.
Published: (2025)
"It's a Fair Game", or Is It? Examining How Users Navigate Disclosure Risks and Benefits When Using LLM-Based Conversational Agents
by: Zhang, Zhiping, et al.
Published: (2023)
by: Zhang, Zhiping, et al.
Published: (2023)
Careful About What App Promotion Ads Recommend! Detecting and Explaining Malware Promotion via App Promotion Graph
by: Ma, Shang, et al.
Published: (2024)
by: Ma, Shang, et al.
Published: (2024)
From Awareness to Action: Exploring End-User Empowerment Interventions for Dark Patterns in UX
by: Lu, Yuwen, et al.
Published: (2023)
by: Lu, Yuwen, et al.
Published: (2023)
TrustAgent: Towards Safe and Trustworthy LLM-based Agents
by: Hua, Wenyue, et al.
Published: (2024)
by: Hua, Wenyue, et al.
Published: (2024)
Agentic LLMs as Powerful Deanonymizers: Re-identification of Participants in the Anthropic Interviewer Dataset
by: Li, Tianshi
Published: (2026)
by: Li, Tianshi
Published: (2026)
Applying LLM-Powered Virtual Humans to Child Interviews in Child-Centered Design
by: Li, Linshi, et al.
Published: (2025)
by: Li, Linshi, et al.
Published: (2025)
VisualTrap: A Stealthy Backdoor Attack on GUI Agents via Visual Grounding Manipulation
by: Ye, Ziang, et al.
Published: (2025)
by: Ye, Ziang, et al.
Published: (2025)
Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining
by: Xiong, Weimin, et al.
Published: (2026)
by: Xiong, Weimin, et al.
Published: (2026)
A Survey on (M)LLM-Based GUI Agents
by: Tang, Fei, et al.
Published: (2025)
by: Tang, Fei, et al.
Published: (2025)
Similar Items
-
The Obvious Invisible Threat: LLM-Powered GUI Agents' Vulnerability to Fine-Print Injections
by: Chen, Chaoran, et al.
Published: (2025) -
Dark Patterns Meet GUI Agents: LLM Agent Susceptibility to Manipulative Interfaces and the Role of Human Oversight
by: Tang, Jingyu, et al.
Published: (2025) -
Comparing Human Oversight Strategies for Computer-Use Agents
by: Chen, Chaoran, et al.
Published: (2026) -
Beyond Permissions: Investigating Mobile Personalization with Simulated Personas
by: Khalilov, Ibrahim, et al.
Published: (2025) -
PriviSense: A Frida-Based Framework for Multi-Sensor Spoofing on Android
by: Khalilov, Ibrahim, et al.
Published: (2026)