Towards a Design Guideline for RPA Evaluation: A Survey of Large Language Model-Based Role-Playing Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Chaoran, Yao, Bingsheng, Zou, Ruishi, Hua, Wenyue, Lyu, Weimin, Ye, Yanfang, Li, Toby Jia-Jun, Wang, Dakuo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Human-Human Collaboration to Human-Agent Collaboration: A Vision, Design Philosophy, and an Empirical Framework for Achieving Successful Partnerships Between Humans and LLM Agents
by: Yao, Bingsheng, et al.
Published: (2026)
by: Yao, Bingsheng, et al.
Published: (2026)
Through the Lens of Human-Human Collaboration: A Configurable Research Platform for Exploring Human-Agent Collaboration
by: Yao, Bingsheng, et al.
Published: (2025)
by: Yao, Bingsheng, et al.
Published: (2025)
Striking a Balance: Evaluating How Aggregations of Multiple Forecasts Impact Judgment Under Uncertainty
by: Zou, Ruishi, et al.
Published: (2024)
by: Zou, Ruishi, et al.
Published: (2024)
Dark Patterns Meet GUI Agents: LLM Agent Susceptibility to Manipulative Interfaces and the Role of Human Oversight
by: Tang, Jingyu, et al.
Published: (2025)
by: Tang, Jingyu, et al.
Published: (2025)
CLEAR: Towards Contextual LLM-Empowered Privacy Policy Analysis and Risk Generation for Large Language Model Applications
by: Chen, Chaoran, et al.
Published: (2024)
by: Chen, Chaoran, et al.
Published: (2024)
UXAgent: An LLM Agent-Based Usability Testing Framework for Web Design
by: Lu, Yuxuan, et al.
Published: (2025)
by: Lu, Yuxuan, et al.
Published: (2025)
UXAgent: A System for Simulating Usability Testing of Web Design with LLM Agents
by: Lu, Yuxuan, et al.
Published: (2025)
by: Lu, Yuxuan, et al.
Published: (2025)
My Favorite Streamer is an LLM: Discovering, Bonding, and Co-Creating in AI VTuber Fandom
by: Ye, Jiayi, et al.
Published: (2025)
by: Ye, Jiayi, et al.
Published: (2025)
An Empathy-Based Sandbox Approach to Bridge the Privacy Gap among Attitudes, Goals, Knowledge, and Behaviors
by: Chen, Chaoran, et al.
Published: (2023)
by: Chen, Chaoran, et al.
Published: (2023)
Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents
by: Chen, Chaoran, et al.
Published: (2025)
by: Chen, Chaoran, et al.
Published: (2025)
Why am I seeing this: Democratizing End User Auditing for Online Content Recommendations
by: Chen, Chaoran, et al.
Published: (2024)
by: Chen, Chaoran, et al.
Published: (2024)
Comparing Human Oversight Strategies for Computer-Use Agents
by: Chen, Chaoran, et al.
Published: (2026)
by: Chen, Chaoran, et al.
Published: (2026)
The Obvious Invisible Threat: LLM-Powered GUI Agents' Vulnerability to Fine-Print Injections
by: Chen, Chaoran, et al.
Published: (2025)
by: Chen, Chaoran, et al.
Published: (2025)
User Interaction Patterns and Breakdowns in Conversing with LLM-Powered Voice Assistants
by: Mahmood, Amama, et al.
Published: (2023)
by: Mahmood, Amama, et al.
Published: (2023)
"I Like Sunnie More Than I Expected!": Exploring User Expectation and Perception of an Anthropomorphic LLM-based Conversational Agent for Well-Being Support
by: Wu, Siyi, et al.
Published: (2024)
by: Wu, Siyi, et al.
Published: (2024)
Secret Use of Large Language Model (LLM)
by: Zhang, Zhiping, et al.
Published: (2024)
by: Zhang, Zhiping, et al.
Published: (2024)
Bridging Knowledge Gaps in Clinical AI: An Activity Theory Perspective on Interdisciplinary Data Work for Telehealth
by: Yao, Bingsheng, et al.
Published: (2024)
by: Yao, Bingsheng, et al.
Published: (2024)
LLM Agent Meets Agentic AI: Can LLM Agents Simulate Customers to Evaluate Agentic-AI-based Shopping Assistants?
by: Sun, Lu, et al.
Published: (2025)
by: Sun, Lu, et al.
Published: (2025)
AgentA/B: Automated and Scalable Web A/BTesting with Interactive LLM Agents
by: Lu, Yuxuan, et al.
Published: (2025)
by: Lu, Yuxuan, et al.
Published: (2025)
"I Wish There Were an AI": Challenges and AI Potential in Cancer Patient-Provider Communication
by: Yang, Ziqi, et al.
Published: (2024)
by: Yang, Ziqi, et al.
Published: (2024)
Human and LLM-Based Voice Assistant Interaction: An Analytical Framework for User Verbal and Nonverbal Behaviors
by: Chan, Szeyi, et al.
Published: (2024)
by: Chan, Szeyi, et al.
Published: (2024)
"Mango Mango, How to Let The Lettuce Dry Without A Spinner?": Exploring User Perceptions of Using An LLM-Based Conversational Assistant Toward Cooking Partner
by: Chan, Szeyi, et al.
Published: (2023)
by: Chan, Szeyi, et al.
Published: (2023)
Graphing Inline: Understanding Word-scale Graphics Use in Scientific Papers
by: Lu, Siyu, et al.
Published: (2026)
by: Lu, Siyu, et al.
Published: (2026)
Stayin' Aligned Over Time: Towards Longitudinal Human-LLM Alignment via Contextual Reflection and Privacy-Preserving Behavioral Data
by: Gebreegziabher, Simret Araya, et al.
Published: (2026)
by: Gebreegziabher, Simret Araya, et al.
Published: (2026)
Designing Human-AI System for Legal Research: A Case Study of Precedent Search in Chinese Law
by: Guan, Jiarui, et al.
Published: (2025)
by: Guan, Jiarui, et al.
Published: (2025)
"It's a Fair Game", or Is It? Examining How Users Navigate Disclosure Risks and Benefits When Using LLM-Based Conversational Agents
by: Zhang, Zhiping, et al.
Published: (2023)
by: Zhang, Zhiping, et al.
Published: (2023)
Who Changed the Destiny of Rural Students, and How?: Unpacking ICT-Mediated Remote Education in Rural China
by: Sun, Yuling, et al.
Published: (2024)
by: Sun, Yuling, et al.
Published: (2024)
Beyond Permissions: Investigating Mobile Personalization with Simulated Personas
by: Khalilov, Ibrahim, et al.
Published: (2025)
by: Khalilov, Ibrahim, et al.
Published: (2025)
More Modality, More AI: Exploring Design Opportunities of AI-Based Multi-modal Remote Monitoring Technologies for Early Detection of Mental Health Sequelae in Youth Concussion Patients
by: Yao, Bingsheng, et al.
Published: (2025)
by: Yao, Bingsheng, et al.
Published: (2025)
Exploring Parent's Needs for Children-Centered AI to Support Preschoolers' Interactive Storytelling and Reading Activities
by: Sun, Yuling, et al.
Published: (2024)
by: Sun, Yuling, et al.
Published: (2024)
OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
"It Felt Like I Was Left in the Dark": Exploring Information Needs and Design Opportunities for Family Caregivers of Older Adult Patients in Critical Care Settings
by: Fu, Shihan, et al.
Published: (2025)
by: Fu, Shihan, et al.
Published: (2025)
PriviSense: A Frida-Based Framework for Multi-Sensor Spoofing on Android
by: Khalilov, Ibrahim, et al.
Published: (2026)
by: Khalilov, Ibrahim, et al.
Published: (2026)
The Behavioral Fabric of LLM-Powered GUI Agents: Human Values and Interaction Outcomes
by: Gebreegziabher, Simret Araya, et al.
Published: (2026)
by: Gebreegziabher, Simret Araya, et al.
Published: (2026)
Human-Centered Privacy Research in the Age of Large Language Models
by: Li, Tianshi, et al.
Published: (2024)
by: Li, Tianshi, et al.
Published: (2024)
Role-Playing Agents Driven by Large Language Models: Current Status, Challenges, and Future Trends
by: Wang, Ye, et al.
Published: (2026)
by: Wang, Ye, et al.
Published: (2026)
GistVis: Automatic Generation of Word-scale Visualizations from Data-rich Documents
by: Zou, Ruishi, et al.
Published: (2025)
by: Zou, Ruishi, et al.
Published: (2025)
CardioAI: A Multimodal AI-based System to Support Symptom Monitoring and Risk Detection of Cancer Treatment-Induced Cardiotoxicity
by: Wu, Siyi, et al.
Published: (2024)
by: Wu, Siyi, et al.
Published: (2024)
Characterizing LLM-Empowered Personalized Story-Reading and Interaction for Children: Insights from Multi-Stakeholder Perspectives
by: Chen, Jiaju, et al.
Published: (2025)
by: Chen, Jiaju, et al.
Published: (2025)
Clinical Challenges and AI Opportunities in Decision-Making for Cancer Treatment-Induced Cardiotoxicity
by: Wu, Siyi, et al.
Published: (2024)
by: Wu, Siyi, et al.
Published: (2024)
Similar Items
-
From Human-Human Collaboration to Human-Agent Collaboration: A Vision, Design Philosophy, and an Empirical Framework for Achieving Successful Partnerships Between Humans and LLM Agents
by: Yao, Bingsheng, et al.
Published: (2026) -
Through the Lens of Human-Human Collaboration: A Configurable Research Platform for Exploring Human-Agent Collaboration
by: Yao, Bingsheng, et al.
Published: (2025) -
Striking a Balance: Evaluating How Aggregations of Multiple Forecasts Impact Judgment Under Uncertainty
by: Zou, Ruishi, et al.
Published: (2024) -
Dark Patterns Meet GUI Agents: LLM Agent Susceptibility to Manipulative Interfaces and the Role of Human Oversight
by: Tang, Jingyu, et al.
Published: (2025) -
CLEAR: Towards Contextual LLM-Empowered Privacy Policy Analysis and Risk Generation for Large Language Model Applications
by: Chen, Chaoran, et al.
Published: (2024)