Beyond Benchmarks: How Users Evaluate AI Chat Assistants
Fuente:
arXiv
Guardado en:
| Autores principales: | Awan, Moiz Sadiq, Noor, Muhammad Haris, Munaf, Muhammad Salman |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Acceptability of AI Assistants for Privacy: Perceptions of Experts and Users on Personalized Privacy Assistants
por: Xu, Meihe, et al.
Publicado: (2025)
por: Xu, Meihe, et al.
Publicado: (2025)
Exploring Culturally Informed AI Assistants: A Comparative Study of ChatBlackGPT and ChatGPT
por: Egede, Lisa, et al.
Publicado: (2025)
por: Egede, Lisa, et al.
Publicado: (2025)
OfficeMate: Pilot Evaluation of an Office Assistant Robot
por: Pan, Jiahe, et al.
Publicado: (2025)
por: Pan, Jiahe, et al.
Publicado: (2025)
Beyond the Hype: Mapping Uncertainty and Gratification in AI Assistant Use
por: Joy, Karen, et al.
Publicado: (2025)
por: Joy, Karen, et al.
Publicado: (2025)
Evaluation and Continual Improvement for an Enterprise AI Assistant
por: Maharaj, Akash V., et al.
Publicado: (2024)
por: Maharaj, Akash V., et al.
Publicado: (2024)
Generating User Experience Based on Personas with AI Assistants
por: Huang, Yutan
Publicado: (2024)
por: Huang, Yutan
Publicado: (2024)
From Use to Oversight: How Mental Models Influence User Behavior and Output in AI Writing Assistants
por: Rismani, Shalaleh, et al.
Publicado: (2026)
por: Rismani, Shalaleh, et al.
Publicado: (2026)
LLM-Glasses: GenAI-driven Glasses with Haptic Feedback for Navigation of Visually Impaired People
por: Tokmurziyev, Issatay, et al.
Publicado: (2025)
por: Tokmurziyev, Issatay, et al.
Publicado: (2025)
MUIAnno: An Expert-Annotated Dataset and Evaluation Benchmark for Mobile UI Understanding
por: Parvez, Athar, et al.
Publicado: (2026)
por: Parvez, Athar, et al.
Publicado: (2026)
Thinking Assistants: LLM-Based Conversational Assistants that Help Users Think By Asking rather than Answering
por: Park, Soya, et al.
Publicado: (2023)
por: Park, Soya, et al.
Publicado: (2023)
Bowling with ChatGPT: On the Evolving User Interactions with Conversational AI Systems
por: Karnam, Sai Keerthana, et al.
Publicado: (2026)
por: Karnam, Sai Keerthana, et al.
Publicado: (2026)
Users' Perception on Appropriateness of Robotic Coaching Assistant's Disclosure Behaviors
por: Nilgar, Atikkhan Faridkhan, et al.
Publicado: (2024)
por: Nilgar, Atikkhan Faridkhan, et al.
Publicado: (2024)
Learning to Live with AI: How Students Develop AI Literacy Through Naturalistic ChatGPT Interaction
por: Ammari, Tawfiq, et al.
Publicado: (2026)
por: Ammari, Tawfiq, et al.
Publicado: (2026)
Exploring the Impact of Word Prediction Assistive Features on Smartphone Keyboards for Blind Users
por: Alnfiai, Mrim M., et al.
Publicado: (2024)
por: Alnfiai, Mrim M., et al.
Publicado: (2024)
Trust in Transparency: How Explainable AI Shapes User Perceptions
por: Sunny, Allen Daniel
Publicado: (2025)
por: Sunny, Allen Daniel
Publicado: (2025)
User Interaction Patterns and Breakdowns in Conversing with LLM-Powered Voice Assistants
por: Mahmood, Amama, et al.
Publicado: (2023)
por: Mahmood, Amama, et al.
Publicado: (2023)
Design and Evaluation of Generative Agent-based Platform for Human-Assistant Interaction Research: A Tale of 10 User Studies
por: Xuan, Ziyi, et al.
Publicado: (2025)
por: Xuan, Ziyi, et al.
Publicado: (2025)
The Illusion of Agreement with ChatGPT: Sycophancy and Beyond
por: Noshin, Kazi, et al.
Publicado: (2026)
por: Noshin, Kazi, et al.
Publicado: (2026)
How University Disability Services Professionals Write Image Descriptions for HCI Figures Using Generative AI
por: Raees, Muhammad, et al.
Publicado: (2026)
por: Raees, Muhammad, et al.
Publicado: (2026)
Walkthrough of Anthropomorphic Features in AI Assistant Tools
por: Maeda, Takuya
Publicado: (2025)
por: Maeda, Takuya
Publicado: (2025)
Rethinking AI Evaluation in Education: The TEACH-AI Framework and Benchmark for Generative AI Assistants
por: Ding, Shi, et al.
Publicado: (2025)
por: Ding, Shi, et al.
Publicado: (2025)
Beyond Explicit and Implicit: How Users Provide Feedback to Shape Personalized Recommendation Content
por: Li, Wenqi, et al.
Publicado: (2025)
por: Li, Wenqi, et al.
Publicado: (2025)
"I Use ChatGPT to Humanize My Words": Affordances and Risks of ChatGPT to Autistic Users
por: Ma, Renkai, et al.
Publicado: (2026)
por: Ma, Renkai, et al.
Publicado: (2026)
From Trust to Appropriate Reliance: Measurement Constructs in Human-AI Decision-Making
por: Raees, Muhammad, et al.
Publicado: (2026)
por: Raees, Muhammad, et al.
Publicado: (2026)
ChatBench: From Static Benchmarks to Human-AI Evaluation
por: Chang, Serina, et al.
Publicado: (2025)
por: Chang, Serina, et al.
Publicado: (2025)
Comparing How a Chatbot References User Utterances from Previous Chatting Sessions: An Investigation of Users' Privacy Concerns and Perceptions
por: Cox, Samuel Rhys, et al.
Publicado: (2023)
por: Cox, Samuel Rhys, et al.
Publicado: (2023)
MHDash: An Online Platform for Benchmarking Mental Health-Aware AI Assistants
por: Zhang, Yihe, et al.
Publicado: (2026)
por: Zhang, Yihe, et al.
Publicado: (2026)
User Intent Recognition and Satisfaction with Large Language Models: A User Study with ChatGPT
por: Bodonhelyi, Anna, et al.
Publicado: (2024)
por: Bodonhelyi, Anna, et al.
Publicado: (2024)
Decoding User Concerns in AI Health Chatbots: An Exploration of Security and Privacy in App Reviews
por: Hassan, Muhammad, et al.
Publicado: (2025)
por: Hassan, Muhammad, et al.
Publicado: (2025)
Campus AI vs. Commercial AI: Comparing How Students and Employees Perceive their University's LLM Chatbot vs. ChatGPT
por: Hannig, Leon, et al.
Publicado: (2025)
por: Hannig, Leon, et al.
Publicado: (2025)
MultiVox: A Benchmark for Evaluating Voice Assistants for Multimodal Interactions
por: Selvakumar, Ramaneswaran, et al.
Publicado: (2025)
por: Selvakumar, Ramaneswaran, et al.
Publicado: (2025)
Beyond Tools: Understanding How Heavy Users Integrate LLMs into Everyday Tasks and Decision-Making
por: Kim, Eunhye, et al.
Publicado: (2025)
por: Kim, Eunhye, et al.
Publicado: (2025)
SARA: Smart AI Reading Assistant for Reading Comprehension
por: Thaqi, Enkeleda, et al.
Publicado: (2024)
por: Thaqi, Enkeleda, et al.
Publicado: (2024)
Need Help? Designing Proactive AI Assistants for Programming
por: Chen, Valerie, et al.
Publicado: (2024)
por: Chen, Valerie, et al.
Publicado: (2024)
Evaluation and Incident Prevention in an Enterprise AI Assistant
por: Maharaj, Akash V., et al.
Publicado: (2025)
por: Maharaj, Akash V., et al.
Publicado: (2025)
User-Assistant Bias in LLMs
por: Pan, Xu, et al.
Publicado: (2025)
por: Pan, Xu, et al.
Publicado: (2025)
Patterns of Creativity: How User Input Shapes AI-Generated Visual Diversity
por: Palmini, Maria-Teresa De Rosa, et al.
Publicado: (2024)
por: Palmini, Maria-Teresa De Rosa, et al.
Publicado: (2024)
Musinger: Communication of Music over a Distance with Wearable Haptic Display and Touch Sensitive Surface
por: Cabrera, Miguel Altamirano, et al.
Publicado: (2024)
por: Cabrera, Miguel Altamirano, et al.
Publicado: (2024)
Estimating Perceptual Attributes of Haptic Textures Using Visuo-Tactile Data
por: Awan, Mudassir Ibrahim, et al.
Publicado: (2025)
por: Awan, Mudassir Ibrahim, et al.
Publicado: (2025)
"Mango Mango, How to Let The Lettuce Dry Without A Spinner?": Exploring User Perceptions of Using An LLM-Based Conversational Assistant Toward Cooking Partner
por: Chan, Szeyi, et al.
Publicado: (2023)
por: Chan, Szeyi, et al.
Publicado: (2023)
Ejemplares similares
-
Acceptability of AI Assistants for Privacy: Perceptions of Experts and Users on Personalized Privacy Assistants
por: Xu, Meihe, et al.
Publicado: (2025) -
Exploring Culturally Informed AI Assistants: A Comparative Study of ChatBlackGPT and ChatGPT
por: Egede, Lisa, et al.
Publicado: (2025) -
OfficeMate: Pilot Evaluation of an Office Assistant Robot
por: Pan, Jiahe, et al.
Publicado: (2025) -
Beyond the Hype: Mapping Uncertainty and Gratification in AI Assistant Use
por: Joy, Karen, et al.
Publicado: (2025) -
Evaluation and Continual Improvement for an Enterprise AI Assistant
por: Maharaj, Akash V., et al.
Publicado: (2024)