SPHERE: An Evaluation Card for Human-AI Systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Ma, Qianou, Zhao, Dora, Zhao, Xinran, Si, Chenglei, Yang, Chenyang, Louie, Ryan, Reiter, Ehud, Yang, Diyi, Wu, Tongshuang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
What Should We Engineer in Prompts? Training Humans in Requirement-Driven LLM Use
di: Ma, Qianou, et al.
Pubblicazione: (2024)
di: Ma, Qianou, et al.
Pubblicazione: (2024)
Not Everyone Wins with LLMs: Behavioral Patterns and Pedagogical Implications for AI Literacy in Programmatic Data Science
di: Ma, Qianou, et al.
Pubblicazione: (2025)
di: Ma, Qianou, et al.
Pubblicazione: (2025)
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
di: Si, Chenglei, et al.
Pubblicazione: (2025)
di: Si, Chenglei, et al.
Pubblicazione: (2025)
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
di: Si, Chenglei, et al.
Pubblicazione: (2024)
di: Si, Chenglei, et al.
Pubblicazione: (2024)
How to Teach Programming in the AI Era? Using LLMs as a Teachable Agent for Debugging
di: Ma, Qianou, et al.
Pubblicazione: (2023)
di: Ma, Qianou, et al.
Pubblicazione: (2023)
RECAP: An End-to-End Platform for Capturing, Replaying, and Analyzing AI-Assisted Programming Interactions
di: He, Keyu, et al.
Pubblicazione: (2026)
di: He, Keyu, et al.
Pubblicazione: (2026)
Knoll: Creating a Knowledge Ecosystem for Large Language Models
di: Zhao, Dora, et al.
Pubblicazione: (2025)
di: Zhao, Dora, et al.
Pubblicazione: (2025)
Evaluation of Human-Understandability of Global Model Explanations using Decision Tree
di: Sivaprasad, Adarsa, et al.
Pubblicazione: (2023)
di: Sivaprasad, Adarsa, et al.
Pubblicazione: (2023)
The Rise of AI Companions: Interaction with AI Companions and Psychological Well-being
di: Zhang, Yutong, et al.
Pubblicazione: (2025)
di: Zhang, Yutong, et al.
Pubblicazione: (2025)
From Prompts to Reflection: Designing Reflective Play for GenAI Literacy
di: Ma, Qianou, et al.
Pubblicazione: (2025)
di: Ma, Qianou, et al.
Pubblicazione: (2025)
Large Language Models Help Humans Verify Truthfulness -- Except When They Are Convincingly Wrong
di: Si, Chenglei, et al.
Pubblicazione: (2023)
di: Si, Chenglei, et al.
Pubblicazione: (2023)
Behavior Latticing: Inferring User Motivations from Unstructured Interactions
di: Zhao, Dora, et al.
Pubblicazione: (2026)
di: Zhao, Dora, et al.
Pubblicazione: (2026)
Can LLM-Simulated Practice and Feedback Upskill Human Counselors? A Randomized Study with 90+ Novice Counselors
di: Louie, Ryan, et al.
Pubblicazione: (2025)
di: Louie, Ryan, et al.
Pubblicazione: (2025)
Collaborative Gym: A Framework for Enabling and Evaluating Human-Agent Collaboration
di: Shao, Yijia, et al.
Pubblicazione: (2024)
di: Shao, Yijia, et al.
Pubblicazione: (2024)
Orbit: A Framework for Designing and Evaluating Multi-objective Rankers
di: Yang, Chenyang, et al.
Pubblicazione: (2024)
di: Yang, Chenyang, et al.
Pubblicazione: (2024)
Wikibench: Community-Driven Data Curation for AI Evaluation on Wikipedia
di: Kuo, Tzu-Sheng, et al.
Pubblicazione: (2024)
di: Kuo, Tzu-Sheng, et al.
Pubblicazione: (2024)
Model Cards for AI Teammates: Comparing Human-AI Team Familiarization Methods for High-Stakes Environments
di: Bowers, Ryan, et al.
Pubblicazione: (2025)
di: Bowers, Ryan, et al.
Pubblicazione: (2025)
How Do AI Agents Do Human Work? Comparing AI and Human Workflows Across Diverse Occupations
di: Wang, Zora Zhiruo, et al.
Pubblicazione: (2025)
di: Wang, Zora Zhiruo, et al.
Pubblicazione: (2025)
Mapping the Spiral of Silence: Surveying Unspoken Opinions in Online Communities
di: Zhao, Dora, et al.
Pubblicazione: (2025)
di: Zhao, Dora, et al.
Pubblicazione: (2025)
When LLMs Can't Help: Real-World Evaluation of LLMs in Nutrition
di: Li, Karen Jia-Hui, et al.
Pubblicazione: (2025)
di: Li, Karen Jia-Hui, et al.
Pubblicazione: (2025)
Position: Towards Bidirectional Human-AI Alignment
di: Shen, Hua, et al.
Pubblicazione: (2024)
di: Shen, Hua, et al.
Pubblicazione: (2024)
Roleplay-doh: Enabling Domain-Experts to Create LLM-simulated Patients via Eliciting and Adhering to Principles
di: Louie, Ryan, et al.
Pubblicazione: (2024)
di: Louie, Ryan, et al.
Pubblicazione: (2024)
When Models Know More Than They Can Explain: Quantifying Knowledge Transfer in Human-AI Collaboration
di: Shi, Quan, et al.
Pubblicazione: (2025)
di: Shi, Quan, et al.
Pubblicazione: (2025)
Human-AI Interaction Design Standards
di: Zhao, Chaoyi, et al.
Pubblicazione: (2025)
di: Zhao, Chaoyi, et al.
Pubblicazione: (2025)
AI-Induced Human Responsibility (AIHR) in AI-Human teams
di: Nyilasy, Greg, et al.
Pubblicazione: (2026)
di: Nyilasy, Greg, et al.
Pubblicazione: (2026)
Towards Human-AI Deliberation: Design and Evaluation of LLM-Empowered Deliberative AI for AI-Assisted Decision-Making
di: Ma, Shuai, et al.
Pubblicazione: (2024)
di: Ma, Shuai, et al.
Pubblicazione: (2024)
The AI-DEC: A Card-based Design Method for User-centered AI Explanations
di: Lee, Christine P, et al.
Pubblicazione: (2024)
di: Lee, Christine P, et al.
Pubblicazione: (2024)
SparkMe: Adaptive Semi-Structured Interviewing for Qualitative Insight Discovery
di: Anugraha, David, et al.
Pubblicazione: (2026)
di: Anugraha, David, et al.
Pubblicazione: (2026)
Generative Experiences for Digital Mental Health Interventions: Evidence from a Randomized Study
di: Bhattacharjee, Ananya, et al.
Pubblicazione: (2026)
di: Bhattacharjee, Ananya, et al.
Pubblicazione: (2026)
Learning to Trust: How Humans Mentally Recalibrate AI Confidence Signals
di: Li, ZhaoBin, et al.
Pubblicazione: (2026)
di: Li, ZhaoBin, et al.
Pubblicazione: (2026)
Classifying Epistemic Relationships in Human-AI Interaction: An Exploratory Approach
di: Yang, Shengnan, et al.
Pubblicazione: (2025)
di: Yang, Shengnan, et al.
Pubblicazione: (2025)
A Two-Phase Visualization System for Continuous Human-AI Collaboration in Sequelae Analysis and Modeling
di: Ouyang, Yang, et al.
Pubblicazione: (2024)
di: Ouyang, Yang, et al.
Pubblicazione: (2024)
As Confidence Aligns: Exploring the Effect of AI Confidence on Human Self-confidence in Human-AI Decision Making
di: Li, Jingshu, et al.
Pubblicazione: (2025)
di: Li, Jingshu, et al.
Pubblicazione: (2025)
Whose Knowledge Counts? Co-Designing Community-Centered AI Auditing Tools with Educators in Hawai`i
di: Zhao, Dora, et al.
Pubblicazione: (2026)
di: Zhao, Dora, et al.
Pubblicazione: (2026)
Multi-Level Feedback Generation with Large Language Models for Empowering Novice Peer Counselors
di: Chaszczewicz, Alicja, et al.
Pubblicazione: (2024)
di: Chaszczewicz, Alicja, et al.
Pubblicazione: (2024)
Human-AI Interaction Alignment: Designing, Evaluating, and Evolving Value-Centered AI For Reciprocal Human-AI Futures
di: Shen, Hua, et al.
Pubblicazione: (2025)
di: Shen, Hua, et al.
Pubblicazione: (2025)
Human-AI Co-Evolution and Epistemic Collapse: A Dynamical Systems Perspective
di: Wu, Xuening, et al.
Pubblicazione: (2026)
di: Wu, Xuening, et al.
Pubblicazione: (2026)
Selenite: Scaffolding Online Sensemaking with Comprehensive Overviews Elicited from Large Language Models
di: Liu, Michael Xieyang, et al.
Pubblicazione: (2023)
di: Liu, Michael Xieyang, et al.
Pubblicazione: (2023)
Interoceptive Divergence in Aesthetic Evaluation and Implications for Human-AI Alignment
di: Abe, Yoshia, et al.
Pubblicazione: (2026)
di: Abe, Yoshia, et al.
Pubblicazione: (2026)
Evaluating Human-AI Collaboration: A Review and Methodological Framework
di: Fragiadakis, George, et al.
Pubblicazione: (2024)
di: Fragiadakis, George, et al.
Pubblicazione: (2024)
Documenti analoghi
-
What Should We Engineer in Prompts? Training Humans in Requirement-Driven LLM Use
di: Ma, Qianou, et al.
Pubblicazione: (2024) -
Not Everyone Wins with LLMs: Behavioral Patterns and Pedagogical Implications for AI Literacy in Programmatic Data Science
di: Ma, Qianou, et al.
Pubblicazione: (2025) -
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
di: Si, Chenglei, et al.
Pubblicazione: (2025) -
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
di: Si, Chenglei, et al.
Pubblicazione: (2024) -
How to Teach Programming in the AI Era? Using LLMs as a Teachable Agent for Debugging
di: Ma, Qianou, et al.
Pubblicazione: (2023)