Evaluating Human Trust in LLM-Based Planners: A Preliminary Study
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Shenghui, Yang, Yunhao, Boggess, Kayla, Heo, Seongkook, Feng, Lu, Topcu, Ufuk |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Human-Agent Cooperation in Games under Incomplete Information through Natural Language Communication
by: Chen, Shenghui, et al.
Published: (2024)
by: Chen, Shenghui, et al.
Published: (2024)
Human-Agent Coordination in Games under Incomplete Information via Multi-Step Intent
by: Chen, Shenghui, et al.
Published: (2024)
by: Chen, Shenghui, et al.
Published: (2024)
ViSIL: Unified Evaluation of Information Loss in Multimodal Video Captioning
by: Li, Po-han, et al.
Published: (2026)
by: Li, Po-han, et al.
Published: (2026)
Preliminary Quantitative Study on Explainability and Trust in AI Systems
by: Sunny, Allen Daniel
Published: (2025)
by: Sunny, Allen Daniel
Published: (2025)
When Trust Collides: Decoding Human-LLM Cooperation Dynamics through the Prisoner's Dilemma
by: Jiang, Guanxuan, et al.
Published: (2025)
by: Jiang, Guanxuan, et al.
Published: (2025)
Engaging with AI: How Interface Design Shapes Human-AI Collaboration in High-Stakes Decision-Making
by: Chen, Zichen, et al.
Published: (2025)
by: Chen, Zichen, et al.
Published: (2025)
Human and AI Trust: Trust Attitude Measurement Instrument
by: Larasati, Retno
Published: (2025)
by: Larasati, Retno
Published: (2025)
Evaluating Trust in AI, Human, and Co-produced Feedback Among Undergraduate Students
by: Zhang, Audrey, et al.
Published: (2025)
by: Zhang, Audrey, et al.
Published: (2025)
Thing2Reality: Transforming 2D Content into Conditioned Multiviews and 3D Gaussian Objects for XR Communication
by: Hu, Erzhen, et al.
Published: (2024)
by: Hu, Erzhen, et al.
Published: (2024)
An Epistemic Human-Aware Task Planner which Anticipates Human Beliefs and Decisions
by: Shekhar, Shashank, et al.
Published: (2024)
by: Shekhar, Shashank, et al.
Published: (2024)
VIBE: Annotation-Free Video-to-Text Information Bottleneck Evaluation for TL;DR
by: Chen, Shenghui, et al.
Published: (2025)
by: Chen, Shenghui, et al.
Published: (2025)
An Empirical Exploration of Trust Dynamics in LLM Supply Chains
by: Balayn, Agathe, et al.
Published: (2024)
by: Balayn, Agathe, et al.
Published: (2024)
Anthropomorphism and Trust in Human-Large Language Model interactions
by: Kadambi, Akila, et al.
Published: (2026)
by: Kadambi, Akila, et al.
Published: (2026)
Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human Preferences
by: Shankar, Shreya, et al.
Published: (2024)
by: Shankar, Shreya, et al.
Published: (2024)
Predicting Trust Dynamics with Dynamic SEM in Human-AI Cooperation
by: Kaneko, Sota, et al.
Published: (2024)
by: Kaneko, Sota, et al.
Published: (2024)
"Trust me on this" Explaining Agent Behavior to a Human Terminator
by: Menkes, Uri, et al.
Published: (2025)
by: Menkes, Uri, et al.
Published: (2025)
Human vs. LLM-Based Thematic Analysis for Digital Mental Health Research: Proof-of-Concept Comparative Study
by: Parkington, Karisa, et al.
Published: (2025)
by: Parkington, Karisa, et al.
Published: (2025)
Assessing the Human-Likeness of LLM-Driven Digital Twins in Simulating Health Care System Trust
by: Wu, Yuzhou, et al.
Published: (2025)
by: Wu, Yuzhou, et al.
Published: (2025)
Learning to Trust: How Humans Mentally Recalibrate AI Confidence Signals
by: Li, ZhaoBin, et al.
Published: (2026)
by: Li, ZhaoBin, et al.
Published: (2026)
The Moral Turing Test: Evaluating Human-LLM Alignment in Moral Decision-Making
by: Garcia, Basile, et al.
Published: (2024)
by: Garcia, Basile, et al.
Published: (2024)
Generate, Evaluate, Iterate: Synthetic Data for Human-in-the-Loop Refinement of LLM Judges
by: Do, Hyo Jin, et al.
Published: (2025)
by: Do, Hyo Jin, et al.
Published: (2025)
Understanding Human-AI Trust in Education
by: Pitts, Griffin, et al.
Published: (2025)
by: Pitts, Griffin, et al.
Published: (2025)
Towards Human-AI Deliberation: Design and Evaluation of LLM-Empowered Deliberative AI for AI-Assisted Decision-Making
by: Ma, Shuai, et al.
Published: (2024)
by: Ma, Shuai, et al.
Published: (2024)
The Persuasion Paradox: When LLM Explanations Fail to Improve Human-AI Team Performance
by: Cohen, Ruth, et al.
Published: (2026)
by: Cohen, Ruth, et al.
Published: (2026)
From Verification Burden to Trusted Collaboration: Design Goals for LLM-Assisted Literature Reviews
by: Nogueira, Brenda, et al.
Published: (2025)
by: Nogueira, Brenda, et al.
Published: (2025)
Generative Students: Using LLM-Simulated Student Profiles to Support Question Item Evaluation
by: Lu, Xinyi, et al.
Published: (2024)
by: Lu, Xinyi, et al.
Published: (2024)
Not All Trust is the Same: Effects of Decision Workflow and Explanations in Human-AI Decision Making
by: Spillner, Laura, et al.
Published: (2026)
by: Spillner, Laura, et al.
Published: (2026)
Assessing the Quality of Mental Health Support in LLM Responses through Multi-Attribute Human Evaluation
by: Badawi, Abeer, et al.
Published: (2026)
by: Badawi, Abeer, et al.
Published: (2026)
Beyond correlation: The Impact of Human Uncertainty in Measuring the Effectiveness of Automatic Evaluation and LLM-as-a-Judge
by: Elangovan, Aparna, et al.
Published: (2024)
by: Elangovan, Aparna, et al.
Published: (2024)
SPHERE: An Evaluation Card for Human-AI Systems
by: Ma, Qianou, et al.
Published: (2025)
by: Ma, Qianou, et al.
Published: (2025)
Human-Centered Evaluation of an LLM-Based Process Modeling Copilot: A Mixed-Methods Study with Domain Experts
by: Lauer, Chantale, et al.
Published: (2026)
by: Lauer, Chantale, et al.
Published: (2026)
Enhancing Human Experience in Human-Agent Collaboration: A Human-Centered Modeling Approach Based on Positive Human Gain
by: Gao, Yiming, et al.
Published: (2024)
by: Gao, Yiming, et al.
Published: (2024)
Building Trust in Mental Health Chatbots: Safety Metrics and LLM-Based Evaluation Tools
by: Park, Jung In, et al.
Published: (2024)
by: Park, Jung In, et al.
Published: (2024)
AgentLens: Visual Analysis for Agent Behaviors in LLM-based Autonomous Systems
by: Lu, Jiaying, et al.
Published: (2024)
by: Lu, Jiaying, et al.
Published: (2024)
A Multi-Layered Research Framework for Human-Centered AI: Defining the Path to Explainability and Trust
by: De Silva, Chameera, et al.
Published: (2025)
by: De Silva, Chameera, et al.
Published: (2025)
Designing and Evaluating Multi-Chatbot Interface for Human-AI Communication: Preliminary Findings from a Persuasion Task
by: Yoon, Sion, et al.
Published: (2024)
by: Yoon, Sion, et al.
Published: (2024)
Key Considerations for Domain Expert Involvement in LLM Design and Evaluation: An Ethnographic Study
by: Szymanski, Annalisa, et al.
Published: (2026)
by: Szymanski, Annalisa, et al.
Published: (2026)
VizTrust: A Visual Analytics Tool for Capturing User Trust Dynamics in Human-AI Communication
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
Can Large Language Model Agents Simulate Human Trust Behavior?
by: Xie, Chengxing, et al.
Published: (2024)
by: Xie, Chengxing, et al.
Published: (2024)
InFerActive: Interactive Tree-Based Exploration of LLM Sampling for Safety Evaluation
by: Hwangbo, Junhyeong, et al.
Published: (2025)
by: Hwangbo, Junhyeong, et al.
Published: (2025)
Similar Items
-
Human-Agent Cooperation in Games under Incomplete Information through Natural Language Communication
by: Chen, Shenghui, et al.
Published: (2024) -
Human-Agent Coordination in Games under Incomplete Information via Multi-Step Intent
by: Chen, Shenghui, et al.
Published: (2024) -
ViSIL: Unified Evaluation of Information Loss in Multimodal Video Captioning
by: Li, Po-han, et al.
Published: (2026) -
Preliminary Quantitative Study on Explainability and Trust in AI Systems
by: Sunny, Allen Daniel
Published: (2025) -
When Trust Collides: Decoding Human-LLM Cooperation Dynamics through the Prisoner's Dilemma
by: Jiang, Guanxuan, et al.
Published: (2025)