The Trust Paradox: How CS Researchers Engage LLM Leaderboards
Fuente:
arXiv
Saved in:
| Main Authors: | Sadeghi, Pouya, Crisan, Anamaria, Lin, Jimmy |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Probing the Visualization Literacy of Vision Language Models: the Good, the Bad, and the Ugly
by: Dong, Lianghan, et al.
Published: (2025)
by: Dong, Lianghan, et al.
Published: (2025)
Linting is People! Exploring the Potential of Human Computation as a Sociotechnical Linter of Data Visualizations
by: Crisan, Anamaria, et al.
Published: (2025)
by: Crisan, Anamaria, et al.
Published: (2025)
AInsight: Augmenting Expert Decision-Making with On-the-Fly Insights Grounded in Historical Data
by: Abolnejadian, Mohammad, et al.
Published: (2025)
by: Abolnejadian, Mohammad, et al.
Published: (2025)
Social Welfare Function Leaderboard: When LLM Agents Allocate Social Welfare
by: Shi, Zhengliang, et al.
Published: (2025)
by: Shi, Zhengliang, et al.
Published: (2025)
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
by: Badawi, Abeer, et al.
Published: (2025)
by: Badawi, Abeer, et al.
Published: (2025)
A Scoping Review of Mixed Initiative Visual Analytics in the Automation Renaissance
by: Monadjemi, Shayan, et al.
Published: (2025)
by: Monadjemi, Shayan, et al.
Published: (2025)
From Text to Trust: Empowering AI-assisted Decision Making with Adaptive LLM-powered Analysis
by: Li, Zhuoyan, et al.
Published: (2025)
by: Li, Zhuoyan, et al.
Published: (2025)
Plan-Then-Execute: An Empirical Study of User Trust and Team Performance When Using LLM Agents As A Daily Assistant
by: He, Gaole, et al.
Published: (2025)
by: He, Gaole, et al.
Published: (2025)
The Persuasion Paradox: When LLM Explanations Fail to Improve Human-AI Team Performance
by: Cohen, Ruth, et al.
Published: (2026)
by: Cohen, Ruth, et al.
Published: (2026)
A Computational Framework for Behavioral Assessment of LLM Therapists
by: Chiu, Yu Ying, et al.
Published: (2024)
by: Chiu, Yu Ying, et al.
Published: (2024)
LOGOS: LLM-driven End-to-End Grounded Theory Development and Schema Induction for Qualitative Research
by: Pi, Xinyu, et al.
Published: (2025)
by: Pi, Xinyu, et al.
Published: (2025)
Aligning Human-AI-Interaction Trust for Mental Health Support: Survey and Position for Multi-Stakeholders
by: Sun, Xin, et al.
Published: (2026)
by: Sun, Xin, et al.
Published: (2026)
Be Friendly, Not Friends: How LLM Sycophancy Shapes User Trust
by: Sun, Yuan, et al.
Published: (2025)
by: Sun, Yuan, et al.
Published: (2025)
Decomposed Prompting to Answer Questions on a Course Discussion Board
by: Jaipersaud, Brandon, et al.
Published: (2024)
by: Jaipersaud, Brandon, et al.
Published: (2024)
Towards Human-like Multimodal Conversational Agent by Generating Engaging Speech
by: Kim, Taesoo, et al.
Published: (2025)
by: Kim, Taesoo, et al.
Published: (2025)
How to Enable Effective Cooperation Between Humans and NLP Models: A Survey of Principles, Formalizations, and Beyond
by: Huang, Chen, et al.
Published: (2025)
by: Huang, Chen, et al.
Published: (2025)
Adjust for Trust: Mitigating Trust-Induced Inappropriate Reliance on AI Assistance
by: Srinivasan, Tejas, et al.
Published: (2025)
by: Srinivasan, Tejas, et al.
Published: (2025)
Trust in AI among Middle Eastern CS Students: Investigating Students' Trust and Usage Patterns Across Saudi Arabia, Kuwait and Jordan
by: Alkhamees, Saleh, et al.
Published: (2026)
by: Alkhamees, Saleh, et al.
Published: (2026)
When Robots Say No: Temporal Trust Recovery Through Explanation
by: Webb, Nicola, et al.
Published: (2025)
by: Webb, Nicola, et al.
Published: (2025)
The Generative AI Paradox on Evaluation: What It Can Solve, It May Not Evaluate
by: Oh, Juhyun, et al.
Published: (2024)
by: Oh, Juhyun, et al.
Published: (2024)
From Prompts to Constructs: A Dual-Validity Framework for LLM Research in Psychology
by: Lin, Zhicheng
Published: (2025)
by: Lin, Zhicheng
Published: (2025)
QuaLLM: An LLM-based Framework to Extract Quantitative Insights from Online Forums
by: Rao, Varun Nagaraj, et al.
Published: (2024)
by: Rao, Varun Nagaraj, et al.
Published: (2024)
"Like a Nesting Doll": Analyzing Recursion Analogies Generated by CS Students using Large Language Models
by: Bernstein, Seth, et al.
Published: (2024)
by: Bernstein, Seth, et al.
Published: (2024)
A Survey on LLM-based Conversational User Simulation
by: Ni, Bo, et al.
Published: (2026)
by: Ni, Bo, et al.
Published: (2026)
VizTrust: A Visual Analytics Tool for Capturing User Trust Dynamics in Human-AI Communication
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
LLM-TA: An LLM-Enhanced Thematic Analysis Pipeline for Transcripts from Parents of Children with Congenital Heart Disease
by: Raza, Muhammad Zain, et al.
Published: (2025)
by: Raza, Muhammad Zain, et al.
Published: (2025)
Redefining Research Crowdsourcing: Incorporating Human Feedback with LLM-Powered Digital Twins
by: Chan, Amanda, et al.
Published: (2025)
by: Chan, Amanda, et al.
Published: (2025)
Simulating Classroom Education with LLM-Empowered Agents
by: Zhang, Zheyuan, et al.
Published: (2024)
by: Zhang, Zheyuan, et al.
Published: (2024)
A-MEM: Agentic Memory for LLM Agents
by: Xu, Wujiang, et al.
Published: (2025)
by: Xu, Wujiang, et al.
Published: (2025)
LLM Agent Meets Agentic AI: Can LLM Agents Simulate Customers to Evaluate Agentic-AI-based Shopping Assistants?
by: Sun, Lu, et al.
Published: (2025)
by: Sun, Lu, et al.
Published: (2025)
MEGAnno+: A Human-LLM Collaborative Annotation System
by: Kim, Hannah, et al.
Published: (2024)
by: Kim, Hannah, et al.
Published: (2024)
Navigating Rifts in Human-LLM Grounding: Study and Benchmark
by: Shaikh, Omar, et al.
Published: (2025)
by: Shaikh, Omar, et al.
Published: (2025)
ACE: A LLM-based Negotiation Coaching System
by: Shea, Ryan, et al.
Published: (2024)
by: Shea, Ryan, et al.
Published: (2024)
Contrastive Learning for Task-Independent SpeechLLM-Pretraining
by: Züfle, Maike, et al.
Published: (2024)
by: Züfle, Maike, et al.
Published: (2024)
Supporting the Digital Autonomy of Elders Through LLM Assistance
by: Roberts, Jesse, et al.
Published: (2024)
by: Roberts, Jesse, et al.
Published: (2024)
Authorship Drift: How Self-Efficacy and Trust Evolve During LLM-Assisted Writing
by: Park, Yeon Su, et al.
Published: (2026)
by: Park, Yeon Su, et al.
Published: (2026)
LLM-Augmented Semantic Steering of Text Embedding Projection Spaces
by: Liu, Wei, et al.
Published: (2026)
by: Liu, Wei, et al.
Published: (2026)
Human-LLM Collaborative Construction of a Cantonese Emotion Lexicon
by: Zhang, Yusong, et al.
Published: (2024)
by: Zhang, Yusong, et al.
Published: (2024)
PRAISE: Enhancing Product Descriptions with LLM-Driven Structured Insights
by: Qidwai, Adnan, et al.
Published: (2025)
by: Qidwai, Adnan, et al.
Published: (2025)
DiscussLLM: Teaching Large Language Models When to Speak
by: Patel, Deep Anil, et al.
Published: (2025)
by: Patel, Deep Anil, et al.
Published: (2025)
Similar Items
-
Probing the Visualization Literacy of Vision Language Models: the Good, the Bad, and the Ugly
by: Dong, Lianghan, et al.
Published: (2025) -
Linting is People! Exploring the Potential of Human Computation as a Sociotechnical Linter of Data Visualizations
by: Crisan, Anamaria, et al.
Published: (2025) -
AInsight: Augmenting Expert Decision-Making with On-the-Fly Insights Grounded in Historical Data
by: Abolnejadian, Mohammad, et al.
Published: (2025) -
Social Welfare Function Leaderboard: When LLM Agents Allocate Social Welfare
by: Shi, Zhengliang, et al.
Published: (2025) -
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
by: Badawi, Abeer, et al.
Published: (2025)