LLAMADRS: Evaluating Open-Source LLMs on Real Clinical Interviews--To Reason or Not to Reason?
Fuente:
arXiv
Saved in:
| Main Authors: | Kebe, Gaoussou Youssouf, Girard, Jeffrey M., Liebenthal, Einat, Baker, Justin, De la Torre, Fernando, Morency, Louis-Philippe |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Social Caption: Evaluating Social Understanding in Multimodal Models
by: Thumu, Bhaavanaa, et al.
Published: (2026)
by: Thumu, Bhaavanaa, et al.
Published: (2026)
Advancing Social Intelligence in AI Agents: Technical Challenges and Open Questions
by: Mathur, Leena, et al.
Published: (2024)
by: Mathur, Leena, et al.
Published: (2024)
Toward Beginner-Friendly LLMs for Language Learning: Controlling Difficulty in Conversation
by: Jin, Meiqing, et al.
Published: (2025)
by: Jin, Meiqing, et al.
Published: (2025)
Social Human Robot Embodied Conversation (SHREC) Dataset: Benchmarking Foundational Models' Social Reasoning
by: Lee, Dong Won, et al.
Published: (2025)
by: Lee, Dong Won, et al.
Published: (2025)
Can LLMs Reason About Attention? Towards Zero-Shot Analysis of Multimodal Classroom Behavior
by: Platt, Nolan, et al.
Published: (2026)
by: Platt, Nolan, et al.
Published: (2026)
Who Sees What? Structured Thought-Action Sequences for Epistemic Reasoning in LLMs
by: Annese, Luca, et al.
Published: (2025)
by: Annese, Luca, et al.
Published: (2025)
Therapy as an NLP Task: Psychologists' Comparison of LLMs and Human Peers in CBT
by: Iftikhar, Zainab, et al.
Published: (2024)
by: Iftikhar, Zainab, et al.
Published: (2024)
Adaptive Interviewing for Persona Simulation in LLMs: Evidence-Grounded Reasoning Improves Decision Alignment
by: Su, Ruoxi, et al.
Published: (2026)
by: Su, Ruoxi, et al.
Published: (2026)
Can LLMs faithfully generate their layperson-understandable 'self'?: A Case Study in High-Stakes Domains
by: Das, Arion, et al.
Published: (2024)
by: Das, Arion, et al.
Published: (2024)
LLMs and people both learn to form conventions -- just not with each other
by: Jones, Cameron R., et al.
Published: (2026)
by: Jones, Cameron R., et al.
Published: (2026)
Can LLMs and humans be friends? Uncovering factors affecting human-AI intimacy formation
by: Hong, Yeseon, et al.
Published: (2025)
by: Hong, Yeseon, et al.
Published: (2025)
Real-Time World Crafting: Generating Structured Game Behaviors from Natural Language with Large Language Models
by: Drake, Austin, et al.
Published: (2025)
by: Drake, Austin, et al.
Published: (2025)
Functional Flexibility in Generative AI Interfaces: Text Editing with LLMs through Conversations, Toolbars, and Prompts
by: Lehmann, Florian, et al.
Published: (2024)
by: Lehmann, Florian, et al.
Published: (2024)
MemPal: Leveraging Multimodal AI and LLMs for Voice-Activated Object Retrieval in Homes of Older Adults
by: Maniar, Natasha, et al.
Published: (2025)
by: Maniar, Natasha, et al.
Published: (2025)
Helping Johnny Make Sense of Privacy Policies with LLMs
by: Freiberger, Vincent, et al.
Published: (2025)
by: Freiberger, Vincent, et al.
Published: (2025)
Improving Dialogue Agents by Decomposing One Global Explicit Annotation with Local Implicit Multimodal Feedback
by: Lee, Dong Won, et al.
Published: (2024)
by: Lee, Dong Won, et al.
Published: (2024)
Limited Ability of LLMs to Simulate Human Psychological Behaviours: a Psychometric Analysis
by: Petrov, Nikolay B, et al.
Published: (2024)
by: Petrov, Nikolay B, et al.
Published: (2024)
AI Conversational Interviewing: Transforming Surveys with LLMs as Adaptive Interviewers
by: Wuttke, Alexander, et al.
Published: (2024)
by: Wuttke, Alexander, et al.
Published: (2024)
TAMA: A Human-AI Collaborative Thematic Analysis Framework Using Multi-Agent LLMs for Clinical Interviews
by: Xu, Huimin, et al.
Published: (2025)
by: Xu, Huimin, et al.
Published: (2025)
Interview AI-ssistant: Designing for Real-Time Human-AI Collaboration in Interview Preparation and Execution
by: Liu, Zhe
Published: (2025)
by: Liu, Zhe
Published: (2025)
Complementary Human-AI Clinical Reasoning in Ophthalmology
by: Sevgi, Mertcan, et al.
Published: (2025)
by: Sevgi, Mertcan, et al.
Published: (2025)
Achieving Tool Calling Functionality in LLMs Using Only Prompt Engineering Without Fine-Tuning
by: He, Shengtao
Published: (2024)
by: He, Shengtao
Published: (2024)
Communication Access Real-Time Translation Through Collaborative Correction of Automatic Speech Recognition
by: Kuhn, Korbinian, et al.
Published: (2025)
by: Kuhn, Korbinian, et al.
Published: (2025)
LLMs achieve adult human performance on higher-order theory of mind tasks
by: Street, Winnie, et al.
Published: (2024)
by: Street, Winnie, et al.
Published: (2024)
Fine-Tuning Open-Weight Language Models to Deliver Cognitive Behavioral Therapy for Depression: A Feasibility Study
by: Tahir, Talha
Published: (2024)
by: Tahir, Talha
Published: (2024)
Voice Interaction With Conversational AI Could Facilitate Thoughtful Reflection and Substantive Revision in Writing
by: Kim, Jiho, et al.
Published: (2025)
by: Kim, Jiho, et al.
Published: (2025)
Can LLMs Compute with Reasons?
by: Sandilya, Harshit, et al.
Published: (2024)
by: Sandilya, Harshit, et al.
Published: (2024)
Virtual Interviewers, Real Results: Exploring AI-Driven Mock Technical Interviews on Student Readiness and Confidence
by: Gomez, Nathalia, et al.
Published: (2025)
by: Gomez, Nathalia, et al.
Published: (2025)
Evaluating AI Counseling in Japanese: Counselor, Client, and Evaluator Roles Assessed by Motivational Interviewing Criteria
by: Kiuchi, Keita, et al.
Published: (2025)
by: Kiuchi, Keita, et al.
Published: (2025)
MAWARITH: A Dataset and Benchmark for Legal Inheritance Reasoning with LLMs
by: Bouchekif, Abdessalam, et al.
Published: (2026)
by: Bouchekif, Abdessalam, et al.
Published: (2026)
Multi-trait User Simulation with Adaptive Decoding for Conversational Task Assistants
by: Ferreira, Rafael, et al.
Published: (2024)
by: Ferreira, Rafael, et al.
Published: (2024)
The Open Source Advantage in Large Language Models (LLMs)
by: Manchanda, Jiya, et al.
Published: (2024)
by: Manchanda, Jiya, et al.
Published: (2024)
Same Voice, Different Lab: On the Homogenization of Frontier LLM Personalities
by: Krishna, Avinash, et al.
Published: (2026)
by: Krishna, Avinash, et al.
Published: (2026)
Visualization of Unstructured Sports Data -- An Example of Cricket Short Text Commentary
by: Behera, Swarup Ranjan, et al.
Published: (2024)
by: Behera, Swarup Ranjan, et al.
Published: (2024)
Development and Benchmarking of a Blended Human-AI Qualitative Research Assistant
by: Matveyenko, Joseph, et al.
Published: (2025)
by: Matveyenko, Joseph, et al.
Published: (2025)
MoodBench 1.0: An Evaluation Benchmark for Emotional Companionship Dialogue Systems
by: Jing, Haifeng, et al.
Published: (2025)
by: Jing, Haifeng, et al.
Published: (2025)
Who Explains Privacy Policies to Me? Embodied and Textual LLM-Powered Privacy Assistants in Virtual Reality
by: Freiberger, Vincent, et al.
Published: (2026)
by: Freiberger, Vincent, et al.
Published: (2026)
Multi-Intent Recognition in Dialogue Understanding: A Comparison Between Smaller Open-Source LLMs
by: Ahmad, Adnan, et al.
Published: (2025)
by: Ahmad, Adnan, et al.
Published: (2025)
Visual Analytics for Causal Reasoning from Real-World Health Data
by: Wang, Arran Zeyu, et al.
Published: (2025)
by: Wang, Arran Zeyu, et al.
Published: (2025)
LLMs Aren't Human: A Critical Perspective on LLM Personality
by: Zierahn, Kim, et al.
Published: (2026)
by: Zierahn, Kim, et al.
Published: (2026)
Similar Items
-
Social Caption: Evaluating Social Understanding in Multimodal Models
by: Thumu, Bhaavanaa, et al.
Published: (2026) -
Advancing Social Intelligence in AI Agents: Technical Challenges and Open Questions
by: Mathur, Leena, et al.
Published: (2024) -
Toward Beginner-Friendly LLMs for Language Learning: Controlling Difficulty in Conversation
by: Jin, Meiqing, et al.
Published: (2025) -
Social Human Robot Embodied Conversation (SHREC) Dataset: Benchmarking Foundational Models' Social Reasoning
by: Lee, Dong Won, et al.
Published: (2025) -
Can LLMs Reason About Attention? Towards Zero-Shot Analysis of Multimodal Classroom Behavior
by: Platt, Nolan, et al.
Published: (2026)