From "Help" to Helpful: A Hierarchical Assessment of LLMs in Mental e-Health Applications
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Steigerwald, Philipp, Albrecht, Jens |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
When LLMs Can't Help: Real-World Evaluation of LLMs in Nutrition
von: Li, Karen Jia-Hui, et al.
Veröffentlicht: (2025)
von: Li, Karen Jia-Hui, et al.
Veröffentlicht: (2025)
LLARS: Enabling Domain Expert & Developer Collaboration for LLM Prompting, Generation and Evaluation
von: Steigerwald, Philipp, et al.
Veröffentlicht: (2026)
von: Steigerwald, Philipp, et al.
Veröffentlicht: (2026)
When "A Helpful Assistant" Is Not Really Helpful: Personas in System Prompts Do Not Improve Performances of Large Language Models
von: Zheng, Mingqian, et al.
Veröffentlicht: (2023)
von: Zheng, Mingqian, et al.
Veröffentlicht: (2023)
Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives
von: Drożdż, Karolina, et al.
Veröffentlicht: (2025)
von: Drożdż, Karolina, et al.
Veröffentlicht: (2025)
PAIR-SAFE: A Paired-Agent Approach for Runtime Auditing and Refining AI-Mediated Mental Health Support
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
An Offline Mobile Conversational Agent for Mental Health Support: Learning from Emotional Dialogues and Psychological Texts with Student-Centered Evaluation
von: A, Vimaleswar, et al.
Veröffentlicht: (2025)
von: A, Vimaleswar, et al.
Veröffentlicht: (2025)
MentalChat16K: A Benchmark Dataset for Conversational Mental Health Assistance
von: Xu, Jia, et al.
Veröffentlicht: (2025)
von: Xu, Jia, et al.
Veröffentlicht: (2025)
Understanding Help-Seeking Behavior of Students Using LLMs vs. Web Search for Writing SQL Queries
von: Kumar, Harsh, et al.
Veröffentlicht: (2024)
von: Kumar, Harsh, et al.
Veröffentlicht: (2024)
A Conditional Companion: Lived Experiences of People with Mental Health Disorders Using LLMs
von: Purohit, Aditya Kumar, et al.
Veröffentlicht: (2026)
von: Purohit, Aditya Kumar, et al.
Veröffentlicht: (2026)
Red Teaming LLMs as Socio-Technical Practice: From Exploration and Data Creation to Evaluation
von: Garcia, Adriana Alvarado, et al.
Veröffentlicht: (2026)
von: Garcia, Adriana Alvarado, et al.
Veröffentlicht: (2026)
Understanding Remote Mental Health Supporters' Help-Seeking in Online Communities
von: Lee, Tuan-He, et al.
Veröffentlicht: (2026)
von: Lee, Tuan-He, et al.
Veröffentlicht: (2026)
Mental Health Impacts of AI Companions: Triangulating Social Media Quasi-Experiments, User Perspectives, and Relational Theory
von: Yuan, Yunhao, et al.
Veröffentlicht: (2025)
von: Yuan, Yunhao, et al.
Veröffentlicht: (2025)
SouLLMate: An Application Enhancing Diverse Mental Health Support with Adaptive LLMs, Prompt Engineering, and RAG Techniques
von: Guo, Qiming, et al.
Veröffentlicht: (2024)
von: Guo, Qiming, et al.
Veröffentlicht: (2024)
EmoAgent: Assessing and Safeguarding Human-AI Interaction for Mental Health Safety
von: Qiu, Jiahao, et al.
Veröffentlicht: (2025)
von: Qiu, Jiahao, et al.
Veröffentlicht: (2025)
A Comprehensive Survey of Bias in LLMs: Current Landscape and Future Directions
von: Ranjan, Rajesh, et al.
Veröffentlicht: (2024)
von: Ranjan, Rajesh, et al.
Veröffentlicht: (2024)
Assessing Risks of Large Language Models in Mental Health Support: A Framework for Automated Clinical AI Red Teaming
von: Steenstra, Ian, et al.
Veröffentlicht: (2026)
von: Steenstra, Ian, et al.
Veröffentlicht: (2026)
The Emotional Spectrum of LLMs: Leveraging Empathy and Emotion-Based Markers for Mental Health Support
von: De Grandi, Alessandro, et al.
Veröffentlicht: (2024)
von: De Grandi, Alessandro, et al.
Veröffentlicht: (2024)
How Persuasive Could LLMs Be? A First Study Combining Linguistic-Rhetorical Analysis and User Experiments
von: Raffini, Daniel, et al.
Veröffentlicht: (2025)
von: Raffini, Daniel, et al.
Veröffentlicht: (2025)
The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness
von: Subedi, Krishna
Veröffentlicht: (2025)
von: Subedi, Krishna
Veröffentlicht: (2025)
What Do LLMs Associate with Your Name? A Human-Centered Black-Box Audit of Personal Data
von: Staufer, Dimitri, et al.
Veröffentlicht: (2026)
von: Staufer, Dimitri, et al.
Veröffentlicht: (2026)
A Checklist for Trustworthy, Safe, and User-Friendly Mental Health Chatbots
von: Haran, Shreya, et al.
Veröffentlicht: (2026)
von: Haran, Shreya, et al.
Veröffentlicht: (2026)
Seeing Like an AI: How LLMs Apply (and Misapply) Wikipedia Neutrality Norms
von: Ashkinaze, Joshua, et al.
Veröffentlicht: (2024)
von: Ashkinaze, Joshua, et al.
Veröffentlicht: (2024)
LLM Agents for Education: Advances and Applications
von: Chu, Zhendong, et al.
Veröffentlicht: (2025)
von: Chu, Zhendong, et al.
Veröffentlicht: (2025)
Empathy Is Not What Changed: Clinical Assessment of Psychological Safety Across GPT Model Generations
von: Keeman, Michael, et al.
Veröffentlicht: (2026)
von: Keeman, Michael, et al.
Veröffentlicht: (2026)
User Perceptions vs. Proxy LLM Judges: Privacy and Helpfulness in LLM Responses to Privacy-Sensitive Scenarios
von: Wu, Xiaoyuan, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyuan, et al.
Veröffentlicht: (2025)
From Prompts to Constructs: A Dual-Validity Framework for LLM Research in Psychology
von: Lin, Zhicheng
Veröffentlicht: (2025)
von: Lin, Zhicheng
Veröffentlicht: (2025)
Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature Review
von: Pang, Rock Yuren, et al.
Veröffentlicht: (2025)
von: Pang, Rock Yuren, et al.
Veröffentlicht: (2025)
A Scoping Review of AI-Driven Digital Interventions in Mental Health Care: Mapping Applications Across Screening, Support, Monitoring, Prevention, and Clinical Education
von: Ni, Yang, et al.
Veröffentlicht: (2026)
von: Ni, Yang, et al.
Veröffentlicht: (2026)
Evaluating the Application of Large Language Models to Generate Feedback in Programming Education
von: Jacobs, Sven, et al.
Veröffentlicht: (2024)
von: Jacobs, Sven, et al.
Veröffentlicht: (2024)
AI for Proactive Mental Health: A Multi-Institutional, Longitudinal, Randomized Controlled Trial
von: Cachia, Julie Y. A., et al.
Veröffentlicht: (2025)
von: Cachia, Julie Y. A., et al.
Veröffentlicht: (2025)
ChatBench: From Static Benchmarks to Human-AI Evaluation
von: Chang, Serina, et al.
Veröffentlicht: (2025)
von: Chang, Serina, et al.
Veröffentlicht: (2025)
From Adoption to Adaption: Tracing the Diffusion of New Emojis on Twitter
von: Zhou, Yuhang, et al.
Veröffentlicht: (2024)
von: Zhou, Yuhang, et al.
Veröffentlicht: (2024)
GenAI Against Humanity: Nefarious Applications of Generative Artificial Intelligence and Large Language Models
von: Ferrara, Emilio
Veröffentlicht: (2023)
von: Ferrara, Emilio
Veröffentlicht: (2023)
LLUMI: Improving LLM Writing Assistance for Mental Health Support with Online Community Feedback
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
When Testing AI Tests Us: Safeguarding Mental Health on the Digital Frontlines
von: Pendse, Sachin R., et al.
Veröffentlicht: (2025)
von: Pendse, Sachin R., et al.
Veröffentlicht: (2025)
The Typing Cure: Experiences with Large Language Model Chatbots for Mental Health Support
von: Song, Inhwa, et al.
Veröffentlicht: (2024)
von: Song, Inhwa, et al.
Veröffentlicht: (2024)
From Google Gemini to OpenAI Q* (Q-Star): A Survey of Reshaping the Generative Artificial Intelligence (AI) Research Landscape
von: McIntosh, Timothy R., et al.
Veröffentlicht: (2023)
von: McIntosh, Timothy R., et al.
Veröffentlicht: (2023)
Comprehensive Study on Sentiment Analysis: From Rule-based to modern LLM based system
von: Gupta, Shailja, et al.
Veröffentlicht: (2024)
von: Gupta, Shailja, et al.
Veröffentlicht: (2024)
From tools to thieves: Measuring and understanding public perceptions of AI through crowdsourced metaphors
von: Cheng, Myra, et al.
Veröffentlicht: (2025)
von: Cheng, Myra, et al.
Veröffentlicht: (2025)
A Study about Distribution and Acceptance of Conversational Agents for Mental Health in Germany: Keep the Human in the Loop?
von: Lukas, Christina
Veröffentlicht: (2025)
von: Lukas, Christina
Veröffentlicht: (2025)
Ähnliche Einträge
-
When LLMs Can't Help: Real-World Evaluation of LLMs in Nutrition
von: Li, Karen Jia-Hui, et al.
Veröffentlicht: (2025) -
LLARS: Enabling Domain Expert & Developer Collaboration for LLM Prompting, Generation and Evaluation
von: Steigerwald, Philipp, et al.
Veröffentlicht: (2026) -
When "A Helpful Assistant" Is Not Really Helpful: Personas in System Prompts Do Not Improve Performances of Large Language Models
von: Zheng, Mingqian, et al.
Veröffentlicht: (2023) -
Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives
von: Drożdż, Karolina, et al.
Veröffentlicht: (2025) -
PAIR-SAFE: A Paired-Agent Approach for Runtime Auditing and Refining AI-Mediated Mental Health Support
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)