Beyond the Single Turn: Reframing Refusals as Dynamic Experiences Embedded in the Context of Mental Health Support Interactions with LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tang, Ningjing, Qian, Alice, Wang, Qiaosi, Howe, Esther, Bullwinkel, Blake, Pedrelli, Paola, Suh, Jina, Heidari, Hoda, Shen, Hong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Assessing Risks of Large Language Models in Mental Health Support: A Framework for Automated Clinical AI Red Teaming
von: Steenstra, Ian, et al.
Veröffentlicht: (2026)
von: Steenstra, Ian, et al.
Veröffentlicht: (2026)
GRP-Obliteration: Unaligning LLMs With a Single Unlabeled Prompt
von: Russinovich, Mark, et al.
Veröffentlicht: (2026)
von: Russinovich, Mark, et al.
Veröffentlicht: (2026)
Effective Automation to Support the Human Infrastructure in AI Red Teaming
von: Zhang, Alice Qian, et al.
Veröffentlicht: (2025)
von: Zhang, Alice Qian, et al.
Veröffentlicht: (2025)
The Agony of Opacity: Foundations for Reflective Interpretability in AI-Mediated Mental Health Support
von: Pendse, Sachin R., et al.
Veröffentlicht: (2025)
von: Pendse, Sachin R., et al.
Veröffentlicht: (2025)
Steering Language Model Refusal with Sparse Autoencoders
von: O'Brien, Kyle, et al.
Veröffentlicht: (2024)
von: O'Brien, Kyle, et al.
Veröffentlicht: (2024)
Navigating Uncertainties: How GenAI Developers Document Their Models on Open-Source Platforms
von: Tang, Ningjing, et al.
Veröffentlicht: (2025)
von: Tang, Ningjing, et al.
Veröffentlicht: (2025)
Seeking Late Night Life Lines: Experiences of Conversational AI Use in Mental Health Crisis
von: Ajmani, Leah Hope, et al.
Veröffentlicht: (2025)
von: Ajmani, Leah Hope, et al.
Veröffentlicht: (2025)
A Closer Look at the Existing Risks of Generative AI: Mapping the Who, What, and How of Real-World Incidents
von: Li, Megan, et al.
Veröffentlicht: (2025)
von: Li, Megan, et al.
Veröffentlicht: (2025)
From Risk Avoidance to User Empowerment in AI Mental Health Crisis Support
von: Kaveladze, Benjamin, et al.
Veröffentlicht: (2026)
von: Kaveladze, Benjamin, et al.
Veröffentlicht: (2026)
MindGuard: Guardrail Classifiers for Multi-Turn Mental Health Support
von: Farinhas, António, et al.
Veröffentlicht: (2026)
von: Farinhas, António, et al.
Veröffentlicht: (2026)
Discourse Diversity in Multi-Turn Empathic Dialogue
von: Zhan, Hongli, et al.
Veröffentlicht: (2026)
von: Zhan, Hongli, et al.
Veröffentlicht: (2026)
Beyond Over-Refusal: Scenario-Based Diagnostics and Post-Hoc Mitigation for Exaggerated Refusals in LLMs
von: Yuan, Shuzhou, et al.
Veröffentlicht: (2025)
von: Yuan, Shuzhou, et al.
Veröffentlicht: (2025)
“Taking Off the Mask”: Experiences of Gender Roles and Mental Health Stigma for Men in Uganda
von: Maureen Kobusingye, et al.
Veröffentlicht: (2025)
von: Maureen Kobusingye, et al.
Veröffentlicht: (2025)
Beyond Continuity: Challenges of Context Switching in Multi-Turn Dialogue with LLMs
von: Sinha, Aditya, et al.
Veröffentlicht: (2026)
von: Sinha, Aditya, et al.
Veröffentlicht: (2026)
When Testing AI Tests Us: Safeguarding Mental Health on the Digital Frontlines
von: Pendse, Sachin R., et al.
Veröffentlicht: (2025)
von: Pendse, Sachin R., et al.
Veröffentlicht: (2025)
Locating Risk: Task Designers and the Challenge of Risk Disclosure in RAI Content Work
von: Qian, Alice, et al.
Veröffentlicht: (2025)
von: Qian, Alice, et al.
Veröffentlicht: (2025)
The Slow Drift of Support: Boundary Failures in Multi-Turn Mental Health LLM Dialogues
von: Cheng, Youyou, et al.
Veröffentlicht: (2026)
von: Cheng, Youyou, et al.
Veröffentlicht: (2026)
Do Reasoning LLMs Refuse What They Infer in Long Contexts?
von: Fu, Yu, et al.
Veröffentlicht: (2026)
von: Fu, Yu, et al.
Veröffentlicht: (2026)
From Measurement to Expertise: Empathetic Expert Adapters for Context-Based Empathy in Conversational AI Agents
von: Shayegani, Erfan, et al.
Veröffentlicht: (2025)
von: Shayegani, Erfan, et al.
Veröffentlicht: (2025)
DISCERN: Designing Decision Support Interfaces to Investigate the Complexities of Workplace Social Decision-Making With Line Managers
von: Khadpe, Pranav, et al.
Veröffentlicht: (2024)
von: Khadpe, Pranav, et al.
Veröffentlicht: (2024)
Persona-Augmented Benchmarking: Evaluating LLMs Across Diverse Writing Styles
von: Truong, Kimberly Le, et al.
Veröffentlicht: (2025)
von: Truong, Kimberly Le, et al.
Veröffentlicht: (2025)
Reasoning Is Not All You Need: Examining LLMs for Multi-Turn Mental Health Conversations
von: Chandra, Mohit, et al.
Veröffentlicht: (2025)
von: Chandra, Mohit, et al.
Veröffentlicht: (2025)
If I Could Turn Back Time: Temporal Reframing as a Historical Reasoning Task for LLMs
von: Bungum, Lars, et al.
Veröffentlicht: (2025)
von: Bungum, Lars, et al.
Veröffentlicht: (2025)
Youth WellTech: A Global Remote Co-Design Sprint for Youth Mental Health Technology
von: Phang, Kenji, et al.
Veröffentlicht: (2024)
von: Phang, Kenji, et al.
Veröffentlicht: (2024)
The Backfiring Effect of Weak AI Safety Regulation
von: Laufer, Benjamin, et al.
Veröffentlicht: (2025)
von: Laufer, Benjamin, et al.
Veröffentlicht: (2025)
Fine-Tuning Games: Bargaining and Adaptation for General-Purpose Models
von: Laufer, Benjamin, et al.
Veröffentlicht: (2023)
von: Laufer, Benjamin, et al.
Veröffentlicht: (2023)
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks
von: Bullwinkel, Blake, et al.
Veröffentlicht: (2025)
von: Bullwinkel, Blake, et al.
Veröffentlicht: (2025)
The Situate AI Guidebook: Co-Designing a Toolkit to Support Multi-Stakeholder Early-stage Deliberations Around Public Sector AI Proposals
von: Kawakami, Anna, et al.
Veröffentlicht: (2024)
von: Kawakami, Anna, et al.
Veröffentlicht: (2024)
OMIND: Framework for Knowledge Grounded Finetuning and Multi-Turn Dialogue Benchmark for Mental Health LLMs
von: Racha, Suraj, et al.
Veröffentlicht: (2026)
von: Racha, Suraj, et al.
Veröffentlicht: (2026)
Experiences of Tobacco Smoking and Quitting Among Mental Health Consumers
von: Helena Roennfeldt, et al.
Veröffentlicht: (2025)
von: Helena Roennfeldt, et al.
Veröffentlicht: (2025)
Reframing Recovery: Revisiting Mental Health Recovery Narratives Framework Through Indian Memoirs
von: Ryno Settrisman, et al.
Veröffentlicht: (2025)
von: Ryno Settrisman, et al.
Veröffentlicht: (2025)
The Typing Cure: Experiences with Large Language Model Chatbots for Mental Health Support
von: Song, Inhwa, et al.
Veröffentlicht: (2024)
von: Song, Inhwa, et al.
Veröffentlicht: (2024)
Refusal and Aporia: At the Limits of Anthropological Knowledge
von: Cory‐Alice André‐Johnson
Veröffentlicht: (2026)
von: Cory‐Alice André‐Johnson
Veröffentlicht: (2026)
Health-ORSC-Bench: A Benchmark for Measuring Over-Refusal and Safety Completion in Health Context
von: Zhang, Zhihao, et al.
Veröffentlicht: (2026)
von: Zhang, Zhihao, et al.
Veröffentlicht: (2026)
Beyond a Single Direction: Chain-of-Thought Disrupts Simple Steering of Refusal
von: Yang, Kia-Jüng, et al.
Veröffentlicht: (2026)
von: Yang, Kia-Jüng, et al.
Veröffentlicht: (2026)
AURA: Amplifying Understanding, Resilience, and Awareness for Responsible AI Content Work
von: Zhang, Alice Qian, et al.
Veröffentlicht: (2024)
von: Zhang, Alice Qian, et al.
Veröffentlicht: (2024)
Worker Discretion Advised: Co-designing Risk Disclosure in Crowdsourced Responsible AI (RAI) Content Work
von: Qian, Alice, et al.
Veröffentlicht: (2025)
von: Qian, Alice, et al.
Veröffentlicht: (2025)
Refusal in LLMs is an Affine Function
von: Marshall, Thomas, et al.
Veröffentlicht: (2024)
von: Marshall, Thomas, et al.
Veröffentlicht: (2024)
AL PASO DE LAS SEMILLAS DE VIDA DEL PUEBLO NASA DE TIERRADENTRO Y LA CONSTRUCCIÓN DE LAS NUEVAS IDENTIDADES
von: Jina Paola Guegia Hurtado
Veröffentlicht: (2017)
von: Jina Paola Guegia Hurtado
Veröffentlicht: (2017)
Unpacking College Students' Mental Health Discourse through YouTube Comments: Insights from Topic Modeling and Sentiment Analysis
von: Heejun Kim, et al.
Veröffentlicht: (2025)
von: Heejun Kim, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Assessing Risks of Large Language Models in Mental Health Support: A Framework for Automated Clinical AI Red Teaming
von: Steenstra, Ian, et al.
Veröffentlicht: (2026) -
GRP-Obliteration: Unaligning LLMs With a Single Unlabeled Prompt
von: Russinovich, Mark, et al.
Veröffentlicht: (2026) -
Effective Automation to Support the Human Infrastructure in AI Red Teaming
von: Zhang, Alice Qian, et al.
Veröffentlicht: (2025) -
The Agony of Opacity: Foundations for Reflective Interpretability in AI-Mediated Mental Health Support
von: Pendse, Sachin R., et al.
Veröffentlicht: (2025) -
Steering Language Model Refusal with Sparse Autoencoders
von: O'Brien, Kyle, et al.
Veröffentlicht: (2024)