PAIR-SAFE: A Paired-Agent Approach for Runtime Auditing and Refining AI-Mediated Mental Health Support
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Jiwon, Rodriguez, Violeta J., Yoo, Dong Whi, Chandrasekharan, Eshwar, Saha, Koustuv |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLUMI: Improving LLM Writing Assistance for Mental Health Support with Online Community Feedback
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
AI Chatbots for Mental Health: Values and Harms from Lived Experiences of Depression
von: Yoo, Dong Whi, et al.
Veröffentlicht: (2025)
von: Yoo, Dong Whi, et al.
Veröffentlicht: (2025)
Mapping Caregiver Needs to AI Chatbot Design: Strengths and Gaps in Mental Health Support for Alzheimer's and Dementia Caregivers
von: Shi, Jiayue Melissa, et al.
Veröffentlicht: (2025)
von: Shi, Jiayue Melissa, et al.
Veröffentlicht: (2025)
Does Positive Reinforcement Work?: A Quasi-Experimental Study of the Effects of Positive Feedback on Reddit
von: Lambert, Charlotte, et al.
Veröffentlicht: (2024)
von: Lambert, Charlotte, et al.
Veröffentlicht: (2024)
A Checklist for Trustworthy, Safe, and User-Friendly Mental Health Chatbots
von: Haran, Shreya, et al.
Veröffentlicht: (2026)
von: Haran, Shreya, et al.
Veröffentlicht: (2026)
Examining Algorithmic Curation on Social Media: An Empirical Audit of Reddit's r/popular Feed
von: Chan, Jackie, et al.
Veröffentlicht: (2025)
von: Chan, Jackie, et al.
Veröffentlicht: (2025)
The Ranking Effect: How Algorithmic Rank Influences Attention on Social Media
von: Chan, Jackie, et al.
Veröffentlicht: (2025)
von: Chan, Jackie, et al.
Veröffentlicht: (2025)
Balancing Caregiving and Self-Care: Exploring Mental Health Needs of Alzheimer's and Dementia Caregivers
von: Shi, Jiayue Melissa, et al.
Veröffentlicht: (2025)
von: Shi, Jiayue Melissa, et al.
Veröffentlicht: (2025)
VASTU: Value-Aligned Social Toolkit for Online Content Curation
von: Goyal, Agam, et al.
Veröffentlicht: (2026)
von: Goyal, Agam, et al.
Veröffentlicht: (2026)
Journeys of Parents with LGBTQ+ Children: How Trauma and Healing Reshape Identity and (Mis)Informating Practices
von: Kwon, Soonho, et al.
Veröffentlicht: (2026)
von: Kwon, Soonho, et al.
Veröffentlicht: (2026)
Algorithmic Cultivation: How Social Media Feeds Shape User Language
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
The Hidden Toll of Social Media News: Causal Effects on Psychosocial Wellbeing
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
"In my defense, only three hours on Instagram": Designing Toward Digital Self-Awareness and Wellbeing
von: Bhat, Karthik S., et al.
Veröffentlicht: (2025)
von: Bhat, Karthik S., et al.
Veröffentlicht: (2025)
Designing KRIYA: An AI Companion for Wellbeing Self-Reflection
von: Zhu, Shanshan, et al.
Veröffentlicht: (2026)
von: Zhu, Shanshan, et al.
Veröffentlicht: (2026)
Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
Towards a Better Modqueue: Designing for Diversity Across Moderator Objectives and Workflows
von: Bajpai, Tanvi, et al.
Veröffentlicht: (2024)
von: Bajpai, Tanvi, et al.
Veröffentlicht: (2024)
"Think about it like you're a firefighter": Understanding How Reddit Moderators Use the Modqueue
von: Bajpai, Tanvi, et al.
Veröffentlicht: (2025)
von: Bajpai, Tanvi, et al.
Veröffentlicht: (2025)
Designing Usable Controls for Customizable Social Media Feeds
von: Choi, Frederick, et al.
Veröffentlicht: (2025)
von: Choi, Frederick, et al.
Veröffentlicht: (2025)
Interpersonal Theory of Suicide as a Lens to Examine Suicidal Ideation in Online Spaces
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
AI Fortune-Teller: Juxtaposing Shaman and AI to Reveal Human Agency in the Age of AI
von: Kwon, Soonho, et al.
Veröffentlicht: (2026)
von: Kwon, Soonho, et al.
Veröffentlicht: (2026)
Linguistic Comparison of AI- and Human-Written Responses to Online Mental Health Queries
von: Saha, Koustuv, et al.
Veröffentlicht: (2025)
von: Saha, Koustuv, et al.
Veröffentlicht: (2025)
Mind Your Ps and Qs: Supporting Positive Reinforcement in Moderation Through a Positive Queue
von: Lambert, Charlotte, et al.
Veröffentlicht: (2025)
von: Lambert, Charlotte, et al.
Veröffentlicht: (2025)
Designing Beyond Language: Sociotechnical Barriers in AI Health Technologies for Limited English Proficiency
von: Huang, Michelle, et al.
Veröffentlicht: (2025)
von: Huang, Michelle, et al.
Veröffentlicht: (2025)
The Language of Approval: Identifying the Drivers of Positive Feedback Online
von: Goyal, Agam, et al.
Veröffentlicht: (2025)
von: Goyal, Agam, et al.
Veröffentlicht: (2025)
Needling Through the Threads: A Visualization Tool for Navigating Threaded Online Discussions
von: Liu, Yijun, et al.
Veröffentlicht: (2025)
von: Liu, Yijun, et al.
Veröffentlicht: (2025)
Value-Sensitive AI for Prayer: Balancing the Agencies Between Human and AI Agents in Spiritual Context
von: Kwon, Soonho, et al.
Veröffentlicht: (2026)
von: Kwon, Soonho, et al.
Veröffentlicht: (2026)
RubRIX: Rubric-Driven Risk Mitigation in Caregiver-AI Interactions
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
AI Psychosis: Does Conversational AI Amplify Delusion-Related Language?
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2026)
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2026)
Toxic HallucinAItions: Perturbing Prompts and Tracing LLM Circuits
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2026)
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2026)
The Agony of Opacity: Foundations for Reflective Interpretability in AI-Mediated Mental Health Support
von: Pendse, Sachin R., et al.
Veröffentlicht: (2025)
von: Pendse, Sachin R., et al.
Veröffentlicht: (2025)
Mental Health Impacts of AI Companions: Triangulating Social Media Quasi-Experiments, User Perspectives, and Relational Theory
von: Yuan, Yunhao, et al.
Veröffentlicht: (2025)
von: Yuan, Yunhao, et al.
Veröffentlicht: (2025)
Exploring the Role of LLMs for Supporting Older Adults: Opportunities and Concerns
von: Kaliappan, Sidharth, et al.
Veröffentlicht: (2024)
von: Kaliappan, Sidharth, et al.
Veröffentlicht: (2024)
AI Content Moderation in Therapy Conversations
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
Creator Hearts: Investigating the Impact Positive Signals from YouTube Creators in Shaping Comment Section Behavior
von: Choi, Frederick, et al.
Veröffentlicht: (2024)
von: Choi, Frederick, et al.
Veröffentlicht: (2024)
Bystanders of Online Moderation: Examining the Effects of Witnessing Post-Removal Explanations
von: Jhaver, Shagun, et al.
Veröffentlicht: (2023)
von: Jhaver, Shagun, et al.
Veröffentlicht: (2023)
Uncovering the Internet's Hidden Values: An Empirical Study of Desirable Behavior Using Highly-Upvoted Content on Reddit
von: Goyal, Agam, et al.
Veröffentlicht: (2024)
von: Goyal, Agam, et al.
Veröffentlicht: (2024)
Do We Know What They Know We Know? Calibrating Student Trust in AI and Human Responses Through Mutual Theory of Mind
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
Causal Stories from Sensor Traces: Auditing Epistemic Overreach in LLM-Generated Personal Sensing Explanations
von: Zhu, Shanshan, et al.
Veröffentlicht: (2026)
von: Zhu, Shanshan, et al.
Veröffentlicht: (2026)
From Risk Avoidance to User Empowerment in AI Mental Health Crisis Support
von: Kaveladze, Benjamin, et al.
Veröffentlicht: (2026)
von: Kaveladze, Benjamin, et al.
Veröffentlicht: (2026)
A Multi-Agent Dual Dialogue System to Support Mental Health Care Providers
von: Kampman, Onno P., et al.
Veröffentlicht: (2024)
von: Kampman, Onno P., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
LLUMI: Improving LLM Writing Assistance for Mental Health Support with Online Community Feedback
von: Kim, Jiwon, et al.
Veröffentlicht: (2026) -
AI Chatbots for Mental Health: Values and Harms from Lived Experiences of Depression
von: Yoo, Dong Whi, et al.
Veröffentlicht: (2025) -
Mapping Caregiver Needs to AI Chatbot Design: Strengths and Gaps in Mental Health Support for Alzheimer's and Dementia Caregivers
von: Shi, Jiayue Melissa, et al.
Veröffentlicht: (2025) -
Does Positive Reinforcement Work?: A Quasi-Experimental Study of the Effects of Positive Feedback on Reddit
von: Lambert, Charlotte, et al.
Veröffentlicht: (2024) -
A Checklist for Trustworthy, Safe, and User-Friendly Mental Health Chatbots
von: Haran, Shreya, et al.
Veröffentlicht: (2026)