AI Content Moderation in Therapy Conversations
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Jiwon, Wang, Claire, Yoon, Taeung, Huang, Sabelle, Saha, Koustuv |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AI Psychosis: Does Conversational AI Amplify Delusion-Related Language?
by: Shimgekar, Soorya Ram, et al.
Published: (2026)
by: Shimgekar, Soorya Ram, et al.
Published: (2026)
LLUMI: Improving LLM Writing Assistance for Mental Health Support with Online Community Feedback
by: Kim, Jiwon, et al.
Published: (2026)
by: Kim, Jiwon, et al.
Published: (2026)
Algorithmic Cultivation: How Social Media Feeds Shape User Language
by: Pal, Olivia, et al.
Published: (2026)
by: Pal, Olivia, et al.
Published: (2026)
The Hidden Toll of Social Media News: Causal Effects on Psychosocial Wellbeing
by: Pal, Olivia, et al.
Published: (2026)
by: Pal, Olivia, et al.
Published: (2026)
Sima AIunty: Caste Audit in LLM-Driven Matchmaking
by: Naik, Atharva, et al.
Published: (2026)
by: Naik, Atharva, et al.
Published: (2026)
Interpersonal Theory of Suicide as a Lens to Examine Suicidal Ideation in Online Spaces
by: Shimgekar, Soorya Ram, et al.
Published: (2025)
by: Shimgekar, Soorya Ram, et al.
Published: (2025)
LLM Content Moderation and User Satisfaction: Evidence from Response Refusals in Chatbot Arena
by: Pasch, Stefan
Published: (2025)
by: Pasch, Stefan
Published: (2025)
Linguistic Comparison of AI- and Human-Written Responses to Online Mental Health Queries
by: Saha, Koustuv, et al.
Published: (2025)
by: Saha, Koustuv, et al.
Published: (2025)
Detecting Early and Implicit Suicidal Ideation via Longitudinal and Information Environment Signals on Social Media
by: Shimgekar, Soorya Ram, et al.
Published: (2025)
by: Shimgekar, Soorya Ram, et al.
Published: (2025)
PAIR-SAFE: A Paired-Agent Approach for Runtime Auditing and Refining AI-Mediated Mental Health Support
by: Kim, Jiwon, et al.
Published: (2026)
by: Kim, Jiwon, et al.
Published: (2026)
RubRIX: Rubric-Driven Risk Mitigation in Caregiver-AI Interactions
by: Goel, Drishti, et al.
Published: (2026)
by: Goel, Drishti, et al.
Published: (2026)
Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles
by: Goel, Drishti, et al.
Published: (2026)
by: Goel, Drishti, et al.
Published: (2026)
TUX: Measuring Human--AI Tacit Understanding
by: Li, Yueshen, et al.
Published: (2026)
by: Li, Yueshen, et al.
Published: (2026)
Toxicity in Online Platforms and AI Systems: A Survey of Needs, Challenges, Mitigations, and Future Directions
by: Khapre, Smita, et al.
Published: (2025)
by: Khapre, Smita, et al.
Published: (2025)
CPG-EVAL: A Multi-Tiered Benchmark for Evaluating the Chinese Pedagogical Grammar Competence of Large Language Models
by: Wang, Dong
Published: (2025)
by: Wang, Dong
Published: (2025)
The DSA Transparency Database: Auditing Self-reported Moderation Actions by Social Media
by: Trujillo, Amaury, et al.
Published: (2023)
by: Trujillo, Amaury, et al.
Published: (2023)
Designing KRIYA: An AI Companion for Wellbeing Self-Reflection
by: Zhu, Shanshan, et al.
Published: (2026)
by: Zhu, Shanshan, et al.
Published: (2026)
Watch Your Language: Investigating Content Moderation with Large Language Models
by: Kumar, Deepak, et al.
Published: (2023)
by: Kumar, Deepak, et al.
Published: (2023)
Mental Health Impacts of AI Companions: Triangulating Social Media Quasi-Experiments, User Perspectives, and Relational Theory
by: Yuan, Yunhao, et al.
Published: (2025)
by: Yuan, Yunhao, et al.
Published: (2025)
Can LLMs Assist Annotators in Identifying Morality Frames? -- Case Study on Vaccination Debate on Social Media
by: Islam, Tunazzina, et al.
Published: (2025)
by: Islam, Tunazzina, et al.
Published: (2025)
Hidden in Plain Text: Measuring LLM Deception Quality Against Human Baselines Using Social Deduction Games
by: Kao, Christopher, et al.
Published: (2026)
by: Kao, Christopher, et al.
Published: (2026)
Does AI Coaching Prepare us for Workplace Negotiations?
by: Duddu, Veda, et al.
Published: (2025)
by: Duddu, Veda, et al.
Published: (2025)
Not My Truce: Personality Differences in AI-Mediated Workplace Negotiation
by: Duddu, Veda, et al.
Published: (2026)
by: Duddu, Veda, et al.
Published: (2026)
Designing Beyond Language: Sociotechnical Barriers in AI Health Technologies for Limited English Proficiency
by: Huang, Michelle, et al.
Published: (2025)
by: Huang, Michelle, et al.
Published: (2025)
Causal Stories from Sensor Traces: Auditing Epistemic Overreach in LLM-Generated Personal Sensing Explanations
by: Zhu, Shanshan, et al.
Published: (2026)
by: Zhu, Shanshan, et al.
Published: (2026)
LLMs Among Us: Generative AI Participating in Digital Discourse
by: Radivojevic, Kristina, et al.
Published: (2024)
by: Radivojevic, Kristina, et al.
Published: (2024)
Mind the Gap! Pathways Towards Unifying AI Safety and Ethics Research
by: Roytburg, Dani, et al.
Published: (2025)
by: Roytburg, Dani, et al.
Published: (2025)
OpenClaw AI Agents as Informal Learners at Moltbook: Characterizing an Emergent Learning Community at Scale
by: Chen, Eason, et al.
Published: (2026)
by: Chen, Eason, et al.
Published: (2026)
When AI Agents Teach Each Other: Discourse Patterns Resembling Peer Learning in the Moltbook Community
by: Chen, Eason, et al.
Published: (2026)
by: Chen, Eason, et al.
Published: (2026)
New contexts, old heuristics: How young people in India and the US trust online content in the age of generative AI
by: Xu, Rachel, et al.
Published: (2024)
by: Xu, Rachel, et al.
Published: (2024)
Socio-Emotional Response Generation: A Human Evaluation Protocol for LLM-Based Conversational Systems
by: Vanel, Lorraine, et al.
Published: (2024)
by: Vanel, Lorraine, et al.
Published: (2024)
Safe Spaces or Toxic Places? Content Moderation and Social Dynamics of Online Eating Disorder Communities
by: Lerman, Kristina, et al.
Published: (2024)
by: Lerman, Kristina, et al.
Published: (2024)
Exposure to Content Written by Large Language Models Can Reduce Stigma Around Opioid Use Disorder in Online Communities
by: Mittal, Shravika, et al.
Published: (2025)
by: Mittal, Shravika, et al.
Published: (2025)
Toxic HallucinAItions: Perturbing Prompts and Tracing LLM Circuits
by: Shimgekar, Soorya Ram, et al.
Published: (2026)
by: Shimgekar, Soorya Ram, et al.
Published: (2026)
Large Language Models Can Infer Psychological Dispositions of Social Media Users
by: Peters, Heinrich, et al.
Published: (2023)
by: Peters, Heinrich, et al.
Published: (2023)
VERA-MH Concept Paper
by: Belli, Luca, et al.
Published: (2025)
by: Belli, Luca, et al.
Published: (2025)
Reranking partisan animosity in algorithmic social media feeds alters affective polarization
by: Piccardi, Tiziano, et al.
Published: (2024)
by: Piccardi, Tiziano, et al.
Published: (2024)
From Future of Work to Future of Workers: Addressing Asymptomatic AI Harms for Dignified Human-AI Interaction
by: Ehsan, Upol, et al.
Published: (2026)
by: Ehsan, Upol, et al.
Published: (2026)
The Levers of Political Persuasion with Conversational AI
by: Hackenburg, Kobi, et al.
Published: (2025)
by: Hackenburg, Kobi, et al.
Published: (2025)
A Checklist for Trustworthy, Safe, and User-Friendly Mental Health Chatbots
by: Haran, Shreya, et al.
Published: (2026)
by: Haran, Shreya, et al.
Published: (2026)
Similar Items
-
AI Psychosis: Does Conversational AI Amplify Delusion-Related Language?
by: Shimgekar, Soorya Ram, et al.
Published: (2026) -
LLUMI: Improving LLM Writing Assistance for Mental Health Support with Online Community Feedback
by: Kim, Jiwon, et al.
Published: (2026) -
Algorithmic Cultivation: How Social Media Feeds Shape User Language
by: Pal, Olivia, et al.
Published: (2026) -
The Hidden Toll of Social Media News: Causal Effects on Psychosocial Wellbeing
by: Pal, Olivia, et al.
Published: (2026) -
Sima AIunty: Caste Audit in LLM-Driven Matchmaking
by: Naik, Atharva, et al.
Published: (2026)