LLM Content Moderation and User Satisfaction: Evidence from Response Refusals in Chatbot Arena
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Pasch, Stefan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AI Content Moderation in Therapy Conversations
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals
von: Pasch, Stefan
Veröffentlicht: (2025)
von: Pasch, Stefan
Veröffentlicht: (2025)
Bottom-Up Perspectives on AI Governance: Insights from User Reviews of AI Products
von: Pasch, Stefan
Veröffentlicht: (2025)
von: Pasch, Stefan
Veröffentlicht: (2025)
Human-AI Interaction and User Satisfaction: Empirical Evidence from Online Reviews of AI Products
von: Pasch, Stefan, et al.
Veröffentlicht: (2025)
von: Pasch, Stefan, et al.
Veröffentlicht: (2025)
Operationalizing AI: Empirical Evidence on MLOps Practices, User Satisfaction, and Organizational Context
von: Pasch, Stefan
Veröffentlicht: (2025)
von: Pasch, Stefan
Veröffentlicht: (2025)
Do Ethical AI Principles Matter to Users? A Large-Scale Analysis of User Sentiment and Satisfaction
von: Pasch, Stefan, et al.
Veröffentlicht: (2025)
von: Pasch, Stefan, et al.
Veröffentlicht: (2025)
Algorithmic Cultivation: How Social Media Feeds Shape User Language
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
Sima AIunty: Caste Audit in LLM-Driven Matchmaking
von: Naik, Atharva, et al.
Veröffentlicht: (2026)
von: Naik, Atharva, et al.
Veröffentlicht: (2026)
Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
LLUMI: Improving LLM Writing Assistance for Mental Health Support with Online Community Feedback
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
Hidden in Plain Text: Measuring LLM Deception Quality Against Human Baselines Using Social Deduction Games
von: Kao, Christopher, et al.
Veröffentlicht: (2026)
von: Kao, Christopher, et al.
Veröffentlicht: (2026)
The DSA Transparency Database: Auditing Self-reported Moderation Actions by Social Media
von: Trujillo, Amaury, et al.
Veröffentlicht: (2023)
von: Trujillo, Amaury, et al.
Veröffentlicht: (2023)
Watch Your Language: Investigating Content Moderation with Large Language Models
von: Kumar, Deepak, et al.
Veröffentlicht: (2023)
von: Kumar, Deepak, et al.
Veröffentlicht: (2023)
Large Language Models Can Infer Psychological Dispositions of Social Media Users
von: Peters, Heinrich, et al.
Veröffentlicht: (2023)
von: Peters, Heinrich, et al.
Veröffentlicht: (2023)
Toxicity in Online Platforms and AI Systems: A Survey of Needs, Challenges, Mitigations, and Future Directions
von: Khapre, Smita, et al.
Veröffentlicht: (2025)
von: Khapre, Smita, et al.
Veröffentlicht: (2025)
Can LLMs Assist Annotators in Identifying Morality Frames? -- Case Study on Vaccination Debate on Social Media
von: Islam, Tunazzina, et al.
Veröffentlicht: (2025)
von: Islam, Tunazzina, et al.
Veröffentlicht: (2025)
Detecting Early and Implicit Suicidal Ideation via Longitudinal and Information Environment Signals on Social Media
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
CPG-EVAL: A Multi-Tiered Benchmark for Evaluating the Chinese Pedagogical Grammar Competence of Large Language Models
von: Wang, Dong
Veröffentlicht: (2025)
von: Wang, Dong
Veröffentlicht: (2025)
Interpersonal Theory of Suicide as a Lens to Examine Suicidal Ideation in Online Spaces
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
AI Psychosis: Does Conversational AI Amplify Delusion-Related Language?
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2026)
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2026)
The Hidden Toll of Social Media News: Causal Effects on Psychosocial Wellbeing
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
Socio-Emotional Response Generation: A Human Evaluation Protocol for LLM-Based Conversational Systems
von: Vanel, Lorraine, et al.
Veröffentlicht: (2024)
von: Vanel, Lorraine, et al.
Veröffentlicht: (2024)
Dialect vs Demographics: Quantifying LLM Bias from Implicit Linguistic Signals vs. Explicit User Profiles
von: Haq, Irti, et al.
Veröffentlicht: (2026)
von: Haq, Irti, et al.
Veröffentlicht: (2026)
Safe Spaces or Toxic Places? Content Moderation and Social Dynamics of Online Eating Disorder Communities
von: Lerman, Kristina, et al.
Veröffentlicht: (2024)
von: Lerman, Kristina, et al.
Veröffentlicht: (2024)
Linguistic Comparison of AI- and Human-Written Responses to Online Mental Health Queries
von: Saha, Koustuv, et al.
Veröffentlicht: (2025)
von: Saha, Koustuv, et al.
Veröffentlicht: (2025)
Exposure to Content Written by Large Language Models Can Reduce Stigma Around Opioid Use Disorder in Online Communities
von: Mittal, Shravika, et al.
Veröffentlicht: (2025)
von: Mittal, Shravika, et al.
Veröffentlicht: (2025)
Oyster-I: Beyond Refusal -- Constructive Safety Alignment for Responsible Language Models
von: Duan, Ranjie, et al.
Veröffentlicht: (2025)
von: Duan, Ranjie, et al.
Veröffentlicht: (2025)
RubRIX: Rubric-Driven Risk Mitigation in Caregiver-AI Interactions
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
Limited Effectiveness of LLM-based Data Augmentation for COVID-19 Misinformation Stance Detection
von: Choi, Eun Cheol, et al.
Veröffentlicht: (2025)
von: Choi, Eun Cheol, et al.
Veröffentlicht: (2025)
CogErgLLM: Exploring Large Language Model Systems Design Perspective Using Cognitive Ergonomics
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
VERA-MH Concept Paper
von: Belli, Luca, et al.
Veröffentlicht: (2025)
von: Belli, Luca, et al.
Veröffentlicht: (2025)
Mind the Gap! Pathways Towards Unifying AI Safety and Ethics Research
von: Roytburg, Dani, et al.
Veröffentlicht: (2025)
von: Roytburg, Dani, et al.
Veröffentlicht: (2025)
LLMs Among Us: Generative AI Participating in Digital Discourse
von: Radivojevic, Kristina, et al.
Veröffentlicht: (2024)
von: Radivojevic, Kristina, et al.
Veröffentlicht: (2024)
OpenClaw AI Agents as Informal Learners at Moltbook: Characterizing an Emergent Learning Community at Scale
von: Chen, Eason, et al.
Veröffentlicht: (2026)
von: Chen, Eason, et al.
Veröffentlicht: (2026)
Reranking partisan animosity in algorithmic social media feeds alters affective polarization
von: Piccardi, Tiziano, et al.
Veröffentlicht: (2024)
von: Piccardi, Tiziano, et al.
Veröffentlicht: (2024)
When AI Agents Teach Each Other: Discourse Patterns Resembling Peer Learning in the Moltbook Community
von: Chen, Eason, et al.
Veröffentlicht: (2026)
von: Chen, Eason, et al.
Veröffentlicht: (2026)
New contexts, old heuristics: How young people in India and the US trust online content in the age of generative AI
von: Xu, Rachel, et al.
Veröffentlicht: (2024)
von: Xu, Rachel, et al.
Veröffentlicht: (2024)
VIRENA: Virtual Arena for Research, Education, and Democratic Innovation
von: Hoes, Emma, et al.
Veröffentlicht: (2026)
von: Hoes, Emma, et al.
Veröffentlicht: (2026)
Mapping the Scholarship of Dark Pattern Regulation: A Systematic Review of Concepts, Regulatory Paradigms, and Solutions from an Interdisciplinary Perspective
von: Yi, Weiwei, et al.
Veröffentlicht: (2024)
von: Yi, Weiwei, et al.
Veröffentlicht: (2024)
Mind the Style: Impact of Communication Style on Human-Chatbot Interaction
von: Derner, Erik, et al.
Veröffentlicht: (2026)
von: Derner, Erik, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
AI Content Moderation in Therapy Conversations
von: Kim, Jiwon, et al.
Veröffentlicht: (2026) -
AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals
von: Pasch, Stefan
Veröffentlicht: (2025) -
Bottom-Up Perspectives on AI Governance: Insights from User Reviews of AI Products
von: Pasch, Stefan
Veröffentlicht: (2025) -
Human-AI Interaction and User Satisfaction: Empirical Evidence from Online Reviews of AI Products
von: Pasch, Stefan, et al.
Veröffentlicht: (2025) -
Operationalizing AI: Empirical Evidence on MLOps Practices, User Satisfaction, and Organizational Context
von: Pasch, Stefan
Veröffentlicht: (2025)