Longitudinal Monitoring of LLM Content Moderation of Social Issues
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dai, Yunlang, Lurie, Emma, Metaxa, Danaé, Friedler, Sorelle A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Identity-related Speech Suppression in Generative AI Content Moderation
von: Proebsting, Grace, et al.
Veröffentlicht: (2024)
von: Proebsting, Grace, et al.
Veröffentlicht: (2024)
The Hidden Language of Harm: Examining the Role of Emojis in Harmful Online Communication and Content Moderation
von: Zhou, Yuhang, et al.
Veröffentlicht: (2025)
von: Zhou, Yuhang, et al.
Veröffentlicht: (2025)
Youth as Peer Auditors: Engaging Teenagers with Algorithm Auditing of Machine Learning Applications
von: Morales-Navarro, Luis, et al.
Veröffentlicht: (2024)
von: Morales-Navarro, Luis, et al.
Veröffentlicht: (2024)
Content Moderation Futures
von: Blackwell, Lindsay
Veröffentlicht: (2025)
von: Blackwell, Lindsay
Veröffentlicht: (2025)
Lower Quantity, Higher Quality: Auditing News Content and User Perceptions on Twitter/X Algorithmic versus Chronological Timelines
von: Wang, Stephanie, et al.
Veröffentlicht: (2024)
von: Wang, Stephanie, et al.
Veröffentlicht: (2024)
Watching the Watchers: A Comparative Fairness Audit of Cloud-based Content Moderation Services
von: Hartmann, David, et al.
Veröffentlicht: (2024)
von: Hartmann, David, et al.
Veröffentlicht: (2024)
Content Moderation Justice and Fairness on Social Media: Comparisons Across Different Contexts and Platforms
von: Cai, Jie, et al.
Veröffentlicht: (2024)
von: Cai, Jie, et al.
Veröffentlicht: (2024)
Generative AI and Perceptual Harms: Who's Suspected of using LLMs?
von: Kadoma, Kowe, et al.
Veröffentlicht: (2024)
von: Kadoma, Kowe, et al.
Veröffentlicht: (2024)
LLM Content Moderation and User Satisfaction: Evidence from Response Refusals in Chatbot Arena
von: Pasch, Stefan
Veröffentlicht: (2025)
von: Pasch, Stefan
Veröffentlicht: (2025)
Safeguarding Decentralized Social Media: LLM Agents for Automating Community Rule Compliance
von: La Cava, Lucio, et al.
Veröffentlicht: (2024)
von: La Cava, Lucio, et al.
Veröffentlicht: (2024)
AI Content Moderation in Therapy Conversations
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
Overreliance on AI in Information-seeking from Video Content
von: Møller, Anders Giovanni, et al.
Veröffentlicht: (2026)
von: Møller, Anders Giovanni, et al.
Veröffentlicht: (2026)
Improving the TENOR of Labeling: Re-evaluating Topic Models for Content Analysis
von: Li, Zongxia, et al.
Veröffentlicht: (2024)
von: Li, Zongxia, et al.
Veröffentlicht: (2024)
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
von: Samory, Mattia, et al.
Veröffentlicht: (2025)
von: Samory, Mattia, et al.
Veröffentlicht: (2025)
Critical Challenges in Content Moderation for People Who Use Drugs (PWUD): Insights into Online Harm Reduction Practices from Moderators
von: Wang, Kaixuan, et al.
Veröffentlicht: (2025)
von: Wang, Kaixuan, et al.
Veröffentlicht: (2025)
Emotionally Aware Moderation: The Potential of Emotion Monitoring in Shaping Healthier Social Media Conversations
von: Su, Xiaotian, et al.
Veröffentlicht: (2025)
von: Su, Xiaotian, et al.
Veröffentlicht: (2025)
Social Welfare Function Leaderboard: When LLM Agents Allocate Social Welfare
von: Shi, Zhengliang, et al.
Veröffentlicht: (2025)
von: Shi, Zhengliang, et al.
Veröffentlicht: (2025)
SOTOPIA-$Ω$: Dynamic Strategy Injection Learning and Social Instruction Following Evaluation for Social Agents
von: Zhang, Wenyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Wenyuan, et al.
Veröffentlicht: (2025)
AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals
von: Pasch, Stefan
Veröffentlicht: (2025)
von: Pasch, Stefan
Veröffentlicht: (2025)
LLM Social Simulations Are a Promising Research Method
von: Anthis, Jacy Reese, et al.
Veröffentlicht: (2025)
von: Anthis, Jacy Reese, et al.
Veröffentlicht: (2025)
DiMA: An LLM-Powered Ride-Hailing Assistant at DiDi
von: Ning, Yansong, et al.
Veröffentlicht: (2025)
von: Ning, Yansong, et al.
Veröffentlicht: (2025)
Exploring the Human-LLM Synergy in Advancing Theory-driven Qualitative Analysis
von: Meng, Han, et al.
Veröffentlicht: (2024)
von: Meng, Han, et al.
Veröffentlicht: (2024)
Lost in Moderation: How Commercial Content Moderation APIs Over- and Under-Moderate Group-Targeted Hate Speech and Linguistic Variations
von: Hartmann, David, et al.
Veröffentlicht: (2025)
von: Hartmann, David, et al.
Veröffentlicht: (2025)
Redefining Research Crowdsourcing: Incorporating Human Feedback with LLM-Powered Digital Twins
von: Chan, Amanda, et al.
Veröffentlicht: (2025)
von: Chan, Amanda, et al.
Veröffentlicht: (2025)
"Would You Want an AI Tutor?" Understanding Stakeholder Perceptions of LLM-based Systems in the Classroom
von: Fuligni, Caterina, et al.
Veröffentlicht: (2025)
von: Fuligni, Caterina, et al.
Veröffentlicht: (2025)
Auditing African Content Moderators' Working Conditions by Using the European General Data Protection Regulation (GDPR)
von: Tighanimine, Mariame, et al.
Veröffentlicht: (2026)
von: Tighanimine, Mariame, et al.
Veröffentlicht: (2026)
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
von: Badawi, Abeer, et al.
Veröffentlicht: (2025)
von: Badawi, Abeer, et al.
Veröffentlicht: (2025)
StereoDetect: Detecting Stereotypes and Anti-stereotypes the Correct Way Using Social Psychological Underpinnings
von: Shejole, Kaustubh Shivshankar, et al.
Veröffentlicht: (2025)
von: Shejole, Kaustubh Shivshankar, et al.
Veröffentlicht: (2025)
Socially Constructed Treatment Plans: Analyzing Online Peer Interactions to Understand How Patients Navigate Complex Medical Conditions
von: Basak, Madhusudan, et al.
Veröffentlicht: (2025)
von: Basak, Madhusudan, et al.
Veröffentlicht: (2025)
Artificial Leviathan: Exploring Social Evolution of LLM Agents Through the Lens of Hobbesian Social Contract Theory
von: Dai, Gordon, et al.
Veröffentlicht: (2024)
von: Dai, Gordon, et al.
Veröffentlicht: (2024)
Not Just Novelty: A Longitudinal Study on Utility and Customization of an AI Workflow
von: Long, Tao, et al.
Veröffentlicht: (2024)
von: Long, Tao, et al.
Veröffentlicht: (2024)
Wisdom of the LLM Crowd: A Large Scale Benchmark of Multi-Label U.S. Election-Related Harmful Social Media Content
von: Wang, Qile, et al.
Veröffentlicht: (2026)
von: Wang, Qile, et al.
Veröffentlicht: (2026)
Does Writing with Language Models Reduce Content Diversity?
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2023)
von: Padmakumar, Vishakh, et al.
Veröffentlicht: (2023)
"You Can Actually Do Something'': Shifts in High School Computer Science Teachers' Conceptions of AI/ML Systems and Algorithmic Justice
von: Noh, Daniel J., et al.
Veröffentlicht: (2026)
von: Noh, Daniel J., et al.
Veröffentlicht: (2026)
Throwaway Accounts and Moderation on Reddit
von: Guo, Cheng, et al.
Veröffentlicht: (2025)
von: Guo, Cheng, et al.
Veröffentlicht: (2025)
Personal Moderation Configurations on Facebook: Exploring the Role of FoMO, Social Media Addiction, Norms, and Platform Trust
von: Jhaver, Shagun
Veröffentlicht: (2024)
von: Jhaver, Shagun
Veröffentlicht: (2024)
Safe Spaces or Toxic Places? Content Moderation and Social Dynamics of Online Eating Disorder Communities
von: Lerman, Kristina, et al.
Veröffentlicht: (2024)
von: Lerman, Kristina, et al.
Veröffentlicht: (2024)
Understanding the Perceptions of Trigger Warning and Content Warning on Social Media Platforms in the U.S
von: Zhang, Xinyi, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyi, et al.
Veröffentlicht: (2025)
Governance of AI-Generated Content: A Case Study on Social Media Platforms
von: Gao, Lan, et al.
Veröffentlicht: (2026)
von: Gao, Lan, et al.
Veröffentlicht: (2026)
LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation
von: Ghosh, Himel, et al.
Veröffentlicht: (2026)
von: Ghosh, Himel, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Identity-related Speech Suppression in Generative AI Content Moderation
von: Proebsting, Grace, et al.
Veröffentlicht: (2024) -
The Hidden Language of Harm: Examining the Role of Emojis in Harmful Online Communication and Content Moderation
von: Zhou, Yuhang, et al.
Veröffentlicht: (2025) -
Youth as Peer Auditors: Engaging Teenagers with Algorithm Auditing of Machine Learning Applications
von: Morales-Navarro, Luis, et al.
Veröffentlicht: (2024) -
Content Moderation Futures
von: Blackwell, Lindsay
Veröffentlicht: (2025) -
Lower Quantity, Higher Quality: Auditing News Content and User Perceptions on Twitter/X Algorithmic versus Chronological Timelines
von: Wang, Stephanie, et al.
Veröffentlicht: (2024)