Identity-related Speech Suppression in Generative AI Content Moderation
Fuente:
arXiv
Guardado en:
| Autores principales: | Proebsting, Grace, Anigboro, Oghenefejiro Isaacs, Crawford, Charlie M., Metaxa, Danaé, Friedler, Sorelle A. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Longitudinal Monitoring of LLM Content Moderation of Social Issues
por: Dai, Yunlang, et al.
Publicado: (2025)
por: Dai, Yunlang, et al.
Publicado: (2025)
Generative AI and Perceptual Harms: Who's Suspected of using LLMs?
por: Kadoma, Kowe, et al.
Publicado: (2024)
por: Kadoma, Kowe, et al.
Publicado: (2024)
The Hidden Language of Harm: Examining the Role of Emojis in Harmful Online Communication and Content Moderation
por: Zhou, Yuhang, et al.
Publicado: (2025)
por: Zhou, Yuhang, et al.
Publicado: (2025)
Cooperative Speech, Semantic Competence, and AI
por: Almotahari, Mahrad
Publicado: (2025)
por: Almotahari, Mahrad
Publicado: (2025)
AI Content Moderation in Therapy Conversations
por: Kim, Jiwon, et al.
Publicado: (2026)
por: Kim, Jiwon, et al.
Publicado: (2026)
Overreliance on AI in Information-seeking from Video Content
por: Møller, Anders Giovanni, et al.
Publicado: (2026)
por: Møller, Anders Giovanni, et al.
Publicado: (2026)
Youth as Peer Auditors: Engaging Teenagers with Algorithm Auditing of Machine Learning Applications
por: Morales-Navarro, Luis, et al.
Publicado: (2024)
por: Morales-Navarro, Luis, et al.
Publicado: (2024)
Content Moderation Futures
por: Blackwell, Lindsay
Publicado: (2025)
por: Blackwell, Lindsay
Publicado: (2025)
Lower Quantity, Higher Quality: Auditing News Content and User Perceptions on Twitter/X Algorithmic versus Chronological Timelines
por: Wang, Stephanie, et al.
Publicado: (2024)
por: Wang, Stephanie, et al.
Publicado: (2024)
Watching the Watchers: A Comparative Fairness Audit of Cloud-based Content Moderation Services
por: Hartmann, David, et al.
Publicado: (2024)
por: Hartmann, David, et al.
Publicado: (2024)
Lost in Moderation: How Commercial Content Moderation APIs Over- and Under-Moderate Group-Targeted Hate Speech and Linguistic Variations
por: Hartmann, David, et al.
Publicado: (2025)
por: Hartmann, David, et al.
Publicado: (2025)
Learning AI Auditing: A Case Study of Teenagers Auditing a Generative AI Model
por: Morales-Navarro, Luis, et al.
Publicado: (2025)
por: Morales-Navarro, Luis, et al.
Publicado: (2025)
Eskwai for Students: Generative AI Assistant for Legal Education in Ghana
por: Boateng, George, et al.
Publicado: (2026)
por: Boateng, George, et al.
Publicado: (2026)
Improving the TENOR of Labeling: Re-evaluating Topic Models for Content Analysis
por: Li, Zongxia, et al.
Publicado: (2024)
por: Li, Zongxia, et al.
Publicado: (2024)
Auditing African Content Moderators' Working Conditions by Using the European General Data Protection Regulation (GDPR)
por: Tighanimine, Mariame, et al.
Publicado: (2026)
por: Tighanimine, Mariame, et al.
Publicado: (2026)
Human Capital Visualization using Speech Amount during Meetings
por: Hashimoto, Ekai, et al.
Publicado: (2025)
por: Hashimoto, Ekai, et al.
Publicado: (2025)
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
por: Samory, Mattia, et al.
Publicado: (2025)
por: Samory, Mattia, et al.
Publicado: (2025)
Critical Challenges in Content Moderation for People Who Use Drugs (PWUD): Insights into Online Harm Reduction Practices from Moderators
por: Wang, Kaixuan, et al.
Publicado: (2025)
por: Wang, Kaixuan, et al.
Publicado: (2025)
Kwame 2.0: Human-in-the-Loop Generative AI Teaching Assistant for Large Scale Online Coding Education in Africa
por: Boateng, George, et al.
Publicado: (2026)
por: Boateng, George, et al.
Publicado: (2026)
Content Moderation Justice and Fairness on Social Media: Comparisons Across Different Contexts and Platforms
por: Cai, Jie, et al.
Publicado: (2024)
por: Cai, Jie, et al.
Publicado: (2024)
Can AI writing be salvaged? Mitigating Idiosyncrasies and Improving Human-AI Alignment in the Writing Process through Edits
por: Chakrabarty, Tuhin, et al.
Publicado: (2024)
por: Chakrabarty, Tuhin, et al.
Publicado: (2024)
AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals
por: Pasch, Stefan
Publicado: (2025)
por: Pasch, Stefan
Publicado: (2025)
Effects of a Prompt Engineering Intervention on Undergraduate Students' AI Self-Efficacy, AI Knowledge and Prompt Engineering Ability: A Mixed Methods Study
por: Woo, David James, et al.
Publicado: (2024)
por: Woo, David James, et al.
Publicado: (2024)
MONAL: Model Autophagy Analysis for Modeling Human-AI Interactions
por: Yang, Shu, et al.
Publicado: (2024)
por: Yang, Shu, et al.
Publicado: (2024)
Voices of Freelance Professional Writers on AI: Limitations, Expectations, and Fears
por: Ivanova, Anastasiia, et al.
Publicado: (2025)
por: Ivanova, Anastasiia, et al.
Publicado: (2025)
"You Can Actually Do Something'': Shifts in High School Computer Science Teachers' Conceptions of AI/ML Systems and Algorithmic Justice
por: Noh, Daniel J., et al.
Publicado: (2026)
por: Noh, Daniel J., et al.
Publicado: (2026)
The Impact and Feasibility of Self-Confidence Shaping for AI-Assisted Decision-Making
por: Takayanagi, Takehiro, et al.
Publicado: (2025)
por: Takayanagi, Takehiro, et al.
Publicado: (2025)
Governance of AI-Generated Content: A Case Study on Social Media Platforms
por: Gao, Lan, et al.
Publicado: (2026)
por: Gao, Lan, et al.
Publicado: (2026)
Threefold model for AI Readiness: A Case Study with Finnish Healthcare SMEs
por: Alnajjar, Mohammed, et al.
Publicado: (2025)
por: Alnajjar, Mohammed, et al.
Publicado: (2025)
Leveraging AI to Advance Science and Computing Education across Africa: Challenges, Progress and Opportunities
por: Boateng, George
Publicado: (2024)
por: Boateng, George
Publicado: (2024)
Exploring EFL Secondary Students' AI-generated Text Editing While Composition Writing
por: Woo, David James, et al.
Publicado: (2025)
por: Woo, David James, et al.
Publicado: (2025)
Generative AI Perceptions: A Survey to Measure the Perceptions of Faculty, Staff, and Students on Generative AI Tools in Academia
por: Amani, Sara, et al.
Publicado: (2023)
por: Amani, Sara, et al.
Publicado: (2023)
"Would You Want an AI Tutor?" Understanding Stakeholder Perceptions of LLM-based Systems in the Classroom
por: Fuligni, Caterina, et al.
Publicado: (2025)
por: Fuligni, Caterina, et al.
Publicado: (2025)
When Algorithms Meet Artists: Semantic Compression of Artists' Concerns in the Public AI-Art Debate
por: Mukherjee-Gandhi, Ariya, et al.
Publicado: (2025)
por: Mukherjee-Gandhi, Ariya, et al.
Publicado: (2025)
AI as a deliberative partner fosters intercultural empathy for Americans but fails for Latin American participants
por: Villanueva, Isabel, et al.
Publicado: (2025)
por: Villanueva, Isabel, et al.
Publicado: (2025)
Human-centered NLP Fact-checking: Co-Designing with Fact-checkers using Matchmaking for AI
por: Liu, Houjiang, et al.
Publicado: (2023)
por: Liu, Houjiang, et al.
Publicado: (2023)
Deconstructing Depression Stigma: Integrating AI-driven Data Collection and Analysis with Causal Knowledge Graphs
por: Meng, Han, et al.
Publicado: (2025)
por: Meng, Han, et al.
Publicado: (2025)
LLM Content Moderation and User Satisfaction: Evidence from Response Refusals in Chatbot Arena
por: Pasch, Stefan
Publicado: (2025)
por: Pasch, Stefan
Publicado: (2025)
AI-Generated Slides: Are They Good? Can Students Tell?
por: Leinonen, Juho, et al.
Publicado: (2026)
por: Leinonen, Juho, et al.
Publicado: (2026)
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
por: Arita, Takaya, et al.
Publicado: (2025)
por: Arita, Takaya, et al.
Publicado: (2025)
Ejemplares similares
-
Longitudinal Monitoring of LLM Content Moderation of Social Issues
por: Dai, Yunlang, et al.
Publicado: (2025) -
Generative AI and Perceptual Harms: Who's Suspected of using LLMs?
por: Kadoma, Kowe, et al.
Publicado: (2024) -
The Hidden Language of Harm: Examining the Role of Emojis in Harmful Online Communication and Content Moderation
por: Zhou, Yuhang, et al.
Publicado: (2025) -
Cooperative Speech, Semantic Competence, and AI
por: Almotahari, Mahrad
Publicado: (2025) -
AI Content Moderation in Therapy Conversations
por: Kim, Jiwon, et al.
Publicado: (2026)