HateModerate: Testing Hate Speech Detectors against Content Moderation Policies
Fuente:
arXiv
Salvato in:
| Autori principali: | Zheng, Jiangrui, Liu, Xueqing, Yang, Guanqun, Haque, Mirazul, Qian, Xing, Rathnasuriya, Ravishka, Yang, Wei, Budhrani, Girish |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On-the-Fly Input Adaptation for Reliable Code Intelligence
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2026)
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2026)
When to Answer and When to Defer: A Decision Framework for Reliable Code Predictions
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2026)
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2026)
Framework for On the Fly Input Refinement for Deep Learning Models
di: Rathnasuriya, Ravishka
Pubblicazione: (2025)
di: Rathnasuriya, Ravishka
Pubblicazione: (2025)
CodeImprove: Program Adaptation for Deep Code Models
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2025)
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2025)
Fast and Accurate Silent Vulnerability Fix Retrieval
di: Liu, Xueqing, et al.
Pubblicazione: (2025)
di: Liu, Xueqing, et al.
Pubblicazione: (2025)
Characterizing Real-World Bugs in Tile Programs for Automated Bug Detection
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2026)
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2026)
Can You Mimic Me? Exploring the Use of Android Record & Replay Tools in Debugging
di: Song, Zihe, et al.
Pubblicazione: (2025)
di: Song, Zihe, et al.
Pubblicazione: (2025)
Can Highlighting Help GitHub Maintainers Track Security Fixes?
di: Liu, Xueqing, et al.
Pubblicazione: (2024)
di: Liu, Xueqing, et al.
Pubblicazione: (2024)
Metamorphic Testing for Audio Content Moderation Software
di: Wang, Wenxuan, et al.
Pubblicazione: (2025)
di: Wang, Wenxuan, et al.
Pubblicazione: (2025)
Towards Effectively Leveraging Execution Traces for Program Repair with Code LLMs
di: Haque, Mirazul, et al.
Pubblicazione: (2025)
di: Haque, Mirazul, et al.
Pubblicazione: (2025)
From Reviewers' Lens: Understanding Bug Bounty Report Invalid Reasons with LLMs
di: Zheng, Jiangrui, et al.
Pubblicazione: (2025)
di: Zheng, Jiangrui, et al.
Pubblicazione: (2025)
Do Autonomous Agents Contribute Test Code? A Study of Tests in Agentic Pull Requests
di: Haque, Sabrina, et al.
Pubblicazione: (2026)
di: Haque, Sabrina, et al.
Pubblicazione: (2026)
Moderately Mighty: To What Extent Can Internal Software Metrics Predict App Popularity at Launch?
di: Opu, Md Nahidul Islam, et al.
Pubblicazione: (2025)
di: Opu, Md Nahidul Islam, et al.
Pubblicazione: (2025)
Log-based, Business-aware REST API Testing
di: Yang, Ding, et al.
Pubblicazione: (2026)
di: Yang, Ding, et al.
Pubblicazione: (2026)
Automated Unit Test Refactoring
di: Gao, Yi, et al.
Pubblicazione: (2024)
di: Gao, Yi, et al.
Pubblicazione: (2024)
Exploiting Efficiency Vulnerabilities in Dynamic Deep Learning Systems
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2025)
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2025)
xNose: A Test Smell Detector for C#
di: Paul, Partha P., et al.
Pubblicazione: (2024)
di: Paul, Partha P., et al.
Pubblicazione: (2024)
Beyond Binary Moderation: Identifying Fine-Grained Sexist and Misogynistic Behavior on GitHub with Large Language Models
di: Dev, Tanni, et al.
Pubblicazione: (2025)
di: Dev, Tanni, et al.
Pubblicazione: (2025)
Vulnerability-Triggering Test Case Generation from Third-Party Libraries
di: Gao, Yi, et al.
Pubblicazione: (2024)
di: Gao, Yi, et al.
Pubblicazione: (2024)
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns
di: Shen, Xinyue, et al.
Pubblicazione: (2025)
di: Shen, Xinyue, et al.
Pubblicazione: (2025)
How Quantization Impacts Privacy Risk on LLMs for Code?
di: Haque, Md Nazmul, et al.
Pubblicazione: (2025)
di: Haque, Md Nazmul, et al.
Pubblicazione: (2025)
SeqTG: Scalable Combinatorial Test Generation via Sequential Integer Linear Programming
di: Yang, Sitong, et al.
Pubblicazione: (2026)
di: Yang, Sitong, et al.
Pubblicazione: (2026)
HateBuffer: Safeguarding Content Moderators' Mental Well-Being through Hate Speech Content Modification
di: Park, Subin, et al.
Pubblicazione: (2025)
di: Park, Subin, et al.
Pubblicazione: (2025)
Cracking CodeWhisperer: Analyzing Developers' Interactions and Patterns During Programming Tasks
di: Javahar, Jeena, et al.
Pubblicazione: (2025)
di: Javahar, Jeena, et al.
Pubblicazione: (2025)
Efficiency Robustness of Dynamic Deep Learning Systems
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2025)
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2025)
Compiler Optimization Testing Based on Optimization-Guided Equivalence Transformations
di: Wu, Jingwen, et al.
Pubblicazione: (2025)
di: Wu, Jingwen, et al.
Pubblicazione: (2025)
Policy Testing with MDPFuzz (Replicability Study)
di: Mazouni, Quentin, et al.
Pubblicazione: (2025)
di: Mazouni, Quentin, et al.
Pubblicazione: (2025)
DEFT: Differentiable Automatic Test Pattern Generation
di: Li, Wei, et al.
Pubblicazione: (2025)
di: Li, Wei, et al.
Pubblicazione: (2025)
Lifecycle-Aware code generation: Leveraging Software Engineering Phases in LLMs
di: Xing, Xing, et al.
Pubblicazione: (2025)
di: Xing, Xing, et al.
Pubblicazione: (2025)
Reflective Unit Test Generation for Precise Type Error Detection with Large Language Models
di: Yang, Chen, et al.
Pubblicazione: (2025)
di: Yang, Chen, et al.
Pubblicazione: (2025)
Beyond Fixed Tests: Repository-Level Issue Resolution as Coevolution of Code and Behavioral Constraints
di: Li, Kefan, et al.
Pubblicazione: (2026)
di: Li, Kefan, et al.
Pubblicazione: (2026)
BIDO: An Out-Of-Distribution Resistant Image-based Malware Detector
di: Wang, Wei, et al.
Pubblicazione: (2025)
di: Wang, Wei, et al.
Pubblicazione: (2025)
Beyond Accuracy: Policy Invariance as a Reliability Test for LLM Safety Judges
di: Weng, Shihao, et al.
Pubblicazione: (2026)
di: Weng, Shihao, et al.
Pubblicazione: (2026)
Deep Learning Framework Testing via Model Mutation: How Far Are We?
di: Mu, Yanzhou, et al.
Pubblicazione: (2025)
di: Mu, Yanzhou, et al.
Pubblicazione: (2025)
SAFE: Harnessing LLM for Scenario-Driven ADS Testing from Multimodal Crash Data
di: Luo, Siwei, et al.
Pubblicazione: (2025)
di: Luo, Siwei, et al.
Pubblicazione: (2025)
Deep Reinforcement Learning for Automated Web GUI Testing
di: Gu, Zhiyu, et al.
Pubblicazione: (2025)
di: Gu, Zhiyu, et al.
Pubblicazione: (2025)
Cast: Automated Resilience Testing for Production Cloud Service Systems
di: Chen, Zhuangbin, et al.
Pubblicazione: (2026)
di: Chen, Zhuangbin, et al.
Pubblicazione: (2026)
Hate‐UDF: Explainable Hateful Meme Detection With Uncertainty‐Aware Dynamic Fusion
di: Xia Lei, et al.
Pubblicazione: (2024)
di: Xia Lei, et al.
Pubblicazione: (2024)
An Exploratory Study on Build Issue Resolution Among Computer Science Students
di: Huang, Sunzhou, et al.
Pubblicazione: (2025)
di: Huang, Sunzhou, et al.
Pubblicazione: (2025)
CITYWALK: Enhancing LLM-Based C++ Unit Test Generation via Project-Dependency Awareness and Language-Specific Knowledge
di: Zhang, Yuwei, et al.
Pubblicazione: (2025)
di: Zhang, Yuwei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
On-the-Fly Input Adaptation for Reliable Code Intelligence
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2026) -
When to Answer and When to Defer: A Decision Framework for Reliable Code Predictions
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2026) -
Framework for On the Fly Input Refinement for Deep Learning Models
di: Rathnasuriya, Ravishka
Pubblicazione: (2025) -
CodeImprove: Program Adaptation for Deep Code Models
di: Rathnasuriya, Ravishka, et al.
Pubblicazione: (2025) -
Fast and Accurate Silent Vulnerability Fix Retrieval
di: Liu, Xueqing, et al.
Pubblicazione: (2025)