SGHateCheck: Functional Tests for Detecting Hate Speech in Low-Resource Languages of Singapore
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ng, Ri Chi, Prakash, Nirmalendu, Hee, Ming Shan, Choo, Kenny Tsu Wei, Lee, Roy Ka-Wei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SEAHateCheck: Functional Tests for Detecting Hate Speech in Low-Resource Languages of Southeast Asia
von: Ng, Ri Chi, et al.
Veröffentlicht: (2026)
von: Ng, Ri Chi, et al.
Veröffentlicht: (2026)
Demystifying Hateful Content: Leveraging Large Multimodal Models for Hateful Meme Detection with Explainable Decisions
von: Hee, Ming Shan, et al.
Veröffentlicht: (2025)
von: Hee, Ming Shan, et al.
Veröffentlicht: (2025)
Bridging Modalities: Enhancing Cross-Modality Hate Speech Detection with Few-Shot In-Context Learning
von: Hee, Ming Shan, et al.
Veröffentlicht: (2024)
von: Hee, Ming Shan, et al.
Veröffentlicht: (2024)
Toxicity Red-Teaming: Benchmarking LLM Safety in Singapore's Low-Resource Languages
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
Interpreting Bias in Large Language Models: A Feature-Based Approach
von: Prakash, Nirmalendu, et al.
Veröffentlicht: (2024)
von: Prakash, Nirmalendu, et al.
Veröffentlicht: (2024)
"I Said Things I Needed to Hear Myself": Peer Support as an Emotional, Organisational, and Sociotechnical Practice in Singapore
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2025)
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2025)
Human-AI Alignment of Multimodal Large Language Models with Speech-Language Pathologists in Parent-Child Interactions
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
ToxiCloakCN: Evaluating Robustness of Offensive Language Detection in Chinese with Cloaking Perturbations
von: Xiao, Yunze, et al.
Veröffentlicht: (2024)
von: Xiao, Yunze, et al.
Veröffentlicht: (2024)
"Is This Really a Human Peer Supporter?": Misalignments Between Peer Supporters and Experts in LLM-Supported Interactions
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2025)
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2025)
When Drawing Is Not Enough: Exploring Spontaneous Speech with Sketch for Intent Alignment in Multimodal LLMs
von: Shi, Weiyan, et al.
Veröffentlicht: (2026)
von: Shi, Weiyan, et al.
Veröffentlicht: (2026)
Humor in Pixels: Benchmarking Large Multimodal Models Understanding of Online Comics
von: Ryan, Yuriel, et al.
Veröffentlicht: (2025)
von: Ryan, Yuriel, et al.
Veröffentlicht: (2025)
A Taxonomy of Human--MLLM Interaction in Early-Stage Sketch-Based Design Ideation
von: Shi, Weiyan, et al.
Veröffentlicht: (2026)
von: Shi, Weiyan, et al.
Veröffentlicht: (2026)
Towards Aligning Multimodal LLMs with Human Experts: A Focus on Parent-Child Interaction
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
More Than 1v1: Human-AI Alignment in Early Developmental Communities with Multimodal LLMs
von: Shi, Weiyan, et al.
Veröffentlicht: (2026)
von: Shi, Weiyan, et al.
Veröffentlicht: (2026)
Recent Advances in Hate Speech Moderation: Multimodality and the Role of Large Models
von: Hee, Ming Shan, et al.
Veröffentlicht: (2024)
von: Hee, Ming Shan, et al.
Veröffentlicht: (2024)
HateXScore: A Metric Suite for Evaluating Reasoning Quality in Hate Speech Explanations
von: Hu, Yujia, et al.
Veröffentlicht: (2026)
von: Hu, Yujia, et al.
Veröffentlicht: (2026)
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
von: Jin, Yiping, et al.
Veröffentlicht: (2024)
von: Jin, Yiping, et al.
Veröffentlicht: (2024)
TalkSketch: Multimodal Generative AI for Real-time Sketch Ideation with Speech
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
Foreign Domestic Workers' Perspectives on an LLM-Based Emotional Support tool for Caregiving Burden
von: Teng, Shin Shoon Nicholas, et al.
Veröffentlicht: (2026)
von: Teng, Shin Shoon Nicholas, et al.
Veröffentlicht: (2026)
Redistributing Voice and Responsibility: AI in Relationship-Centred Care
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2026)
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2026)
Envisioning an AI-Enhanced Mental Health Ecosystem
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2025)
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2025)
"I'm Not Able to Be There for You": Emotional Labour, Responsibility, and AI in Peer Support
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2026)
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2026)
HateClipSeg: A Segment-Level Annotated Dataset for Fine-Grained Hate Video Detection
von: Wang, Han, et al.
Veröffentlicht: (2025)
von: Wang, Han, et al.
Veröffentlicht: (2025)
Exploring Gaze Pattern Differences Between Autistic and Neurotypical Children: Clustering, Visualisation, and Prediction
von: Shi, Weiyan, et al.
Veröffentlicht: (2024)
von: Shi, Weiyan, et al.
Veröffentlicht: (2024)
Examining Augmented Virtuality Impairment Simulation for Mobile App Accessibility Design
von: Choo, Kenny Tsu Wei, et al.
Veröffentlicht: (2025)
von: Choo, Kenny Tsu Wei, et al.
Veröffentlicht: (2025)
Towards Multimodal Large-Language Models for Parent-Child Interaction: A Focus on Joint Attention
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
Sword and Shield: Uses and Strategies of LLMs in Navigating Disinformation
von: Lim, Gionnieve, et al.
Veröffentlicht: (2025)
von: Lim, Gionnieve, et al.
Veröffentlicht: (2025)
Modularized Networks for Few-shot Hateful Meme Detection
von: Cao, Rui, et al.
Veröffentlicht: (2024)
von: Cao, Rui, et al.
Veröffentlicht: (2024)
A Survey on Automatic Online Hate Speech Detection in Low-Resource Languages
von: Das, Susmita, et al.
Veröffentlicht: (2024)
von: Das, Susmita, et al.
Veröffentlicht: (2024)
Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal
von: Prakash, Nirmalendu, et al.
Veröffentlicht: (2025)
von: Prakash, Nirmalendu, et al.
Veröffentlicht: (2025)
Understanding Refusal in Language Models with Sparse Autoencoders
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2025)
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2025)
MultiHateClip: A Multilingual Benchmark Dataset for Hateful Video Detection on YouTube and Bilibili
von: Wang, Han, et al.
Veröffentlicht: (2024)
von: Wang, Han, et al.
Veröffentlicht: (2024)
When Hate Meets Facts: LLMs-in-the-Loop for Check-worthiness Detection in Hate Speech
von: Ocampo, Nicolás Benjamín, et al.
Veröffentlicht: (2026)
von: Ocampo, Nicolás Benjamín, et al.
Veröffentlicht: (2026)
Towards Understanding Emotions for Engaged Mental Health Conversations
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2024)
von: Sim, Kellie Yu Hui, et al.
Veröffentlicht: (2024)
BLEnD-Vis: Benchmarking Multimodal Cultural Understanding in Vision Language Models
von: Tan, Bryan Chen Zhengyu, et al.
Veröffentlicht: (2025)
von: Tan, Bryan Chen Zhengyu, et al.
Veröffentlicht: (2025)
Cross-Modal Transfer from Memes to Videos: Addressing Data Scarcity in Hateful Video Detection
von: Wang, Han, et al.
Veröffentlicht: (2025)
von: Wang, Han, et al.
Veröffentlicht: (2025)
Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages
von: Prome, Ruhina Tabasshum, et al.
Veröffentlicht: (2025)
von: Prome, Ruhina Tabasshum, et al.
Veröffentlicht: (2025)
LongGenBench: Benchmarking Long-Form Generation in Long Context LLMs
von: Wu, Yuhao, et al.
Veröffentlicht: (2024)
von: Wu, Yuhao, et al.
Veröffentlicht: (2024)
HateModerate: Testing Hate Speech Detectors against Content Moderation Policies
von: Zheng, Jiangrui, et al.
Veröffentlicht: (2023)
von: Zheng, Jiangrui, et al.
Veröffentlicht: (2023)
Multimodal Zero-Shot Framework for Deepfake Hate Speech Detection in Low-Resource Languages
von: Ranjan, Rishabh, et al.
Veröffentlicht: (2025)
von: Ranjan, Rishabh, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SEAHateCheck: Functional Tests for Detecting Hate Speech in Low-Resource Languages of Southeast Asia
von: Ng, Ri Chi, et al.
Veröffentlicht: (2026) -
Demystifying Hateful Content: Leveraging Large Multimodal Models for Hateful Meme Detection with Explainable Decisions
von: Hee, Ming Shan, et al.
Veröffentlicht: (2025) -
Bridging Modalities: Enhancing Cross-Modality Hate Speech Detection with Few-Shot In-Context Learning
von: Hee, Ming Shan, et al.
Veröffentlicht: (2024) -
Toxicity Red-Teaming: Benchmarking LLM Safety in Singapore's Low-Resource Languages
von: Hu, Yujia, et al.
Veröffentlicht: (2025) -
Interpreting Bias in Large Language Models: A Feature-Based Approach
von: Prakash, Nirmalendu, et al.
Veröffentlicht: (2024)