AI vs. Human Moderators: A Comparative Evaluation of Multimodal LLMs in Content Moderation for Brand Safety
Fuente:
arXiv
Saved in:
| Main Authors: | Levi, Adi, Levi, Or, Mishra, Sardhendu, Morra, Jonathan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Native Memes to Global Moderation: Cross-Cultural Evaluation of Vision-Language Models for Hateful Meme Detection
by: Wang, Mo, et al.
Published: (2026)
by: Wang, Mo, et al.
Published: (2026)
Cultural Encoding in Large Language Models: The Existence Gap in AI-Mediated Brand Discovery
by: Junyao, Huang, et al.
Published: (2025)
by: Junyao, Huang, et al.
Published: (2025)
A Grounded Memory System For Smart Personal Assistants
by: Ocker, Felix, et al.
Published: (2025)
by: Ocker, Felix, et al.
Published: (2025)
Aximorphic Perspective Projection Model for Immersive Imagery
by: Fober, Jakub Maksymilian
Published: (2021)
by: Fober, Jakub Maksymilian
Published: (2021)
Deterministic Fuzzy Triage for Legal Compliance Classification and Evidence Retrieval
by: Atri, Rian
Published: (2026)
by: Atri, Rian
Published: (2026)
The Cultural Gene of Large Language Models: A Study on the Impact of Cross-Corpus Training on Model Values and Biases
by: Fenech-Borg, Emanuel Z., et al.
Published: (2025)
by: Fenech-Borg, Emanuel Z., et al.
Published: (2025)
Harmful Terms and Where to Find Them: Measuring and Modeling Unfavorable Financial Terms and Conditions in Shopping Websites at Scale
by: Tsai, Elisa, et al.
Published: (2025)
by: Tsai, Elisa, et al.
Published: (2025)
Bhakti: A Lightweight Vector Database Management System for Endowing Large Language Models with Semantic Search Capabilities and Memory
by: Wu, Zihao
Published: (2025)
by: Wu, Zihao
Published: (2025)
TriAlignGR: Triangular Multitask Alignment with Multimodal Deep Interest Mining for Generative Recommendation
by: Zeng, Yangchen, et al.
Published: (2026)
by: Zeng, Yangchen, et al.
Published: (2026)
Who Leads in the Shadows? ERGM and Centrality Analysis of Congressional Democrats on Bluesky
by: Hew, Gordon, et al.
Published: (2025)
by: Hew, Gordon, et al.
Published: (2025)
Cross-Subreddit Behavior as Open-Source Indicators of Coordinated Influence: A Case Study of r/Sino & r/China
by: Pilaud, Manon, et al.
Published: (2025)
by: Pilaud, Manon, et al.
Published: (2025)
Leveraging OpenFlamingo for Multimodal Embedding Analysis of C2C Car Parts Data
by: Rashid, Maisha Binte, et al.
Published: (2025)
by: Rashid, Maisha Binte, et al.
Published: (2025)
A Scalable and High Availability Solution for Recommending Resolutions to Problem Tickets
by: Saragadam, Harish, et al.
Published: (2025)
by: Saragadam, Harish, et al.
Published: (2025)
Diversification as Risk Minimization
by: Takehi, Rikiya, et al.
Published: (2025)
by: Takehi, Rikiya, et al.
Published: (2025)
A Prompt-Aware Structuring Framework for Reliable Reuse of AI-Generated Content in the Agentic Web
by: Egami, Shusaku, et al.
Published: (2026)
by: Egami, Shusaku, et al.
Published: (2026)
Auditing Preferences for Brands and Cultures in LLMs
by: Rienecker, Jasmine, et al.
Published: (2026)
by: Rienecker, Jasmine, et al.
Published: (2026)
VocSim: A Training-free Benchmark for Zero-shot Content Identity in Single-source Audio
by: Basha, Maris, et al.
Published: (2025)
by: Basha, Maris, et al.
Published: (2025)
IMDMR: An Intelligent Multi-Dimensional Memory Retrieval System for Enhanced Conversational AI
by: Pawar, Tejas, et al.
Published: (2025)
by: Pawar, Tejas, et al.
Published: (2025)
ChronoMedKG: A Temporally-Grounded Biomedical Knowledge Graph and Benchmark for Clinical Reasoning
by: Ahmed, Md Shamim, et al.
Published: (2026)
by: Ahmed, Md Shamim, et al.
Published: (2026)
Real-World En Call Center Transcripts Dataset with PII Redaction
by: Dao, Ha, et al.
Published: (2025)
by: Dao, Ha, et al.
Published: (2025)
Experimenting active and sequential learning in a medieval music manuscript
by: Sharma, Sachin, et al.
Published: (2025)
by: Sharma, Sachin, et al.
Published: (2025)
Evaluating Perspectival Biases in Cross-Modal Retrieval
by: Saengsukhiran, Teerapol, et al.
Published: (2025)
by: Saengsukhiran, Teerapol, et al.
Published: (2025)
Large Language Model for Qualitative Research -- A Systematic Mapping Study
by: Barros, Cauã Ferreira, et al.
Published: (2024)
by: Barros, Cauã Ferreira, et al.
Published: (2024)
BatchBench: Toward a Workload-Aware Benchmark for Autoscaling Policies in Big Data Batch Processing -- A Proposed Framework
by: Budigi, Venkata Krishna Prasanth, et al.
Published: (2026)
by: Budigi, Venkata Krishna Prasanth, et al.
Published: (2026)
Routing End User Queries to Enterprise Databases
by: Sudarshan, Saikrishna, et al.
Published: (2026)
by: Sudarshan, Saikrishna, et al.
Published: (2026)
AuthorityBench: Benchmarking LLM Authority Perception for Reliable Retrieval-Augmented Generation
by: Yao, Zhihui, et al.
Published: (2026)
by: Yao, Zhihui, et al.
Published: (2026)
Learning Joint Denoising, Demosaicing, and Compression from the Raw Natural Image Noise Dataset
by: Brummer, Benoit, et al.
Published: (2025)
by: Brummer, Benoit, et al.
Published: (2025)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
by: Bachar, Or, et al.
Published: (2026)
by: Bachar, Or, et al.
Published: (2026)
Bottleneck-based Encoder-decoder ARchitecture (BEAR) for Learning Unbiased Consumer-to-Consumer Image Representations
by: Rivas, Pablo, et al.
Published: (2024)
by: Rivas, Pablo, et al.
Published: (2024)
Big Help or Big Brother? Auditing Tracking, Profiling, and Personalization in Generative AI Assistants
by: Vekaria, Yash, et al.
Published: (2025)
by: Vekaria, Yash, et al.
Published: (2025)
Deep Interest Mining for Intent-Enriched Semantic IDs in Multimodal Generative Recommendation
by: Zeng, Yangchen, et al.
Published: (2026)
by: Zeng, Yangchen, et al.
Published: (2026)
Playing telephone with generative models: "verification disability," "compelled reliance," and accessibility in data visualization
by: Elavsky, Frank, et al.
Published: (2025)
by: Elavsky, Frank, et al.
Published: (2025)
Understanding Multi-Agent LLM Frameworks: A Unified Benchmark and Experimental Analysis
by: Orogat, Abdelghny, et al.
Published: (2026)
by: Orogat, Abdelghny, et al.
Published: (2026)
Machine Learning for Coding Retail Product Names to Consumer-Price Categories: A Rule-plus-Bag-of-Words Pipeline with Reliability-Weighted Human-in-the-Loop Labeling
by: Beskorovainyi, Vladimir
Published: (2026)
by: Beskorovainyi, Vladimir
Published: (2026)
HOME-KGQA: A Benchmark Dataset for Multimodal Knowledge Graph Question Answering on Household Daily Activities
by: Egami, Shusaku, et al.
Published: (2026)
by: Egami, Shusaku, et al.
Published: (2026)
Towards Robust Retrieval-Augmented Generation Based on Knowledge Graph: A Comparative Analysis
by: Amamou, Hazem, et al.
Published: (2026)
by: Amamou, Hazem, et al.
Published: (2026)
GE-Chat: A Graph Enhanced RAG Framework for Evidential Response Generation of LLMs
by: Da, Longchao, et al.
Published: (2025)
by: Da, Longchao, et al.
Published: (2025)
Agent-Based User-Adaptive Filtering for Categorized Harassing Communication
by: Rahaman, Zenefa, et al.
Published: (2026)
by: Rahaman, Zenefa, et al.
Published: (2026)
SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
by: Park, Sungho, et al.
Published: (2026)
by: Park, Sungho, et al.
Published: (2026)
Gender and Race Bias in Consumer Product Recommendations by Large Language Models
by: Xu, Ke, et al.
Published: (2026)
by: Xu, Ke, et al.
Published: (2026)
Similar Items
-
From Native Memes to Global Moderation: Cross-Cultural Evaluation of Vision-Language Models for Hateful Meme Detection
by: Wang, Mo, et al.
Published: (2026) -
Cultural Encoding in Large Language Models: The Existence Gap in AI-Mediated Brand Discovery
by: Junyao, Huang, et al.
Published: (2025) -
A Grounded Memory System For Smart Personal Assistants
by: Ocker, Felix, et al.
Published: (2025) -
Aximorphic Perspective Projection Model for Immersive Imagery
by: Fober, Jakub Maksymilian
Published: (2021) -
Deterministic Fuzzy Triage for Legal Compliance Classification and Evidence Retrieval
by: Atri, Rian
Published: (2026)