ICM-Assistant: Instruction-tuning Multimodal Large Language Models for Rule-based Explainable Image Content Moderation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Mengyang, Zhao, Yuzhi, Cao, Jialun, Xu, Mingjie, Jiang, Zhongming, Wang, Xuehui, Li, Qinbin, Hu, Guangneng, Qin, Shengchao, Fu, Chi-Wing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evian: Towards Explainable Visual Instruction-tuning Data Auditing
von: Jia, Zimu, et al.
Veröffentlicht: (2026)
von: Jia, Zimu, et al.
Veröffentlicht: (2026)
Let Community Rules Be Reflected in Online Content Moderation
von: Xin, Wangjiaxuan, et al.
Veröffentlicht: (2024)
von: Xin, Wangjiaxuan, et al.
Veröffentlicht: (2024)
VP-Bench: A Comprehensive Benchmark for Visual Prompting in Multimodal Large Language Models
von: Xu, Mingjie, et al.
Veröffentlicht: (2025)
von: Xu, Mingjie, et al.
Veröffentlicht: (2025)
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations
von: Xu, Mingjie, et al.
Veröffentlicht: (2024)
von: Xu, Mingjie, et al.
Veröffentlicht: (2024)
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
von: Samory, Mattia, et al.
Veröffentlicht: (2025)
von: Samory, Mattia, et al.
Veröffentlicht: (2025)
Towards Real-World Adverse Weather Image Restoration: Enhancing Clearness and Semantics with Vision-Language Models
von: Xu, Jiaqi, et al.
Veröffentlicht: (2024)
von: Xu, Jiaqi, et al.
Veröffentlicht: (2024)
Beyond Hate: Differentiating Uncivil and Intolerant Speech in Multimodal Content Moderation
von: Herrmann, Nils A., et al.
Veröffentlicht: (2026)
von: Herrmann, Nils A., et al.
Veröffentlicht: (2026)
VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
von: Wang, Ke, et al.
Veröffentlicht: (2025)
von: Wang, Ke, et al.
Veröffentlicht: (2025)
A Browser-based Open Source Assistant for Multimodal Content Verification
von: Milner, Rosanna, et al.
Veröffentlicht: (2026)
von: Milner, Rosanna, et al.
Veröffentlicht: (2026)
Content Moderation Futures
von: Blackwell, Lindsay
Veröffentlicht: (2025)
von: Blackwell, Lindsay
Veröffentlicht: (2025)
TRUST-VL: An Explainable News Assistant for General Multimodal Misinformation Detection
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
Exploring the Boundaries of Content Moderation in Text-to-Image Generation
von: Riccio, Piera, et al.
Veröffentlicht: (2024)
von: Riccio, Piera, et al.
Veröffentlicht: (2024)
Multi-Modal Explainable Medical AI Assistant for Trustworthy Human-AI Collaboration
von: Yang, Honglong, et al.
Veröffentlicht: (2025)
von: Yang, Honglong, et al.
Veröffentlicht: (2025)
Embedding-based Retrieval in Multimodal Content Moderation
von: Liang, Hanzhong, et al.
Veröffentlicht: (2025)
von: Liang, Hanzhong, et al.
Veröffentlicht: (2025)
Learning or Self-aligning? Rethinking Instruction Fine-tuning
von: Ren, Mengjie, et al.
Veröffentlicht: (2024)
von: Ren, Mengjie, et al.
Veröffentlicht: (2024)
Instruction-tuned Self-Questioning Framework for Multimodal Reasoning
von: Jang, You-Won, et al.
Veröffentlicht: (2025)
von: Jang, You-Won, et al.
Veröffentlicht: (2025)
ShieldGemma 2: Robust and Tractable Image Content Moderation
von: Zeng, Wenjun, et al.
Veröffentlicht: (2025)
von: Zeng, Wenjun, et al.
Veröffentlicht: (2025)
C3LLM: Conditional Multimodal Content Generation Using Large Language Models
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
Compositional Image Retrieval via Instruction-Aware Contrastive Learning
von: Zhong, Wenliang, et al.
Veröffentlicht: (2024)
von: Zhong, Wenliang, et al.
Veröffentlicht: (2024)
PixWizard: Versatile Image-to-Image Visual Assistant with Open-Language Instructions
von: Lin, Weifeng, et al.
Veröffentlicht: (2024)
von: Lin, Weifeng, et al.
Veröffentlicht: (2024)
Enhanced Multimodal Content Moderation of Children's Videos using Audiovisual Fusion
von: Ahmed, Syed Hammad, et al.
Veröffentlicht: (2024)
von: Ahmed, Syed Hammad, et al.
Veröffentlicht: (2024)
Multimodal Guidance Network for Missing-Modality Inference in Content Moderation
von: Zhao, Zhuokai, et al.
Veröffentlicht: (2023)
von: Zhao, Zhuokai, et al.
Veröffentlicht: (2023)
Fine-tuning Multimodal LLMs to Follow Zero-shot Demonstrative Instructions
von: Li, Juncheng, et al.
Veröffentlicht: (2023)
von: Li, Juncheng, et al.
Veröffentlicht: (2023)
AI Content Moderation in Therapy Conversations
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
SEFE: Superficial and Essential Forgetting Eliminator for Multimodal Continual Instruction Tuning
von: Chen, Jinpeng, et al.
Veröffentlicht: (2025)
von: Chen, Jinpeng, et al.
Veröffentlicht: (2025)
Longitudinal Monitoring of LLM Content Moderation of Social Issues
von: Dai, Yunlang, et al.
Veröffentlicht: (2025)
von: Dai, Yunlang, et al.
Veröffentlicht: (2025)
BLM-Guard: Explainable Multimodal Ad Moderation with Chain-of-Thought and Policy-Aligned Rewards
von: Yang, Yiran, et al.
Veröffentlicht: (2026)
von: Yang, Yiran, et al.
Veröffentlicht: (2026)
Tool-MCoT: Tool Augmented Multimodal Chain-of-Thought for Content Safety Moderation
von: Zhang, Shutong, et al.
Veröffentlicht: (2026)
von: Zhang, Shutong, et al.
Veröffentlicht: (2026)
AesBiasBench: Evaluating Bias and Alignment in Multimodal Language Models for Personalized Image Aesthetic Assessment
von: Li, Kun, et al.
Veröffentlicht: (2025)
von: Li, Kun, et al.
Veröffentlicht: (2025)
Large Continual Instruction Assistant
von: Qiao, Jingyang, et al.
Veröffentlicht: (2024)
von: Qiao, Jingyang, et al.
Veröffentlicht: (2024)
Identity-related Speech Suppression in Generative AI Content Moderation
von: Proebsting, Grace, et al.
Veröffentlicht: (2024)
von: Proebsting, Grace, et al.
Veröffentlicht: (2024)
Ideology-Based LLMs for Content Moderation
von: Civelli, Stefano, et al.
Veröffentlicht: (2025)
von: Civelli, Stefano, et al.
Veröffentlicht: (2025)
Lost in Moderation: How Commercial Content Moderation APIs Over- and Under-Moderate Group-Targeted Hate Speech and Linguistic Variations
von: Hartmann, David, et al.
Veröffentlicht: (2025)
von: Hartmann, David, et al.
Veröffentlicht: (2025)
From Informal to Formal -- Incorporating and Evaluating LLMs on Natural Language Requirements to Verifiable Formal Proofs
von: Cao, Jialun, et al.
Veröffentlicht: (2025)
von: Cao, Jialun, et al.
Veröffentlicht: (2025)
MedHallTune: An Instruction-Tuning Benchmark for Mitigating Medical Hallucination in Vision-Language Models
von: Yan, Qiao, et al.
Veröffentlicht: (2025)
von: Yan, Qiao, et al.
Veröffentlicht: (2025)
CvhSlicer 2.0: Immersive and Interactive Visualization of Chinese Visible Human Data in XR Environments
von: Qiu, Yue, et al.
Veröffentlicht: (2025)
von: Qiu, Yue, et al.
Veröffentlicht: (2025)
PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media
von: Kachwala, Zoher, et al.
Veröffentlicht: (2026)
von: Kachwala, Zoher, et al.
Veröffentlicht: (2026)
Multimodal Generation of Animatable 3D Human Models with AvatarForge
von: Liu, Xinhang, et al.
Veröffentlicht: (2025)
von: Liu, Xinhang, et al.
Veröffentlicht: (2025)
Explainable Token-level Noise Filtering for LLM Fine-tuning Datasets
von: Yang, Yuchen, et al.
Veröffentlicht: (2026)
von: Yang, Yuchen, et al.
Veröffentlicht: (2026)
Enhancing Question Answering Precision with Optimized Vector Retrieval and Instructions
von: Yang, Lixiao, et al.
Veröffentlicht: (2024)
von: Yang, Lixiao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Evian: Towards Explainable Visual Instruction-tuning Data Auditing
von: Jia, Zimu, et al.
Veröffentlicht: (2026) -
Let Community Rules Be Reflected in Online Content Moderation
von: Xin, Wangjiaxuan, et al.
Veröffentlicht: (2024) -
VP-Bench: A Comprehensive Benchmark for Visual Prompting in Multimodal Large Language Models
von: Xu, Mingjie, et al.
Veröffentlicht: (2025) -
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations
von: Xu, Mingjie, et al.
Veröffentlicht: (2024) -
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
von: Samory, Mattia, et al.
Veröffentlicht: (2025)