ICM-Assistant: Instruction-tuning Multimodal Large Language Models for Rule-based Explainable Image Content Moderation
Fuente:
arXiv
Guardado en:
| Autores principales: | Wu, Mengyang, Zhao, Yuzhi, Cao, Jialun, Xu, Mingjie, Jiang, Zhongming, Wang, Xuehui, Li, Qinbin, Hu, Guangneng, Qin, Shengchao, Fu, Chi-Wing |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Evian: Towards Explainable Visual Instruction-tuning Data Auditing
por: Jia, Zimu, et al.
Publicado: (2026)
por: Jia, Zimu, et al.
Publicado: (2026)
Let Community Rules Be Reflected in Online Content Moderation
por: Xin, Wangjiaxuan, et al.
Publicado: (2024)
por: Xin, Wangjiaxuan, et al.
Publicado: (2024)
VP-Bench: A Comprehensive Benchmark for Visual Prompting in Multimodal Large Language Models
por: Xu, Mingjie, et al.
Publicado: (2025)
por: Xu, Mingjie, et al.
Publicado: (2025)
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations
por: Xu, Mingjie, et al.
Publicado: (2024)
por: Xu, Mingjie, et al.
Publicado: (2024)
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
por: Samory, Mattia, et al.
Publicado: (2025)
por: Samory, Mattia, et al.
Publicado: (2025)
Towards Real-World Adverse Weather Image Restoration: Enhancing Clearness and Semantics with Vision-Language Models
por: Xu, Jiaqi, et al.
Publicado: (2024)
por: Xu, Jiaqi, et al.
Publicado: (2024)
Beyond Hate: Differentiating Uncivil and Intolerant Speech in Multimodal Content Moderation
por: Herrmann, Nils A., et al.
Publicado: (2026)
por: Herrmann, Nils A., et al.
Publicado: (2026)
VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
por: Wang, Ke, et al.
Publicado: (2025)
por: Wang, Ke, et al.
Publicado: (2025)
A Browser-based Open Source Assistant for Multimodal Content Verification
por: Milner, Rosanna, et al.
Publicado: (2026)
por: Milner, Rosanna, et al.
Publicado: (2026)
Content Moderation Futures
por: Blackwell, Lindsay
Publicado: (2025)
por: Blackwell, Lindsay
Publicado: (2025)
TRUST-VL: An Explainable News Assistant for General Multimodal Misinformation Detection
por: Yan, Zehong, et al.
Publicado: (2025)
por: Yan, Zehong, et al.
Publicado: (2025)
Exploring the Boundaries of Content Moderation in Text-to-Image Generation
por: Riccio, Piera, et al.
Publicado: (2024)
por: Riccio, Piera, et al.
Publicado: (2024)
Multi-Modal Explainable Medical AI Assistant for Trustworthy Human-AI Collaboration
por: Yang, Honglong, et al.
Publicado: (2025)
por: Yang, Honglong, et al.
Publicado: (2025)
Embedding-based Retrieval in Multimodal Content Moderation
por: Liang, Hanzhong, et al.
Publicado: (2025)
por: Liang, Hanzhong, et al.
Publicado: (2025)
Learning or Self-aligning? Rethinking Instruction Fine-tuning
por: Ren, Mengjie, et al.
Publicado: (2024)
por: Ren, Mengjie, et al.
Publicado: (2024)
Instruction-tuned Self-Questioning Framework for Multimodal Reasoning
por: Jang, You-Won, et al.
Publicado: (2025)
por: Jang, You-Won, et al.
Publicado: (2025)
ShieldGemma 2: Robust and Tractable Image Content Moderation
por: Zeng, Wenjun, et al.
Publicado: (2025)
por: Zeng, Wenjun, et al.
Publicado: (2025)
C3LLM: Conditional Multimodal Content Generation Using Large Language Models
por: Wang, Zixuan, et al.
Publicado: (2024)
por: Wang, Zixuan, et al.
Publicado: (2024)
Compositional Image Retrieval via Instruction-Aware Contrastive Learning
por: Zhong, Wenliang, et al.
Publicado: (2024)
por: Zhong, Wenliang, et al.
Publicado: (2024)
PixWizard: Versatile Image-to-Image Visual Assistant with Open-Language Instructions
por: Lin, Weifeng, et al.
Publicado: (2024)
por: Lin, Weifeng, et al.
Publicado: (2024)
Enhanced Multimodal Content Moderation of Children's Videos using Audiovisual Fusion
por: Ahmed, Syed Hammad, et al.
Publicado: (2024)
por: Ahmed, Syed Hammad, et al.
Publicado: (2024)
Multimodal Guidance Network for Missing-Modality Inference in Content Moderation
por: Zhao, Zhuokai, et al.
Publicado: (2023)
por: Zhao, Zhuokai, et al.
Publicado: (2023)
Fine-tuning Multimodal LLMs to Follow Zero-shot Demonstrative Instructions
por: Li, Juncheng, et al.
Publicado: (2023)
por: Li, Juncheng, et al.
Publicado: (2023)
AI Content Moderation in Therapy Conversations
por: Kim, Jiwon, et al.
Publicado: (2026)
por: Kim, Jiwon, et al.
Publicado: (2026)
SEFE: Superficial and Essential Forgetting Eliminator for Multimodal Continual Instruction Tuning
por: Chen, Jinpeng, et al.
Publicado: (2025)
por: Chen, Jinpeng, et al.
Publicado: (2025)
Longitudinal Monitoring of LLM Content Moderation of Social Issues
por: Dai, Yunlang, et al.
Publicado: (2025)
por: Dai, Yunlang, et al.
Publicado: (2025)
BLM-Guard: Explainable Multimodal Ad Moderation with Chain-of-Thought and Policy-Aligned Rewards
por: Yang, Yiran, et al.
Publicado: (2026)
por: Yang, Yiran, et al.
Publicado: (2026)
Tool-MCoT: Tool Augmented Multimodal Chain-of-Thought for Content Safety Moderation
por: Zhang, Shutong, et al.
Publicado: (2026)
por: Zhang, Shutong, et al.
Publicado: (2026)
AesBiasBench: Evaluating Bias and Alignment in Multimodal Language Models for Personalized Image Aesthetic Assessment
por: Li, Kun, et al.
Publicado: (2025)
por: Li, Kun, et al.
Publicado: (2025)
Large Continual Instruction Assistant
por: Qiao, Jingyang, et al.
Publicado: (2024)
por: Qiao, Jingyang, et al.
Publicado: (2024)
Identity-related Speech Suppression in Generative AI Content Moderation
por: Proebsting, Grace, et al.
Publicado: (2024)
por: Proebsting, Grace, et al.
Publicado: (2024)
Ideology-Based LLMs for Content Moderation
por: Civelli, Stefano, et al.
Publicado: (2025)
por: Civelli, Stefano, et al.
Publicado: (2025)
From Informal to Formal -- Incorporating and Evaluating LLMs on Natural Language Requirements to Verifiable Formal Proofs
por: Cao, Jialun, et al.
Publicado: (2025)
por: Cao, Jialun, et al.
Publicado: (2025)
Lost in Moderation: How Commercial Content Moderation APIs Over- and Under-Moderate Group-Targeted Hate Speech and Linguistic Variations
por: Hartmann, David, et al.
Publicado: (2025)
por: Hartmann, David, et al.
Publicado: (2025)
MedHallTune: An Instruction-Tuning Benchmark for Mitigating Medical Hallucination in Vision-Language Models
por: Yan, Qiao, et al.
Publicado: (2025)
por: Yan, Qiao, et al.
Publicado: (2025)
CvhSlicer 2.0: Immersive and Interactive Visualization of Chinese Visible Human Data in XR Environments
por: Qiu, Yue, et al.
Publicado: (2025)
por: Qiu, Yue, et al.
Publicado: (2025)
PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media
por: Kachwala, Zoher, et al.
Publicado: (2026)
por: Kachwala, Zoher, et al.
Publicado: (2026)
Multimodal Generation of Animatable 3D Human Models with AvatarForge
por: Liu, Xinhang, et al.
Publicado: (2025)
por: Liu, Xinhang, et al.
Publicado: (2025)
Explainable Token-level Noise Filtering for LLM Fine-tuning Datasets
por: Yang, Yuchen, et al.
Publicado: (2026)
por: Yang, Yuchen, et al.
Publicado: (2026)
Enhancing Question Answering Precision with Optimized Vector Retrieval and Instructions
por: Yang, Lixiao, et al.
Publicado: (2024)
por: Yang, Lixiao, et al.
Publicado: (2024)
Ejemplares similares
-
Evian: Towards Explainable Visual Instruction-tuning Data Auditing
por: Jia, Zimu, et al.
Publicado: (2026) -
Let Community Rules Be Reflected in Online Content Moderation
por: Xin, Wangjiaxuan, et al.
Publicado: (2024) -
VP-Bench: A Comprehensive Benchmark for Visual Prompting in Multimodal Large Language Models
por: Xu, Mingjie, et al.
Publicado: (2025) -
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations
por: Xu, Mingjie, et al.
Publicado: (2024) -
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
por: Samory, Mattia, et al.
Publicado: (2025)