Towards a Robust Framework for Multimodal Hate Detection: A Study on Video vs. Image-based Content
Fuente:
arXiv
Guardado en:
| Autores principales: | Koushik, Girish A., Kanojia, Diptesh, Treharne, Helen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DROID: Dual Representation for Out-of-Scope Intent Detection
por: Rashwan, Wael, et al.
Publicado: (2025)
por: Rashwan, Wael, et al.
Publicado: (2025)
Correspondence of high-dimensional emotion structures elicited by video clips between humans and Multimodal LLMs
por: Asanuma, Haruka, et al.
Publicado: (2025)
por: Asanuma, Haruka, et al.
Publicado: (2025)
InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information
por: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Publicado: (2025)
por: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Publicado: (2025)
FAME: Feature Activation Map Explanation on Image Classification and Face Recognition
por: Zhang, Xinyi, et al.
Publicado: (2026)
por: Zhang, Xinyi, et al.
Publicado: (2026)
Evaluation Before Generation: A Paradigm for Robust Multimodal Sentiment Analysis with Missing Modalities
por: Chen, Rongfei, et al.
Publicado: (2026)
por: Chen, Rongfei, et al.
Publicado: (2026)
MIMIC-SR-ICD11: A Dataset for Narrative-Based Diagnosis
por: Wu, Yuexin, et al.
Publicado: (2025)
por: Wu, Yuexin, et al.
Publicado: (2025)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
por: Semenov, Andrei, et al.
Publicado: (2024)
por: Semenov, Andrei, et al.
Publicado: (2024)
Logits-Constrained Framework with RoBERTa for Ancient Chinese NER
por: Hua, Wenjie, et al.
Publicado: (2025)
por: Hua, Wenjie, et al.
Publicado: (2025)
SigLino: Efficient Multi-Teacher Distillation for Agglomerative Vision Foundation Models
por: Chaybouti, Sofian, et al.
Publicado: (2025)
por: Chaybouti, Sofian, et al.
Publicado: (2025)
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
por: Gonzalez, Alberto Andres Valdes
Publicado: (2026)
por: Gonzalez, Alberto Andres Valdes
Publicado: (2026)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
por: Reddy, Sandeep, et al.
Publicado: (2025)
por: Reddy, Sandeep, et al.
Publicado: (2025)
Efficient Hate Speech Detection: Evaluating 38 Models from Traditional Methods to Transformers
por: Abusaqer, Mahmoud, et al.
Publicado: (2025)
por: Abusaqer, Mahmoud, et al.
Publicado: (2025)
A Lightweight Approach to Detection of AI-Generated Texts Using Stylometric Features
por: Aityan, Sergey K., et al.
Publicado: (2025)
por: Aityan, Sergey K., et al.
Publicado: (2025)
Transparent but Powerful: Explainability, Accuracy, and Generalizability in ADHD Detection from Social Media Data
por: Wiechmann, D., et al.
Publicado: (2024)
por: Wiechmann, D., et al.
Publicado: (2024)
Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection Theory
por: Cacioli, Jon-Paul
Publicado: (2026)
por: Cacioli, Jon-Paul
Publicado: (2026)
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
por: Banthia, Saumya, et al.
Publicado: (2020)
por: Banthia, Saumya, et al.
Publicado: (2020)
Semantic Reconstruction of Adversarial Plagiarism: A Context-Aware Framework for Detecting and Restoring "Tortured Phrases" in Scientific Literature
por: Maiti, Agniva, et al.
Publicado: (2025)
por: Maiti, Agniva, et al.
Publicado: (2025)
Tversky Neural Networks: Psychologically Plausible Deep Learning with Differentiable Tversky Similarity
por: Doumbouya, Moussa Koulako Bala, et al.
Publicado: (2025)
por: Doumbouya, Moussa Koulako Bala, et al.
Publicado: (2025)
Accelerating Language Model Workflows with Prompt Choreography
por: Bai, TJ, et al.
Publicado: (2025)
por: Bai, TJ, et al.
Publicado: (2025)
Predicting When to Trust Vision-Language Models for Spatial Reasoning
por: Imran, Muhammad, et al.
Publicado: (2026)
por: Imran, Muhammad, et al.
Publicado: (2026)
A Semantic Approach to Negation Detection and Word Disambiguation with Natural Language Processing
por: Okpala, Izunna, et al.
Publicado: (2023)
por: Okpala, Izunna, et al.
Publicado: (2023)
Analyzing Quality, Bias, and Performance in Text-to-Image Generative Models
por: Masrourisaadat, Nila, et al.
Publicado: (2024)
por: Masrourisaadat, Nila, et al.
Publicado: (2024)
MANGO: Learning Disentangled Image Transformation Manifolds with Grouped Operators
por: Ancelin, Brighton, et al.
Publicado: (2024)
por: Ancelin, Brighton, et al.
Publicado: (2024)
ZeShot-VQA: Zero-Shot Visual Question Answering Framework with Answer Mapping for Natural Disaster Damage Assessment
por: Karimi, Ehsan, et al.
Publicado: (2025)
por: Karimi, Ehsan, et al.
Publicado: (2025)
Efficient Strategy for Improving Large Language Model (LLM) Capabilities
por: Gutiérrez, Julián Camilo Velandia
Publicado: (2025)
por: Gutiérrez, Julián Camilo Velandia
Publicado: (2025)
When is dataset cartography ineffective? Using training dynamics does not improve robustness against Adversarial SQuAD
por: Mandal, Paul K.
Publicado: (2025)
por: Mandal, Paul K.
Publicado: (2025)
Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages
por: Schneider, Felix, et al.
Publicado: (2026)
por: Schneider, Felix, et al.
Publicado: (2026)
Co-Training with Active Contrastive Learning and Meta-Pseudo-Labeling on 2D Projections for Deep Semi-Supervised Learning
por: Aparco-Cardenas, David, et al.
Publicado: (2025)
por: Aparco-Cardenas, David, et al.
Publicado: (2025)
Feature-Augmented Deep Networks for Multiscale Building Segmentation in High-Resolution UAV and Satellite Imagery
por: Maniyar, Chintan B., et al.
Publicado: (2025)
por: Maniyar, Chintan B., et al.
Publicado: (2025)
Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning
por: Ji, Binbin, et al.
Publicado: (2025)
por: Ji, Binbin, et al.
Publicado: (2025)
Applied Explainability for Large Language Models: A Comparative Study
por: Kancharla, Venkata Abhinandan
Publicado: (2026)
por: Kancharla, Venkata Abhinandan
Publicado: (2026)
A Multi-Encoder Frozen-Decoder Approach for Fine-Tuning Large Language Models
por: Dhole, Kaustubh D.
Publicado: (2025)
por: Dhole, Kaustubh D.
Publicado: (2025)
Adversarially Probing Cross-Family Sound Symbolism in 27 Languages
por: Sharma, Anika, et al.
Publicado: (2025)
por: Sharma, Anika, et al.
Publicado: (2025)
Neural Encoding for Image Recall: Human-Like Memory
por: Foussereau, Virgile, et al.
Publicado: (2024)
por: Foussereau, Virgile, et al.
Publicado: (2024)
Cytoarchitecture in Words: Weakly Supervised Vision-Language Modeling for Human Brain Microscopy
por: Sutton, Matthew, et al.
Publicado: (2026)
por: Sutton, Matthew, et al.
Publicado: (2026)
Large Language Models Are Not Strong Abstract Reasoners
por: Gendron, Gaël, et al.
Publicado: (2023)
por: Gendron, Gaël, et al.
Publicado: (2023)
Lost in Latent Space: Disentangled Models and the Challenge of Combinatorial Generalisation
por: Montero, Milton L., et al.
Publicado: (2022)
por: Montero, Milton L., et al.
Publicado: (2022)
Suicide Risk Assessment Using Multimodal Speech Features: A Study on the SW1 Challenge Dataset
por: Marie, Ambre, et al.
Publicado: (2025)
por: Marie, Ambre, et al.
Publicado: (2025)
Kolmogorov-Arnold Attention: Is Learnable Attention Better For Vision Transformers?
por: Maity, Subhajit, et al.
Publicado: (2025)
por: Maity, Subhajit, et al.
Publicado: (2025)
A Human-Machine Collaboration Framework for the Development of Schemas
por: Isaak, Nicos
Publicado: (2024)
por: Isaak, Nicos
Publicado: (2024)
Ejemplares similares
-
DROID: Dual Representation for Out-of-Scope Intent Detection
por: Rashwan, Wael, et al.
Publicado: (2025) -
Correspondence of high-dimensional emotion structures elicited by video clips between humans and Multimodal LLMs
por: Asanuma, Haruka, et al.
Publicado: (2025) -
InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information
por: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Publicado: (2025) -
FAME: Feature Activation Map Explanation on Image Classification and Face Recognition
por: Zhang, Xinyi, et al.
Publicado: (2026) -
Evaluation Before Generation: A Paradigm for Robust Multimodal Sentiment Analysis with Missing Modalities
por: Chen, Rongfei, et al.
Publicado: (2026)