Focal Inferential Infusion Coupled with Tractable Density Discrimination for Implicit Hate Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Masud, Sarah, Bajpai, Ashutosh, Chakraborty, Tanmoy |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hate Personified: Investigating the role of LLMs in content moderation
por: Masud, Sarah, et al.
Publicado: (2024)
por: Masud, Sarah, et al.
Publicado: (2024)
SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes
por: Nandi, Palash, et al.
Publicado: (2024)
por: Nandi, Palash, et al.
Publicado: (2024)
Information Anxiety in Large Language Models
por: Bajpai, Prasoon, et al.
Publicado: (2024)
por: Bajpai, Prasoon, et al.
Publicado: (2024)
Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?
por: Bajpai, Ashutosh, et al.
Publicado: (2025)
por: Bajpai, Ashutosh, et al.
Publicado: (2025)
Tox-BART: Leveraging Toxicity Attributes for Explanation Generation of Implicit Hate Speech
por: Yadav, Neemesh, et al.
Publicado: (2024)
por: Yadav, Neemesh, et al.
Publicado: (2024)
Probing Critical Learning Dynamics of PLMs for Hate Speech Detection
por: Masud, Sarah, et al.
Publicado: (2024)
por: Masud, Sarah, et al.
Publicado: (2024)
Multilingual LLMs Inherently Reward In-Language Time-Sensitive Semantic Alignment for Low-Resource Languages
por: Bajpai, Ashutosh, et al.
Publicado: (2024)
por: Bajpai, Ashutosh, et al.
Publicado: (2024)
Decoding Memes: Benchmarking Narrative Role Classification across Multilingual and Multimodal Models
por: Sharma, Shivam, et al.
Publicado: (2025)
por: Sharma, Shivam, et al.
Publicado: (2025)
Rethinking Hate Speech Detection on Social Media: Can LLMs Replace Traditional Models?
por: Singh, Daman Deep, et al.
Publicado: (2025)
por: Singh, Daman Deep, et al.
Publicado: (2025)
AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection
por: Lee, Yejin, et al.
Publicado: (2025)
por: Lee, Yejin, et al.
Publicado: (2025)
Multilingual Test-Time Scaling via Initial Thought Transfer
por: Bajpai, Prasoon, et al.
Publicado: (2025)
por: Bajpai, Prasoon, et al.
Publicado: (2025)
Temporally Consistent Factuality Probing for Large Language Models
por: Bajpai, Ashutosh, et al.
Publicado: (2024)
por: Bajpai, Ashutosh, et al.
Publicado: (2024)
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
por: Jin, Yiping, et al.
Publicado: (2024)
por: Jin, Yiping, et al.
Publicado: (2024)
MemeMQA: Multimodal Question Answering for Memes via Rationale-Based Inferencing
por: Agarwal, Siddhant, et al.
Publicado: (2024)
por: Agarwal, Siddhant, et al.
Publicado: (2024)
The Psychology of Falsehood: A Human-Centric Survey of Misinformation Detection
por: Nandi, Arghodeep, et al.
Publicado: (2025)
por: Nandi, Arghodeep, et al.
Publicado: (2025)
The Unseen Targets of Hate -- A Systematic Review of Hateful Communication Datasets
por: Yu, Zehui, et al.
Publicado: (2024)
por: Yu, Zehui, et al.
Publicado: (2024)
Few-shot Hate Speech Detection Based on the MindSpore Framework
por: Qin, Zhenkai, et al.
Publicado: (2025)
por: Qin, Zhenkai, et al.
Publicado: (2025)
Algorithmic Fairness in NLP: Persona-Infused LLMs for Human-Centric Hate Speech Detection
por: Gajewska, Ewelina, et al.
Publicado: (2025)
por: Gajewska, Ewelina, et al.
Publicado: (2025)
SpatialMath: Spatial Comprehension-Infused Symbolic Reasoning for Mathematical Problem-Solving
por: Bajpai, Ashutosh, et al.
Publicado: (2026)
por: Bajpai, Ashutosh, et al.
Publicado: (2026)
Towards Fairness Assessment of Dutch Hate Speech Detection
por: Bauer, Julie, et al.
Publicado: (2025)
por: Bauer, Julie, et al.
Publicado: (2025)
SUKHSANDESH: An Avatar Therapeutic Question Answering Platform for Sexual Education in Rural India
por: Singh, Salam Michael, et al.
Publicado: (2024)
por: Singh, Salam Michael, et al.
Publicado: (2024)
The Enforcement and Feasibility of Hate Speech Moderation on Twitter
por: Tonneau, Manuel, et al.
Publicado: (2026)
por: Tonneau, Manuel, et al.
Publicado: (2026)
Subjective $\textit{Isms}$? On the Danger of Conflating Hate and Offence in Abusive Language Detection
por: Curry, Amanda Cercas, et al.
Publicado: (2024)
por: Curry, Amanda Cercas, et al.
Publicado: (2024)
Deep Learning Approaches for Detecting Adversarial Cyberbullying and Hate Speech in Social Networks
por: Azumah, Sylvia Worlali, et al.
Publicado: (2024)
por: Azumah, Sylvia Worlali, et al.
Publicado: (2024)
Navigating Dialectal Bias and Ethical Complexities in Levantine Arabic Hate Speech Detection
por: Ahmed, Ahmed Haj, et al.
Publicado: (2024)
por: Ahmed, Ahmed Haj, et al.
Publicado: (2024)
Can LLMs reason over extended multilingual contexts? Towards long-context evaluation beyond retrieval and haystacks
por: Hengle, Amey, et al.
Publicado: (2025)
por: Hengle, Amey, et al.
Publicado: (2025)
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
por: Borah, Angana, et al.
Publicado: (2024)
por: Borah, Angana, et al.
Publicado: (2024)
Multilingualism, Transnationality, and K-pop in the Online #StopAsianHate Movement
por: Masis, Tessa, et al.
Publicado: (2025)
por: Masis, Tessa, et al.
Publicado: (2025)
Beyond Hate: Differentiating Uncivil and Intolerant Speech in Multimodal Content Moderation
por: Herrmann, Nils A., et al.
Publicado: (2026)
por: Herrmann, Nils A., et al.
Publicado: (2026)
Data-Efficient Hate Speech Detection via Cross-Lingual Nearest Neighbor Retrieval with Limited Labeled Data
por: Ghorbanpour, Faeze, et al.
Publicado: (2025)
por: Ghorbanpour, Faeze, et al.
Publicado: (2025)
Can Prompting LLMs Unlock Hate Speech Detection across Languages? A Zero-shot and Few-shot Study
por: Ghorbanpour, Faeze, et al.
Publicado: (2025)
por: Ghorbanpour, Faeze, et al.
Publicado: (2025)
Measuring Online Hate on 4chan using Pre-trained Deep Learning Models
por: Bermudez-Villalva, Adrian, et al.
Publicado: (2025)
por: Bermudez-Villalva, Adrian, et al.
Publicado: (2025)
CSEval: Towards Automated, Multi-Dimensional, and Reference-Free Counterspeech Evaluation using Auto-Calibrated LLMs
por: Hengle, Amey, et al.
Publicado: (2025)
por: Hengle, Amey, et al.
Publicado: (2025)
Waking Up Blind: Cold-Start Optimization of Supervision-Free Agentic Trajectories for Grounded Visual Perception
por: Bajpai, Ashutosh, et al.
Publicado: (2026)
por: Bajpai, Ashutosh, et al.
Publicado: (2026)
EtiCor++: Towards Understanding Etiquettical Bias in LLMs
por: Dwivedi, Ashutosh, et al.
Publicado: (2025)
por: Dwivedi, Ashutosh, et al.
Publicado: (2025)
A Modular Taxonomy for Hate Speech Definitions and Its Impact on Zero-Shot LLM Classification Performance
por: Melis, Matteo, et al.
Publicado: (2025)
por: Melis, Matteo, et al.
Publicado: (2025)
Multilingual Needle in a Haystack: Investigating Long-Context Behavior of Multilingual Large Language Models
por: Hengle, Amey, et al.
Publicado: (2024)
por: Hengle, Amey, et al.
Publicado: (2024)
Can LLMs replace Neil deGrasse Tyson? Evaluating the Reliability of LLMs as Science Communicators
por: Bajpai, Prasoon, et al.
Publicado: (2024)
por: Bajpai, Prasoon, et al.
Publicado: (2024)
Towards Weakly-Supervised Hate Speech Classification Across Datasets
por: Jin, Yiping, et al.
Publicado: (2023)
por: Jin, Yiping, et al.
Publicado: (2023)
PRIDE -- Parameter-Efficient Reduction of Identity Discrimination for Equality in LLMs
por: Menke, Maluna, et al.
Publicado: (2025)
por: Menke, Maluna, et al.
Publicado: (2025)
Ejemplares similares
-
Hate Personified: Investigating the role of LLMs in content moderation
por: Masud, Sarah, et al.
Publicado: (2024) -
SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes
por: Nandi, Palash, et al.
Publicado: (2024) -
Information Anxiety in Large Language Models
por: Bajpai, Prasoon, et al.
Publicado: (2024) -
Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?
por: Bajpai, Ashutosh, et al.
Publicado: (2025) -
Tox-BART: Leveraging Toxicity Attributes for Explanation Generation of Implicit Hate Speech
por: Yadav, Neemesh, et al.
Publicado: (2024)