Dealing with Annotator Disagreement in Hate Speech Classification
Fuente:
arXiv
Salvato in:
| Autori principali: | Dehghan, Somaiyeh, Sen, Mehmet Umut, Yanikoglu, Berrin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Distilling Knowledge from Large Language Models: A Concept Bottleneck Model for Hate and Counter Speech Recognition
di: Labadie-Tamayo, Roberto, et al.
Pubblicazione: (2025)
di: Labadie-Tamayo, Roberto, et al.
Pubblicazione: (2025)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
di: Fadli, Samih
Pubblicazione: (2025)
di: Fadli, Samih
Pubblicazione: (2025)
Large Language Model (LLM) Bias Index -- LLMBI
di: Oketunji, Abiodun Finbarrs, et al.
Pubblicazione: (2023)
di: Oketunji, Abiodun Finbarrs, et al.
Pubblicazione: (2023)
DISCO: DISCovering Overfittings as Causal Rules for Text Classification Models
di: Zhang, Zijian, et al.
Pubblicazione: (2024)
di: Zhang, Zijian, et al.
Pubblicazione: (2024)
Enhancing Burmese News Classification with Kolmogorov-Arnold Network Head Fine-tuning
di: Aung, Thura, et al.
Pubblicazione: (2025)
di: Aung, Thura, et al.
Pubblicazione: (2025)
In-Context Fixation: When Demonstrated Labels Override Semantics in Few-Shot Classification
di: Liu, Ming
Pubblicazione: (2026)
di: Liu, Ming
Pubblicazione: (2026)
ObfusQAte: A Proposed Framework to Evaluate LLM Robustness on Obfuscated Factual Question Answering
di: Ghosh, Shubhra, et al.
Pubblicazione: (2025)
di: Ghosh, Shubhra, et al.
Pubblicazione: (2025)
Improving Discrete Diffusion Unmasking Policies Beyond Explicit Reference Policies
di: Hong, Chunsan, et al.
Pubblicazione: (2025)
di: Hong, Chunsan, et al.
Pubblicazione: (2025)
Listen to the Layers: Mitigating Hallucinations with Inter-Layer Disagreement
di: Subbalakshmi, Koduvayur, et al.
Pubblicazione: (2026)
di: Subbalakshmi, Koduvayur, et al.
Pubblicazione: (2026)
Enhancing Hate Speech Detection on Social Media: A Comparative Analysis of Machine Learning Models and Text Transformation Approaches
di: Mishra, Saurabh, et al.
Pubblicazione: (2026)
di: Mishra, Saurabh, et al.
Pubblicazione: (2026)
Part-of-Speech Tagger for Bodo Language using Deep Learning approach
di: Pathak, Dhrubajyoti, et al.
Pubblicazione: (2024)
di: Pathak, Dhrubajyoti, et al.
Pubblicazione: (2024)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
di: Saji, Alan, et al.
Pubblicazione: (2025)
di: Saji, Alan, et al.
Pubblicazione: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025)
di: Peters, Sydney, et al.
Pubblicazione: (2025)
Grokking in the Wild: Data Augmentation for Real-World Multi-Hop Reasoning with Transformers
di: Abramov, Roman, et al.
Pubblicazione: (2025)
di: Abramov, Roman, et al.
Pubblicazione: (2025)
Zero-Shot Cross-Lingual Transfer using Prefix-Based Adaptation
di: A, Snegha, et al.
Pubblicazione: (2025)
di: A, Snegha, et al.
Pubblicazione: (2025)
MultiMatch: Multihead Consistency Regularization Matching for Semi-Supervised Text Classification
di: Sirbu, Iustin, et al.
Pubblicazione: (2025)
di: Sirbu, Iustin, et al.
Pubblicazione: (2025)
Concept Navigation and Classification via Open-Source Large Language Model Processing
di: Kubli, Maël
Pubblicazione: (2025)
di: Kubli, Maël
Pubblicazione: (2025)
A Survey of the State of Explainable AI for Natural Language Processing
di: Danilevsky, Marina, et al.
Pubblicazione: (2020)
di: Danilevsky, Marina, et al.
Pubblicazione: (2020)
Evaluation of Hate Speech Detection Using Large Language Models and Geographical Contextualization
di: Zahid, Anwar Hossain, et al.
Pubblicazione: (2025)
di: Zahid, Anwar Hossain, et al.
Pubblicazione: (2025)
Enhancing Text Classification through LLM-Driven Active Learning and Human Annotation
di: Rouzegar, Hamidreza, et al.
Pubblicazione: (2024)
di: Rouzegar, Hamidreza, et al.
Pubblicazione: (2024)
Confident, Calibrated, or Complicit: Safety Alignment and Ideological Bias in LLM Hate Speech Detection
di: Selvaganapathy, Sanjeeevan, et al.
Pubblicazione: (2025)
di: Selvaganapathy, Sanjeeevan, et al.
Pubblicazione: (2025)
LLMs in the Loop: Leveraging Large Language Model Annotations for Active Learning in Low-Resource Languages
di: Kholodna, Nataliia, et al.
Pubblicazione: (2024)
di: Kholodna, Nataliia, et al.
Pubblicazione: (2024)
ConciseRL: Conciseness-Guided Reinforcement Learning for Efficient Reasoning Models
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2025)
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2025)
Cognitive Load Limits in Large Language Models: Benchmarking Multi-Hop Reasoning
di: Adapala, Sai Teja Reddy
Pubblicazione: (2025)
di: Adapala, Sai Teja Reddy
Pubblicazione: (2025)
Beyond Pass@k: Breadth-Depth Metrics for Reasoning Boundaries
di: Dragoi, Marius, et al.
Pubblicazione: (2025)
di: Dragoi, Marius, et al.
Pubblicazione: (2025)
Revisiting Intermediate-Layer Matching in Knowledge Distillation: Layer-Selection Strategy Doesn't Matter (Much)
di: Yu, Zony, et al.
Pubblicazione: (2025)
di: Yu, Zony, et al.
Pubblicazione: (2025)
Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models
di: Cui, Sasha, et al.
Pubblicazione: (2025)
di: Cui, Sasha, et al.
Pubblicazione: (2025)
A Llama walks into the 'Bar': Efficient Supervised Fine-Tuning for Legal Reasoning in the Multi-state Bar Exam
di: Fernandes, Rean, et al.
Pubblicazione: (2025)
di: Fernandes, Rean, et al.
Pubblicazione: (2025)
MaPPO: Maximum a Posteriori Preference Optimization with Prior Knowledge
di: Lan, Guangchen, et al.
Pubblicazione: (2025)
di: Lan, Guangchen, et al.
Pubblicazione: (2025)
NRR-Core: Non-Resolution Reasoning as a Computational Framework for Contextual Identity and Ambiguity Preservation
di: Saito, Kei
Pubblicazione: (2025)
di: Saito, Kei
Pubblicazione: (2025)
Sycophancy as compositions of Atomic Psychometric Traits
di: Jain, Shreyans, et al.
Pubblicazione: (2025)
di: Jain, Shreyans, et al.
Pubblicazione: (2025)
CorrSteer: Generation-Time LLM Steering via Correlated Sparse Autoencoder Features
di: Cho, Seonglae, et al.
Pubblicazione: (2025)
di: Cho, Seonglae, et al.
Pubblicazione: (2025)
CopySpec: Accelerating LLMs with Speculative Copy-and-Paste Without Compromising Quality
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2025)
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2025)
Bounding Hallucinations: Information-Theoretic Guarantees for RAG Systems via Merlin-Arthur Protocols
di: Deiseroth, Björn, et al.
Pubblicazione: (2025)
di: Deiseroth, Björn, et al.
Pubblicazione: (2025)
Harnessing Negative Signals: Reinforcement Distillation from Teacher Data for LLM Reasoning
di: Xu, Shuyao, et al.
Pubblicazione: (2025)
di: Xu, Shuyao, et al.
Pubblicazione: (2025)
EasyMath: A 0-shot Math Benchmark for SLMs
di: Karki, Drishya, et al.
Pubblicazione: (2025)
di: Karki, Drishya, et al.
Pubblicazione: (2025)
DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory
di: Zhou, Wenxuan, et al.
Pubblicazione: (2025)
di: Zhou, Wenxuan, et al.
Pubblicazione: (2025)
Contextual Integrity in LLMs via Reasoning and Reinforcement Learning
di: Lan, Guangchen, et al.
Pubblicazione: (2025)
di: Lan, Guangchen, et al.
Pubblicazione: (2025)
RWKV-7 "Goose" with Expressive Dynamic State Evolution
di: Peng, Bo, et al.
Pubblicazione: (2025)
di: Peng, Bo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Distilling Knowledge from Large Language Models: A Concept Bottleneck Model for Hate and Counter Speech Recognition
di: Labadie-Tamayo, Roberto, et al.
Pubblicazione: (2025) -
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023) -
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
di: Fadli, Samih
Pubblicazione: (2025) -
Large Language Model (LLM) Bias Index -- LLMBI
di: Oketunji, Abiodun Finbarrs, et al.
Pubblicazione: (2023) -
DISCO: DISCovering Overfittings as Causal Rules for Text Classification Models
di: Zhang, Zijian, et al.
Pubblicazione: (2024)