LLMs and Finetuning: Benchmarking cross-domain performance for hate speech detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nasir, Ahmad, Sharma, Aadish, Jaidka, Kokil, Ahmed, Saifuddin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Impact of Decoding Methods on Human Alignment of Conversational LLMs
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
Incivility and Rigidity: Evaluating the Risks of Fine-Tuning LLMs for Political Argumentation
von: Churina, Svetlana, et al.
Veröffentlicht: (2024)
von: Churina, Svetlana, et al.
Veröffentlicht: (2024)
Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2025)
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2025)
Turn-Level Empathy Prediction Using Psychological Indicators
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
The MediaSpin Dataset: Post-Publication News Headline Edits Annotated for Media Bias
von: Verma, Preetika, et al.
Veröffentlicht: (2024)
von: Verma, Preetika, et al.
Veröffentlicht: (2024)
Reading Between the Lines: How Electronic Nonverbal Cues shape Emotion Decoding
von: Kumar, Taara, et al.
Veröffentlicht: (2026)
von: Kumar, Taara, et al.
Veröffentlicht: (2026)
CommunityFact: A Dynamic, Multilingual, Multi-domain Benchmark for Misinformation Detection in the Wild
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2026)
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2026)
"Reasoning" with Rhetoric: On the Style-Evidence Tradeoff in LLM-Generated Counter-Arguments
von: Verma, Preetika, et al.
Veröffentlicht: (2024)
von: Verma, Preetika, et al.
Veröffentlicht: (2024)
Do You Trust Me? Cognitive-Affective Signatures of Trustworthiness in Large Language Models
von: Yeo, Gerard, et al.
Veröffentlicht: (2025)
von: Yeo, Gerard, et al.
Veröffentlicht: (2025)
Conversations: Love Them, Hate Them, Steer Them
von: Chebrolu, Niranjan, et al.
Veröffentlicht: (2025)
von: Chebrolu, Niranjan, et al.
Veröffentlicht: (2025)
Beyond Text: Leveraging Multi-Task Learning and Cognitive Appraisal Theory for Post-Purchase Intention Analysis
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2024)
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2024)
GitSearch: Enhancing Community Notes Generation with Gap-Informed Targeted Search
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2026)
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2026)
Learning Through Dialogue: Engagement and Efficacy Matter More Than Explanations
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2026)
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2026)
On the Limitations of Steering in Language Model Alignment
von: Niranjan, Chebrolu, et al.
Veröffentlicht: (2025)
von: Niranjan, Chebrolu, et al.
Veröffentlicht: (2025)
From Passive to Persuasive: Localized Activation Injection for Empathy and Negotiation
von: Chebrolu, Niranjan, et al.
Veröffentlicht: (2025)
von: Chebrolu, Niranjan, et al.
Veröffentlicht: (2025)
Labels or Input? Rethinking Augmentation in Multimodal Hate Detection
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2025)
von: Singh, Sahajpreet, et al.
Veröffentlicht: (2025)
What is the social benefit of hate speech detection research? A Systematic Review
von: Wong, Sidney Gig-Jan
Veröffentlicht: (2024)
von: Wong, Sidney Gig-Jan
Veröffentlicht: (2024)
Predicting Sentence-Level Factuality of News and Bias of Media Outlets
von: Vargas, Francielle, et al.
Veröffentlicht: (2023)
von: Vargas, Francielle, et al.
Veröffentlicht: (2023)
Disentangling Codemixing in Chats: The NUS ABC Codemixed Corpus
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
Classification is a RAG problem: A case study on hate speech detection
von: Willats, Richard, et al.
Veröffentlicht: (2025)
von: Willats, Richard, et al.
Veröffentlicht: (2025)
Exploring the topics, sentiments and hate speech in the Spanish information environment
von: LOPEZ, ALEJANDRO BUITRAGO, et al.
Veröffentlicht: (2024)
von: LOPEZ, ALEJANDRO BUITRAGO, et al.
Veröffentlicht: (2024)
AtteSTNet -- An attention and subword tokenization based approach for code-switched text hate speech detection
von: Shingi, Geet, et al.
Veröffentlicht: (2021)
von: Shingi, Geet, et al.
Veröffentlicht: (2021)
Ensemble of pre-trained language models and data augmentation for hate speech detection from Arabic tweets
von: Daouadi, Kheir Eddine, et al.
Veröffentlicht: (2024)
von: Daouadi, Kheir Eddine, et al.
Veröffentlicht: (2024)
Using LLMs to discover emerging coded antisemitic hate-speech in extremist social media
von: Kikkisetti, Dhanush, et al.
Veröffentlicht: (2024)
von: Kikkisetti, Dhanush, et al.
Veröffentlicht: (2024)
TuPy-E: detecting hate speech in Brazilian Portuguese social media with a novel dataset and comprehensive analysis of models
von: Oliveira, Felipe, et al.
Veröffentlicht: (2023)
von: Oliveira, Felipe, et al.
Veröffentlicht: (2023)
Thinking Fair and Slow: On the Efficacy of Structured Prompts for Debiasing Language Models
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024)
Hateful Meme Detection through Context-Sensitive Prompting and Fine-Grained Labeling
von: Ouyang, Rongxin, et al.
Veröffentlicht: (2024)
von: Ouyang, Rongxin, et al.
Veröffentlicht: (2024)
A multilingual dataset for offensive language and hate speech detection for hausa, yoruba and igbo languages
von: Aliyu, Saminu Mohammad, et al.
Veröffentlicht: (2024)
von: Aliyu, Saminu Mohammad, et al.
Veröffentlicht: (2024)
Althea: Human-AI Collaboration for Fact-Checking and Critical Reasoning
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
OMIND: Framework for Knowledge Grounded Finetuning and Multi-Turn Dialogue Benchmark for Mental Health LLMs
von: Racha, Suraj, et al.
Veröffentlicht: (2026)
von: Racha, Suraj, et al.
Veröffentlicht: (2026)
Digital Guardians: Can GPT-4, Perspective API, and Moderation API reliably detect hate speech in reader comments of German online newspapers?
von: Weber, Manuel, et al.
Veröffentlicht: (2025)
von: Weber, Manuel, et al.
Veröffentlicht: (2025)
Understanding the Effects of Domain Finetuning on LLMs
von: Tanwar, Eshaan, et al.
Veröffentlicht: (2025)
von: Tanwar, Eshaan, et al.
Veröffentlicht: (2025)
Locking Down the Finetuned LLMs Safety
von: Zhu, Minjun, et al.
Veröffentlicht: (2024)
von: Zhu, Minjun, et al.
Veröffentlicht: (2024)
Finetuning LLMs for Comparative Assessment Tasks
von: Raina, Vatsal, et al.
Veröffentlicht: (2024)
von: Raina, Vatsal, et al.
Veröffentlicht: (2024)
PHAnToM: Persona-based Prompting Has An Effect on Theory-of-Mind Reasoning in Large Language Models
von: Tan, Fiona Anting, et al.
Veröffentlicht: (2024)
von: Tan, Fiona Anting, et al.
Veröffentlicht: (2024)
Bridging the gap in online hate speech detection: a comparative analysis of BERT and traditional models for homophobic content identification on X/Twitter
von: McGiff, Josh, et al.
Veröffentlicht: (2024)
von: McGiff, Josh, et al.
Veröffentlicht: (2024)
Synthesizing Privacy-Preserving Text Data via Finetuning without Finetuning Billion-Scale LLMs
von: Tan, Bowen, et al.
Veröffentlicht: (2025)
von: Tan, Bowen, et al.
Veröffentlicht: (2025)
cantnlp@DravidianLangTech 2026: organic domain adaptation improves multi-class hope speech detection in Tulu
von: Li, Andrew, et al.
Veröffentlicht: (2026)
von: Li, Andrew, et al.
Veröffentlicht: (2026)
Improving code-mixed hate detection by native sample mixing: A case study for Hindi-English code-mixed scenario
von: Mazumder, Debajyoti, et al.
Veröffentlicht: (2024)
von: Mazumder, Debajyoti, et al.
Veröffentlicht: (2024)
Evaluation of Finetuned LLMs in AMR Parsing
von: Ho, Shu Han
Veröffentlicht: (2025)
von: Ho, Shu Han
Veröffentlicht: (2025)
Ähnliche Einträge
-
Impact of Decoding Methods on Human Alignment of Conversational LLMs
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024) -
Incivility and Rigidity: Evaluating the Risks of Fine-Tuning LLMs for Political Argumentation
von: Churina, Svetlana, et al.
Veröffentlicht: (2024) -
Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2025) -
Turn-Level Empathy Prediction Using Psychological Indicators
von: Furniturewala, Shaz, et al.
Veröffentlicht: (2024) -
The MediaSpin Dataset: Post-Publication News Headline Edits Annotated for Media Bias
von: Verma, Preetika, et al.
Veröffentlicht: (2024)