Efficient Models for the Detection of Hate, Abuse and Profanity
Fuente:
arXiv
Saved in:
| Main Authors: | Tillmann, Christoph, Trivedi, Aashka, Bhattacharjee, Bishwaranjan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Human and LLM Biases in Hate Speech Annotations: A Socio-Demographic Analysis of Annotators and Targets
by: Giorgi, Tommaso, et al.
Published: (2024)
by: Giorgi, Tommaso, et al.
Published: (2024)
Large Language Models for Automatic Milestone Detection in Group Discussions
by: Duan, Zhuoxu, et al.
Published: (2024)
by: Duan, Zhuoxu, et al.
Published: (2024)
Can Large Language Models Detect Verbal Indicators of Romantic Attraction?
by: Matz, Sandra C., et al.
Published: (2024)
by: Matz, Sandra C., et al.
Published: (2024)
Explaining News Bias Detection: A Comparative SHAP Analysis of Transformer Model Decision Mechanisms
by: Ghosh, Himel
Published: (2025)
by: Ghosh, Himel
Published: (2025)
Efficient Machine Translation Corpus Generation: Integrating Human-in-the-Loop Post-Editing with Large Language Models
by: Yuksel, Kamer Ali, et al.
Published: (2025)
by: Yuksel, Kamer Ali, et al.
Published: (2025)
Assessing LLM Response Quality in the Context of Technology-Facilitated Abuse
by: Prakash, Vijay, et al.
Published: (2026)
by: Prakash, Vijay, et al.
Published: (2026)
To Bias or Not to Bias: Detecting bias in News with bias-detector
by: Ghosh, Himel, et al.
Published: (2025)
by: Ghosh, Himel, et al.
Published: (2025)
Cognition Chain for Explainable Psychological Stress Detection on Social Media
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
The BS-meter: A ChatGPT-Trained Instrument to Detect Sloppy Language-Games
by: Trevisan, Alessandro, et al.
Published: (2024)
by: Trevisan, Alessandro, et al.
Published: (2024)
Planning Ahead with RSA: Efficient Signalling in Dynamic Environments by Projecting User Awareness across Future Timesteps
by: Das, Anwesha, et al.
Published: (2025)
by: Das, Anwesha, et al.
Published: (2025)
Using Machine Learning to Enhance the Detection of Obfuscated Abusive Words in Swahili: A Focus on Child Safety
by: Nabangi, Phyllis, et al.
Published: (2026)
by: Nabangi, Phyllis, et al.
Published: (2026)
Prompt2DeModel: Declarative Neuro-Symbolic Modeling with Natural Language
by: Faghihi, Hossein Rajaby, et al.
Published: (2024)
by: Faghihi, Hossein Rajaby, et al.
Published: (2024)
Chain of Empathy: Enhancing Empathetic Response of Large Language Models Based on Psychotherapy Models
by: Lee, Yoon Kyung, et al.
Published: (2023)
by: Lee, Yoon Kyung, et al.
Published: (2023)
Generative Interfaces for Language Models
by: Chen, Jiaqi, et al.
Published: (2025)
by: Chen, Jiaqi, et al.
Published: (2025)
Status Hierarchies in Language Models
by: Barkett, Emilio
Published: (2026)
by: Barkett, Emilio
Published: (2026)
Aligning Model Evaluations with Human Preferences: Mitigating Token Count Bias in Language Model Assessments
by: Daynauth, Roland, et al.
Published: (2024)
by: Daynauth, Roland, et al.
Published: (2024)
Epistemic Integrity in Large Language Models
by: Ghafouri, Bijean, et al.
Published: (2024)
by: Ghafouri, Bijean, et al.
Published: (2024)
An Evaluation of Estimative Uncertainty in Large Language Models
by: Tang, Zhisheng, et al.
Published: (2024)
by: Tang, Zhisheng, et al.
Published: (2024)
Evaluating the Prompt Steerability of Large Language Models
by: Miehling, Erik, et al.
Published: (2024)
by: Miehling, Erik, et al.
Published: (2024)
The Art of Saying No: Contextual Noncompliance in Language Models
by: Brahman, Faeze, et al.
Published: (2024)
by: Brahman, Faeze, et al.
Published: (2024)
Autonomous Prompt Engineering in Large Language Models
by: Kepel, Daan, et al.
Published: (2024)
by: Kepel, Daan, et al.
Published: (2024)
Accounting for Sycophancy in Language Model Uncertainty Estimation
by: Sicilia, Anthony, et al.
Published: (2024)
by: Sicilia, Anthony, et al.
Published: (2024)
Inertia in Moral and Value Judgments of Large Language Models
by: Lee, Bruce W., et al.
Published: (2024)
by: Lee, Bruce W., et al.
Published: (2024)
Large Language Models and Games: A Survey and Roadmap
by: Gallotta, Roberto, et al.
Published: (2024)
by: Gallotta, Roberto, et al.
Published: (2024)
(Ir)rationality and Cognitive Biases in Large Language Models
by: Macmillan-Scott, Olivia, et al.
Published: (2024)
by: Macmillan-Scott, Olivia, et al.
Published: (2024)
Evaluating Large Language Models in Analysing Classroom Dialogue
by: Long, Yun, et al.
Published: (2024)
by: Long, Yun, et al.
Published: (2024)
A Scalable Framework for Evaluating Health Language Models
by: Mallinar, Neil, et al.
Published: (2025)
by: Mallinar, Neil, et al.
Published: (2025)
Spontaneous Persuasion: An Audit of Model Persuasiveness in Everyday Conversations
by: Poungpeth, Nalin, et al.
Published: (2026)
by: Poungpeth, Nalin, et al.
Published: (2026)
Creating General User Models from Computer Use
by: Shaikh, Omar, et al.
Published: (2025)
by: Shaikh, Omar, et al.
Published: (2025)
Automated Interpretability and Feature Discovery in Language Models with Agents
by: Marin-Llobet, Arnau, et al.
Published: (2026)
by: Marin-Llobet, Arnau, et al.
Published: (2026)
Large Language Model Use Impact Locus of Control
by: Fu, Jenny Xiyu, et al.
Published: (2025)
by: Fu, Jenny Xiyu, et al.
Published: (2025)
Empowering Private Tutoring by Chaining Large Language Models
by: Chen, Yulin, et al.
Published: (2023)
by: Chen, Yulin, et al.
Published: (2023)
RNR: Teaching Large Language Models to Follow Roles and Rules
by: Wang, Kuan, et al.
Published: (2024)
by: Wang, Kuan, et al.
Published: (2024)
Automatic Histograms: Leveraging Language Models for Text Dataset Exploration
by: Reif, Emily, et al.
Published: (2024)
by: Reif, Emily, et al.
Published: (2024)
Meta-Prompting: Enhancing Language Models with Task-Agnostic Scaffolding
by: Suzgun, Mirac, et al.
Published: (2024)
by: Suzgun, Mirac, et al.
Published: (2024)
Language Models in Dialogue: Conversational Maxims for Human-AI Interactions
by: Miehling, Erik, et al.
Published: (2024)
by: Miehling, Erik, et al.
Published: (2024)
Large Language Model-Brained GUI Agents: A Survey
by: Zhang, Chaoyun, et al.
Published: (2024)
by: Zhang, Chaoyun, et al.
Published: (2024)
Understanding the Dataset Practitioners Behind Large Language Model Development
by: Qian, Crystal, et al.
Published: (2024)
by: Qian, Crystal, et al.
Published: (2024)
A Survey on Human-AI Collaboration with Large Foundation Models
by: Vats, Vanshika, et al.
Published: (2024)
by: Vats, Vanshika, et al.
Published: (2024)
Relying on the Unreliable: The Impact of Language Models' Reluctance to Express Uncertainty
by: Zhou, Kaitlyn, et al.
Published: (2024)
by: Zhou, Kaitlyn, et al.
Published: (2024)
Similar Items
-
Human and LLM Biases in Hate Speech Annotations: A Socio-Demographic Analysis of Annotators and Targets
by: Giorgi, Tommaso, et al.
Published: (2024) -
Large Language Models for Automatic Milestone Detection in Group Discussions
by: Duan, Zhuoxu, et al.
Published: (2024) -
Can Large Language Models Detect Verbal Indicators of Romantic Attraction?
by: Matz, Sandra C., et al.
Published: (2024) -
Explaining News Bias Detection: A Comparative SHAP Analysis of Transformer Model Decision Mechanisms
by: Ghosh, Himel
Published: (2025) -
Efficient Machine Translation Corpus Generation: Integrating Human-in-the-Loop Post-Editing with Large Language Models
by: Yuksel, Kamer Ali, et al.
Published: (2025)