Distilling Knowledge from Large Language Models: A Concept Bottleneck Model for Hate and Counter Speech Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Labadie-Tamayo, Roberto, Slijepčević, Djordje, Chen, Xihui, Böck, Adrian Jaques, Babic, Andreas, Freimann, Liz, Zeppelzauer, Christiane Atzmüller Matthias |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FHSTP@EXIST 2025 Benchmark: Sexism Detection with Transparent Speech Concept Bottleneck Models
by: Labadie-Tamayo, Roberto, et al.
Published: (2025)
by: Labadie-Tamayo, Roberto, et al.
Published: (2025)
Exploring the Plausibility of Hate and Counter Speech Detectors with Explainable AI
by: Böck, Adrian Jaques, et al.
Published: (2024)
by: Böck, Adrian Jaques, et al.
Published: (2024)
Explanatory Interactive Machine Learning for Bias Mitigation in Visual Gender Classification
by: Satriani, Nathanya, et al.
Published: (2026)
by: Satriani, Nathanya, et al.
Published: (2026)
COT: A Generative Approach for Hate Speech Counter-Narratives via Contrastive Optimal Transport
by: Zhang, Linhao, et al.
Published: (2024)
by: Zhang, Linhao, et al.
Published: (2024)
Interactive Discovery and Exploration of Visual Bias in Generative Text-to-Image Models
by: Eschner, Johannes, et al.
Published: (2025)
by: Eschner, Johannes, et al.
Published: (2025)
Arabic Hate Speech Identification and Masking in Social Media using Deep Learning Models and Pre-trained Models Fine-tuning
by: Doghmash, Salam Thabet, et al.
Published: (2025)
by: Doghmash, Salam Thabet, et al.
Published: (2025)
Seeing Hate Differently: Hate Subspace Modeling for Culture-Aware Hate Speech Detection
by: Cai, Weibin, et al.
Published: (2025)
by: Cai, Weibin, et al.
Published: (2025)
Machine Learning in Biomechanics: Key Applications and Limitations in Walking, Running, and Sports Movements
by: Dindorf, Carlo, et al.
Published: (2025)
by: Dindorf, Carlo, et al.
Published: (2025)
Enhancing Hate Speech Detection on Social Media: A Comparative Analysis of Machine Learning Models and Text Transformation Approaches
by: Mishra, Saurabh, et al.
Published: (2026)
by: Mishra, Saurabh, et al.
Published: (2026)
Evaluation of Hate Speech Detection Using Large Language Models and Geographical Contextualization
by: Zahid, Anwar Hossain, et al.
Published: (2025)
by: Zahid, Anwar Hossain, et al.
Published: (2025)
Dealing with Annotator Disagreement in Hate Speech Classification
by: Dehghan, Somaiyeh, et al.
Published: (2025)
by: Dehghan, Somaiyeh, et al.
Published: (2025)
Boosting Accuracy and Interpretability in Multilingual Hate Speech Detection Through Layer Freezing and Explainable AI
by: Bilehsavar, Meysam Shirdel, et al.
Published: (2026)
by: Bilehsavar, Meysam Shirdel, et al.
Published: (2026)
Efficient Hate Speech Detection: Evaluating 38 Models from Traditional Methods to Transformers
by: Abusaqer, Mahmoud, et al.
Published: (2025)
by: Abusaqer, Mahmoud, et al.
Published: (2025)
Bielik-Minitron-7B: Compressing Large Language Models via Structured Pruning and Knowledge Distillation for the Polish Language
by: Kinas, Remigiusz, et al.
Published: (2026)
by: Kinas, Remigiusz, et al.
Published: (2026)
Analysis of Hybrid Compositions in Animation Film with Weakly Supervised Learning
by: Portos, Mónica Apellaniz, et al.
Published: (2024)
by: Portos, Mónica Apellaniz, et al.
Published: (2024)
RV-HATE: Reinforced Multi-Module Voting for Implicit Hate Speech Detection
by: Lee, Yejin, et al.
Published: (2025)
by: Lee, Yejin, et al.
Published: (2025)
LLM-Based Multi-Task Bangla Hate Speech Detection: Type, Severity, and Target
by: Hasan, Md Arid, et al.
Published: (2025)
by: Hasan, Md Arid, et al.
Published: (2025)
Confident, Calibrated, or Complicit: Safety Alignment and Ideological Bias in LLM Hate Speech Detection
by: Selvaganapathy, Sanjeeevan, et al.
Published: (2025)
by: Selvaganapathy, Sanjeeevan, et al.
Published: (2025)
Synthetic Voice Data for Automatic Speech Recognition in African Languages
by: DeRenzi, Brian, et al.
Published: (2025)
by: DeRenzi, Brian, et al.
Published: (2025)
Distilled HuBERT for Mobile Speech Emotion Recognition: A Cross-Corpus Validation Study
by: Ismail, Saifelden M.
Published: (2025)
by: Ismail, Saifelden M.
Published: (2025)
BlasBench: An Open Benchmark for Irish Speech Recognition
by: Raj, Jyoutir, et al.
Published: (2026)
by: Raj, Jyoutir, et al.
Published: (2026)
Can Large Language Models Imitate Human Speech for Clinical Assessment? LLM-Driven Data Augmentation for Cognitive Score Prediction
by: Ketir, Si-Belkacem Yamine, et al.
Published: (2026)
by: Ketir, Si-Belkacem Yamine, et al.
Published: (2026)
AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection
by: Lee, Yejin, et al.
Published: (2025)
by: Lee, Yejin, et al.
Published: (2025)
Synergy: End-to-end Concept Model
by: Zheng, Keli, et al.
Published: (2025)
by: Zheng, Keli, et al.
Published: (2025)
Outcome-Constrained Large Language Models for Countering Hate Speech
by: Hong, Lingzi, et al.
Published: (2024)
by: Hong, Lingzi, et al.
Published: (2024)
Preventing and Countering Online Hate Speech
by: Palermo, Rosa
Published: (2024)
by: Palermo, Rosa
Published: (2024)
Comparing Complex Concepts with Transformers: Matching Patent Claims Against Natural Language Text
by: Blume, Matthias, et al.
Published: (2024)
by: Blume, Matthias, et al.
Published: (2024)
Named Entity Recognition for Address Extraction in Speech-to-Text Transcriptions Using Synthetic Data
by: Lajčinová, Bibiána, et al.
Published: (2024)
by: Lajčinová, Bibiána, et al.
Published: (2024)
Improving Speech Recognition Accuracy Using Custom Language Models with the Vosk Toolkit
by: Soni, Aniket Abhishek
Published: (2025)
by: Soni, Aniket Abhishek
Published: (2025)
Robust Long-Form Bangla Speech Processing: Automatic Speech Recognition and Speaker Diarization
by: Chowdhury, MD. Sagor, et al.
Published: (2026)
by: Chowdhury, MD. Sagor, et al.
Published: (2026)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
by: Yang, Yibo
Published: (2025)
by: Yang, Yibo
Published: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
by: Ashuach, Tomer, et al.
Published: (2025)
by: Ashuach, Tomer, et al.
Published: (2025)
Developing Acoustic Models for Automatic Speech Recognition in Swedish
by: Salvi, Giampiero
Published: (2024)
by: Salvi, Giampiero
Published: (2024)
Align-to-Distill: Trainable Attention Alignment for Knowledge Distillation in Neural Machine Translation
by: Jin, Heegon, et al.
Published: (2024)
by: Jin, Heegon, et al.
Published: (2024)
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
by: Pan, Leyi, et al.
Published: (2025)
by: Pan, Leyi, et al.
Published: (2025)
Measuring the Accuracy of Automatic Speech Recognition Solutions
by: Kuhn, Korbinian, et al.
Published: (2024)
by: Kuhn, Korbinian, et al.
Published: (2024)
The Curious Case of Visual Grounding: Different Effects for Speech- and Text-based Language Encoders
by: Sauter, Adrian, et al.
Published: (2025)
by: Sauter, Adrian, et al.
Published: (2025)
Convex Low-resource Accent-Robust Language Detection in Speech Recognition
by: Feng, Miria, et al.
Published: (2026)
by: Feng, Miria, et al.
Published: (2026)
Multi-View Multi-Task Modeling with Speech Foundation Models for Speech Forensic Tasks
by: Phukan, Orchid Chetia, et al.
Published: (2024)
by: Phukan, Orchid Chetia, et al.
Published: (2024)
Automatic Speech Recognition (ASR) for the Diagnosis of pronunciation of Speech Sound Disorders in Korean children
by: Ahn, Taekyung, et al.
Published: (2024)
by: Ahn, Taekyung, et al.
Published: (2024)
Similar Items
-
FHSTP@EXIST 2025 Benchmark: Sexism Detection with Transparent Speech Concept Bottleneck Models
by: Labadie-Tamayo, Roberto, et al.
Published: (2025) -
Exploring the Plausibility of Hate and Counter Speech Detectors with Explainable AI
by: Böck, Adrian Jaques, et al.
Published: (2024) -
Explanatory Interactive Machine Learning for Bias Mitigation in Visual Gender Classification
by: Satriani, Nathanya, et al.
Published: (2026) -
COT: A Generative Approach for Hate Speech Counter-Narratives via Contrastive Optimal Transport
by: Zhang, Linhao, et al.
Published: (2024) -
Interactive Discovery and Exploration of Visual Bias in Generative Text-to-Image Models
by: Eschner, Johannes, et al.
Published: (2025)