LLM generated responses to mitigate the impact of hate speech
Fuente:
arXiv
Guardado en:
| Autores principales: | Podolak, Jakub, Łukasik, Szymon, Balawender, Paweł, Ossowski, Jan, Piotrowski, Jan, Bąkowicz, Katarzyna, Sankowski, Piotr |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PL-Guard: Benchmarking Language Model Safety for Polish
por: Krasnodębska, Aleksandra, et al.
Publicado: (2025)
por: Krasnodębska, Aleksandra, et al.
Publicado: (2025)
Large Language Models for Biomedical Article Classification
por: Proboszcz, Jakub, et al.
Publicado: (2026)
por: Proboszcz, Jakub, et al.
Publicado: (2026)
Large Language Model (LLM) Bias Index -- LLMBI
por: Oketunji, Abiodun Finbarrs, et al.
Publicado: (2023)
por: Oketunji, Abiodun Finbarrs, et al.
Publicado: (2023)
Digital Guardians: Can GPT-4, Perspective API, and Moderation API reliably detect hate speech in reader comments of German online newspapers?
por: Weber, Manuel, et al.
Publicado: (2025)
por: Weber, Manuel, et al.
Publicado: (2025)
Is Textual Similarity Invariant under Machine Translation? Evidence Based on the Political Manifesto Corpus
por: Boratyn, Daria, et al.
Publicado: (2026)
por: Boratyn, Daria, et al.
Publicado: (2026)
Auditing Meta-Cognitive Hallucinations in Reasoning Large Language Models
por: Lu, Haolang, et al.
Publicado: (2025)
por: Lu, Haolang, et al.
Publicado: (2025)
Evaluating the Clinical Safety of LLMs in Response to High-Risk Mental Health Disclosures
por: Shah, Siddharth, et al.
Publicado: (2025)
por: Shah, Siddharth, et al.
Publicado: (2025)
An ethical study of generative AI from the Actor-Network Theory perspective
por: Li, Yuying, et al.
Publicado: (2024)
por: Li, Yuying, et al.
Publicado: (2024)
Value Lens: Using Large Language Models to Understand Human Values
por: Fernández, Eduardo de la Cruz, et al.
Publicado: (2025)
por: Fernández, Eduardo de la Cruz, et al.
Publicado: (2025)
A multilingual dataset for offensive language and hate speech detection for hausa, yoruba and igbo languages
por: Aliyu, Saminu Mohammad, et al.
Publicado: (2024)
por: Aliyu, Saminu Mohammad, et al.
Publicado: (2024)
Mobile Phone Sensor-based Nigerian Driving Dataset to Detect Alcohol-influenced Behaviours
por: Thompson, Iniakpokeikiye Peter, et al.
Publicado: (2025)
por: Thompson, Iniakpokeikiye Peter, et al.
Publicado: (2025)
Engineering A Large Language Model From Scratch
por: Oketunji, Abiodun Finbarrs
Publicado: (2024)
por: Oketunji, Abiodun Finbarrs
Publicado: (2024)
Variance-Aware LLM Annotation for Strategy Research: Sources, Diagnostics, and a Protocol for Reliable Measurement
por: Camuffo, Arnaldo, et al.
Publicado: (2025)
por: Camuffo, Arnaldo, et al.
Publicado: (2025)
Statistical Scouting Finds Debate-Safe but Not Debate-Useful Cases: A Matched-Ceiling Study of Open-Weight LLM Reasoning Protocols
por: Hu, Julia, et al.
Publicado: (2026)
por: Hu, Julia, et al.
Publicado: (2026)
Supervised Semantic Differential for Cross-Cultural Concept Analysis: A Case Study of Human Affect
por: Sikora, Jan, et al.
Publicado: (2026)
por: Sikora, Jan, et al.
Publicado: (2026)
Conversations with Andrea: Visitors' Opinions on Android Robots in a Museum
por: Heisler, Marcel, et al.
Publicado: (2025)
por: Heisler, Marcel, et al.
Publicado: (2025)
Toward Architecture-Aware Evaluation Metrics for LLM Agents
por: Souza, Débora, et al.
Publicado: (2026)
por: Souza, Débora, et al.
Publicado: (2026)
Text Finder Application for Android
por: Godase, Milind, et al.
Publicado: (2023)
por: Godase, Milind, et al.
Publicado: (2023)
Fane at SemEval-2025 Task 10: Zero-Shot Entity Framing with Large Language Models
por: Fane, Enfa, et al.
Publicado: (2025)
por: Fane, Enfa, et al.
Publicado: (2025)
Towards Red Teaming in Multimodal and Multilingual Translation
por: Ropers, Christophe, et al.
Publicado: (2024)
por: Ropers, Christophe, et al.
Publicado: (2024)
How Well Do LLMs Imitate Human Writing Style?
por: Jemama, Rebira, et al.
Publicado: (2025)
por: Jemama, Rebira, et al.
Publicado: (2025)
Random Silicon Sampling: Simulating Human Sub-Population Opinion Using a Large Language Model Based on Group-Level Demographic Information
por: Sun, Seungjong, et al.
Publicado: (2024)
por: Sun, Seungjong, et al.
Publicado: (2024)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
por: Ashuach, Tomer, et al.
Publicado: (2025)
por: Ashuach, Tomer, et al.
Publicado: (2025)
A Benchmark for Audio Reasoning Capabilities of Multimodal Large Language Models
por: Christop, Iwona, et al.
Publicado: (2026)
por: Christop, Iwona, et al.
Publicado: (2026)
Are Generative Models Underconfident? Better Quality Estimation with Boosted Model Probability
por: Dinh, Tu Anh, et al.
Publicado: (2025)
por: Dinh, Tu Anh, et al.
Publicado: (2025)
Sigmoid Head for Quality Estimation under Language Ambiguity
por: Dinh, Tu Anh, et al.
Publicado: (2026)
por: Dinh, Tu Anh, et al.
Publicado: (2026)
Social and Ethical Risks Posed by General-Purpose LLMs for Settling Newcomers in Canada
por: Nejadgholi, Isar, et al.
Publicado: (2024)
por: Nejadgholi, Isar, et al.
Publicado: (2024)
The Impact of Generative Artificial Intelligence on Ideation and the performance of Innovation Teams (Preprint)
por: Gindert, Michael, et al.
Publicado: (2024)
por: Gindert, Michael, et al.
Publicado: (2024)
MKJ at SemEval-2026 Task 9: A Comparative Study of Generalist, Specialist, and Ensemble Strategies for Multilingual Polarization
por: Jouneghani, Maziar Kianimoghadam
Publicado: (2026)
por: Jouneghani, Maziar Kianimoghadam
Publicado: (2026)
A Faceted Proposal for Transparent Attribution of AI-Assisted Text Production
por: Xexéo, Geraldo
Publicado: (2026)
por: Xexéo, Geraldo
Publicado: (2026)
propella-1: Multi-Property Document Annotation for LLM Data Curation at Scale
por: Idahl, Maximilian, et al.
Publicado: (2026)
por: Idahl, Maximilian, et al.
Publicado: (2026)
Beyond the Cloud: Assessing the Benefits and Drawbacks of Local LLM Deployment for Translators
por: Sandrini, Peter
Publicado: (2025)
por: Sandrini, Peter
Publicado: (2025)
Synthetic Student Responses: LLM-Extracted Features for IRT Difficulty Parameter Estimation
por: Hoyl, Matias
Publicado: (2026)
por: Hoyl, Matias
Publicado: (2026)
Self-hosted Lecture-to-Quiz: Local LLM MCQ Generation with Deterministic Quality Control
por: Shintani, Seine A.
Publicado: (2026)
por: Shintani, Seine A.
Publicado: (2026)
Quantifying Algorithmic Friction in Automated Resume Screening Systems
por: Fofanah, Ibrahim Denis
Publicado: (2026)
por: Fofanah, Ibrahim Denis
Publicado: (2026)
Methodological Foundations for AI-Driven Survey Question Generation
por: Mburu, Ted K., et al.
Publicado: (2025)
por: Mburu, Ted K., et al.
Publicado: (2025)
Chatbot Deployment Considerations for Application-Agnostic Human-Machine Dialogues
por: Rivas, Pablo, et al.
Publicado: (2025)
por: Rivas, Pablo, et al.
Publicado: (2025)
Beyond Replacement or Augmentation: How Creative Workers Reconfigure Division of Labor with Generative AI
por: Clarke, Michael, et al.
Publicado: (2025)
por: Clarke, Michael, et al.
Publicado: (2025)
Quality Estimation with $k$-nearest Neighbors and Automatic Evaluation for Model-specific Quality Estimation
por: Dinh, Tu Anh, et al.
Publicado: (2024)
por: Dinh, Tu Anh, et al.
Publicado: (2024)
How Large Language Models Are Changing MOOC Essay Answers: A Comparison of Pre- and Post-LLM Responses
por: Leppänen, Leo, et al.
Publicado: (2025)
por: Leppänen, Leo, et al.
Publicado: (2025)
Ejemplares similares
-
PL-Guard: Benchmarking Language Model Safety for Polish
por: Krasnodębska, Aleksandra, et al.
Publicado: (2025) -
Large Language Models for Biomedical Article Classification
por: Proboszcz, Jakub, et al.
Publicado: (2026) -
Large Language Model (LLM) Bias Index -- LLMBI
por: Oketunji, Abiodun Finbarrs, et al.
Publicado: (2023) -
Digital Guardians: Can GPT-4, Perspective API, and Moderation API reliably detect hate speech in reader comments of German online newspapers?
por: Weber, Manuel, et al.
Publicado: (2025) -
Is Textual Similarity Invariant under Machine Translation? Evidence Based on the Political Manifesto Corpus
por: Boratyn, Daria, et al.
Publicado: (2026)