A Collaborative Content Moderation Framework for Toxicity Detection based on Conformalized Estimates of Annotation Disagreement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Villate-Castillo, Guillermo, Del Ser, Javier, Sanz, Borja |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Probabilistic Question Answering Over Tabular Data
von: Shen, Chen, et al.
Veröffentlicht: (2025)
von: Shen, Chen, et al.
Veröffentlicht: (2025)
Experimentation in Content Moderation using RWKV
von: Yildirim, Umut, et al.
Veröffentlicht: (2024)
von: Yildirim, Umut, et al.
Veröffentlicht: (2024)
Low-Resource Neural Machine Translation Using Recurrent Neural Networks and Transfer Learning: A Case Study on English-to-Igbo
von: Ekle, Ocheme Anthony, et al.
Veröffentlicht: (2025)
von: Ekle, Ocheme Anthony, et al.
Veröffentlicht: (2025)
Prompt Engineering and the Effectiveness of Large Language Models in Enhancing Human Productivity
von: Anam, Rizal Khoirul
Veröffentlicht: (2025)
von: Anam, Rizal Khoirul
Veröffentlicht: (2025)
What is lost in Normalization? Exploring Pitfalls in Multilingual ASR Model Evaluations
von: Manohar, Kavya, et al.
Veröffentlicht: (2024)
von: Manohar, Kavya, et al.
Veröffentlicht: (2024)
Towards Conversational AI for Human-Machine Collaborative MLOps
von: Fatouros, George, et al.
Veröffentlicht: (2025)
von: Fatouros, George, et al.
Veröffentlicht: (2025)
Value Lens: Using Large Language Models to Understand Human Values
von: Fernández, Eduardo de la Cruz, et al.
Veröffentlicht: (2025)
von: Fernández, Eduardo de la Cruz, et al.
Veröffentlicht: (2025)
An Epidemiological Knowledge Graph extracted from the World Health Organization's Disease Outbreak News
von: Consoli, Sergio, et al.
Veröffentlicht: (2025)
von: Consoli, Sergio, et al.
Veröffentlicht: (2025)
MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare
von: Wang, Yihao, et al.
Veröffentlicht: (2026)
von: Wang, Yihao, et al.
Veröffentlicht: (2026)
Understanding the Uncertainty of LLM Explanations: A Perspective Based on Reasoning Topology
von: Da, Longchao, et al.
Veröffentlicht: (2025)
von: Da, Longchao, et al.
Veröffentlicht: (2025)
LLM-Assisted Crisis Management: Building Advanced LLM Platforms for Effective Emergency Response and Public Collaboration
von: Otal, Hakan T., et al.
Veröffentlicht: (2024)
von: Otal, Hakan T., et al.
Veröffentlicht: (2024)
Exploring the Structure of AI-Induced Language Change in Scientific English
von: Galpin, Riley, et al.
Veröffentlicht: (2025)
von: Galpin, Riley, et al.
Veröffentlicht: (2025)
An Iterative Optimizing Framework for Radiology Report Summarization with ChatGPT
von: Ma, Chong, et al.
Veröffentlicht: (2023)
von: Ma, Chong, et al.
Veröffentlicht: (2023)
Humor Mechanics: Advancing Humor Generation with Multistep Reasoning
von: Tikhonov, Alexey, et al.
Veröffentlicht: (2024)
von: Tikhonov, Alexey, et al.
Veröffentlicht: (2024)
Constitution or Collapse? Exploring Constitutional AI with Llama 3-8B
von: Zhang, Xue
Veröffentlicht: (2025)
von: Zhang, Xue
Veröffentlicht: (2025)
The Knesset Corpus: An Annotated Corpus of Hebrew Parliamentary Proceedings
von: Goldin, Gili, et al.
Veröffentlicht: (2024)
von: Goldin, Gili, et al.
Veröffentlicht: (2024)
Disambiguation of Emotion Annotations by Contextualizing Events in Plausible Narratives
von: Schäfer, Johannes, et al.
Veröffentlicht: (2025)
von: Schäfer, Johannes, et al.
Veröffentlicht: (2025)
PerkwE_COQA: Enhanced Persian Conversational Question Answering by combining contextual keyword extraction with Large Language Models
von: Moradbeiki, Pardis, et al.
Veröffentlicht: (2024)
von: Moradbeiki, Pardis, et al.
Veröffentlicht: (2024)
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
von: Zhang, Li, et al.
Veröffentlicht: (2026)
von: Zhang, Li, et al.
Veröffentlicht: (2026)
JAM: Controllable and Responsible Text Generation via Causal Reasoning and Latent Vector Manipulation
von: Huang, Yingbing, et al.
Veröffentlicht: (2025)
von: Huang, Yingbing, et al.
Veröffentlicht: (2025)
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
von: Haque, Md. Asraful, et al.
Veröffentlicht: (2026)
von: Haque, Md. Asraful, et al.
Veröffentlicht: (2026)
Triplètoile: Extraction of Knowledge from Microblogging Text
von: Zavarella, Vanni, et al.
Veröffentlicht: (2024)
von: Zavarella, Vanni, et al.
Veröffentlicht: (2024)
The Pursuit of Empathy: Evaluating Small Language Models for PTSD Dialogue Support
von: BN, Suhas, et al.
Veröffentlicht: (2025)
von: BN, Suhas, et al.
Veröffentlicht: (2025)
How much do LLMs learn from negative examples?
von: Hamdan, Shadi, et al.
Veröffentlicht: (2025)
von: Hamdan, Shadi, et al.
Veröffentlicht: (2025)
D-SMART: Enhancing LLM Dialogue Consistency via Dynamic Structured Memory And Reasoning Tree
von: Lei, Xiang, et al.
Veröffentlicht: (2025)
von: Lei, Xiang, et al.
Veröffentlicht: (2025)
NOTAI.AI: Explainable Detection of Machine-Generated Text via Curvature and Feature Attribution
von: Breneur, Oleksandr Marchenko, et al.
Veröffentlicht: (2026)
von: Breneur, Oleksandr Marchenko, et al.
Veröffentlicht: (2026)
Rethinking the Multilingual Reasoning Gap with Layer Swap
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2026)
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2026)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
von: Yang, Yibo
Veröffentlicht: (2025)
von: Yang, Yibo
Veröffentlicht: (2025)
AI Agents-as-Judge: Automated Assessment of Accuracy, Consistency, Completeness and Clarity for Enterprise Documents
von: Dasgupta, Sudip, et al.
Veröffentlicht: (2025)
von: Dasgupta, Sudip, et al.
Veröffentlicht: (2025)
Large Language Models are Inconsistent and Biased Evaluators
von: Stureborg, Rickard, et al.
Veröffentlicht: (2024)
von: Stureborg, Rickard, et al.
Veröffentlicht: (2024)
NERCat: Fine-Tuning for Enhanced Named Entity Recognition in Catalan
von: Ferreres, Guillem Cadevall, et al.
Veröffentlicht: (2025)
von: Ferreres, Guillem Cadevall, et al.
Veröffentlicht: (2025)
TIS-DPO: Token-level Importance Sampling for Direct Preference Optimization With Estimated Weights
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
XAutoLM: Efficient Fine-Tuning of Language Models via Meta-Learning and AutoML
von: Estevanell-Valladares, Ernesto L., et al.
Veröffentlicht: (2025)
von: Estevanell-Valladares, Ernesto L., et al.
Veröffentlicht: (2025)
Tailoring Vaccine Messaging with Common-Ground Opinions
von: Stureborg, Rickard, et al.
Veröffentlicht: (2024)
von: Stureborg, Rickard, et al.
Veröffentlicht: (2024)
Diagnosing Multi-step Reasoning Failures in Black-box LLMs via Stepwise Confidence Attribution
von: Liu, Xiaoou, et al.
Veröffentlicht: (2026)
von: Liu, Xiaoou, et al.
Veröffentlicht: (2026)
CognitiveArm: Enabling Real-Time EEG-Controlled Prosthetic Arm Using Embodied Machine Learning
von: Basit, Abdul, et al.
Veröffentlicht: (2025)
von: Basit, Abdul, et al.
Veröffentlicht: (2025)
Biomedical Visual Instruction Tuning with Clinician Preference Alignment
von: Cui, Hejie, et al.
Veröffentlicht: (2024)
von: Cui, Hejie, et al.
Veröffentlicht: (2024)
Large Language Models for Propaganda Span Annotation
von: Hasanain, Maram, et al.
Veröffentlicht: (2023)
von: Hasanain, Maram, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Towards Probabilistic Question Answering Over Tabular Data
von: Shen, Chen, et al.
Veröffentlicht: (2025) -
Experimentation in Content Moderation using RWKV
von: Yildirim, Umut, et al.
Veröffentlicht: (2024) -
Low-Resource Neural Machine Translation Using Recurrent Neural Networks and Transfer Learning: A Case Study on English-to-Igbo
von: Ekle, Ocheme Anthony, et al.
Veröffentlicht: (2025) -
Prompt Engineering and the Effectiveness of Large Language Models in Enhancing Human Productivity
von: Anam, Rizal Khoirul
Veröffentlicht: (2025) -
What is lost in Normalization? Exploring Pitfalls in Multilingual ASR Model Evaluations
von: Manohar, Kavya, et al.
Veröffentlicht: (2024)