Saved in:
| Main Authors: | Gupta, Soumyajit, Kovatchev, Venelin, Das, Anubrata, De-Arteaga, Maria, Lease, Matthew |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2204.07661 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmark Transparency: Measuring the Impact of Data on Evaluation
by: Kovatchev, Venelin, et al.
Published: (2024)
by: Kovatchev, Venelin, et al.
Published: (2024)
Fairness-Aware Multi-Group Target Detection in Online Discussion
by: Gupta, Soumyajit, et al.
Published: (2024)
by: Gupta, Soumyajit, et al.
Published: (2024)
Capturing Classic Authorial Style in Long-Form Story Generation with GRPO Fine-Tuning
by: Liu, Jinlong, et al.
Published: (2025)
by: Liu, Jinlong, et al.
Published: (2025)
Human-centered NLP Fact-checking: Co-Designing with Fact-checkers using Matchmaking for AI
by: Liu, Houjiang, et al.
Published: (2023)
by: Liu, Houjiang, et al.
Published: (2023)
Diverse, but Divisive: LLMs Can Exaggerate Gender Differences in Opinion Related to Harms of Misinformation
by: Neumann, Terrence, et al.
Published: (2024)
by: Neumann, Terrence, et al.
Published: (2024)
LLM REgression with a Latent Iterative State Head
by: Su, Yiheng, et al.
Published: (2026)
by: Su, Yiheng, et al.
Published: (2026)
Promoting Constructive Deliberation: Reframing for Receptiveness
by: Kambhatla, Gauri, et al.
Published: (2024)
by: Kambhatla, Gauri, et al.
Published: (2024)
Linear Representation Transferability Hypothesis: Leveraging Small Models to Steer Large Models
by: Bello, Femi, et al.
Published: (2025)
by: Bello, Femi, et al.
Published: (2025)
The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification
by: Bui, Minh Duc, et al.
Published: (2024)
by: Bui, Minh Duc, et al.
Published: (2024)
Navigating the Alignment-Calibration Trade-off: A Pareto-Superior Frontier via Model Merging
by: Hu, Tiancheng, et al.
Published: (2025)
by: Hu, Tiancheng, et al.
Published: (2025)
Exploring Accuracy-Fairness Trade-off in Large Language Models
by: Zhang, Qingquan, et al.
Published: (2024)
by: Zhang, Qingquan, et al.
Published: (2024)
On Bias and Fairness in NLP: Investigating the Impact of Bias and Debiasing in Language Models on the Fairness of Toxicity Detection
by: Elsafoury, Fatma, et al.
Published: (2023)
by: Elsafoury, Fatma, et al.
Published: (2023)
Exploring Multimodal Challenges in Toxic Chinese Detection: Taxonomy, Benchmark, and Findings
by: Yang, Shujian, et al.
Published: (2025)
by: Yang, Shujian, et al.
Published: (2025)
Graph Neural Networks for Misinformation Detection: Performance-Efficiency Trade-offs
by: Kuntur, Soveatin, et al.
Published: (2026)
by: Kuntur, Soveatin, et al.
Published: (2026)
ToXCL: A Unified Framework for Toxic Speech Detection and Explanation
by: Hoang, Nhat M., et al.
Published: (2024)
by: Hoang, Nhat M., et al.
Published: (2024)
On the Interplay of Human-AI Alignment,Fairness, and Performance Trade-offs in Medical Imaging
by: Luo, Haozhe, et al.
Published: (2025)
by: Luo, Haozhe, et al.
Published: (2025)
Understanding Fairness-Accuracy Trade-offs in Machine Learning Models: Does Promoting Fairness Undermine Performance?
by: Liu, Junhua, et al.
Published: (2024)
by: Liu, Junhua, et al.
Published: (2024)
PIE: Performance Interval Estimation for Free-Form Generation Tasks
by: Hsu, Chi-Yang, et al.
Published: (2025)
by: Hsu, Chi-Yang, et al.
Published: (2025)
ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances
by: Do, Huy Ba, et al.
Published: (2025)
by: Do, Huy Ba, et al.
Published: (2025)
Towards Fairness Assessment of Dutch Hate Speech Detection
by: Bauer, Julie, et al.
Published: (2025)
by: Bauer, Julie, et al.
Published: (2025)
When to Invoke: Refining LLM Fairness with Toxicity Assessment
by: Ren, Jing, et al.
Published: (2026)
by: Ren, Jing, et al.
Published: (2026)
Hate Speech Detection with Generalizable Target-aware Fairness
by: Chen, Tong, et al.
Published: (2024)
by: Chen, Tong, et al.
Published: (2024)
Accurate and Data-Efficient Toxicity Prediction when Annotators Disagree
by: Jaggi, Harbani, et al.
Published: (2024)
by: Jaggi, Harbani, et al.
Published: (2024)
ID-XCB: Data-independent Debiasing for Fair and Accurate Transformer-based Cyberbullying Detection
by: Yi, Peiling, et al.
Published: (2024)
by: Yi, Peiling, et al.
Published: (2024)
Algospeak, Hiding in the Open: The Trade-off Between Legible Meaning and Detection Avoidance
by: Fillies, Jan, et al.
Published: (2026)
by: Fillies, Jan, et al.
Published: (2026)
SpeechT: Findings of the First Mentorship in Speech Translation
by: Moslem, Yasmin, et al.
Published: (2025)
by: Moslem, Yasmin, et al.
Published: (2025)
Enhancing Multilingual Voice Toxicity Detection with Speech-Text Alignment
by: Liu, Joseph, et al.
Published: (2024)
by: Liu, Joseph, et al.
Published: (2024)
An Effective, Robust and Fairness-aware Hate Speech Detection Framework
by: Mou, Guanyi, et al.
Published: (2024)
by: Mou, Guanyi, et al.
Published: (2024)
Bridging Fairness and Explainability: Can Input-Based Explanations Promote Fairness in Hate Speech Detection?
by: Wang, Yifan, et al.
Published: (2025)
by: Wang, Yifan, et al.
Published: (2025)
Toxicity Detection for Free
by: Hu, Zhanhao, et al.
Published: (2024)
by: Hu, Zhanhao, et al.
Published: (2024)
Improving the Distributional Alignment of LLMs using Supervision
by: Kambhatla, Gauri, et al.
Published: (2025)
by: Kambhatla, Gauri, et al.
Published: (2025)
Probabilistic Interval Analysis of Unreliable Programs
by: Das, Dibyendu, et al.
Published: (2024)
by: Das, Dibyendu, et al.
Published: (2024)
Who Owns Creativity and Who Does the Work? Trade-offs in LLM-Supported Research Ideation
by: Liu, Houjiang, et al.
Published: (2026)
by: Liu, Houjiang, et al.
Published: (2026)
On the Role of Speech Data in Reducing Toxicity Detection Bias
by: Bell, Samuel J., et al.
Published: (2024)
by: Bell, Samuel J., et al.
Published: (2024)
Fundamental Safety-Capability Trade-offs in Fine-tuning Large Language Models
by: Chen, Pin-Yu, et al.
Published: (2025)
by: Chen, Pin-Yu, et al.
Published: (2025)
WaterJudge: Quality-Detection Trade-off when Watermarking Large Language Models
by: Molenda, Piotr, et al.
Published: (2024)
by: Molenda, Piotr, et al.
Published: (2024)
Algorithmic Fairness in NLP: Persona-Infused LLMs for Human-Centric Hate Speech Detection
by: Gajewska, Ewelina, et al.
Published: (2025)
by: Gajewska, Ewelina, et al.
Published: (2025)
FFSplit: Split Feed-Forward Network For Optimizing Accuracy-Efficiency Trade-off in Language Model Inference
by: Liu, Zirui, et al.
Published: (2024)
by: Liu, Zirui, et al.
Published: (2024)
WaterSearch: Exploring Seed Pooling for Improving the Quality-Detectability Trade-off in LLM Watermarking
by: Lin, Yukang, et al.
Published: (2025)
by: Lin, Yukang, et al.
Published: (2025)
Responsible Benchmarking of Fairness for Automatic Speech Recognition
by: Herron, Felix, et al.
Published: (2026)
by: Herron, Felix, et al.
Published: (2026)
Similar Items
-
Benchmark Transparency: Measuring the Impact of Data on Evaluation
by: Kovatchev, Venelin, et al.
Published: (2024) -
Fairness-Aware Multi-Group Target Detection in Online Discussion
by: Gupta, Soumyajit, et al.
Published: (2024) -
Capturing Classic Authorial Style in Long-Form Story Generation with GRPO Fine-Tuning
by: Liu, Jinlong, et al.
Published: (2025) -
Human-centered NLP Fact-checking: Co-Designing with Fact-checkers using Matchmaking for AI
by: Liu, Houjiang, et al.
Published: (2023) -
Diverse, but Divisive: LLMs Can Exaggerate Gender Differences in Opinion Related to Harms of Misinformation
by: Neumann, Terrence, et al.
Published: (2024)