DetoxBench: Benchmarking Large Language Models for Multitask Fraud & Abuse Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chakraborty, Joymallya, Xia, Wei, Majumder, Anirban, Ma, Dan, Chaabene, Walid, Janvekar, Naveed |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
When Less is More: On the Value of "Co-training" for Semi-Supervised Software Defect Predictors
von: Majumder, Suvodeep, et al.
Veröffentlicht: (2022)
von: Majumder, Suvodeep, et al.
Veröffentlicht: (2022)
UniDetox: Universal Detoxification of Large Language Models via Dataset Distillation
von: Lu, Huimin, et al.
Veröffentlicht: (2025)
von: Lu, Huimin, et al.
Veröffentlicht: (2025)
Small Wins Big: Comparing Large Language Models and Domain Fine-Tuned Models for Sarcasm Detection in Code-Mixed Hinglish Text
von: Majumder, Bitan, et al.
Veröffentlicht: (2026)
von: Majumder, Bitan, et al.
Veröffentlicht: (2026)
PRECISE: Reducing the Bias of LLM Evaluations Using Prediction-Powered Ranking Estimation
von: Divekar, Abhishek, et al.
Veröffentlicht: (2026)
von: Divekar, Abhishek, et al.
Veröffentlicht: (2026)
OutSafe-Bench: A Benchmark for Multimodal Offensive Content Detection in Large Language Models
von: Yan, Yuping, et al.
Veröffentlicht: (2025)
von: Yan, Yuping, et al.
Veröffentlicht: (2025)
DetoxLLM: A Framework for Detoxification with Explanations
von: Khondaker, Md Tawkat Islam, et al.
Veröffentlicht: (2024)
von: Khondaker, Md Tawkat Islam, et al.
Veröffentlicht: (2024)
AlignBench: Benchmarking Chinese Alignment of Large Language Models
von: Liu, Xiao, et al.
Veröffentlicht: (2023)
von: Liu, Xiao, et al.
Veröffentlicht: (2023)
CoBa: Convergence Balancer for Multitask Finetuning of Large Language Models
von: Gong, Zi, et al.
Veröffentlicht: (2024)
von: Gong, Zi, et al.
Veröffentlicht: (2024)
TritonBench: Benchmarking Large Language Model Capabilities for Generating Triton Operators
von: Li, Jianling, et al.
Veröffentlicht: (2025)
von: Li, Jianling, et al.
Veröffentlicht: (2025)
DiscoveryBench: Towards Data-Driven Discovery with Large Language Models
von: Majumder, Bodhisattwa Prasad, et al.
Veröffentlicht: (2024)
von: Majumder, Bodhisattwa Prasad, et al.
Veröffentlicht: (2024)
EvasionBench: A Large-Scale Benchmark for Detecting Managerial Evasion in Earnings Call Q&A
von: Ma, Shijian, et al.
Veröffentlicht: (2026)
von: Ma, Shijian, et al.
Veröffentlicht: (2026)
FairBalance: How to Achieve Equalized Odds With Data Pre-processing
von: Yu, Zhe, et al.
Veröffentlicht: (2021)
von: Yu, Zhe, et al.
Veröffentlicht: (2021)
DebateBench: A Challenging Long Context Reasoning Benchmark For Large Language Models
von: Tiwari, Utkarsh, et al.
Veröffentlicht: (2025)
von: Tiwari, Utkarsh, et al.
Veröffentlicht: (2025)
Deep Prompt Multi-task Network for Abuse Language Detection
von: Zhu, Jian, et al.
Veröffentlicht: (2024)
von: Zhu, Jian, et al.
Veröffentlicht: (2024)
Multilingual Needle in a Haystack: Investigating Long-Context Behavior of Multilingual Large Language Models
von: Hengle, Amey, et al.
Veröffentlicht: (2024)
von: Hengle, Amey, et al.
Veröffentlicht: (2024)
AssertBench: A Benchmark for Evaluating Self-Assertion in Large Language Models
von: Lee, Jaeho, et al.
Veröffentlicht: (2025)
von: Lee, Jaeho, et al.
Veröffentlicht: (2025)
EquiBench: Benchmarking Large Language Models' Reasoning about Program Semantics via Equivalence Checking
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
MMSciBench: Benchmarking Language Models on Chinese Multimodal Scientific Problems
von: Ye, Xinwu, et al.
Veröffentlicht: (2025)
von: Ye, Xinwu, et al.
Veröffentlicht: (2025)
GemDetox at TextDetox CLEF 2025: Enhancing a Massively Multilingual Model for Text Detoxification on Low-resource Languages
von: Dang, Trung Duc Anh, et al.
Veröffentlicht: (2025)
von: Dang, Trung Duc Anh, et al.
Veröffentlicht: (2025)
On the Quantization Robustness of Diffusion Language Models in Coding Benchmarks
von: Gupta, Aarav, et al.
Veröffentlicht: (2026)
von: Gupta, Aarav, et al.
Veröffentlicht: (2026)
LLM Surgery: Efficient Knowledge Unlearning and Editing in Large Language Models
von: Veldanda, Akshaj Kumar, et al.
Veröffentlicht: (2024)
von: Veldanda, Akshaj Kumar, et al.
Veröffentlicht: (2024)
Enhancing Next-Generation Language Models with Knowledge Graphs: Extending Claude, Mistral IA, and GPT-4 via KG-BERT
von: Chaabene, Nour El Houda Ben, et al.
Veröffentlicht: (2025)
von: Chaabene, Nour El Houda Ben, et al.
Veröffentlicht: (2025)
LoraxBench: A Multitask, Multilingual Benchmark Suite for 20 Indonesian Languages
von: Aji, Alham Fikri, et al.
Veröffentlicht: (2025)
von: Aji, Alham Fikri, et al.
Veröffentlicht: (2025)
Large Language Models in the Abuse Detection Pipeline
von: Kath, Suraj, et al.
Veröffentlicht: (2026)
von: Kath, Suraj, et al.
Veröffentlicht: (2026)
Bayesian Mixture of Experts For Large Language Models
von: Dialameh, Maryam, et al.
Veröffentlicht: (2025)
von: Dialameh, Maryam, et al.
Veröffentlicht: (2025)
Understanding Structured Financial Data with LLMs: A Case Study on Fraud Detection
von: Tan, Xuwei, et al.
Veröffentlicht: (2025)
von: Tan, Xuwei, et al.
Veröffentlicht: (2025)
Overview of the 2023 ICON Shared Task on Gendered Abuse Detection in Indic Languages
von: Vaidya, Aatman, et al.
Veröffentlicht: (2024)
von: Vaidya, Aatman, et al.
Veröffentlicht: (2024)
Hallucination Detox: Sensitivity Dropout (SenD) for Large Language Model Training
von: Mohammadzadeh, Shahrad, et al.
Veröffentlicht: (2024)
von: Mohammadzadeh, Shahrad, et al.
Veröffentlicht: (2024)
SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
AIR-Bench: Benchmarking Large Audio-Language Models via Generative Comprehension
von: Yang, Qian, et al.
Veröffentlicht: (2024)
von: Yang, Qian, et al.
Veröffentlicht: (2024)
SALAD-Bench: A Hierarchical and Comprehensive Safety Benchmark for Large Language Models
von: Li, Lijun, et al.
Veröffentlicht: (2024)
von: Li, Lijun, et al.
Veröffentlicht: (2024)
CS-Bench: A Comprehensive Benchmark for Large Language Models towards Computer Science Mastery
von: Song, Xiaoshuai, et al.
Veröffentlicht: (2024)
von: Song, Xiaoshuai, et al.
Veröffentlicht: (2024)
VMMU: A Vietnamese Multitask Multimodal Understanding and Reasoning Benchmark
von: Dang, Vy Tuong, et al.
Veröffentlicht: (2025)
von: Dang, Vy Tuong, et al.
Veröffentlicht: (2025)
Evaluation and Improvement of Fault Detection for Large Language Models
von: Hu, Qiang, et al.
Veröffentlicht: (2024)
von: Hu, Qiang, et al.
Veröffentlicht: (2024)
CliBench: A Multifaceted and Multigranular Evaluation of Large Language Models for Clinical Decision Making
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2024)
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2024)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Explaining Large Language Models with gSMILE
von: Dehghani, Zeinab, et al.
Veröffentlicht: (2025)
von: Dehghani, Zeinab, et al.
Veröffentlicht: (2025)
Reinterpreting 'the Company a Word Keeps': Towards Explainable and Ontologically Grounded Language Models
von: Saba, Walid S.
Veröffentlicht: (2024)
von: Saba, Walid S.
Veröffentlicht: (2024)
Mechanistic Anomaly Detection for "Quirky" Language Models
von: Johnston, David O., et al.
Veröffentlicht: (2025)
von: Johnston, David O., et al.
Veröffentlicht: (2025)
AraLingBench A Human-Annotated Benchmark for Evaluating Arabic Linguistic Capabilities of Large Language Models
von: Zbeeb, Mohammad, et al.
Veröffentlicht: (2025)
von: Zbeeb, Mohammad, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
When Less is More: On the Value of "Co-training" for Semi-Supervised Software Defect Predictors
von: Majumder, Suvodeep, et al.
Veröffentlicht: (2022) -
UniDetox: Universal Detoxification of Large Language Models via Dataset Distillation
von: Lu, Huimin, et al.
Veröffentlicht: (2025) -
Small Wins Big: Comparing Large Language Models and Domain Fine-Tuned Models for Sarcasm Detection in Code-Mixed Hinglish Text
von: Majumder, Bitan, et al.
Veröffentlicht: (2026) -
PRECISE: Reducing the Bias of LLM Evaluations Using Prediction-Powered Ranking Estimation
von: Divekar, Abhishek, et al.
Veröffentlicht: (2026) -
OutSafe-Bench: A Benchmark for Multimodal Offensive Content Detection in Large Language Models
von: Yan, Yuping, et al.
Veröffentlicht: (2025)