Realistic Evaluation of Toxicity in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Luong, Tinh Son, Le, Thanh-Thien, Van, Linh Ngo, Nguyen, Thien Huu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ToVo: Toxicity Taxonomy via Voting
by: Luong, Tinh Son, et al.
Published: (2024)
by: Luong, Tinh Son, et al.
Published: (2024)
Zero-shot Cross-lingual Transfer Learning with Multiple Source and Target Languages for Information Extraction: Language Selection and Adversarial Training
by: Ngo, Nghia Trung, et al.
Published: (2024)
by: Ngo, Nghia Trung, et al.
Published: (2024)
Preserving Generalization of Language models in Few-shot Continual Relation Extraction
by: Tran, Quyen, et al.
Published: (2024)
by: Tran, Quyen, et al.
Published: (2024)
Comprehensive and Practical Evaluation of Retrieval-Augmented Generation Systems for Medical Question Answering
by: Ngo, Nghia Trung, et al.
Published: (2024)
by: Ngo, Nghia Trung, et al.
Published: (2024)
mSCoRe: a $M$ultilingual and Scalable Benchmark for $S$kill-based $Co$mmonsense $Re$asoning
by: Ngo, Nghia Trung, et al.
Published: (2025)
by: Ngo, Nghia Trung, et al.
Published: (2025)
Few-Shot, No Problem: Descriptive Continual Relation Extraction
by: Thanh, Nguyen Xuan, et al.
Published: (2025)
by: Thanh, Nguyen Xuan, et al.
Published: (2025)
Lifelong Event Detection via Optimal Transport
by: Dao, Viet, et al.
Published: (2024)
by: Dao, Viet, et al.
Published: (2024)
Adaptive Prompting for Continual Relation Extraction: A Within-Task Variance Perspective
by: Le, Minh, et al.
Published: (2024)
by: Le, Minh, et al.
Published: (2024)
BERT-VBD: Vietnamese Multi-Document Summarization Framework
by: Vuong, Tuan-Cuong, et al.
Published: (2024)
by: Vuong, Tuan-Cuong, et al.
Published: (2024)
Selective Off-Policy Reference Tuning with Plan Guidance
by: Le, Duc Anh, et al.
Published: (2026)
by: Le, Duc Anh, et al.
Published: (2026)
NeuroMax: Enhancing Neural Topic Modeling via Maximizing Mutual Information and Group Topic Regularization
by: Pham, Duy-Tung, et al.
Published: (2024)
by: Pham, Duy-Tung, et al.
Published: (2024)
Few-shot Continual Relation Extraction via Open Information Extraction
by: Nguyen, Thiem, et al.
Published: (2025)
by: Nguyen, Thiem, et al.
Published: (2025)
BIMA: Bijective Maximum Likelihood Learning Approach to Hallucination Prediction and Mitigation in Large Vision-Language Models
by: Tran, Huu-Thien, et al.
Published: (2025)
by: Tran, Huu-Thien, et al.
Published: (2025)
GloCTM: Cross-Lingual Topic Modeling via a Global Context Space
by: Phat, Nguyen Tien, et al.
Published: (2026)
by: Phat, Nguyen Tien, et al.
Published: (2026)
ULLME: A Unified Framework for Large Language Model Embeddings with Generation-Augmented Learning
by: Man, Hieu, et al.
Published: (2024)
by: Man, Hieu, et al.
Published: (2024)
Taipan: Efficient and Expressive State Space Language Models with Selective Attention
by: Van Nguyen, Chien, et al.
Published: (2024)
by: Van Nguyen, Chien, et al.
Published: (2024)
CoT2Align: Cross-Chain of Thought Distillation via Optimal Transport Alignment for Language Models with Different Tokenizers
by: Le, Anh Duc, et al.
Published: (2025)
by: Le, Anh Duc, et al.
Published: (2025)
GloCOM: A Short Text Neural Topic Model via Global Clustering Context
by: Nguyen, Quang Duc, et al.
Published: (2024)
by: Nguyen, Quang Duc, et al.
Published: (2024)
PhoWhisper: Automatic Speech Recognition for Vietnamese
by: Le, Thanh-Thien, et al.
Published: (2024)
by: Le, Thanh-Thien, et al.
Published: (2024)
Householder Pseudo-Rotation: A Novel Approach to Activation Editing in LLMs with Direction-Magnitude Perspective
by: Pham, Van-Cuong, et al.
Published: (2024)
by: Pham, Van-Cuong, et al.
Published: (2024)
ZeFaV: Boosting Large Language Models for Zero-shot Fact Verification
by: Luu, Son T., et al.
Published: (2024)
by: Luu, Son T., et al.
Published: (2024)
Medalyze: Lightweight Medical Report Summarization Application Using FLAN-T5-Large
by: Nguyen, Van-Tinh, et al.
Published: (2025)
by: Nguyen, Van-Tinh, et al.
Published: (2025)
Brain Tumor Segmentation in MRI Images with 3D U-Net and Contextual Transformer
by: Nguyen, Thien-Qua T., et al.
Published: (2024)
by: Nguyen, Thien-Qua T., et al.
Published: (2024)
LIBMoE: A Library for comprehensive benchmarking Mixture of Experts in Large Language Models
by: Nguyen, Nam V., et al.
Published: (2024)
by: Nguyen, Nam V., et al.
Published: (2024)
On Effects of Steering Latent Representation for Large Language Model Unlearning
by: Huu-Tien, Dang, et al.
Published: (2024)
by: Huu-Tien, Dang, et al.
Published: (2024)
TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching
by: Nguyen, Truong, et al.
Published: (2026)
by: Nguyen, Truong, et al.
Published: (2026)
Evaluating Large Language Model Capability in Vietnamese Fact-Checking Data Generation
by: To, Long Truong, et al.
Published: (2024)
by: To, Long Truong, et al.
Published: (2024)
NOWJ@COLIEE 2025: A Multi-stage Framework Integrating Embedding Models and Large Language Models for Legal Retrieval and Entailment
by: Nguyen, Hoang-Trung, et al.
Published: (2025)
by: Nguyen, Hoang-Trung, et al.
Published: (2025)
ViLLM-Eval: A Comprehensive Evaluation Suite for Vietnamese Large Language Models
by: Nguyen, Trong-Hieu, et al.
Published: (2024)
by: Nguyen, Trong-Hieu, et al.
Published: (2024)
CURATRON: Complete and Robust Preference Data for Rigorous Alignment of Large Language Models
by: Nguyen, Son The, et al.
Published: (2024)
by: Nguyen, Son The, et al.
Published: (2024)
BERT-based model for Vietnamese Fact Verification Dataset
by: Tran, Bao, et al.
Published: (2025)
by: Tran, Bao, et al.
Published: (2025)
Stepwise Alignment for Constrained Language Model Policy Optimization
by: Wachi, Akifumi, et al.
Published: (2024)
by: Wachi, Akifumi, et al.
Published: (2024)
Rethinking Toxicity Evaluation in Large Language Models: A Multi-Label Perspective
by: Kou, Zhiqiang, et al.
Published: (2025)
by: Kou, Zhiqiang, et al.
Published: (2025)
Misinformation Detection using Large Language Models with Explainability
by: Patel, Jainee, et al.
Published: (2025)
by: Patel, Jainee, et al.
Published: (2025)
Vulnerability Mitigation for Safety-Aligned Language Models via Debiasing
by: Tran, Thien Q., et al.
Published: (2025)
by: Tran, Thien Q., et al.
Published: (2025)
MDToC: Metacognitive Dynamic Tree of Concepts for Boosting Mathematical Problem-Solving of Large Language Models
by: Ta, Tung Duong, et al.
Published: (2025)
by: Ta, Tung Duong, et al.
Published: (2025)
Detection of Illicit Content on Online Marketplaces using Large Language Models
by: Tran, Quoc Khoa, et al.
Published: (2026)
by: Tran, Quoc Khoa, et al.
Published: (2026)
Characterising Toxicity in Generative Large Language Models
by: Zhang, Zhiyao, et al.
Published: (2026)
by: Zhang, Zhiyao, et al.
Published: (2026)
ViRanker: A BGE-M3 & Blockwise Parallel Transformer Cross-Encoder for Vietnamese Reranking
by: Dang, Phuong-Nam, et al.
Published: (2025)
by: Dang, Phuong-Nam, et al.
Published: (2025)
LUSIFER: Language Universal Space Integration for Enhanced Multilingual Embeddings with Large Language Models
by: Man, Hieu, et al.
Published: (2025)
by: Man, Hieu, et al.
Published: (2025)
Similar Items
-
ToVo: Toxicity Taxonomy via Voting
by: Luong, Tinh Son, et al.
Published: (2024) -
Zero-shot Cross-lingual Transfer Learning with Multiple Source and Target Languages for Information Extraction: Language Selection and Adversarial Training
by: Ngo, Nghia Trung, et al.
Published: (2024) -
Preserving Generalization of Language models in Few-shot Continual Relation Extraction
by: Tran, Quyen, et al.
Published: (2024) -
Comprehensive and Practical Evaluation of Retrieval-Augmented Generation Systems for Medical Question Answering
by: Ngo, Nghia Trung, et al.
Published: (2024) -
mSCoRe: a $M$ultilingual and Scalable Benchmark for $S$kill-based $Co$mmonsense $Re$asoning
by: Ngo, Nghia Trung, et al.
Published: (2025)