Expert-Guided Extinction of Toxic Tokens for Debiased Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Xueyao, Shi, Kaize, Tang, Haoran, Xu, Guandong, Li, Qing |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Compressing Long Context for Enhancing RAG with AMR-based Concept Distillation
by: Shi, Kaize, et al.
Published: (2024)
by: Shi, Kaize, et al.
Published: (2024)
CoCR-RAG: Enhancing Retrieval-Augmented Generation in Web Q&A via Concept-oriented Context Reconstruction
by: Shi, Kaize, et al.
Published: (2026)
by: Shi, Kaize, et al.
Published: (2026)
LLaMA-E: Empowering E-commerce Authoring with Object-Interleaved Instruction Following
by: Shi, Kaize, et al.
Published: (2023)
by: Shi, Kaize, et al.
Published: (2023)
Concept than Document: Context Compression via AMR-based Conceptual Entropy
by: Shi, Kaize, et al.
Published: (2025)
by: Shi, Kaize, et al.
Published: (2025)
APLe: Token-Wise Adaptive for Multi-Modal Prompt Learning
by: Cao, Guiming, et al.
Published: (2024)
by: Cao, Guiming, et al.
Published: (2024)
RVISA: Reasoning and Verification for Implicit Sentiment Analysis
by: Lai, Wenna, et al.
Published: (2024)
by: Lai, Wenna, et al.
Published: (2024)
Multi-Task Learning with LLMs for Implicit Sentiment Analysis: Data-level and Task-level Automatic Weight Learning
by: Lai, Wenna, et al.
Published: (2024)
by: Lai, Wenna, et al.
Published: (2024)
STAR: Stepwise Task Augmentation with Relation Learning for Aspect Sentiment Quad Prediction
by: Lai, Wenna, et al.
Published: (2025)
by: Lai, Wenna, et al.
Published: (2025)
General Phrase Debiaser: Debiasing Masked Language Models at a Multi-Token Level
by: Shi, Bingkang, et al.
Published: (2023)
by: Shi, Bingkang, et al.
Published: (2023)
When LLMs Team Up: The Emergence of Collaborative Affective Computing
by: Lai, Wenna, et al.
Published: (2025)
by: Lai, Wenna, et al.
Published: (2025)
Listwise Preference Optimization with Element-wise Confusions for Aspect Sentiment Quad Prediction
by: Lai, Wenna, et al.
Published: (2025)
by: Lai, Wenna, et al.
Published: (2025)
Fine-grained Verification via Diagnostic Reasoning Supervision for Aspect Sentiment Triplet Extraction
by: Lai, Wenna, et al.
Published: (2026)
by: Lai, Wenna, et al.
Published: (2026)
ExpertFlow: Efficient Mixture-of-Experts Inference via Predictive Expert Caching and Token Scheduling
by: He, Xin, et al.
Published: (2024)
by: He, Xin, et al.
Published: (2024)
Information Gain-Guided Causal Intervention for Autonomous Debiasing Large Language Models
by: Sun, Zhouhao, et al.
Published: (2025)
by: Sun, Zhouhao, et al.
Published: (2025)
Causal-Guided Active Learning for Debiasing Large Language Models
by: Du, Li, et al.
Published: (2024)
by: Du, Li, et al.
Published: (2024)
An Expert is Worth One Token: Synergizing Multiple Expert LLMs as Generalist via Expert Token Routing
by: Chai, Ziwei, et al.
Published: (2024)
by: Chai, Ziwei, et al.
Published: (2024)
On Bias and Fairness in NLP: Investigating the Impact of Bias and Debiasing in Language Models on the Fairness of Toxicity Detection
by: Elsafoury, Fatma, et al.
Published: (2023)
by: Elsafoury, Fatma, et al.
Published: (2023)
Teacher-Student Training for Debiasing: General Permutation Debiasing for Large Language Models
by: Liusie, Adian, et al.
Published: (2024)
by: Liusie, Adian, et al.
Published: (2024)
Reasoning Factual Knowledge in Structured Data with Large Language Models
by: Huang, Sirui, et al.
Published: (2024)
by: Huang, Sirui, et al.
Published: (2024)
HyperG: Hypergraph-Enhanced LLMs for Structured Knowledge
by: Huang, Sirui, et al.
Published: (2025)
by: Huang, Sirui, et al.
Published: (2025)
Take its Essence, Discard its Dross! Debiasing for Toxic Language Detection via Counterfactual Causal Effect
by: Lu, Junyu, et al.
Published: (2024)
by: Lu, Junyu, et al.
Published: (2024)
GNN-as-Judge: Unleashing the Power of LLMs for Graph Learning with GNN Feedback
by: Xu, Ruiyao, et al.
Published: (2026)
by: Xu, Ruiyao, et al.
Published: (2026)
HopRank: Self-Supervised LLM Preference-Tuning on Graphs for Few-Shot Node Classification
by: Wang, Ziqing, et al.
Published: (2026)
by: Wang, Ziqing, et al.
Published: (2026)
Mixture-of-Retrieval Experts for Reasoning-Guided Multimodal Knowledge Exploitation
by: Peng, Chunyi, et al.
Published: (2025)
by: Peng, Chunyi, et al.
Published: (2025)
Redefining Experts: Interpretable Decomposition of Language Models for Toxicity Mitigation
by: Shaik, Zuhair Hasan, et al.
Published: (2025)
by: Shaik, Zuhair Hasan, et al.
Published: (2025)
MaskMoE: Boosting Token-Level Learning via Routing Mask in Mixture-of-Experts
by: Su, Zhenpeng, et al.
Published: (2024)
by: Su, Zhenpeng, et al.
Published: (2024)
Avoiding Copyright Infringement via Large Language Model Unlearning
by: Dou, Guangyao, et al.
Published: (2024)
by: Dou, Guangyao, et al.
Published: (2024)
TokenRec: Learning to Tokenize ID for LLM-based Generative Recommendation
by: Qu, Haohao, et al.
Published: (2024)
by: Qu, Haohao, et al.
Published: (2024)
Enhancing Multilingual RAG Systems with Debiased Language Preference-Guided Query Fusion
by: Park, Jeonghyun, et al.
Published: (2026)
by: Park, Jeonghyun, et al.
Published: (2026)
Expert-Token Resonance MoE: Bidirectional Routing with Efficiency Affinity-Driven Active Selection
by: Li, Jing, et al.
Published: (2024)
by: Li, Jing, et al.
Published: (2024)
TEAM: Temporal-Spatial Consistency Guided Expert Activation for MoE Diffusion Language Model Acceleration
by: Wei, Linye, et al.
Published: (2026)
by: Wei, Linye, et al.
Published: (2026)
A Survey of Large Language Models for Text-Guided Molecular Discovery: from Molecule Generation to Optimization
by: Wang, Ziqing, et al.
Published: (2025)
by: Wang, Ziqing, et al.
Published: (2025)
FairJudge: An Adaptive, Debiased, and Consistent LLM-as-a-Judge
by: Yang, Bo, et al.
Published: (2026)
by: Yang, Bo, et al.
Published: (2026)
Whispering Experts: Neural Interventions for Toxicity Mitigation in Language Models
by: Suau, Xavier, et al.
Published: (2024)
by: Suau, Xavier, et al.
Published: (2024)
On the Salience of Low-Probability Tokens for AI-Generated Text Detection: A Multiscale Uncertainty Perspective
by: Guo, Yikai, et al.
Published: (2026)
by: Guo, Yikai, et al.
Published: (2026)
Token-Guard: Towards Token-Level Hallucination Control via Self-Checking Decoding
by: Zhu, Yifan, et al.
Published: (2026)
by: Zhu, Yifan, et al.
Published: (2026)
A Debiased Nearest Neighbors Framework for Multi-Label Text Classification
by: Cheng, Zifeng, et al.
Published: (2024)
by: Cheng, Zifeng, et al.
Published: (2024)
MambaFormer: Token-Level Guided Routing Mixture-of-Experts for Accurate and Efficient Clinical Assistance
by: Khan, Hamad, et al.
Published: (2026)
by: Khan, Hamad, et al.
Published: (2026)
Towards Universal Debiasing for Language Models-based Tabular Data Generation
by: Li, Tianchun, et al.
Published: (2025)
by: Li, Tianchun, et al.
Published: (2025)
Creative Convergence or Imitation? Genre-Specific Homogeneity in LLM-Generated Chinese Literature
by: Ma, Yuanchi, et al.
Published: (2026)
by: Ma, Yuanchi, et al.
Published: (2026)
Similar Items
-
Compressing Long Context for Enhancing RAG with AMR-based Concept Distillation
by: Shi, Kaize, et al.
Published: (2024) -
CoCR-RAG: Enhancing Retrieval-Augmented Generation in Web Q&A via Concept-oriented Context Reconstruction
by: Shi, Kaize, et al.
Published: (2026) -
LLaMA-E: Empowering E-commerce Authoring with Object-Interleaved Instruction Following
by: Shi, Kaize, et al.
Published: (2023) -
Concept than Document: Context Compression via AMR-based Conceptual Entropy
by: Shi, Kaize, et al.
Published: (2025) -
APLe: Token-Wise Adaptive for Multi-Modal Prompt Learning
by: Cao, Guiming, et al.
Published: (2024)