Transparent Neighborhood Approximation for Text Classifier Explanation
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Yi, Zimek, Arthur, Ntoutsi, Eirini, Wunder, Gerhard |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mitigating Semantic Drift: Evaluating LLMs' Efficacy in Psychotherapy through MI Dialogue Summarization
by: Kumar, Vivek, et al.
Published: (2025)
by: Kumar, Vivek, et al.
Published: (2025)
Exploring Fusion Techniques in Multimodal AI-Based Recruitment: Insights from FairCVdb
by: Swati, Swati, et al.
Published: (2024)
by: Swati, Swati, et al.
Published: (2024)
Achieving Hilbert-Schmidt Independence Under Rényi Differential Privacy for Fair and Private Data Generation
by: Hyrup, Tobias, et al.
Published: (2025)
by: Hyrup, Tobias, et al.
Published: (2025)
TABCF: Counterfactual Explanations for Tabular Data Using a Transformer-Based VAE
by: Panagiotou, Emmanouil, et al.
Published: (2024)
by: Panagiotou, Emmanouil, et al.
Published: (2024)
On Gradient-like Explanation under a Black-box Setting: When Black-box Explanations Become as Good as White-box
by: Cai, Yi, et al.
Published: (2023)
by: Cai, Yi, et al.
Published: (2023)
Exploring the Trade-off Between Model Performance and Explanation Plausibility of Text Classifiers Using Human Rationales
by: Resck, Lucas E., et al.
Published: (2024)
by: Resck, Lucas E., et al.
Published: (2024)
MMM-fair: An Interactive Toolkit for Exploring and Operationalizing Multi-Fairness Trade-offs
by: Swati, Swati, et al.
Published: (2025)
by: Swati, Swati, et al.
Published: (2025)
FairBranch: Mitigating Bias Transfer in Fair Multi-task Learning
by: Roy, Arjun, et al.
Published: (2023)
by: Roy, Arjun, et al.
Published: (2023)
Rethinking Explanation Evaluation under the Retraining Scheme
by: Cai, Yi, et al.
Published: (2025)
by: Cai, Yi, et al.
Published: (2025)
The Impact of Quantization on the Robustness of Transformer-based Text Classifiers
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
Fairness Evaluation with Item Response Theory
by: Xu, Ziqi, et al.
Published: (2024)
by: Xu, Ziqi, et al.
Published: (2024)
DISCERN: Decoding Systematic Errors in Natural Language for Text Classifiers
by: Menon, Rakesh R., et al.
Published: (2024)
by: Menon, Rakesh R., et al.
Published: (2024)
Explaining Text Classifiers with Counterfactual Representations
by: Lemberger, Pirmin, et al.
Published: (2024)
by: Lemberger, Pirmin, et al.
Published: (2024)
An Evaluation of Explanation Methods for Black-Box Detectors of Machine-Generated Text
by: Schoenegger, Loris, et al.
Published: (2024)
by: Schoenegger, Loris, et al.
Published: (2024)
ylmmcl at Multilingual Text Detoxification 2025: Lexicon-Guided Detoxification and Classifier-Gated Rewriting
by: Lai-Lopez, Nicole, et al.
Published: (2025)
by: Lai-Lopez, Nicole, et al.
Published: (2025)
Synthetic Tabular Data Generation for Class Imbalance and Fairness: A Comparative Study
by: Panagiotou, Emmanouil, et al.
Published: (2024)
by: Panagiotou, Emmanouil, et al.
Published: (2024)
TT-XAI: Trustworthy Clinical Text Explanations via Keyword Distillation and LLM Reasoning
by: Miok, Kristian, et al.
Published: (2025)
by: Miok, Kristian, et al.
Published: (2025)
FSEVAL: Feature Selection Evaluation Toolbox and Dashboard
by: Rajabinasab, Muhammad, et al.
Published: (2026)
by: Rajabinasab, Muhammad, et al.
Published: (2026)
Explanations as Bias Detectors: A Critical Study of Local Post-hoc XAI Methods for Fairness Exploration
by: Papanikou, Vasiliki, et al.
Published: (2025)
by: Papanikou, Vasiliki, et al.
Published: (2025)
Topic Modelling: Going Beyond Token Outputs
by: Williams, Lowri, et al.
Published: (2024)
by: Williams, Lowri, et al.
Published: (2024)
Enhanced Labeling Technique for Reddit Text and Fine-Tuned Longformer Models for Classifying Depression Severity in English and Luganda
by: Kimera, Richard, et al.
Published: (2024)
by: Kimera, Richard, et al.
Published: (2024)
GRILL: Restoring Gradient Signal in Ill-Conditioned Layers for More Effective Adversarial Attacks on Autoencoders
by: Ramanaik, Chethan Krishnamurthy, et al.
Published: (2025)
by: Ramanaik, Chethan Krishnamurthy, et al.
Published: (2025)
Debiasing Text Safety Classifiers through a Fairness-Aware Ensemble
by: Sturman, Olivia, et al.
Published: (2024)
by: Sturman, Olivia, et al.
Published: (2024)
Autonomous Data Selection with Zero-shot Generative Classifiers for Mathematical Texts
by: Zhang, Yifan, et al.
Published: (2024)
by: Zhang, Yifan, et al.
Published: (2024)
Large Language Models in the Task of Automatic Validation of Text Classifier Predictions
by: Tsymbalov, Aleksandr, et al.
Published: (2025)
by: Tsymbalov, Aleksandr, et al.
Published: (2025)
Towards LLM-guided Causal Explainability for Black-box Text Classifiers
by: Bhattacharjee, Amrita, et al.
Published: (2023)
by: Bhattacharjee, Amrita, et al.
Published: (2023)
Isotropy, Clusters, and Classifiers
by: Mickus, Timothee, et al.
Published: (2024)
by: Mickus, Timothee, et al.
Published: (2024)
Classifying and Addressing the Diversity of Errors in Retrieval-Augmented Generation Systems
by: Leung, Kin Kwan, et al.
Published: (2025)
by: Leung, Kin Kwan, et al.
Published: (2025)
Learning without training: The implicit dynamics of in-context learning
by: Dherin, Benoit, et al.
Published: (2025)
by: Dherin, Benoit, et al.
Published: (2025)
GONE: Structural Knowledge Unlearning via Neighborhood-Expanded Distribution Shaping
by: Dahal, Chahana, et al.
Published: (2026)
by: Dahal, Chahana, et al.
Published: (2026)
Adversarial Robustness of VAEs across Intersectional Subgroups
by: Ramanaik, Chethan Krishnamurthy, et al.
Published: (2024)
by: Ramanaik, Chethan Krishnamurthy, et al.
Published: (2024)
Contrast-CAT: Contrasting Activations for Enhanced Interpretability in Transformer-based Text Classifiers
by: Han, Sungmin, et al.
Published: (2025)
by: Han, Sungmin, et al.
Published: (2025)
Rule2Text: Natural Language Explanation of Logical Rules in Knowledge Graphs
by: Shirvani-Mahdavi, Nasim, et al.
Published: (2025)
by: Shirvani-Mahdavi, Nasim, et al.
Published: (2025)
Selective Explanations
by: Paes, Lucas Monteiro, et al.
Published: (2024)
by: Paes, Lucas Monteiro, et al.
Published: (2024)
Individual Text Corpora Predict Openness, Interests, Knowledge and Level of Education
by: Hofmann, Markus J., et al.
Published: (2024)
by: Hofmann, Markus J., et al.
Published: (2024)
StreetMath: Study of LLMs' Approximation Behaviors
by: Tseng, Chiung-Yi, et al.
Published: (2025)
by: Tseng, Chiung-Yi, et al.
Published: (2025)
Diffusion Explainer: Visual Explanation for Text-to-image Stable Diffusion
by: Lee, Seongmin, et al.
Published: (2023)
by: Lee, Seongmin, et al.
Published: (2023)
GECOBench: A Gender-Controlled Text Dataset and Benchmark for Quantifying Biases in Explanations
by: Wilming, Rick, et al.
Published: (2024)
by: Wilming, Rick, et al.
Published: (2024)
The Hidden Cost of Modeling P(X): Vulnerability to Membership Inference Attacks in Generative Text Classifiers
by: Makroo, Owais, et al.
Published: (2025)
by: Makroo, Owais, et al.
Published: (2025)
A Multi-Task Text Classification Pipeline with Natural Language Explanations: A User-Centric Evaluation in Sentiment Analysis and Offensive Language Identification in Greek Tweets
by: Mylonas, Nikolaos, et al.
Published: (2024)
by: Mylonas, Nikolaos, et al.
Published: (2024)
Similar Items
-
Mitigating Semantic Drift: Evaluating LLMs' Efficacy in Psychotherapy through MI Dialogue Summarization
by: Kumar, Vivek, et al.
Published: (2025) -
Exploring Fusion Techniques in Multimodal AI-Based Recruitment: Insights from FairCVdb
by: Swati, Swati, et al.
Published: (2024) -
Achieving Hilbert-Schmidt Independence Under Rényi Differential Privacy for Fair and Private Data Generation
by: Hyrup, Tobias, et al.
Published: (2025) -
TABCF: Counterfactual Explanations for Tabular Data Using a Transformer-Based VAE
by: Panagiotou, Emmanouil, et al.
Published: (2024) -
On Gradient-like Explanation under a Black-box Setting: When Black-box Explanations Become as Good as White-box
by: Cai, Yi, et al.
Published: (2023)