Investigating the Relationship Between Debiasing and Artifact Removal using Saliency Maps
Fuente:
arXiv
Saved in:
| Main Authors: | Sztukiewicz, Lukasz, Stępka, Ignacy, Wiliński, Michał, Stefanowski, Jerzy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DetoxAI: a Python Toolkit for Debiasing Deep Learning Models in Computer Vision
by: Stępka, Ignacy, et al.
Published: (2025)
by: Stępka, Ignacy, et al.
Published: (2025)
Explaining Concept Drift through the Evolution of Group Counterfactuals
by: Stępka, Ignacy, et al.
Published: (2025)
by: Stępka, Ignacy, et al.
Published: (2025)
A multi-criteria approach for selecting an explanation from the set of counterfactuals produced by an ensemble of explainers
by: Stępka, Ignacy, et al.
Published: (2024)
by: Stępka, Ignacy, et al.
Published: (2024)
Counterfactual Explanations with Probabilistic Guarantees on their Robustness to Model Change
by: Stępka, Ignacy, et al.
Published: (2024)
by: Stępka, Ignacy, et al.
Published: (2024)
Mitigating Persistent Client Dropout in Asynchronous Decentralized Federated Learning
by: Stępka, Ignacy, et al.
Published: (2025)
by: Stępka, Ignacy, et al.
Published: (2025)
The Problem of Coherence in Natural Language Explanations of Recommendations
by: Raczyński, Jakub, et al.
Published: (2023)
by: Raczyński, Jakub, et al.
Published: (2023)
Exploring Loss Design Techniques For Decision Tree Robustness To Label Noise
by: Sztukiewicz, Lukasz, et al.
Published: (2024)
by: Sztukiewicz, Lukasz, et al.
Published: (2024)
Unifying Perspectives: Plausible Counterfactual Explanations on Global, Group-wise, and Local Levels
by: Furman, Oleksii, et al.
Published: (2024)
by: Furman, Oleksii, et al.
Published: (2024)
A Probabilistic Consensus-Driven Approach for Robust Counterfactual Explanations
by: Kostrzewa, Marcin, et al.
Published: (2026)
by: Kostrzewa, Marcin, et al.
Published: (2026)
Towards Differentiating Between Failures and Domain Shifts in Industrial Data Streams
by: Wojak-Strzelecka, Natalia, et al.
Published: (2026)
by: Wojak-Strzelecka, Natalia, et al.
Published: (2026)
A Multi-LLM Debiasing Framework
by: Owens, Deonna M., et al.
Published: (2024)
by: Owens, Deonna M., et al.
Published: (2024)
On the Effectiveness and Generalization of Race Representations for Debiasing High-Stakes Decisions
by: Nguyen, Dang, et al.
Published: (2025)
by: Nguyen, Dang, et al.
Published: (2025)
LLM-Assisted Content Conditional Debiasing for Fair Text Embedding
by: Deng, Wenlong, et al.
Published: (2024)
by: Deng, Wenlong, et al.
Published: (2024)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes
by: Gallegos, Isabel O., et al.
Published: (2024)
by: Gallegos, Isabel O., et al.
Published: (2024)
AXOLOTL: Fairness through Assisted Self-Debiasing of Large Language Model Outputs
by: Ebrahimi, Sana, et al.
Published: (2024)
by: Ebrahimi, Sana, et al.
Published: (2024)
Probabilistically Plausible Counterfactual Explanations with Normalizing Flows
by: Wielopolski, Patryk, et al.
Published: (2024)
by: Wielopolski, Patryk, et al.
Published: (2024)
Transit for All: Mapping Equitable Bike2Subway Connection using Region Representation Learning
by: Namgung, Min, et al.
Published: (2025)
by: Namgung, Min, et al.
Published: (2025)
Inference-Time Rule Eraser: Fair Recognition via Distilling and Removing Biased Rules
by: Zhang, Yi, et al.
Published: (2024)
by: Zhang, Yi, et al.
Published: (2024)
Debiasing surgeon: fantastic weights and how to find them
by: Nahon, Rémi, et al.
Published: (2024)
by: Nahon, Rémi, et al.
Published: (2024)
A SAT-based approach to rigorous verification of Bayesian networks
by: Stępka, Ignacy, et al.
Published: (2024)
by: Stępka, Ignacy, et al.
Published: (2024)
MICA: Multivariate Infini Compressive Attention for Time Series Forecasting
by: Potosnak, Willa, et al.
Published: (2026)
by: Potosnak, Willa, et al.
Published: (2026)
Mapping the Potential of Explainable AI for Fairness Along the AI Lifecycle
by: Deck, Luca, et al.
Published: (2024)
by: Deck, Luca, et al.
Published: (2024)
Between Randomness and Arbitrariness: Some Lessons for Reliable Machine Learning at Scale
by: Cooper, A. Feder
Published: (2024)
by: Cooper, A. Feder
Published: (2024)
Investigating Trustworthiness of Nonparametric Deep Survival Models for Alzheimer's Disease Progression Analysis
by: Thrasher, Jacob, et al.
Published: (2026)
by: Thrasher, Jacob, et al.
Published: (2026)
Mapping the Media Landscape: Predicting Factual Reporting and Political Bias Through Web Interactions
by: Sánchez-Cortés, Dairazalia, et al.
Published: (2024)
by: Sánchez-Cortés, Dairazalia, et al.
Published: (2024)
Leveraging Machine Learning Techniques to Investigate Media and Information Literacy Competence in Tackling Disinformation
by: Alcalde-Llergo, José Manuel, et al.
Published: (2026)
by: Alcalde-Llergo, José Manuel, et al.
Published: (2026)
Machine Unlearning Fails to Remove Data Poisoning Attacks
by: Pawelczyk, Martin, et al.
Published: (2024)
by: Pawelczyk, Martin, et al.
Published: (2024)
Who Evaluates AI's Social Impacts? Mapping Coverage and Gaps in First and Third Party Evaluations
by: Reuel, Anka, et al.
Published: (2025)
by: Reuel, Anka, et al.
Published: (2025)
Investigating Thematic Patterns and User Preferences in LLM Interactions using BERTopic
by: Bhandarkar, Abhay, et al.
Published: (2025)
by: Bhandarkar, Abhay, et al.
Published: (2025)
Can You Trust an LLM with Your Life-Changing Decision? An Investigation into AI High-Stakes Responses
by: Cahyono, Joshua Adrian, et al.
Published: (2025)
by: Cahyono, Joshua Adrian, et al.
Published: (2025)
Risk Analysis in Customer Relationship Management via Quantile Region Convolutional Neural Network-Long Short-Term Memory and Cross-Attention Mechanism
by: Huang, Yaowen, et al.
Published: (2024)
by: Huang, Yaowen, et al.
Published: (2024)
Debiasing Methods for Fairer Neural Models in Vision and Language Research: A Survey
by: Parraga, Otávio, et al.
Published: (2022)
by: Parraga, Otávio, et al.
Published: (2022)
The Personality Illusion: Revealing Dissociation Between Self-Reports & Behavior in LLMs
by: Han, Pengrui, et al.
Published: (2025)
by: Han, Pengrui, et al.
Published: (2025)
Emissions and Performance Trade-off Between Small and Large Language Models
by: Garg, Anandita, et al.
Published: (2025)
by: Garg, Anandita, et al.
Published: (2025)
On Defining Smart Cities using Transformer Neural Networks
by: Khurshudov, Andrei
Published: (2024)
by: Khurshudov, Andrei
Published: (2024)
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods
by: Jang, Yeonwoo, et al.
Published: (2025)
by: Jang, Yeonwoo, et al.
Published: (2025)
Properties of fairness measures in the context of varying class imbalance and protected group ratios
by: Brzezinski, Dariusz, et al.
Published: (2024)
by: Brzezinski, Dariusz, et al.
Published: (2024)
Effective Controllable Bias Mitigation for Classification and Retrieval using Gate Adapters
by: Masoudian, Shahed, et al.
Published: (2024)
by: Masoudian, Shahed, et al.
Published: (2024)
Are clinicians ethically obligated to disclose their use of medical machine learning systems to patients?
by: Hatherley, Joshua
Published: (2025)
by: Hatherley, Joshua
Published: (2025)
Similar Items
-
DetoxAI: a Python Toolkit for Debiasing Deep Learning Models in Computer Vision
by: Stępka, Ignacy, et al.
Published: (2025) -
Explaining Concept Drift through the Evolution of Group Counterfactuals
by: Stępka, Ignacy, et al.
Published: (2025) -
A multi-criteria approach for selecting an explanation from the set of counterfactuals produced by an ensemble of explainers
by: Stępka, Ignacy, et al.
Published: (2024) -
Counterfactual Explanations with Probabilistic Guarantees on their Robustness to Model Change
by: Stępka, Ignacy, et al.
Published: (2024) -
Mitigating Persistent Client Dropout in Asynchronous Decentralized Federated Learning
by: Stępka, Ignacy, et al.
Published: (2025)