Controlled Model Debiasing through Minimal and Interpretable Updates
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Di Gennaro, Federico, Laugel, Thibault, Grari, Vincent, Detyniecki, Marcin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Post-processing fairness with minimal changes
von: Di Gennaro, Federico, et al.
Veröffentlicht: (2024)
von: Di Gennaro, Federico, et al.
Veröffentlicht: (2024)
SAKE: Steering Activations for Knowledge Editing
von: Scialanga, Marco, et al.
Veröffentlicht: (2025)
von: Scialanga, Marco, et al.
Veröffentlicht: (2025)
When mitigating bias is unfair: multiplicity and arbitrariness in algorithmic group fairness
von: Krco, Natasa, et al.
Veröffentlicht: (2023)
von: Krco, Natasa, et al.
Veröffentlicht: (2023)
ACT: Agentic Classification Tree
von: Grari, Vincent, et al.
Veröffentlicht: (2025)
von: Grari, Vincent, et al.
Veröffentlicht: (2025)
OptiGrad: A Fair and more Efficient Price Elasticity Optimization via a Gradient Based Learning
von: Grari, Vincent, et al.
Veröffentlicht: (2024)
von: Grari, Vincent, et al.
Veröffentlicht: (2024)
Understanding Prediction Discrepancies in Machine Learning Classifiers
von: Renard, Xavier, et al.
Veröffentlicht: (2021)
von: Renard, Xavier, et al.
Veröffentlicht: (2021)
Agentic Adversarial QA for Improving Domain-Specific LLMs
von: Grari, Vincent, et al.
Veröffentlicht: (2026)
von: Grari, Vincent, et al.
Veröffentlicht: (2026)
Why do explanations fail? A typology and discussion on failures in XAI
von: Bove, Clara, et al.
Veröffentlicht: (2024)
von: Bove, Clara, et al.
Veröffentlicht: (2024)
Energy Scaling Laws for Diffusion Models: Quantifying Compute in Image Generation
von: Iyengar, Aniketh, et al.
Veröffentlicht: (2025)
von: Iyengar, Aniketh, et al.
Veröffentlicht: (2025)
From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes
von: Garzón, Rubén, et al.
Veröffentlicht: (2026)
von: Garzón, Rubén, et al.
Veröffentlicht: (2026)
Regret-Optimized Portfolio Enhancement through Deep Reinforcement Learning and Future Looking Rewards
von: Karzanov, Daniil, et al.
Veröffentlicht: (2025)
von: Karzanov, Daniil, et al.
Veröffentlicht: (2025)
Metric assessment protocol in the context of answer fluctuation on MCQ tasks
von: Goliakova, Ekaterina, et al.
Veröffentlicht: (2025)
von: Goliakova, Ekaterina, et al.
Veröffentlicht: (2025)
Alignment Reduces Expressed but Not Encoded Gender Bias: A Unified Framework and Study
von: Bouchouchi, Nour, et al.
Veröffentlicht: (2026)
von: Bouchouchi, Nour, et al.
Veröffentlicht: (2026)
LAYA: Layer-wise Attention Aggregation for Interpretable Depth-Aware Neural Networks
von: Vessio, Gennaro
Veröffentlicht: (2025)
von: Vessio, Gennaro
Veröffentlicht: (2025)
Efficient Logistic Regression with Mixture of Sigmoids
von: Di Gennaro, Federico, et al.
Veröffentlicht: (2026)
von: Di Gennaro, Federico, et al.
Veröffentlicht: (2026)
Instance-Dependent Regret Bounds for Nonstochastic Linear Partial Monitoring
von: Di Gennaro, Federico, et al.
Veröffentlicht: (2025)
von: Di Gennaro, Federico, et al.
Veröffentlicht: (2025)
Debiasing Algorithm through Model Adaptation
von: Limisiewicz, Tomasz, et al.
Veröffentlicht: (2023)
von: Limisiewicz, Tomasz, et al.
Veröffentlicht: (2023)
Beyond the Black Box: Identifiable Interpretation and Control in Generative Models via Causal Minimality
von: Kong, Lingjing, et al.
Veröffentlicht: (2025)
von: Kong, Lingjing, et al.
Veröffentlicht: (2025)
Debiased Model-based Representations for Sample-efficient Continuous Control
von: Lyu, Jiafei, et al.
Veröffentlicht: (2026)
von: Lyu, Jiafei, et al.
Veröffentlicht: (2026)
Direct Debiased Machine Learning via Bregman Divergence Minimization
von: Kato, Masahiro
Veröffentlicht: (2025)
von: Kato, Masahiro
Veröffentlicht: (2025)
Continual Learning through Control Minimization
von: de Haan, Sander, et al.
Veröffentlicht: (2026)
von: de Haan, Sander, et al.
Veröffentlicht: (2026)
CausalTAD: Causal Implicit Generative Model for Debiased Online Trajectory Anomaly Detection
von: Li, Wenbin, et al.
Veröffentlicht: (2024)
von: Li, Wenbin, et al.
Veröffentlicht: (2024)
Debiasing Synthetic Data Generated by Deep Generative Models
von: Decruyenaere, Alexander, et al.
Veröffentlicht: (2024)
von: Decruyenaere, Alexander, et al.
Veröffentlicht: (2024)
Debiasing Kernel-Based Generative Models
von: Qin, Tian, et al.
Veröffentlicht: (2025)
von: Qin, Tian, et al.
Veröffentlicht: (2025)
Debiased Model-based Interactive Recommendation
von: Li, Zijian, et al.
Veröffentlicht: (2024)
von: Li, Zijian, et al.
Veröffentlicht: (2024)
DoubleGen: Debiased Generative Modeling of Counterfactuals
von: Luedtke, Alex, et al.
Veröffentlicht: (2025)
von: Luedtke, Alex, et al.
Veröffentlicht: (2025)
Anchor Pair Selection in TDOA Positioning Systems by Door Transition Error Minimization
von: Kolakowski, Marcin, et al.
Veröffentlicht: (2024)
von: Kolakowski, Marcin, et al.
Veröffentlicht: (2024)
Looking at Model Debiasing through the Lens of Anomaly Detection
von: Pastore, Vito Paolo, et al.
Veröffentlicht: (2024)
von: Pastore, Vito Paolo, et al.
Veröffentlicht: (2024)
Fast Debiasing of the LASSO Estimator
von: Banerjee, Shuvayan, et al.
Veröffentlicht: (2025)
von: Banerjee, Shuvayan, et al.
Veröffentlicht: (2025)
Efficient Sparse Selective-Update RNNs for Long-Range Sequence Modeling
von: Yin, Bojian, et al.
Veröffentlicht: (2026)
von: Yin, Bojian, et al.
Veröffentlicht: (2026)
The "Huh?" Button: Improving Understanding in Educational Videos with Large Language Models
von: Ruf, Boris, et al.
Veröffentlicht: (2024)
von: Ruf, Boris, et al.
Veröffentlicht: (2024)
Debiasing Reward Models by Representation Learning with Guarantees
von: Ng, Ignavier, et al.
Veröffentlicht: (2025)
von: Ng, Ignavier, et al.
Veröffentlicht: (2025)
Probing Information Distribution in Transformer Architectures through Entropy Analysis
von: Buonanno, Amedeo, et al.
Veröffentlicht: (2025)
von: Buonanno, Amedeo, et al.
Veröffentlicht: (2025)
Adversarial Debiasing for Unbiased Parameter Recovery
von: Sanford, Luke C, et al.
Veröffentlicht: (2025)
von: Sanford, Luke C, et al.
Veröffentlicht: (2025)
Debiased neural operators for estimating functionals
von: Hess, Konstantin, et al.
Veröffentlicht: (2026)
von: Hess, Konstantin, et al.
Veröffentlicht: (2026)
Debiased Distribution Compression
von: Li, Lingxiao, et al.
Veröffentlicht: (2024)
von: Li, Lingxiao, et al.
Veröffentlicht: (2024)
SEM: Sparse Embedding Modulation for Post-Hoc Debiasing of Vision-Language Models
von: Guimard, Quentin, et al.
Veröffentlicht: (2026)
von: Guimard, Quentin, et al.
Veröffentlicht: (2026)
Model Debiasing by Learnable Data Augmentation
von: Morerio, Pietro, et al.
Veröffentlicht: (2024)
von: Morerio, Pietro, et al.
Veröffentlicht: (2024)
Value Iteration for Learning Concurrently Executable Robotic Control Tasks
von: Tahmid, Sheikh A., et al.
Veröffentlicht: (2025)
von: Tahmid, Sheikh A., et al.
Veröffentlicht: (2025)
Inference Time Debiasing Concepts in Diffusion Models
von: Kupssinskü, Lucas S., et al.
Veröffentlicht: (2025)
von: Kupssinskü, Lucas S., et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Post-processing fairness with minimal changes
von: Di Gennaro, Federico, et al.
Veröffentlicht: (2024) -
SAKE: Steering Activations for Knowledge Editing
von: Scialanga, Marco, et al.
Veröffentlicht: (2025) -
When mitigating bias is unfair: multiplicity and arbitrariness in algorithmic group fairness
von: Krco, Natasa, et al.
Veröffentlicht: (2023) -
ACT: Agentic Classification Tree
von: Grari, Vincent, et al.
Veröffentlicht: (2025) -
OptiGrad: A Fair and more Efficient Price Elasticity Optimization via a Gradient Based Learning
von: Grari, Vincent, et al.
Veröffentlicht: (2024)