Tracing Stereotypes in Pre-trained Transformers: From Biased Neurons to Fairer Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Voria, Gianmario, Openja, Moses, Khomh, Foutse, Catolino, Gemma, Palomba, Fabio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Data Preparation for Fairness-Performance Trade-Offs: A Practitioner-Friendly Alternative?
von: Voria, Gianmario, et al.
Veröffentlicht: (2024)
von: Voria, Gianmario, et al.
Veröffentlicht: (2024)
SCOPE: A Dataset of Stereotyped Prompts for Counterfactual Fairness Assessment of LLMs
von: Parziale, Alessandra, et al.
Veröffentlicht: (2026)
von: Parziale, Alessandra, et al.
Veröffentlicht: (2026)
FairFLRep: Fairness aware fault localization and repair of Deep Neural Networks
von: Openja, Moses, et al.
Veröffentlicht: (2025)
von: Openja, Moses, et al.
Veröffentlicht: (2025)
RECOVER: Toward Requirements Generation from Stakeholders' Conversations
von: Voria, Gianmario, et al.
Veröffentlicht: (2024)
von: Voria, Gianmario, et al.
Veröffentlicht: (2024)
A Catalog of Fairness-Aware Practices in Machine Learning Engineering
von: Voria, Gianmario, et al.
Veröffentlicht: (2024)
von: Voria, Gianmario, et al.
Veröffentlicht: (2024)
From Expectation to Habit: Why Do Software Practitioners Adopt Fairness Toolkits?
von: Voria, Gianmario, et al.
Veröffentlicht: (2024)
von: Voria, Gianmario, et al.
Veröffentlicht: (2024)
Contextual Fairness-Aware Practices in ML: A Cost-Effective Empirical Evaluation
von: Parziale, Alessandra, et al.
Veröffentlicht: (2025)
von: Parziale, Alessandra, et al.
Veröffentlicht: (2025)
Toward Systematic Counterfactual Fairness Evaluation of Large Language Models: The CAFFE Framework
von: Parziale, Alessandra, et al.
Veröffentlicht: (2025)
von: Parziale, Alessandra, et al.
Veröffentlicht: (2025)
Bias Ahead: Sensitive Prompts as Early Warnings for Fairness in Large Language Models
von: Voria, Gianmario, et al.
Veröffentlicht: (2026)
von: Voria, Gianmario, et al.
Veröffentlicht: (2026)
An empirical study of testing machine learning in the wild
von: Openja, Moses, et al.
Veröffentlicht: (2023)
von: Openja, Moses, et al.
Veröffentlicht: (2023)
Once Upon a Team: Investigating Bias in LLM-Driven Software Team Composition and Task Allocation
von: Parziale, Alessandra, et al.
Veröffentlicht: (2026)
von: Parziale, Alessandra, et al.
Veröffentlicht: (2026)
Trained Without My Consent: Detecting Code Inclusion In Language Models Trained on Code
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
On the Effectiveness of Log Representation for Log-based Anomaly Detection
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
GIST: Generated Inputs Sets Transferability in Deep Learning
von: Tambon, Florian, et al.
Veröffentlicht: (2023)
von: Tambon, Florian, et al.
Veröffentlicht: (2023)
What Information Contributes to Log-based Anomaly Detection? Insights from a Configurable Transformer-Based Approach
von: Wu, Xingfang, et al.
Veröffentlicht: (2024)
von: Wu, Xingfang, et al.
Veröffentlicht: (2024)
An Efficient Model Maintenance Approach for MLOps
von: Majidi, Forough, et al.
Veröffentlicht: (2024)
von: Majidi, Forough, et al.
Veröffentlicht: (2024)
Fault Localization in Deep Learning-based Software: A System-level Approach
von: Morovati, Mohammad Mehdi, et al.
Veröffentlicht: (2024)
von: Morovati, Mohammad Mehdi, et al.
Veröffentlicht: (2024)
Machine Learning Robustness: A Primer
von: Braiek, Houssem Ben, et al.
Veröffentlicht: (2024)
von: Braiek, Houssem Ben, et al.
Veröffentlicht: (2024)
Motivations, Challenges, Best Practices, and Benefits for Bots and Conversational Agents in Software Engineering: A Multivocal Literature Review
von: Lambiase, Stefano, et al.
Veröffentlicht: (2024)
von: Lambiase, Stefano, et al.
Veröffentlicht: (2024)
Mining Action Rules for Defect Reduction Planning
von: Oueslati, Khouloud, et al.
Veröffentlicht: (2024)
von: Oueslati, Khouloud, et al.
Veröffentlicht: (2024)
Towards Enhancing the Reproducibility of Deep Learning Bugs: An Empirical Study
von: Shah, Mehil B., et al.
Veröffentlicht: (2024)
von: Shah, Mehil B., et al.
Veröffentlicht: (2024)
Prism: Dynamic and Flexible Benchmarking of LLMs Code Generation with Monte Carlo Tree Search
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2025)
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2025)
How Do Communities of ML-Enabled Systems Smell? A Cross-Sectional Study on the Prevalence of Community Smells
von: Annunziata, Giusy, et al.
Veröffentlicht: (2025)
von: Annunziata, Giusy, et al.
Veröffentlicht: (2025)
Exploring Individual Factors in the Adoption of LLMs for Specific Software Engineering Purposes
von: Lambiase, Stefano, et al.
Veröffentlicht: (2025)
von: Lambiase, Stefano, et al.
Veröffentlicht: (2025)
Imitation Game: Reproducing Deep Learning Bugs Leveraging an Intelligent Agent
von: Shah, Mehil B, et al.
Veröffentlicht: (2025)
von: Shah, Mehil B, et al.
Veröffentlicht: (2025)
Investigating the Role of Cultural Values in Adopting Large Language Models for Software Engineering
von: Lambiase, Stefano, et al.
Veröffentlicht: (2024)
von: Lambiase, Stefano, et al.
Veröffentlicht: (2024)
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
von: Taraghi, Mina, et al.
Veröffentlicht: (2024)
von: Taraghi, Mina, et al.
Veröffentlicht: (2024)
Refining GPT-3 Embeddings with a Siamese Structure for Technical Post Duplicate Detection
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
Tracing Optimization for Performance Modeling and Regression Detection
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2024)
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2024)
Software Engineering Principles for Fairer Systems: Experiments with GroupCART
von: Peng, Kewen, et al.
Veröffentlicht: (2025)
von: Peng, Kewen, et al.
Veröffentlicht: (2025)
DeepCodeProbe: Towards Understanding What Models Trained on Code Learn
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
Characterizing and Classifying Developer Forum Posts with their Intentions
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
From Technical Excellence to Practical Adoption: Lessons Learned Building an ML-Enhanced Trace Analysis Tool
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2025)
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2025)
Quality Issues in Machine Learning Software Systems
von: Côté, Pierre-Olivier, et al.
Veröffentlicht: (2023)
von: Côté, Pierre-Olivier, et al.
Veröffentlicht: (2023)
Trimming the Risk: Towards Reliable Continuous Training for Deep Learning Inspection Systems
von: Abbassi, Altaf Allah, et al.
Veröffentlicht: (2024)
von: Abbassi, Altaf Allah, et al.
Veröffentlicht: (2024)
Toward Debugging Deep Reinforcement Learning Programs with RLExplorer
von: Bouchoucha, Rached, et al.
Veröffentlicht: (2024)
von: Bouchoucha, Rached, et al.
Veröffentlicht: (2024)
LEANCODE: Understanding Models Better for Code Simplification of Pre-trained Large Language Models
von: Wang, Yan, et al.
Veröffentlicht: (2025)
von: Wang, Yan, et al.
Veröffentlicht: (2025)
Real Faults in Model Context Protocol (MCP) Software: a Comprehensive Taxonomy
von: Taraghi, Mina, et al.
Veröffentlicht: (2026)
von: Taraghi, Mina, et al.
Veröffentlicht: (2026)
Towards Assessing Deep Learning Test Input Generators
von: Mzoughi, Seif, et al.
Veröffentlicht: (2025)
von: Mzoughi, Seif, et al.
Veröffentlicht: (2025)
Protecting Privacy in Software Logs: What Should Be Anonymized?
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2024)
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Data Preparation for Fairness-Performance Trade-Offs: A Practitioner-Friendly Alternative?
von: Voria, Gianmario, et al.
Veröffentlicht: (2024) -
SCOPE: A Dataset of Stereotyped Prompts for Counterfactual Fairness Assessment of LLMs
von: Parziale, Alessandra, et al.
Veröffentlicht: (2026) -
FairFLRep: Fairness aware fault localization and repair of Deep Neural Networks
von: Openja, Moses, et al.
Veröffentlicht: (2025) -
RECOVER: Toward Requirements Generation from Stakeholders' Conversations
von: Voria, Gianmario, et al.
Veröffentlicht: (2024) -
A Catalog of Fairness-Aware Practices in Machine Learning Engineering
von: Voria, Gianmario, et al.
Veröffentlicht: (2024)