Fairness Mediator: Neutralize Stereotype Associations to Mitigate Bias in Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Xiao, Yisong, Liu, Aishan, Liang, Siyuan, Liu, Xianglong, Tao, Dacheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Latent Imitator: Generating Natural Individual Discriminatory Instances for Black-Box Fairness Testing
di: Xiao, Yisong, et al.
Pubblicazione: (2023)
di: Xiao, Yisong, et al.
Pubblicazione: (2023)
BDefects4NN: A Backdoor Defect Database for Controlled Localization Studies in Neural Networks
di: Xiao, Yisong, et al.
Pubblicazione: (2024)
di: Xiao, Yisong, et al.
Pubblicazione: (2024)
Software Development Life Cycle Perspective: A Survey of Benchmarks for Code Large Language Models and Agents
di: Wang, Kaixin, et al.
Pubblicazione: (2025)
di: Wang, Kaixin, et al.
Pubblicazione: (2025)
Detoxifying Large Language Models via Autoregressive Reward Guided Representation Editing
di: Xiao, Yisong, et al.
Pubblicazione: (2025)
di: Xiao, Yisong, et al.
Pubblicazione: (2025)
Bias Ahead: Sensitive Prompts as Early Warnings for Fairness in Large Language Models
di: Voria, Gianmario, et al.
Pubblicazione: (2026)
di: Voria, Gianmario, et al.
Pubblicazione: (2026)
From Context to Intent: Reasoning-Guided Function-Level Code Completion
di: Li, Yanzhou, et al.
Pubblicazione: (2025)
di: Li, Yanzhou, et al.
Pubblicazione: (2025)
Fairness Is Not Just Ethical: Performance Trade-Off via Data Correlation Tuning to Mitigate Bias in ML Software
di: Xiao, Ying, et al.
Pubblicazione: (2025)
di: Xiao, Ying, et al.
Pubblicazione: (2025)
CodeChemist: Functional Knowledge Transfer for Low-Resource Code Generation via Test-Time Scaling
di: Wang, Kaixin, et al.
Pubblicazione: (2025)
di: Wang, Kaixin, et al.
Pubblicazione: (2025)
SCOPE: A Dataset of Stereotyped Prompts for Counterfactual Fairness Assessment of LLMs
di: Parziale, Alessandra, et al.
Pubblicazione: (2026)
di: Parziale, Alessandra, et al.
Pubblicazione: (2026)
Mitigating Gender Bias in Code Large Language Models via Model Editing
di: Qin, Zhanyue, et al.
Pubblicazione: (2024)
di: Qin, Zhanyue, et al.
Pubblicazione: (2024)
Meta-Fair: AI-Assisted Fairness Testing of Large Language Models
di: Romero-Arjona, Miguel, et al.
Pubblicazione: (2025)
di: Romero-Arjona, Miguel, et al.
Pubblicazione: (2025)
Investigating Training Data Detection in AI Coders
di: Li, Tianlin, et al.
Pubblicazione: (2025)
di: Li, Tianlin, et al.
Pubblicazione: (2025)
CodeMorph: Mitigating Data Leakage in Large Language Model Assessment
di: Rao, Hongzhou, et al.
Pubblicazione: (2025)
di: Rao, Hongzhou, et al.
Pubblicazione: (2025)
GenFair: Systematic Test Generation for Fairness Fault Detection in Large Language Models
di: Srinivasan, Madhusudan, et al.
Pubblicazione: (2025)
di: Srinivasan, Madhusudan, et al.
Pubblicazione: (2025)
Identifying and Mitigating API Misuse in Large Language Models
di: Zhuo, Terry Yue, et al.
Pubblicazione: (2025)
di: Zhuo, Terry Yue, et al.
Pubblicazione: (2025)
Generating Mitigations for Downstream Projects to Neutralize Upstream Library Vulnerability
di: Chen, Zirui, et al.
Pubblicazione: (2025)
di: Chen, Zirui, et al.
Pubblicazione: (2025)
From Bias To Improved Prompts: A Case Study of Bias Mitigation of Clone Detection Models
di: Chen, QiHong, et al.
Pubblicazione: (2025)
di: Chen, QiHong, et al.
Pubblicazione: (2025)
Where Is Self-admitted Code Generated by Large Language Models on GitHub?
di: Yu, Xiao, et al.
Pubblicazione: (2024)
di: Yu, Xiao, et al.
Pubblicazione: (2024)
Efficient Fairness Testing in Large Language Models: Prioritizing Metamorphic Relations for Bias Detection
di: Giramata, Suavis, et al.
Pubblicazione: (2025)
di: Giramata, Suavis, et al.
Pubblicazione: (2025)
Toward Systematic Counterfactual Fairness Evaluation of Large Language Models: The CAFFE Framework
di: Parziale, Alessandra, et al.
Pubblicazione: (2025)
di: Parziale, Alessandra, et al.
Pubblicazione: (2025)
Lifting the Veil on Composition, Risks, and Mitigations of the Large Language Model Supply Chain
di: Huang, Kaifeng, et al.
Pubblicazione: (2024)
di: Huang, Kaifeng, et al.
Pubblicazione: (2024)
Analyzing and Mitigating Surface Bias in Code Evaluation Metrics
di: Dristi, Simantika Bhattacharjee, et al.
Pubblicazione: (2025)
di: Dristi, Simantika Bhattacharjee, et al.
Pubblicazione: (2025)
Mitigating Omitted Variable Bias in Empirical Software Engineering
di: Furia, Carlo A., et al.
Pubblicazione: (2025)
di: Furia, Carlo A., et al.
Pubblicazione: (2025)
Large Language Models for Unit Testing: A Systematic Literature Review
di: Zhang, Quanjun, et al.
Pubblicazione: (2025)
di: Zhang, Quanjun, et al.
Pubblicazione: (2025)
FlexFL: Flexible and Effective Fault Localization with Open-Source Large Language Models
di: Xu, Chuyang, et al.
Pubblicazione: (2024)
di: Xu, Chuyang, et al.
Pubblicazione: (2024)
GenderBias-\emph{VL}: Benchmarking Gender Bias in Vision Language Models via Counterfactual Probing
di: Xiao, Yisong, et al.
Pubblicazione: (2024)
di: Xiao, Yisong, et al.
Pubblicazione: (2024)
ROCODE: Integrating Backtracking Mechanism and Program Analysis in Large Language Models for Code Generation
di: Jiang, Xue, et al.
Pubblicazione: (2024)
di: Jiang, Xue, et al.
Pubblicazione: (2024)
TickIt: Leveraging Large Language Models for Automated Ticket Escalation
di: Liu, Fengrui, et al.
Pubblicazione: (2025)
di: Liu, Fengrui, et al.
Pubblicazione: (2025)
Optimizing Case-Based Reasoning System for Functional Test Script Generation with Large Language Models
di: Guo, Siyuan, et al.
Pubblicazione: (2025)
di: Guo, Siyuan, et al.
Pubblicazione: (2025)
Narrowing the Complexity Gap in the Evaluation of Large Language Models
di: Chen, Yang, et al.
Pubblicazione: (2026)
di: Chen, Yang, et al.
Pubblicazione: (2026)
Instructive Code Retriever: Learn from Large Language Model's Feedback for Code Intelligence Tasks
di: Lu, Jiawei, et al.
Pubblicazione: (2024)
di: Lu, Jiawei, et al.
Pubblicazione: (2024)
On the Evaluation of Large Language Models in Unit Test Generation
di: Yang, Lin, et al.
Pubblicazione: (2024)
di: Yang, Lin, et al.
Pubblicazione: (2024)
Evolution of Kernels: Automated RISC-V Kernel Optimization with Large Language Models
di: Chen, Siyuan, et al.
Pubblicazione: (2025)
di: Chen, Siyuan, et al.
Pubblicazione: (2025)
Nigerian Software Engineer or American Data Scientist? GitHub Profile Recruitment Bias in Large Language Models
di: Nakano, Takashi, et al.
Pubblicazione: (2024)
di: Nakano, Takashi, et al.
Pubblicazione: (2024)
Efficient Function Orchestration for Large Language Models
di: Liu, Xiaoxia, et al.
Pubblicazione: (2025)
di: Liu, Xiaoxia, et al.
Pubblicazione: (2025)
A Tool for In-depth Analysis of Code Execution Reasoning of Large Language Models
di: Liu, Changshu, et al.
Pubblicazione: (2025)
di: Liu, Changshu, et al.
Pubblicazione: (2025)
Code2Bench: Scaling Source and Rigor for Dynamic Benchmark Construction
di: Zhang, Zhe, et al.
Pubblicazione: (2025)
di: Zhang, Zhe, et al.
Pubblicazione: (2025)
Unveiling Bias in Fairness Evaluations of Large Language Models: A Critical Literature Review of Music and Movie Recommendation Systems
di: Sah, Chandan Kumar, et al.
Pubblicazione: (2024)
di: Sah, Chandan Kumar, et al.
Pubblicazione: (2024)
Assessing Coherency and Consistency of Code Execution Reasoning by Large Language Models
di: Liu, Changshu, et al.
Pubblicazione: (2025)
di: Liu, Changshu, et al.
Pubblicazione: (2025)
Exploring the Potential and Limitations of Large Language Models for Novice Program Fault Localization
di: Xu, Hexiang, et al.
Pubblicazione: (2025)
di: Xu, Hexiang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Latent Imitator: Generating Natural Individual Discriminatory Instances for Black-Box Fairness Testing
di: Xiao, Yisong, et al.
Pubblicazione: (2023) -
BDefects4NN: A Backdoor Defect Database for Controlled Localization Studies in Neural Networks
di: Xiao, Yisong, et al.
Pubblicazione: (2024) -
Software Development Life Cycle Perspective: A Survey of Benchmarks for Code Large Language Models and Agents
di: Wang, Kaixin, et al.
Pubblicazione: (2025) -
Detoxifying Large Language Models via Autoregressive Reward Guided Representation Editing
di: Xiao, Yisong, et al.
Pubblicazione: (2025) -
Bias Ahead: Sensitive Prompts as Early Warnings for Fairness in Large Language Models
di: Voria, Gianmario, et al.
Pubblicazione: (2026)