Evaluating Gender Bias Transfer between Pre-trained and Prompt-Adapted Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Mackraz, Natalie, Sivakumar, Nivedha, Khorshidi, Samira, Patel, Krishna, Theobald, Barry-John, Zappella, Luca, Apostoloff, Nicholas |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bias after Prompting: Persistent Discrimination in Large Language Models
by: Sivakumar, Nivedha, et al.
Published: (2025)
by: Sivakumar, Nivedha, et al.
Published: (2025)
Fairness Dynamics During Training
by: Patel, Krishna, et al.
Published: (2025)
by: Patel, Krishna, et al.
Published: (2025)
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
by: Wang, Yinong Oliver, et al.
Published: (2025)
by: Wang, Yinong Oliver, et al.
Published: (2025)
DSO: Direct Steering Optimization for Bias Mitigation
by: Paes, Lucas Monteiro, et al.
Published: (2025)
by: Paes, Lucas Monteiro, et al.
Published: (2025)
Investigating Intersectional Bias in Large Language Models using Confidence Disparities in Coreference Resolution
by: Khan, Falaah Arif, et al.
Published: (2025)
by: Khan, Falaah Arif, et al.
Published: (2025)
Theoretical Limits of Language Model Alignment
by: Paes, Lucas Monteiro, et al.
Published: (2026)
by: Paes, Lucas Monteiro, et al.
Published: (2026)
PREDICT: Preference Reasoning by Evaluating Decomposed preferences Inferred from Candidate Trajectories
by: Aroca-Ouellette, Stephane, et al.
Published: (2024)
by: Aroca-Ouellette, Stephane, et al.
Published: (2024)
Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards
by: Metcalf, Katherine, et al.
Published: (2024)
by: Metcalf, Katherine, et al.
Published: (2024)
Aligning LLMs by Predicting Preferences from User Writing Samples
by: Aroca-Ouellette, Stéphane, et al.
Published: (2025)
by: Aroca-Ouellette, Stéphane, et al.
Published: (2025)
Controlling Language and Diffusion Models by Transporting Activations
by: Rodriguez, Pau, et al.
Published: (2024)
by: Rodriguez, Pau, et al.
Published: (2024)
Whispering Experts: Neural Interventions for Toxicity Mitigation in Language Models
by: Suau, Xavier, et al.
Published: (2024)
by: Suau, Xavier, et al.
Published: (2024)
Evaluating Gender Bias of Pre-trained Language Models in Natural Language Inference by Considering All Labels
by: Anantaprayoon, Panatchakorn, et al.
Published: (2023)
by: Anantaprayoon, Panatchakorn, et al.
Published: (2023)
Projective Methods for Mitigating Gender Bias in Pre-trained Language Models
by: Dawkins, Hillary, et al.
Published: (2024)
by: Dawkins, Hillary, et al.
Published: (2024)
The power of Prompts: Evaluating and Mitigating Gender Bias in MT with LLMs
by: Sant, Aleix, et al.
Published: (2024)
by: Sant, Aleix, et al.
Published: (2024)
The Gaps between Pre-train and Downstream Settings in Bias Evaluation and Debiasing
by: Kaneko, Masahiro, et al.
Published: (2024)
by: Kaneko, Masahiro, et al.
Published: (2024)
Revisiting Uncertainty Quantification Evaluation in Language Models: Spurious Interactions with Response Length Bias Results
by: Santilli, Andrea, et al.
Published: (2025)
by: Santilli, Andrea, et al.
Published: (2025)
STRUCTURAL FRAGMENTATION IN INDIAN ENVIRONMENTAL LAW: ENFORCEMENT DEFICITS AND THE IMPERATIVE FOR A UNIFIED ENVIRONMENTAL CODE
by: Joan Nivedha S
Published: (2026)
by: Joan Nivedha S
Published: (2026)
Improving Transfer Learning for Sequence Labeling Tasks by Adapting Pre-trained Neural Language Models
by: Dukić, David
Published: (2025)
by: Dukić, David
Published: (2025)
Aggregate-and-Adapt Natural Language Prompts for Downstream Generalization of CLIP
by: Huang, Chen, et al.
Published: (2024)
by: Huang, Chen, et al.
Published: (2024)
Evaluating Gender Bias in Large Language Models via Chain-of-Thought Prompting
by: Kaneko, Masahiro, et al.
Published: (2024)
by: Kaneko, Masahiro, et al.
Published: (2024)
Evaluating Gender Bias in Large Language Models
by: Döll, Michael, et al.
Published: (2024)
by: Döll, Michael, et al.
Published: (2024)
Efficient Knowledge Probing of Large Language Models by Adapting Pre-trained Embeddings
by: Sharma, Kartik, et al.
Published: (2025)
by: Sharma, Kartik, et al.
Published: (2025)
MIA-Tuner: Adapting Large Language Models as Pre-training Text Detector
by: Fu, Wenjie, et al.
Published: (2024)
by: Fu, Wenjie, et al.
Published: (2024)
PLUM: Adapting Pre-trained Language Models for Industrial-scale Generative Recommendations
by: He, Ruining, et al.
Published: (2025)
by: He, Ruining, et al.
Published: (2025)
Attention Prompt Tuning: Parameter-efficient Adaptation of Pre-trained Models for Spatiotemporal Modeling
by: Bandara, Wele Gedara Chaminda, et al.
Published: (2024)
by: Bandara, Wele Gedara Chaminda, et al.
Published: (2024)
Context-Aware Counterfactual Data Augmentation for Gender Bias Mitigation in Language Models
by: Parihar, Shweta, et al.
Published: (2026)
by: Parihar, Shweta, et al.
Published: (2026)
New Deformity Outline on the Breast Radiation Therapy for diminishing Absorbed Dose Ratio
by: A. Khorshidi
Published: (2023)
by: A. Khorshidi
Published: (2023)
Domain-Adapted Pre-trained Language Models for Implicit Information Extraction in Crash Narratives
by: Wang, Xixi, et al.
Published: (2025)
by: Wang, Xixi, et al.
Published: (2025)
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
by: Danieli, Federico, et al.
Published: (2025)
by: Danieli, Federico, et al.
Published: (2025)
AfroXLMR-Social: Adapting Pre-trained Language Models for African Languages Social Media Text
by: Belay, Tadesse Destaw, et al.
Published: (2025)
by: Belay, Tadesse Destaw, et al.
Published: (2025)
Pre-trained Language Model with Prompts for Temporal Knowledge Graph Completion
by: Xu, Wenjie, et al.
Published: (2023)
by: Xu, Wenjie, et al.
Published: (2023)
Understanding the Multi-modal Prompts of the Pre-trained Vision-Language Model
by: Ma, Shuailei, et al.
Published: (2023)
by: Ma, Shuailei, et al.
Published: (2023)
Adapting Multilingual LLMs to Low-Resource Languages using Continued Pre-training and Synthetic Corpus
by: Joshi, Raviraj, et al.
Published: (2024)
by: Joshi, Raviraj, et al.
Published: (2024)
Adapting Multi-modal Large Language Model to Concept Drift From Pre-training Onwards
by: Yang, Xiaoyu, et al.
Published: (2024)
by: Yang, Xiaoyu, et al.
Published: (2024)
The Effect of Pearl Vortices on the Shape and Position of Néel-Type Skyrmions in Superconductor-Chiral Ferromagnet Heterostructures
by: Apostoloff, S. S., et al.
Published: (2025)
by: Apostoloff, S. S., et al.
Published: (2025)
Bound states and scattering of magnons on a superconducting vortex in ferromagnet-superconductor heterostructures
by: Katkov, D. S., et al.
Published: (2024)
by: Katkov, D. S., et al.
Published: (2024)
Deformation of a Néel-type Skyrmion in a Weak Inhomogeneous Magnetic Field: Magnetization Ansatz and Interaction with a Pearl Vortex
by: Apostoloff, S. S., et al.
Published: (2023)
by: Apostoloff, S. S., et al.
Published: (2023)
Majorana bound states in chiral ferromagnet-superconductor heterostructures revisited
by: Slobodskoi, A. S., et al.
Published: (2026)
by: Slobodskoi, A. S., et al.
Published: (2026)
Choice of PEFT Technique in Continual Learning: Prompt Tuning is Not All You Need
by: Wistuba, Martin, et al.
Published: (2024)
by: Wistuba, Martin, et al.
Published: (2024)
Patronus: Identifying and Mitigating Transferable Backdoors in Pre-trained Language Models
by: Zhao, Tianhang, et al.
Published: (2025)
by: Zhao, Tianhang, et al.
Published: (2025)
Similar Items
-
Bias after Prompting: Persistent Discrimination in Large Language Models
by: Sivakumar, Nivedha, et al.
Published: (2025) -
Fairness Dynamics During Training
by: Patel, Krishna, et al.
Published: (2025) -
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
by: Wang, Yinong Oliver, et al.
Published: (2025) -
DSO: Direct Steering Optimization for Bias Mitigation
by: Paes, Lucas Monteiro, et al.
Published: (2025) -
Investigating Intersectional Bias in Large Language Models using Confidence Disparities in Coreference Resolution
by: Khan, Falaah Arif, et al.
Published: (2025)