LFTF: Locating First and Then Fine-Tuning for Mitigating Gender Bias in Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Qin, Zhanyue, Ding, Yue, Liu, Deyuan, Liu, Qingbin, Cai, Junxian, Chen, Xi, Tu, Zhiying, Chu, Dianhui, Gao, Cuiyun, Sui, Dianbo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Mitigating Gender Bias in Code Large Language Models via Model Editing
di: Qin, Zhanyue, et al.
Pubblicazione: (2024)
di: Qin, Zhanyue, et al.
Pubblicazione: (2024)
Beyond Confidence: The Rhythms of Reasoning in Generative Models
di: Liu, Deyuan, et al.
Pubblicazione: (2026)
di: Liu, Deyuan, et al.
Pubblicazione: (2026)
Pruning via Merging: Compressing LLMs via Manifold Alignment Based Layer Merging
di: Liu, Deyuan, et al.
Pubblicazione: (2024)
di: Liu, Deyuan, et al.
Pubblicazione: (2024)
UNO Arena for Evaluating Sequential Decision-Making Capability of Large Language Models
di: Qin, Zhanyue, et al.
Pubblicazione: (2024)
di: Qin, Zhanyue, et al.
Pubblicazione: (2024)
Checkpoint Merging via Bayesian Optimization in LLM Pretraining
di: Liu, Deyuan, et al.
Pubblicazione: (2024)
di: Liu, Deyuan, et al.
Pubblicazione: (2024)
HBot: A Chatbot for Healthcare Applications in Traditional Chinese Medicine Based on Human Body 3D Visualization
di: Zhang, Bolin, et al.
Pubblicazione: (2024)
di: Zhang, Bolin, et al.
Pubblicazione: (2024)
ScEdit: Script-based Assessment of Knowledge Editing
di: Li, Xinye, et al.
Pubblicazione: (2025)
di: Li, Xinye, et al.
Pubblicazione: (2025)
Enhancing the Code Reasoning Capabilities of LLMs via Consistency-based Reinforcement Learning
di: Qin, Zhanyue, et al.
Pubblicazione: (2026)
di: Qin, Zhanyue, et al.
Pubblicazione: (2026)
Gradients Must Earn Their Influence: Unifying SFT with Generalized Entropic Objectives
di: Wang, Zecheng, et al.
Pubblicazione: (2026)
di: Wang, Zecheng, et al.
Pubblicazione: (2026)
Locating and Mitigating Gender Bias in Large Language Models
di: Cai, Yuchen, et al.
Pubblicazione: (2024)
di: Cai, Yuchen, et al.
Pubblicazione: (2024)
A Survey on Data Selection for LLM Instruction Tuning
di: Zhang, Bolin, et al.
Pubblicazione: (2024)
di: Zhang, Bolin, et al.
Pubblicazione: (2024)
Plug-and-Play Performance Estimation for LLM Services without Relying on Labeled Data
di: Wang, Can, et al.
Pubblicazione: (2024)
di: Wang, Can, et al.
Pubblicazione: (2024)
Catalog, Impact, and Evaluation of Microservice Bad Smells: A Systematic Literature Review
di: Yongchao Xing, et al.
Pubblicazione: (2026)
di: Yongchao Xing, et al.
Pubblicazione: (2026)
EchoReview: Learning Peer Review from the Echoes of Scientific Citations
di: Zhang, Yinuo, et al.
Pubblicazione: (2026)
di: Zhang, Yinuo, et al.
Pubblicazione: (2026)
TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs
di: Wang, Haochuan, et al.
Pubblicazione: (2024)
di: Wang, Haochuan, et al.
Pubblicazione: (2024)
A Framework for Effective Invocation Methods of Various LLM Services
di: Wang, Can, et al.
Pubblicazione: (2024)
di: Wang, Can, et al.
Pubblicazione: (2024)
Prefix Grouper: Efficient GRPO Training through Shared-Prefix Forward
di: Liu, Zikang, et al.
Pubblicazione: (2025)
di: Liu, Zikang, et al.
Pubblicazione: (2025)
VTG-LLM: Integrating Timestamp Knowledge into Video LLMs for Enhanced Video Temporal Grounding
di: Guo, Yongxin, et al.
Pubblicazione: (2024)
di: Guo, Yongxin, et al.
Pubblicazione: (2024)
VRoPE: Rotary Position Embedding for Video Large Language Models
di: Liu, Zikang, et al.
Pubblicazione: (2025)
di: Liu, Zikang, et al.
Pubblicazione: (2025)
To Forget or Not? Towards Practical Knowledge Unlearning for Large Language Models
di: Tian, Bozhong, et al.
Pubblicazione: (2024)
di: Tian, Bozhong, et al.
Pubblicazione: (2024)
Fragile Reconstruction: Adversarial Vulnerability of Reconstruction-Based Detectors for Diffusion-Generated Images
di: Jiang, Haoyang, et al.
Pubblicazione: (2026)
di: Jiang, Haoyang, et al.
Pubblicazione: (2026)
On the Effectiveness of Context Compression for Repository-Level Tasks: An Empirical Investigation
di: Feng, Jia, et al.
Pubblicazione: (2026)
di: Feng, Jia, et al.
Pubblicazione: (2026)
From Detection to Mitigation: Addressing Gender Bias in Chinese Texts via Efficient Tuning and Voting-Based Rebalancing
di: Wu, Chengyan, et al.
Pubblicazione: (2025)
di: Wu, Chengyan, et al.
Pubblicazione: (2025)
Mitigating Gender Bias in Depression Detection via Counterfactual Inference
di: Hu, Mingxuan, et al.
Pubblicazione: (2025)
di: Hu, Mingxuan, et al.
Pubblicazione: (2025)
Laplacian Score Sharpening for Mitigating Hallucination in Diffusion Models
di: C, Barath Chandran., et al.
Pubblicazione: (2025)
di: C, Barath Chandran., et al.
Pubblicazione: (2025)
Beware of Your Po! Measuring and Mitigating AI Safety Risks in Role-Play Fine-Tuning of LLMs
di: Zhao, Weixiang, et al.
Pubblicazione: (2025)
di: Zhao, Weixiang, et al.
Pubblicazione: (2025)
Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning
di: Du, Yanrui, et al.
Pubblicazione: (2024)
di: Du, Yanrui, et al.
Pubblicazione: (2024)
Mitigating Extrinsic Gender Bias for Bangla Classification Tasks
di: Joy, Sajib Kumar Saha, et al.
Pubblicazione: (2024)
di: Joy, Sajib Kumar Saha, et al.
Pubblicazione: (2024)
Disclosure and Mitigation of Gender Bias in LLMs
di: Dong, Xiangjue, et al.
Pubblicazione: (2024)
di: Dong, Xiangjue, et al.
Pubblicazione: (2024)
Reinforcement Fine-Tuning Naturally Mitigates Forgetting in Continual Post-Training
di: Lai, Song, et al.
Pubblicazione: (2025)
di: Lai, Song, et al.
Pubblicazione: (2025)
OFA-Diffusion Compression: Compressing Diffusion Model in One-Shot Manner
di: Jiang, Haoyang, et al.
Pubblicazione: (2026)
di: Jiang, Haoyang, et al.
Pubblicazione: (2026)
LLMSR@XLLM25: An Empirical Study of LLM for Structural Reasoning
di: Li, Xinye, et al.
Pubblicazione: (2025)
di: Li, Xinye, et al.
Pubblicazione: (2025)
Mitigating Premature Discretization with Progressive Quantization for Robust Vector Tokenization
di: Zhao, Wenhao, et al.
Pubblicazione: (2026)
di: Zhao, Wenhao, et al.
Pubblicazione: (2026)
LoCA: Location-Aware Cosine Adaptation for Parameter-Efficient Fine-Tuning
di: Du, Zhekai, et al.
Pubblicazione: (2025)
di: Du, Zhekai, et al.
Pubblicazione: (2025)
Staircase Recognition and Location Based on Polarization Vision
di: Kong, Weifeng, et al.
Pubblicazione: (2025)
di: Kong, Weifeng, et al.
Pubblicazione: (2025)
Mitigating Gender Bias in Contextual Word Embeddings
di: Yarrabelly, Navya, et al.
Pubblicazione: (2024)
di: Yarrabelly, Navya, et al.
Pubblicazione: (2024)
Context-Aware Counterfactual Data Augmentation for Gender Bias Mitigation in Language Models
di: Parihar, Shweta, et al.
Pubblicazione: (2026)
di: Parihar, Shweta, et al.
Pubblicazione: (2026)
Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting
di: Ding, Fei, et al.
Pubblicazione: (2025)
di: Ding, Fei, et al.
Pubblicazione: (2025)
Reasoning Towards Fairness: Mitigating Bias in Language Models through Reasoning-Guided Fine-Tuning
di: Kabra, Sanchit, et al.
Pubblicazione: (2025)
di: Kabra, Sanchit, et al.
Pubblicazione: (2025)
Auto-Search and Refinement: An Automated Framework for Gender Bias Mitigation in Large Language Models
di: Xu, Yue, et al.
Pubblicazione: (2025)
di: Xu, Yue, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Mitigating Gender Bias in Code Large Language Models via Model Editing
di: Qin, Zhanyue, et al.
Pubblicazione: (2024) -
Beyond Confidence: The Rhythms of Reasoning in Generative Models
di: Liu, Deyuan, et al.
Pubblicazione: (2026) -
Pruning via Merging: Compressing LLMs via Manifold Alignment Based Layer Merging
di: Liu, Deyuan, et al.
Pubblicazione: (2024) -
UNO Arena for Evaluating Sequential Decision-Making Capability of Large Language Models
di: Qin, Zhanyue, et al.
Pubblicazione: (2024) -
Checkpoint Merging via Bayesian Optimization in LLM Pretraining
di: Liu, Deyuan, et al.
Pubblicazione: (2024)