Fine-Tuned LLMs are "Time Capsules" for Tracking Societal Bias Through Books
Fuente:
arXiv
Saved in:
| Main Authors: | Madhusudan, Sangmitra, Morabito, Robert, Reid, Skye, Sadr, Nikta Gohari, Emami, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Which Words Matter Most in Zero-Shot Prompts?
by: Sadr, Nikta Gohari, et al.
Published: (2025)
by: Sadr, Nikta Gohari, et al.
Published: (2025)
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
by: Morabito, Robert, et al.
Published: (2024)
by: Morabito, Robert, et al.
Published: (2024)
The Dog the Cat Chased Stumped the Model: Measuring When Language Models Abandon Structure for Shortcuts
by: Madhusudan, Sangmitra, et al.
Published: (2025)
by: Madhusudan, Sangmitra, et al.
Published: (2025)
We Politely Insist: Your LLM Must Learn the Persian Art of Taarof
by: Sadr, Nikta Gohari, et al.
Published: (2025)
by: Sadr, Nikta Gohari, et al.
Published: (2025)
Common to Whom? Regional Cultural Commonsense and LLM Bias in India
by: Madhusudan, Sangmitra, et al.
Published: (2026)
by: Madhusudan, Sangmitra, et al.
Published: (2026)
Confidence Under the Hood: An Investigation into the Confidence-Probability Alignment in Large Language Models
by: Kumar, Abhishek, et al.
Published: (2024)
by: Kumar, Abhishek, et al.
Published: (2024)
Translate With Care: Addressing Gender Bias, Neutrality, and Reasoning in Large Language Model Translations
by: Zahraei, Pardis Sadat, et al.
Published: (2025)
by: Zahraei, Pardis Sadat, et al.
Published: (2025)
Safeguard Fine-Tuned LLMs Through Pre- and Post-Tuning Model Merging
by: Farn, Hua, et al.
Published: (2024)
by: Farn, Hua, et al.
Published: (2024)
Tracking Universal Features Through Fine-Tuning and Model Merging
by: Horn, Niels, et al.
Published: (2024)
by: Horn, Niels, et al.
Published: (2024)
CoBia: Constructed Conversations Can Trigger Otherwise Concealed Societal Biases in LLMs
by: Nikeghbal, Nafiseh, et al.
Published: (2025)
by: Nikeghbal, Nafiseh, et al.
Published: (2025)
Augmented Fine-Tuned LLMs for Enhanced Recruitment Automation
by: Younes, Mohamed T., et al.
Published: (2025)
by: Younes, Mohamed T., et al.
Published: (2025)
Fine-Tuning LLMs for Reliable Medical Question-Answering Services
by: Anaissi, Ali, et al.
Published: (2024)
by: Anaissi, Ali, et al.
Published: (2024)
Trace-of-Thought Prompting: Investigating Prompt-Based Knowledge Distillation Through Question Decomposition
by: McDonald, Tyler, et al.
Published: (2025)
by: McDonald, Tyler, et al.
Published: (2025)
On The Conceptualization and Societal Impact of Cross-Cultural Bias
by: Bhandari, Vitthal
Published: (2025)
by: Bhandari, Vitthal
Published: (2025)
Subtle Biases Need Subtler Measures: Dual Metrics for Evaluating Representative and Affinity Bias in Large Language Models
by: Kumar, Abhishek, et al.
Published: (2024)
by: Kumar, Abhishek, et al.
Published: (2024)
Fine-Tuning LLMs for Low-Resource Dialect Translation: The Case of Lebanese
by: Yakhni, Silvana, et al.
Published: (2025)
by: Yakhni, Silvana, et al.
Published: (2025)
DART: Mitigating Harm Drift in Difference-Aware LLMs via Distill-Audit-Repair Training
by: Pan, Ziwen, et al.
Published: (2026)
by: Pan, Ziwen, et al.
Published: (2026)
Memory Dial: A Training Framework for Controllable Memorization in Language Models
by: Zhang, Xiangbo, et al.
Published: (2026)
by: Zhang, Xiangbo, et al.
Published: (2026)
Position-Aware Parameter Efficient Fine-Tuning Approach for Reducing Positional Bias in LLMs
by: Zhang, Zheng, et al.
Published: (2024)
by: Zhang, Zheng, et al.
Published: (2024)
Metamorphic Testing for Fairness Evaluation in Large Language Models: Identifying Intersectional Bias in LLaMA and GPT
by: Reddy, Harishwar, et al.
Published: (2025)
by: Reddy, Harishwar, et al.
Published: (2025)
KLAAD: Refining Attention Mechanisms to Reduce Societal Bias in Generative Language Models
by: Kim, Seorin, et al.
Published: (2025)
by: Kim, Seorin, et al.
Published: (2025)
EvoGrad: A Dynamic Take on the Winograd Schema Challenge with Human Adversaries
by: Sun, Jing Han, et al.
Published: (2024)
by: Sun, Jing Han, et al.
Published: (2024)
Red-Teaming for Inducing Societal Bias in Large Language Models
by: Luo, Chu Fei, et al.
Published: (2024)
by: Luo, Chu Fei, et al.
Published: (2024)
Order-Independence Without Fine Tuning
by: McIlroy-Young, Reid, et al.
Published: (2024)
by: McIlroy-Young, Reid, et al.
Published: (2024)
QEFT: Quantization for Efficient Fine-Tuning of LLMs
by: Lee, Changhun, et al.
Published: (2024)
by: Lee, Changhun, et al.
Published: (2024)
CALF: Aligning LLMs for Time Series Forecasting via Cross-modal Fine-Tuning
by: Liu, Peiyuan, et al.
Published: (2024)
by: Liu, Peiyuan, et al.
Published: (2024)
Towards Pedagogical LLMs with Supervised Fine Tuning for Computing Education
by: Vassar, Alexandra, et al.
Published: (2024)
by: Vassar, Alexandra, et al.
Published: (2024)
Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
by: Gekhman, Zorik, et al.
Published: (2024)
by: Gekhman, Zorik, et al.
Published: (2024)
Parameter-Efficient Fine-Tuning of LLMs with Mixture of Space Experts
by: Zhang, Buze, et al.
Published: (2026)
by: Zhang, Buze, et al.
Published: (2026)
Dataset Scale and Societal Consistency Mediate Facial Impression Bias in Vision-Language AI
by: Wolfe, Robert, et al.
Published: (2024)
by: Wolfe, Robert, et al.
Published: (2024)
Towards Better Understanding of Cybercrime: The Role of Fine-Tuned LLMs in Translation
by: Valeros, Veronica, et al.
Published: (2024)
by: Valeros, Veronica, et al.
Published: (2024)
Language Bias under Conflicting Information in Multilingual LLMs
by: Östling, Robert, et al.
Published: (2026)
by: Östling, Robert, et al.
Published: (2026)
Supervised Fine-Tuning LLMs to Behave as Pedagogical Agents in Programming Education
by: Ross, Emily, et al.
Published: (2025)
by: Ross, Emily, et al.
Published: (2025)
Controlling Distributional Bias in Multi-Round LLM Generation via KL-Optimized Fine-Tuning
by: Jiang, Yanbei, et al.
Published: (2026)
by: Jiang, Yanbei, et al.
Published: (2026)
Fine-Tuning LLMs with Fine-Grained Human Feedback on Text Spans
by: CH-Wang, Sky, et al.
Published: (2025)
by: CH-Wang, Sky, et al.
Published: (2025)
A Unified Framework and Dataset for Assessing Societal Bias in Vision-Language Models
by: Sathe, Ashutosh, et al.
Published: (2024)
by: Sathe, Ashutosh, et al.
Published: (2024)
Knowledge Capsules: Structured Nonparametric Memory Units for LLMs
by: Ju, Bin, et al.
Published: (2026)
by: Ju, Bin, et al.
Published: (2026)
From 'Showgirls' to 'Performers': Fine-tuning with Gender-inclusive Language for Bias Reduction in LLMs
by: Bartl, Marion, et al.
Published: (2024)
by: Bartl, Marion, et al.
Published: (2024)
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
by: Hanif, Ikhlasul Akmal, et al.
Published: (2026)
by: Hanif, Ikhlasul Akmal, et al.
Published: (2026)
Quantifying and Mitigating Selection Bias in LLMs: A Transferable LoRA Fine-Tuning and Efficient Majority Voting Approach
by: Guda, Blessed, et al.
Published: (2025)
by: Guda, Blessed, et al.
Published: (2025)
Similar Items
-
Which Words Matter Most in Zero-Shot Prompts?
by: Sadr, Nikta Gohari, et al.
Published: (2025) -
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
by: Morabito, Robert, et al.
Published: (2024) -
The Dog the Cat Chased Stumped the Model: Measuring When Language Models Abandon Structure for Shortcuts
by: Madhusudan, Sangmitra, et al.
Published: (2025) -
We Politely Insist: Your LLM Must Learn the Persian Art of Taarof
by: Sadr, Nikta Gohari, et al.
Published: (2025) -
Common to Whom? Regional Cultural Commonsense and LLM Bias in India
by: Madhusudan, Sangmitra, et al.
Published: (2026)