Fine-Tuned LLMs are "Time Capsules" for Tracking Societal Bias Through Books
Fuente:
arXiv
Guardado en:
| Autores principales: | Madhusudan, Sangmitra, Morabito, Robert, Reid, Skye, Sadr, Nikta Gohari, Emami, Ali |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Which Words Matter Most in Zero-Shot Prompts?
por: Sadr, Nikta Gohari, et al.
Publicado: (2025)
por: Sadr, Nikta Gohari, et al.
Publicado: (2025)
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
por: Morabito, Robert, et al.
Publicado: (2024)
por: Morabito, Robert, et al.
Publicado: (2024)
The Dog the Cat Chased Stumped the Model: Measuring When Language Models Abandon Structure for Shortcuts
por: Madhusudan, Sangmitra, et al.
Publicado: (2025)
por: Madhusudan, Sangmitra, et al.
Publicado: (2025)
We Politely Insist: Your LLM Must Learn the Persian Art of Taarof
por: Sadr, Nikta Gohari, et al.
Publicado: (2025)
por: Sadr, Nikta Gohari, et al.
Publicado: (2025)
Common to Whom? Regional Cultural Commonsense and LLM Bias in India
por: Madhusudan, Sangmitra, et al.
Publicado: (2026)
por: Madhusudan, Sangmitra, et al.
Publicado: (2026)
Confidence Under the Hood: An Investigation into the Confidence-Probability Alignment in Large Language Models
por: Kumar, Abhishek, et al.
Publicado: (2024)
por: Kumar, Abhishek, et al.
Publicado: (2024)
Translate With Care: Addressing Gender Bias, Neutrality, and Reasoning in Large Language Model Translations
por: Zahraei, Pardis Sadat, et al.
Publicado: (2025)
por: Zahraei, Pardis Sadat, et al.
Publicado: (2025)
Safeguard Fine-Tuned LLMs Through Pre- and Post-Tuning Model Merging
por: Farn, Hua, et al.
Publicado: (2024)
por: Farn, Hua, et al.
Publicado: (2024)
Tracking Universal Features Through Fine-Tuning and Model Merging
por: Horn, Niels, et al.
Publicado: (2024)
por: Horn, Niels, et al.
Publicado: (2024)
CoBia: Constructed Conversations Can Trigger Otherwise Concealed Societal Biases in LLMs
por: Nikeghbal, Nafiseh, et al.
Publicado: (2025)
por: Nikeghbal, Nafiseh, et al.
Publicado: (2025)
Augmented Fine-Tuned LLMs for Enhanced Recruitment Automation
por: Younes, Mohamed T., et al.
Publicado: (2025)
por: Younes, Mohamed T., et al.
Publicado: (2025)
Fine-Tuning LLMs for Reliable Medical Question-Answering Services
por: Anaissi, Ali, et al.
Publicado: (2024)
por: Anaissi, Ali, et al.
Publicado: (2024)
Trace-of-Thought Prompting: Investigating Prompt-Based Knowledge Distillation Through Question Decomposition
por: McDonald, Tyler, et al.
Publicado: (2025)
por: McDonald, Tyler, et al.
Publicado: (2025)
On The Conceptualization and Societal Impact of Cross-Cultural Bias
por: Bhandari, Vitthal
Publicado: (2025)
por: Bhandari, Vitthal
Publicado: (2025)
Subtle Biases Need Subtler Measures: Dual Metrics for Evaluating Representative and Affinity Bias in Large Language Models
por: Kumar, Abhishek, et al.
Publicado: (2024)
por: Kumar, Abhishek, et al.
Publicado: (2024)
Fine-Tuning LLMs for Low-Resource Dialect Translation: The Case of Lebanese
por: Yakhni, Silvana, et al.
Publicado: (2025)
por: Yakhni, Silvana, et al.
Publicado: (2025)
DART: Mitigating Harm Drift in Difference-Aware LLMs via Distill-Audit-Repair Training
por: Pan, Ziwen, et al.
Publicado: (2026)
por: Pan, Ziwen, et al.
Publicado: (2026)
Memory Dial: A Training Framework for Controllable Memorization in Language Models
por: Zhang, Xiangbo, et al.
Publicado: (2026)
por: Zhang, Xiangbo, et al.
Publicado: (2026)
Position-Aware Parameter Efficient Fine-Tuning Approach for Reducing Positional Bias in LLMs
por: Zhang, Zheng, et al.
Publicado: (2024)
por: Zhang, Zheng, et al.
Publicado: (2024)
Metamorphic Testing for Fairness Evaluation in Large Language Models: Identifying Intersectional Bias in LLaMA and GPT
por: Reddy, Harishwar, et al.
Publicado: (2025)
por: Reddy, Harishwar, et al.
Publicado: (2025)
KLAAD: Refining Attention Mechanisms to Reduce Societal Bias in Generative Language Models
por: Kim, Seorin, et al.
Publicado: (2025)
por: Kim, Seorin, et al.
Publicado: (2025)
EvoGrad: A Dynamic Take on the Winograd Schema Challenge with Human Adversaries
por: Sun, Jing Han, et al.
Publicado: (2024)
por: Sun, Jing Han, et al.
Publicado: (2024)
Red-Teaming for Inducing Societal Bias in Large Language Models
por: Luo, Chu Fei, et al.
Publicado: (2024)
por: Luo, Chu Fei, et al.
Publicado: (2024)
Order-Independence Without Fine Tuning
por: McIlroy-Young, Reid, et al.
Publicado: (2024)
por: McIlroy-Young, Reid, et al.
Publicado: (2024)
QEFT: Quantization for Efficient Fine-Tuning of LLMs
por: Lee, Changhun, et al.
Publicado: (2024)
por: Lee, Changhun, et al.
Publicado: (2024)
CALF: Aligning LLMs for Time Series Forecasting via Cross-modal Fine-Tuning
por: Liu, Peiyuan, et al.
Publicado: (2024)
por: Liu, Peiyuan, et al.
Publicado: (2024)
Towards Pedagogical LLMs with Supervised Fine Tuning for Computing Education
por: Vassar, Alexandra, et al.
Publicado: (2024)
por: Vassar, Alexandra, et al.
Publicado: (2024)
Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
por: Gekhman, Zorik, et al.
Publicado: (2024)
por: Gekhman, Zorik, et al.
Publicado: (2024)
Parameter-Efficient Fine-Tuning of LLMs with Mixture of Space Experts
por: Zhang, Buze, et al.
Publicado: (2026)
por: Zhang, Buze, et al.
Publicado: (2026)
Dataset Scale and Societal Consistency Mediate Facial Impression Bias in Vision-Language AI
por: Wolfe, Robert, et al.
Publicado: (2024)
por: Wolfe, Robert, et al.
Publicado: (2024)
Towards Better Understanding of Cybercrime: The Role of Fine-Tuned LLMs in Translation
por: Valeros, Veronica, et al.
Publicado: (2024)
por: Valeros, Veronica, et al.
Publicado: (2024)
Language Bias under Conflicting Information in Multilingual LLMs
por: Östling, Robert, et al.
Publicado: (2026)
por: Östling, Robert, et al.
Publicado: (2026)
Supervised Fine-Tuning LLMs to Behave as Pedagogical Agents in Programming Education
por: Ross, Emily, et al.
Publicado: (2025)
por: Ross, Emily, et al.
Publicado: (2025)
Controlling Distributional Bias in Multi-Round LLM Generation via KL-Optimized Fine-Tuning
por: Jiang, Yanbei, et al.
Publicado: (2026)
por: Jiang, Yanbei, et al.
Publicado: (2026)
Fine-Tuning LLMs with Fine-Grained Human Feedback on Text Spans
por: CH-Wang, Sky, et al.
Publicado: (2025)
por: CH-Wang, Sky, et al.
Publicado: (2025)
A Unified Framework and Dataset for Assessing Societal Bias in Vision-Language Models
por: Sathe, Ashutosh, et al.
Publicado: (2024)
por: Sathe, Ashutosh, et al.
Publicado: (2024)
Knowledge Capsules: Structured Nonparametric Memory Units for LLMs
por: Ju, Bin, et al.
Publicado: (2026)
por: Ju, Bin, et al.
Publicado: (2026)
From 'Showgirls' to 'Performers': Fine-tuning with Gender-inclusive Language for Bias Reduction in LLMs
por: Bartl, Marion, et al.
Publicado: (2024)
por: Bartl, Marion, et al.
Publicado: (2024)
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
por: Hanif, Ikhlasul Akmal, et al.
Publicado: (2026)
por: Hanif, Ikhlasul Akmal, et al.
Publicado: (2026)
Quantifying and Mitigating Selection Bias in LLMs: A Transferable LoRA Fine-Tuning and Efficient Majority Voting Approach
por: Guda, Blessed, et al.
Publicado: (2025)
por: Guda, Blessed, et al.
Publicado: (2025)
Ejemplares similares
-
Which Words Matter Most in Zero-Shot Prompts?
por: Sadr, Nikta Gohari, et al.
Publicado: (2025) -
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
por: Morabito, Robert, et al.
Publicado: (2024) -
The Dog the Cat Chased Stumped the Model: Measuring When Language Models Abandon Structure for Shortcuts
por: Madhusudan, Sangmitra, et al.
Publicado: (2025) -
We Politely Insist: Your LLM Must Learn the Persian Art of Taarof
por: Sadr, Nikta Gohari, et al.
Publicado: (2025) -
Common to Whom? Regional Cultural Commonsense and LLM Bias in India
por: Madhusudan, Sangmitra, et al.
Publicado: (2026)