Invisible Influences: Investigating Implicit Intersectional Biases through Persona Engineering in Large Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Arimanda, Nandini, Mukund, Achyuth, Muthiah, Sakthi Balan, Sharma, Rajesh |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Entangled in Representations: Mechanistic Investigation of Cultural Biases in Large Language Models
par: Yu, Haeun, et autres
Publié: (2025)
par: Yu, Haeun, et autres
Publié: (2025)
Mapping and Influencing the Political Ideology of Large Language Models using Synthetic Personas
par: Bernardelle, Pietro, et autres
Publié: (2024)
par: Bernardelle, Pietro, et autres
Publié: (2024)
Investigating Intersectional Bias in Large Language Models using Confidence Disparities in Coreference Resolution
par: Khan, Falaah Arif, et autres
Publié: (2025)
par: Khan, Falaah Arif, et autres
Publié: (2025)
Large Language Models are Biased Because They Are Large Language Models
par: Resnik, Philip
Publié: (2024)
par: Resnik, Philip
Publié: (2024)
ImplicitBBQ: Benchmarking Implicit Bias in Large Language Models through Characteristic Based Cues
par: Vedula, Bhaskara Hanuma, et autres
Publié: (2026)
par: Vedula, Bhaskara Hanuma, et autres
Publié: (2026)
Personas within Parameters: Fine-Tuning Small Language Models with Low-Rank Adapters to Mimic User Behaviors
par: Thakur, Himanshu, et autres
Publié: (2025)
par: Thakur, Himanshu, et autres
Publié: (2025)
PersonaLLM: Investigating the Ability of Large Language Models to Express Personality Traits
par: Jiang, Hang, et autres
Publié: (2023)
par: Jiang, Hang, et autres
Publié: (2023)
Investigating Implicit Bias in Large Language Models: A Large-Scale Study of Over 50 LLMs
par: Kumar, Divyanshu, et autres
Publié: (2024)
par: Kumar, Divyanshu, et autres
Publié: (2024)
The African Woman is Rhythmic and Soulful: An Investigation of Implicit Biases in LLM Open-ended Text Generation
par: Lim, Serene, et autres
Publié: (2024)
par: Lim, Serene, et autres
Publié: (2024)
Mitigating Social Biases in Language Models through Unlearning
par: Dige, Omkar, et autres
Publié: (2024)
par: Dige, Omkar, et autres
Publié: (2024)
Evaluating Implicit Biases in LLM Reasoning through Logic Grid Puzzles
par: Jahara, Fatima, et autres
Publié: (2025)
par: Jahara, Fatima, et autres
Publié: (2025)
"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas
par: Ding, Junchen, et autres
Publié: (2025)
par: Ding, Junchen, et autres
Publié: (2025)
Identifying Implicit Social Biases in Vision-Language Models
par: Hamidieh, Kimia, et autres
Publié: (2024)
par: Hamidieh, Kimia, et autres
Publié: (2024)
Multilingual Performance Biases of Large Language Models in Education
par: Gupta, Vansh, et autres
Publié: (2025)
par: Gupta, Vansh, et autres
Publié: (2025)
"The Dentist is an involved parent, the bartender is not": Revealing Implicit Biases in QA with Implicit BBQ
par: Wagh, Aarushi, et autres
Publié: (2025)
par: Wagh, Aarushi, et autres
Publié: (2025)
Whose Journey Matters? Investigating Identity Biases in Large Language Models (LLMs) for Travel Planning Assistance
par: Ren, Ruiping, et autres
Publié: (2024)
par: Ren, Ruiping, et autres
Publié: (2024)
Multi-Persona Thinking for Bias Mitigation in Large Language Models
par: Chen, Yuxing, et autres
Publié: (2026)
par: Chen, Yuxing, et autres
Publié: (2026)
A Systematic Analysis of Biases in Large Language Models
par: Zhang, Xulang, et autres
Publié: (2025)
par: Zhang, Xulang, et autres
Publié: (2025)
The Impact of Steering Large Language Models with Persona Vectors in Educational Applications
par: Wu, Yongchao, et autres
Publié: (2026)
par: Wu, Yongchao, et autres
Publié: (2026)
ImF: Implicit Fingerprint for Large Language Models
par: Wu, Jiaxuan, et autres
Publié: (2025)
par: Wu, Jiaxuan, et autres
Publié: (2025)
DateLogicQA: Benchmarking Temporal Biases in Large Language Models
par: Bhatia, Gagan, et autres
Publié: (2024)
par: Bhatia, Gagan, et autres
Publié: (2024)
Cognitive Biases in Large Language Models: A Survey and Mitigation Experiments
par: Sumita, Yasuaki, et autres
Publié: (2024)
par: Sumita, Yasuaki, et autres
Publié: (2024)
Large Language Models are Geographically Biased
par: Manvi, Rohin, et autres
Publié: (2024)
par: Manvi, Rohin, et autres
Publié: (2024)
Invisible Filters: Cultural Bias in Hiring Evaluations Using Large Language Models
par: Rao, Pooja S. B., et autres
Publié: (2025)
par: Rao, Pooja S. B., et autres
Publié: (2025)
The Persona Paradox: Medical Personas as Behavioral Priors in Clinical Language Models
par: Abdullahi, Tassallah, et autres
Publié: (2026)
par: Abdullahi, Tassallah, et autres
Publié: (2026)
Relative Value Biases in Large Language Models
par: Hayes, William M., et autres
Publié: (2024)
par: Hayes, William M., et autres
Publié: (2024)
Large Language Models are Biased Reinforcement Learners
par: Hayes, William M., et autres
Publié: (2024)
par: Hayes, William M., et autres
Publié: (2024)
Actions Speak Louder than Words: Agent Decisions Reveal Implicit Biases in Language Models
par: Li, Yuxuan, et autres
Publié: (2025)
par: Li, Yuxuan, et autres
Publié: (2025)
PRISM: A Methodology for Auditing Biases in Large Language Models
par: Azzopardi, Leif, et autres
Publié: (2024)
par: Azzopardi, Leif, et autres
Publié: (2024)
Persona-Assigned Large Language Models Exhibit Human-Like Motivated Reasoning
par: Dash, Saloni, et autres
Publié: (2025)
par: Dash, Saloni, et autres
Publié: (2025)
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration
par: Wang, Zhenhailong, et autres
Publié: (2023)
par: Wang, Zhenhailong, et autres
Publié: (2023)
Comparing Human and Large Language Model Interpretation of Implicit Information
par: De Santis, Antonio, et autres
Publié: (2026)
par: De Santis, Antonio, et autres
Publié: (2026)
Implicit Reasoning in Large Language Models: A Comprehensive Survey
par: Li, Jindong, et autres
Publié: (2025)
par: Li, Jindong, et autres
Publié: (2025)
Intersectional Bias in Japanese Large Language Models from a Contextualized Perspective
par: Yanaka, Hitomi, et autres
Publié: (2025)
par: Yanaka, Hitomi, et autres
Publié: (2025)
Unveiling Selection Biases: Exploring Order and Token Sensitivity in Large Language Models
par: Wei, Sheng-Lun, et autres
Publié: (2024)
par: Wei, Sheng-Lun, et autres
Publié: (2024)
Intent-Aware Self-Correction for Mitigating Social Biases in Large Language Models
par: Anantaprayoon, Panatchakorn, et autres
Publié: (2025)
par: Anantaprayoon, Panatchakorn, et autres
Publié: (2025)
(Ir)rationality and Cognitive Biases in Large Language Models
par: Macmillan-Scott, Olivia, et autres
Publié: (2024)
par: Macmillan-Scott, Olivia, et autres
Publié: (2024)
Location Not Found: Exposing Implicit Local and Global Biases in Multilingual LLMs
par: Mor-Lan, Guy, et autres
Publié: (2026)
par: Mor-Lan, Guy, et autres
Publié: (2026)
BILLY: Steering Large Language Models via Merging Persona Vectors for Creative Generation
par: Pai, Tsung-Min, et autres
Publié: (2025)
par: Pai, Tsung-Min, et autres
Publié: (2025)
PICLe: Eliciting Diverse Behaviors from Large Language Models with Persona In-Context Learning
par: Choi, Hyeong Kyu, et autres
Publié: (2024)
par: Choi, Hyeong Kyu, et autres
Publié: (2024)
Documents similaires
-
Entangled in Representations: Mechanistic Investigation of Cultural Biases in Large Language Models
par: Yu, Haeun, et autres
Publié: (2025) -
Mapping and Influencing the Political Ideology of Large Language Models using Synthetic Personas
par: Bernardelle, Pietro, et autres
Publié: (2024) -
Investigating Intersectional Bias in Large Language Models using Confidence Disparities in Coreference Resolution
par: Khan, Falaah Arif, et autres
Publié: (2025) -
Large Language Models are Biased Because They Are Large Language Models
par: Resnik, Philip
Publié: (2024) -
ImplicitBBQ: Benchmarking Implicit Bias in Large Language Models through Characteristic Based Cues
par: Vedula, Bhaskara Hanuma, et autres
Publié: (2026)