Mitigating Social Desirability Bias in Random Silicon Sampling
Fuente:
arXiv
Salvato in:
| Autori principali: | Chapala, Sashank, Mironov, Maksym, Deng, Songgaojun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AdversaRiskQA: An Adversarial Factuality Benchmark for High-Risk Domains
di: Szelestey, Adam, et al.
Pubblicazione: (2026)
di: Szelestey, Adam, et al.
Pubblicazione: (2026)
Automated Item Neutralization for Non-Cognitive Scales: A Large Language Model Approach to Reducing Social-Desirability Bias
di: Wu, Sirui, et al.
Pubblicazione: (2025)
di: Wu, Sirui, et al.
Pubblicazione: (2025)
Leveraging Prototypical Representations for Mitigating Social Bias without Demographic Information
di: Iskander, Shadi, et al.
Pubblicazione: (2024)
di: Iskander, Shadi, et al.
Pubblicazione: (2024)
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
di: Borah, Angana, et al.
Pubblicazione: (2024)
di: Borah, Angana, et al.
Pubblicazione: (2024)
DSO: Direct Steering Optimization for Bias Mitigation
di: Paes, Lucas Monteiro, et al.
Pubblicazione: (2025)
di: Paes, Lucas Monteiro, et al.
Pubblicazione: (2025)
Fairness through Difference Awareness: Measuring Desired Group Discrimination in LLMs
di: Wang, Angelina, et al.
Pubblicazione: (2025)
di: Wang, Angelina, et al.
Pubblicazione: (2025)
Exploring the Linear Subspace Hypothesis in Gender Bias Mitigation
di: Vargas, Francisco, et al.
Pubblicazione: (2020)
di: Vargas, Francisco, et al.
Pubblicazione: (2020)
Beyond Natural Language Plans: Structure-Aware Planning for Query-Focused Table Summarization
di: Zhang, Weijia, et al.
Pubblicazione: (2025)
di: Zhang, Weijia, et al.
Pubblicazione: (2025)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
di: Ma, Mingyu Derek, et al.
Pubblicazione: (2023)
di: Ma, Mingyu Derek, et al.
Pubblicazione: (2023)
Mitigating Gender Bias via Fostering Exploratory Thinking in LLMs
di: Wei, Kangda, et al.
Pubblicazione: (2025)
di: Wei, Kangda, et al.
Pubblicazione: (2025)
Fair Representation in Parliamentary Summaries: Measuring and Mitigating Inclusion Bias
di: Cunningham, Eoghan, et al.
Pubblicazione: (2025)
di: Cunningham, Eoghan, et al.
Pubblicazione: (2025)
Bike Frames: Understanding the Implicit Portrayal of Cyclists in the News
di: Zhao, Xingmeng, et al.
Pubblicazione: (2023)
di: Zhao, Xingmeng, et al.
Pubblicazione: (2023)
Tackling Social Bias against the Poor: A Dataset and Taxonomy on Aporophobia
di: Curto, Georgina, et al.
Pubblicazione: (2025)
di: Curto, Georgina, et al.
Pubblicazione: (2025)
No Free Lunch in Language Model Bias Mitigation? Targeted Bias Reduction Can Exacerbate Unmitigated LLM Biases
di: Chand, Shireen, et al.
Pubblicazione: (2025)
di: Chand, Shireen, et al.
Pubblicazione: (2025)
Stars, Stripes, and Silicon: Unravelling the ChatGPT's All-American, Monochrome, Cis-centric Bias
di: Torrielli, Federico
Pubblicazione: (2024)
di: Torrielli, Federico
Pubblicazione: (2024)
Large Language Models Show Human-like Social Desirability Biases in Survey Responses
di: Salecha, Aadesh, et al.
Pubblicazione: (2024)
di: Salecha, Aadesh, et al.
Pubblicazione: (2024)
VIGIL: An Extensible System for Real-Time Detection and Mitigation of Cognitive Bias Triggers
di: Kang, Bo, et al.
Pubblicazione: (2026)
di: Kang, Bo, et al.
Pubblicazione: (2026)
InsideOut: Measuring and Mitigating Insider-Outsider Bias in Interview Script Generation
di: Wan, Yixin, et al.
Pubblicazione: (2025)
di: Wan, Yixin, et al.
Pubblicazione: (2025)
Inference-Time Reasoning Selectively Reduces Implicit Social Bias in Large Language Models
di: Apsel, Molly, et al.
Pubblicazione: (2026)
di: Apsel, Molly, et al.
Pubblicazione: (2026)
LLM Bias Detection and Mitigation through the Lens of Desired Distributions
di: Shrestha, Ingroj, et al.
Pubblicazione: (2025)
di: Shrestha, Ingroj, et al.
Pubblicazione: (2025)
Social Bias in Popular Question-Answering Benchmarks
di: Kraft, Angelie, et al.
Pubblicazione: (2025)
di: Kraft, Angelie, et al.
Pubblicazione: (2025)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
di: Xu, Xin, et al.
Pubblicazione: (2025)
di: Xu, Xin, et al.
Pubblicazione: (2025)
Covert Bias: The Severity of Social Views' Unalignment in Language Models Towards Implicit and Explicit Opinion
di: Aldayel, Abeer, et al.
Pubblicazione: (2024)
di: Aldayel, Abeer, et al.
Pubblicazione: (2024)
Intrinsic Meets Extrinsic Fairness: Assessing the Downstream Impact of Bias Mitigation in Large Language Models
di: Arzaghi', 'Mina, et al.
Pubblicazione: (2025)
di: Arzaghi', 'Mina, et al.
Pubblicazione: (2025)
Counterspeech for Mitigating the Influence of Media Bias: Comparing Human and LLM-Generated Responses
di: Lin, Luyang, et al.
Pubblicazione: (2025)
di: Lin, Luyang, et al.
Pubblicazione: (2025)
The Lifecycle of "Facts": A Survey of Social Bias in Knowledge Graphs
di: Kraft, Angelie, et al.
Pubblicazione: (2022)
di: Kraft, Angelie, et al.
Pubblicazione: (2022)
Rethinking LLM Bias Probing Using Lessons from the Social Sciences
di: Morehouse, Kirsten N., et al.
Pubblicazione: (2025)
di: Morehouse, Kirsten N., et al.
Pubblicazione: (2025)
AesBiasBench: Evaluating Bias and Alignment in Multimodal Language Models for Personalized Image Aesthetic Assessment
di: Li, Kun, et al.
Pubblicazione: (2025)
di: Li, Kun, et al.
Pubblicazione: (2025)
The Silicon Ceiling: Auditing GPT's Race and Gender Biases in Hiring
di: Armstrong, Lena, et al.
Pubblicazione: (2024)
di: Armstrong, Lena, et al.
Pubblicazione: (2024)
ChatGPT is not A Man but Das Man: Representativeness and Structural Consistency of Silicon Samples Generated by Large Language Models
di: Li, Dai, et al.
Pubblicazione: (2025)
di: Li, Dai, et al.
Pubblicazione: (2025)
On The Conceptualization and Societal Impact of Cross-Cultural Bias
di: Bhandari, Vitthal
Pubblicazione: (2025)
di: Bhandari, Vitthal
Pubblicazione: (2025)
Media Bias Matters: Understanding the Impact of Politically Biased News on Vaccine Attitudes in Social Media
di: Jiang, Bohan, et al.
Pubblicazione: (2024)
di: Jiang, Bohan, et al.
Pubblicazione: (2024)
Evaluating LLMs for Demographic-Targeted Social Bias Detection: A Comprehensive Benchmark Study
di: Majumdar, Ayan, et al.
Pubblicazione: (2025)
di: Majumdar, Ayan, et al.
Pubblicazione: (2025)
Unsupervised Concept Vector Extraction for Bias Control in LLMs
di: Cyberey, Hannah, et al.
Pubblicazione: (2025)
di: Cyberey, Hannah, et al.
Pubblicazione: (2025)
Characterizing Selective Refusal Bias in Large Language Models
di: Khorramrouz, Adel, et al.
Pubblicazione: (2025)
di: Khorramrouz, Adel, et al.
Pubblicazione: (2025)
Gender Bias in Emotion Recognition by Large Language Models
di: Herbert, Maureen, et al.
Pubblicazione: (2025)
di: Herbert, Maureen, et al.
Pubblicazione: (2025)
Theories of "Sexuality" in Natural Language Processing Bias Research
di: Hobbs, Jacob
Pubblicazione: (2025)
di: Hobbs, Jacob
Pubblicazione: (2025)
Counterfactual Probing for the Influence of Affect and Specificity on Intergroup Bias
di: Govindarajan, Venkata S, et al.
Pubblicazione: (2023)
di: Govindarajan, Venkata S, et al.
Pubblicazione: (2023)
The Collapse of Heterogeneity in Silicon Philosophers
di: Shi, Yuanming, et al.
Pubblicazione: (2026)
di: Shi, Yuanming, et al.
Pubblicazione: (2026)
Assessing Judging Bias in Large Reasoning Models: An Empirical Study
di: Wang, Qian, et al.
Pubblicazione: (2025)
di: Wang, Qian, et al.
Pubblicazione: (2025)
Documenti analoghi
-
AdversaRiskQA: An Adversarial Factuality Benchmark for High-Risk Domains
di: Szelestey, Adam, et al.
Pubblicazione: (2026) -
Automated Item Neutralization for Non-Cognitive Scales: A Large Language Model Approach to Reducing Social-Desirability Bias
di: Wu, Sirui, et al.
Pubblicazione: (2025) -
Leveraging Prototypical Representations for Mitigating Social Bias without Demographic Information
di: Iskander, Shadi, et al.
Pubblicazione: (2024) -
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
di: Borah, Angana, et al.
Pubblicazione: (2024) -
DSO: Direct Steering Optimization for Bias Mitigation
di: Paes, Lucas Monteiro, et al.
Pubblicazione: (2025)