Bring Your Own Prompts: Use-Case-Specific Bias and Fairness Evaluation for LLMs
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Bouchard, Dylan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LangFair: A Python Package for Assessing Bias and Fairness in Large Language Model Use Cases
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025)
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025)
Is Escalation Worth It? A Decision-Theoretic Characterization of LLM Cascades
von: Bouchard, Dylan
Veröffentlicht: (2026)
von: Bouchard, Dylan
Veröffentlicht: (2026)
Bring Your Own KG: Self-Supervised Program Synthesis for Zero-Shot KGQA
von: Agarwal, Dhruv, et al.
Veröffentlicht: (2023)
von: Agarwal, Dhruv, et al.
Veröffentlicht: (2023)
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
von: Wang, Yinong Oliver, et al.
Veröffentlicht: (2025)
von: Wang, Yinong Oliver, et al.
Veröffentlicht: (2025)
Position: LLMs Must Use Functor-Based and RAG-Driven Bias Mitigation for Fairness
von: Ranjan, Ravi, et al.
Veröffentlicht: (2026)
von: Ranjan, Ravi, et al.
Veröffentlicht: (2026)
BYOL: Bring Your Own Language Into LLMs
von: Zamir, Syed Waqas, et al.
Veröffentlicht: (2026)
von: Zamir, Syed Waqas, et al.
Veröffentlicht: (2026)
Steering Towards Fairness: Mitigating Political Bias in LLMs
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
Take Care of Your Prompt Bias! Investigating and Mitigating Prompt Bias in Factual Knowledge Extraction
von: Xu, Ziyang, et al.
Veröffentlicht: (2024)
von: Xu, Ziyang, et al.
Veröffentlicht: (2024)
Your AI, Not Your View: The Bias of LLMs in Investment Analysis
von: Lee, Hoyoung, et al.
Veröffentlicht: (2025)
von: Lee, Hoyoung, et al.
Veröffentlicht: (2025)
Where to show Demos in Your Prompt: A Positional Bias of In-Context Learning
von: Cobbina, Kwesi, et al.
Veröffentlicht: (2025)
von: Cobbina, Kwesi, et al.
Veröffentlicht: (2025)
Uncertainty Quantification for Language Models: A Suite of Black-Box, White-Box, LLM Judge, and Ensemble Scorers
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025)
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025)
Addressing Bias in LLMs: Strategies and Application to Fair AI-based Recruitment
von: Peña, Alejandro, et al.
Veröffentlicht: (2025)
von: Peña, Alejandro, et al.
Veröffentlicht: (2025)
Mechanics of Bias and Reasoning: Interpreting the Impact of Chain-of-Thought Prompting on Gender Bias in LLMs
von: Pearman, Edie, et al.
Veröffentlicht: (2026)
von: Pearman, Edie, et al.
Veröffentlicht: (2026)
Do LLMs Benefit From Their Own Words?
von: Huang, Jenny Y., et al.
Veröffentlicht: (2026)
von: Huang, Jenny Y., et al.
Veröffentlicht: (2026)
MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models
von: Yu, Longhui, et al.
Veröffentlicht: (2023)
von: Yu, Longhui, et al.
Veröffentlicht: (2023)
Fairness Evaluation and Inference Level Mitigation in LLMs
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
Prove Your Point!: Bringing Proof-Enhancement Principles to Argumentative Essay Generation
von: Xiao, Ruiyu, et al.
Veröffentlicht: (2024)
von: Xiao, Ruiyu, et al.
Veröffentlicht: (2024)
If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs
von: Fan, Siqi, et al.
Veröffentlicht: (2025)
von: Fan, Siqi, et al.
Veröffentlicht: (2025)
LLM Evaluators Recognize and Favor Their Own Generations
von: Panickssery, Arjun, et al.
Veröffentlicht: (2024)
von: Panickssery, Arjun, et al.
Veröffentlicht: (2024)
Qworld: Question-Specific Evaluation Criteria for LLMs
von: Gao, Shanghua, et al.
Veröffentlicht: (2026)
von: Gao, Shanghua, et al.
Veröffentlicht: (2026)
Enabling Scalable Evaluation of Bias Patterns in Medical LLMs
von: Fayyaz, Hamed, et al.
Veröffentlicht: (2024)
von: Fayyaz, Hamed, et al.
Veröffentlicht: (2024)
From Long to Short: LLMs Excel at Trimming Own Reasoning Chains
von: Han, Wei, et al.
Veröffentlicht: (2025)
von: Han, Wei, et al.
Veröffentlicht: (2025)
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
von: Hanif, Ikhlasul Akmal, et al.
Veröffentlicht: (2026)
von: Hanif, Ikhlasul Akmal, et al.
Veröffentlicht: (2026)
Bias Similarity Measurement: A Black-Box Audit of Fairness Across LLMs
von: Jeong, Hyejun, et al.
Veröffentlicht: (2024)
von: Jeong, Hyejun, et al.
Veröffentlicht: (2024)
Common Sense vs. Morality: The Curious Case of Narrative Focus Bias in LLMs
von: Purkayastha, Saugata, et al.
Veröffentlicht: (2026)
von: Purkayastha, Saugata, et al.
Veröffentlicht: (2026)
Position is Power: System Prompts as a Mechanism of Bias in Large Language Models (LLMs)
von: Neumann, Anna, et al.
Veröffentlicht: (2025)
von: Neumann, Anna, et al.
Veröffentlicht: (2025)
Evaluating Bias in Spoken Dialogue LLMs for Real-World Decisions and Recommendations
von: Wu, Yihao, et al.
Veröffentlicht: (2025)
von: Wu, Yihao, et al.
Veröffentlicht: (2025)
Bias Beyond Borders: Political Ideology Evaluation and Steering in Multilingual LLMs
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2026)
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2026)
When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation
von: Xu, Wenda, et al.
Veröffentlicht: (2025)
von: Xu, Wenda, et al.
Veröffentlicht: (2025)
Graph Your Own Prompt
von: Ding, Xi, et al.
Veröffentlicht: (2025)
von: Ding, Xi, et al.
Veröffentlicht: (2025)
FairTranslate: An English-French Dataset for Gender Bias Evaluation in Machine Translation by Overcoming Gender Binarity
von: Jourdan, Fanny, et al.
Veröffentlicht: (2025)
von: Jourdan, Fanny, et al.
Veröffentlicht: (2025)
Metamorphic Testing for Fairness Evaluation in Large Language Models: Identifying Intersectional Bias in LLaMA and GPT
von: Reddy, Harishwar, et al.
Veröffentlicht: (2025)
von: Reddy, Harishwar, et al.
Veröffentlicht: (2025)
CEA-LIST at CheckThat! 2025: Evaluating LLMs as Detectors of Bias and Opinion in Text
von: Elbouanani, Akram, et al.
Veröffentlicht: (2025)
von: Elbouanani, Akram, et al.
Veröffentlicht: (2025)
Man Made Language Models? Evaluating LLMs' Perpetuation of Masculine Generics Bias
von: Doyen, Enzo, et al.
Veröffentlicht: (2025)
von: Doyen, Enzo, et al.
Veröffentlicht: (2025)
Are Your LLMs Capable of Stable Reasoning?
von: Liu, Junnan, et al.
Veröffentlicht: (2024)
von: Liu, Junnan, et al.
Veröffentlicht: (2024)
Capturing Bias Diversity in LLMs
von: Gosavi, Purva Prasad, et al.
Veröffentlicht: (2024)
von: Gosavi, Purva Prasad, et al.
Veröffentlicht: (2024)
How Well do LLMs Compress Their Own Chain-of-Thought? A Token Complexity Approach
von: Lee, Ayeong, et al.
Veröffentlicht: (2025)
von: Lee, Ayeong, et al.
Veröffentlicht: (2025)
Fine-Grained Uncertainty Quantification for Long-Form Language Model Outputs: A Comparative Study
von: Bouchard, Dylan, et al.
Veröffentlicht: (2026)
von: Bouchard, Dylan, et al.
Veröffentlicht: (2026)
From Guidelines to Guarantees: A Graph-Based Evaluation Harness for Domain-Specific Evaluation of LLMs
von: Lundin, Jessica M., et al.
Veröffentlicht: (2025)
von: Lundin, Jessica M., et al.
Veröffentlicht: (2025)
StealthGraph: Exposing Domain-Specific Risks in LLMs through Knowledge-Graph-Guided Harmful Prompt Generation
von: Zheng, Huawei, et al.
Veröffentlicht: (2026)
von: Zheng, Huawei, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
LangFair: A Python Package for Assessing Bias and Fairness in Large Language Model Use Cases
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025) -
Is Escalation Worth It? A Decision-Theoretic Characterization of LLM Cascades
von: Bouchard, Dylan
Veröffentlicht: (2026) -
Bring Your Own KG: Self-Supervised Program Synthesis for Zero-Shot KGQA
von: Agarwal, Dhruv, et al.
Veröffentlicht: (2023) -
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
von: Wang, Yinong Oliver, et al.
Veröffentlicht: (2025) -
Position: LLMs Must Use Functor-Based and RAG-Driven Bias Mitigation for Fairness
von: Ranjan, Ravi, et al.
Veröffentlicht: (2026)