Enabling Scalable Evaluation of Bias Patterns in Medical LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fayyaz, Hamed, Poulain, Raphael, Beheshti, Rahmatollah |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Aligning (Medical) LLMs for (Counterfactual) Fairness
von: Poulain, Raphael, et al.
Veröffentlicht: (2024)
von: Poulain, Raphael, et al.
Veröffentlicht: (2024)
Bias patterns in the application of LLMs for clinical decision support: A comprehensive study
von: Poulain, Raphael, et al.
Veröffentlicht: (2024)
von: Poulain, Raphael, et al.
Veröffentlicht: (2024)
ConTextual: Improving Clinical Text Summarization in LLMs with Context-preserving Token Filtering and Knowledge Graphs
von: Piya, Fahmida Liza, et al.
Veröffentlicht: (2025)
von: Piya, Fahmida Liza, et al.
Veröffentlicht: (2025)
AgenticSum: An Agentic Inference-Time Framework for Faithful Clinical Text Summarization
von: Piya, Fahmida Liza, et al.
Veröffentlicht: (2026)
von: Piya, Fahmida Liza, et al.
Veröffentlicht: (2026)
Toward Revealing Nuanced Biases in Medical LLMs
von: Adiba, Farzana Islam, et al.
Veröffentlicht: (2025)
von: Adiba, Farzana Islam, et al.
Veröffentlicht: (2025)
A Scalable Entity-Based Framework for Auditing Bias in LLMs
von: Elbouanani, Akram, et al.
Veröffentlicht: (2026)
von: Elbouanani, Akram, et al.
Veröffentlicht: (2026)
Evaluating Human Alignment and Model Faithfulness of LLM Rationale
von: Fayyaz, Mohsen, et al.
Veröffentlicht: (2024)
von: Fayyaz, Mohsen, et al.
Veröffentlicht: (2024)
Structured Outputs Enable General-Purpose LLMs to be Medical Experts
von: Guo, Guangfu, et al.
Veröffentlicht: (2025)
von: Guo, Guangfu, et al.
Veröffentlicht: (2025)
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
von: Hanif, Ikhlasul Akmal, et al.
Veröffentlicht: (2026)
von: Hanif, Ikhlasul Akmal, et al.
Veröffentlicht: (2026)
Evaluating Bias in Spoken Dialogue LLMs for Real-World Decisions and Recommendations
von: Wu, Yihao, et al.
Veröffentlicht: (2025)
von: Wu, Yihao, et al.
Veröffentlicht: (2025)
Bias Beyond Borders: Political Ideology Evaluation and Steering in Multilingual LLMs
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2026)
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2026)
When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation
von: Xu, Wenda, et al.
Veröffentlicht: (2025)
von: Xu, Wenda, et al.
Veröffentlicht: (2025)
CEA-LIST at CheckThat! 2025: Evaluating LLMs as Detectors of Bias and Opinion in Text
von: Elbouanani, Akram, et al.
Veröffentlicht: (2025)
von: Elbouanani, Akram, et al.
Veröffentlicht: (2025)
Man Made Language Models? Evaluating LLMs' Perpetuation of Masculine Generics Bias
von: Doyen, Enzo, et al.
Veröffentlicht: (2025)
von: Doyen, Enzo, et al.
Veröffentlicht: (2025)
Capturing Bias Diversity in LLMs
von: Gosavi, Purva Prasad, et al.
Veröffentlicht: (2024)
von: Gosavi, Purva Prasad, et al.
Veröffentlicht: (2024)
Bring Your Own Prompts: Use-Case-Specific Bias and Fairness Evaluation for LLMs
von: Bouchard, Dylan
Veröffentlicht: (2024)
von: Bouchard, Dylan
Veröffentlicht: (2024)
An Interoperable Machine Learning Pipeline for Pediatric Obesity Risk Estimation
von: Fayyaz, Hamed, et al.
Veröffentlicht: (2024)
von: Fayyaz, Hamed, et al.
Veröffentlicht: (2024)
Evaluate Bias without Manual Test Sets: A Concept Representation Perspective for LLMs
von: Gao, Lang, et al.
Veröffentlicht: (2025)
von: Gao, Lang, et al.
Veröffentlicht: (2025)
Can Large Language Models be Trusted for Evaluation? Scalable Meta-Evaluation of LLMs as Evaluators via Agent Debate
von: Chern, Steffi, et al.
Veröffentlicht: (2024)
von: Chern, Steffi, et al.
Veröffentlicht: (2024)
Cognitive Bias in Decision-Making with LLMs
von: Echterhoff, Jessica, et al.
Veröffentlicht: (2024)
von: Echterhoff, Jessica, et al.
Veröffentlicht: (2024)
Implicit Bias in LLMs: A Survey
von: Lin, Xinru, et al.
Veröffentlicht: (2025)
von: Lin, Xinru, et al.
Veröffentlicht: (2025)
Morphemes Without Borders: Evaluating Root-Pattern Morphology in Arabic Tokenizers and LLMs
von: Alakeel, Yara, et al.
Veröffentlicht: (2026)
von: Alakeel, Yara, et al.
Veröffentlicht: (2026)
Mind the Language Gap: Automated and Augmented Evaluation of Bias in LLMs for High- and Low-Resource Languages
von: Buscemi, Alessio, et al.
Veröffentlicht: (2025)
von: Buscemi, Alessio, et al.
Veröffentlicht: (2025)
From Word to World: Evaluate and Mitigate Culture Bias in LLMs via Word Association Test
von: Dai, Xunlian, et al.
Veröffentlicht: (2025)
von: Dai, Xunlian, et al.
Veröffentlicht: (2025)
Reward Hacking Mitigation using Verifiable Composite Rewards
von: Tarek, Mirza Farhan Bin, et al.
Veröffentlicht: (2025)
von: Tarek, Mirza Farhan Bin, et al.
Veröffentlicht: (2025)
Multi-Agent LLMs for Generating Research Limitations
von: Azher, Ibrahim Al, et al.
Veröffentlicht: (2025)
von: Azher, Ibrahim Al, et al.
Veröffentlicht: (2025)
Mechanics of Bias and Reasoning: Interpreting the Impact of Chain-of-Thought Prompting on Gender Bias in LLMs
von: Pearman, Edie, et al.
Veröffentlicht: (2026)
von: Pearman, Edie, et al.
Veröffentlicht: (2026)
Adapting LLMs for the Medical Domain in Portuguese: A Study on Fine-Tuning and Model Evaluation
von: Paiola, Pedro Henrique, et al.
Veröffentlicht: (2024)
von: Paiola, Pedro Henrique, et al.
Veröffentlicht: (2024)
RAG-based Architectures for Drug Side Effect Retrieval in LLMs
von: Nygren, Shad, et al.
Veröffentlicht: (2025)
von: Nygren, Shad, et al.
Veröffentlicht: (2025)
Are Bias Evaluation Methods Biased ?
von: Berrayana, Lina, et al.
Veröffentlicht: (2025)
von: Berrayana, Lina, et al.
Veröffentlicht: (2025)
Steering Towards Fairness: Mitigating Political Bias in LLMs
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective
von: Chandna, Bhavik, et al.
Veröffentlicht: (2025)
von: Chandna, Bhavik, et al.
Veröffentlicht: (2025)
MedEthicsQA: A Comprehensive Question Answering Benchmark for Medical Ethics Evaluation of LLMs
von: Wei, Jianhui, et al.
Veröffentlicht: (2025)
von: Wei, Jianhui, et al.
Veröffentlicht: (2025)
User-Assistant Bias in LLMs
von: Pan, Xu, et al.
Veröffentlicht: (2025)
von: Pan, Xu, et al.
Veröffentlicht: (2025)
User Behavior Prediction as a Generic, Robust, Scalable, and Low-Cost Evaluation Strategy for Estimating Generalization in LLMs
von: Saha, Sougata, et al.
Veröffentlicht: (2025)
von: Saha, Sougata, et al.
Veröffentlicht: (2025)
BEADs: Bias Evaluation Across Domains
von: Raza, Shaina, et al.
Veröffentlicht: (2024)
von: Raza, Shaina, et al.
Veröffentlicht: (2024)
Bias in the Mirror: Are LLMs opinions robust to their own adversarial attacks ?
von: Rennard, Virgile, et al.
Veröffentlicht: (2024)
von: Rennard, Virgile, et al.
Veröffentlicht: (2024)
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability
von: Raimondi, Bianca, et al.
Veröffentlicht: (2025)
von: Raimondi, Bianca, et al.
Veröffentlicht: (2025)
Framing Political Bias in Multilingual LLMs Across Pakistani Languages
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
von: Nadeem, Afrozah, et al.
Veröffentlicht: (2025)
Reasoning Isn't Enough: Examining Truth-Bias and Sycophancy in LLMs
von: Barkett, Emilio, et al.
Veröffentlicht: (2025)
von: Barkett, Emilio, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Aligning (Medical) LLMs for (Counterfactual) Fairness
von: Poulain, Raphael, et al.
Veröffentlicht: (2024) -
Bias patterns in the application of LLMs for clinical decision support: A comprehensive study
von: Poulain, Raphael, et al.
Veröffentlicht: (2024) -
ConTextual: Improving Clinical Text Summarization in LLMs with Context-preserving Token Filtering and Knowledge Graphs
von: Piya, Fahmida Liza, et al.
Veröffentlicht: (2025) -
AgenticSum: An Agentic Inference-Time Framework for Faithful Clinical Text Summarization
von: Piya, Fahmida Liza, et al.
Veröffentlicht: (2026) -
Toward Revealing Nuanced Biases in Medical LLMs
von: Adiba, Farzana Islam, et al.
Veröffentlicht: (2025)