Saved in:
| Main Authors: | Poulain, Raphael, Fayyaz, Hamed, Beheshti, Rahmatollah |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2408.12055 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bias patterns in the application of LLMs for clinical decision support: A comprehensive study
by: Poulain, Raphael, et al.
Published: (2024)
by: Poulain, Raphael, et al.
Published: (2024)
Enabling Scalable Evaluation of Bias Patterns in Medical LLMs
by: Fayyaz, Hamed, et al.
Published: (2024)
by: Fayyaz, Hamed, et al.
Published: (2024)
Fairness-Optimized Synthetic EHR Generation for Arbitrary Downstream Predictive Tasks
by: Tarek, Mirza Farhan Bin, et al.
Published: (2024)
by: Tarek, Mirza Farhan Bin, et al.
Published: (2024)
Multimodal Sleep Apnea Detection with Missing or Noisy Modalities
by: Fayyaz, Hamed, et al.
Published: (2024)
by: Fayyaz, Hamed, et al.
Published: (2024)
ConTextual: Improving Clinical Text Summarization in LLMs with Context-preserving Token Filtering and Knowledge Graphs
by: Piya, Fahmida Liza, et al.
Published: (2025)
by: Piya, Fahmida Liza, et al.
Published: (2025)
HealthGAT: Node Classifications in Electronic Health Records using Graph Attention Networks
by: Piya, Fahmida Liza, et al.
Published: (2024)
by: Piya, Fahmida Liza, et al.
Published: (2024)
AgenticSum: An Agentic Inference-Time Framework for Faithful Clinical Text Summarization
by: Piya, Fahmida Liza, et al.
Published: (2026)
by: Piya, Fahmida Liza, et al.
Published: (2026)
An Interoperable Machine Learning Pipeline for Pediatric Obesity Risk Estimation
by: Fayyaz, Hamed, et al.
Published: (2024)
by: Fayyaz, Hamed, et al.
Published: (2024)
Reward Hacking Mitigation using Verifiable Composite Rewards
by: Tarek, Mirza Farhan Bin, et al.
Published: (2025)
by: Tarek, Mirza Farhan Bin, et al.
Published: (2025)
Steering MoE LLMs via Expert (De)Activation
by: Fayyaz, Mohsen, et al.
Published: (2025)
by: Fayyaz, Mohsen, et al.
Published: (2025)
Text Mining Analysis of Symptom Patterns in Medical Chatbot Conversations
by: Razavi, Hamed
Published: (2025)
by: Razavi, Hamed
Published: (2025)
HyMaTE: A Hybrid Mamba and Transformer Model for EHR Representation Learning
by: Mottalib, Md Mozaharul, et al.
Published: (2025)
by: Mottalib, Md Mozaharul, et al.
Published: (2025)
The Impossibility of Fair LLMs
by: Anthis, Jacy, et al.
Published: (2024)
by: Anthis, Jacy, et al.
Published: (2024)
Few-Shot Knowledge Distillation of LLMs With Counterfactual Explanations
by: Hamman, Faisal, et al.
Published: (2025)
by: Hamman, Faisal, et al.
Published: (2025)
Aligning LLMs by Predicting Preferences from User Writing Samples
by: Aroca-Ouellette, Stéphane, et al.
Published: (2025)
by: Aroca-Ouellette, Stéphane, et al.
Published: (2025)
ALIEN: Aligned Entropy Head for Improving Uncertainty Estimation of LLMs
by: Zabolotnyi, Artem, et al.
Published: (2025)
by: Zabolotnyi, Artem, et al.
Published: (2025)
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
by: Wang, Yinong Oliver, et al.
Published: (2025)
by: Wang, Yinong Oliver, et al.
Published: (2025)
Toward Revealing Nuanced Biases in Medical LLMs
by: Adiba, Farzana Islam, et al.
Published: (2025)
by: Adiba, Farzana Islam, et al.
Published: (2025)
Group Fairness Meets the Black Box: Enabling Fair Algorithms on Closed LLMs via Post-Processing
by: Xian, Ruicheng, et al.
Published: (2025)
by: Xian, Ruicheng, et al.
Published: (2025)
Counterfactual Evaluation Reveals Hidden Capability Profiles in Clinical LLMs and Agents
by: Turk, Matt
Published: (2026)
by: Turk, Matt
Published: (2026)
BOND: Aligning LLMs with Best-of-N Distillation
by: Sessa, Pier Giuseppe, et al.
Published: (2024)
by: Sessa, Pier Giuseppe, et al.
Published: (2024)
Template-Based Probes Are Imperfect Lenses for Counterfactual Bias Evaluation in LLMs
by: Kohankhaki, Farnaz, et al.
Published: (2024)
by: Kohankhaki, Farnaz, et al.
Published: (2024)
VITAL: A New Dataset for Benchmarking Pluralistic Alignment in Healthcare
by: Shetty, Anudeex, et al.
Published: (2025)
by: Shetty, Anudeex, et al.
Published: (2025)
Counterfactual Generation with Identifiability Guarantees
by: Yan, Hanqi, et al.
Published: (2024)
by: Yan, Hanqi, et al.
Published: (2024)
CALF: Aligning LLMs for Time Series Forecasting via Cross-modal Fine-Tuning
by: Liu, Peiyuan, et al.
Published: (2024)
by: Liu, Peiyuan, et al.
Published: (2024)
Confronting LLMs with Traditional ML: Rethinking the Fairness of Large Language Models in Tabular Classifications
by: Liu, Yanchen, et al.
Published: (2023)
by: Liu, Yanchen, et al.
Published: (2023)
LLMs Don't Know Their Own Decision Boundaries: The Unreliability of Self-Generated Counterfactual Explanations
by: Mayne, Harry, et al.
Published: (2025)
by: Mayne, Harry, et al.
Published: (2025)
Can Public LLMs be used for Self-Diagnosis of Medical Conditions ?
by: Balasubramanian, Nikil Sharan Prabahar, et al.
Published: (2024)
by: Balasubramanian, Nikil Sharan Prabahar, et al.
Published: (2024)
Are the Values of LLMs Structurally Aligned with Humans? A Causal Perspective
by: Kang, Yipeng, et al.
Published: (2024)
by: Kang, Yipeng, et al.
Published: (2024)
Probing Scientific General Intelligence of LLMs with Scientist-Aligned Workflows
by: Xu, Wanghan, et al.
Published: (2025)
by: Xu, Wanghan, et al.
Published: (2025)
Density estimation with LLMs: a geometric investigation of in-context learning trajectories
by: Liu, Toni J. B., et al.
Published: (2024)
by: Liu, Toni J. B., et al.
Published: (2024)
Multilingual Routing in Mixture-of-Experts
by: Bandarkar, Lucas, et al.
Published: (2025)
by: Bandarkar, Lucas, et al.
Published: (2025)
AlignSAE: Concept-Aligned Sparse Autoencoders
by: Yang, Minglai, et al.
Published: (2025)
by: Yang, Minglai, et al.
Published: (2025)
Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs
by: Liu, Hongliang, et al.
Published: (2026)
by: Liu, Hongliang, et al.
Published: (2026)
Compared to What? Baselines and Metrics for Counterfactual Prompting
by: Yang, Zihao, et al.
Published: (2026)
by: Yang, Zihao, et al.
Published: (2026)
Unintended Harms of Value-Aligned LLMs: Psychological and Empirical Insights
by: Choi, Sooyung, et al.
Published: (2025)
by: Choi, Sooyung, et al.
Published: (2025)
Expected Harm: Rethinking Safety Evaluation of (Mis)Aligned LLMs
by: Chen, Yen-Shan, et al.
Published: (2026)
by: Chen, Yen-Shan, et al.
Published: (2026)
H3Fusion: Helpful, Harmless, Honest Fusion of Aligned LLMs
by: Tekin, Selim Furkan, et al.
Published: (2024)
by: Tekin, Selim Furkan, et al.
Published: (2024)
Aligning Frozen LLMs by Reinforcement Learning: An Iterative Reweight-then-Optimize Approach
by: Zhang, Xinnan, et al.
Published: (2025)
by: Zhang, Xinnan, et al.
Published: (2025)
Audit Me If You Can: Query-Efficient Active Fairness Auditing of Black-Box LLMs
by: Hartmann, David, et al.
Published: (2026)
by: Hartmann, David, et al.
Published: (2026)
Similar Items
-
Bias patterns in the application of LLMs for clinical decision support: A comprehensive study
by: Poulain, Raphael, et al.
Published: (2024) -
Enabling Scalable Evaluation of Bias Patterns in Medical LLMs
by: Fayyaz, Hamed, et al.
Published: (2024) -
Fairness-Optimized Synthetic EHR Generation for Arbitrary Downstream Predictive Tasks
by: Tarek, Mirza Farhan Bin, et al.
Published: (2024) -
Multimodal Sleep Apnea Detection with Missing or Noisy Modalities
by: Fayyaz, Hamed, et al.
Published: (2024) -
ConTextual: Improving Clinical Text Summarization in LLMs with Context-preserving Token Filtering and Knowledge Graphs
by: Piya, Fahmida Liza, et al.
Published: (2025)