SAGED: A Holistic Bias-Benchmarking Pipeline for Language Models with Customisable Fairness Calibration
Fuente:
arXiv
Saved in:
| Main Authors: | Guan, Xin, Wang, Ze, Demchak, Nathaniel, Gupta, Saloni, Ertekin Jr., Ediz, Koshiyama, Adriano, Kazim, Emre, Wu, Zekun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Assessing Bias in Metric Models for LLM Open-Ended Generation Bias Benchmarks
by: Demchak, Nathaniel, et al.
Published: (2024)
by: Demchak, Nathaniel, et al.
Published: (2024)
JobFair: A Framework for Benchmarking Gender Hiring Bias in Large Language Models
by: Wang, Ze, et al.
Published: (2024)
by: Wang, Ze, et al.
Published: (2024)
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
by: Jia, Xiao
Published: (2026)
by: Jia, Xiao
Published: (2026)
EQUITRIAGE: A Fairness Audit of Gender Bias in LLM-Based Emergency Department Triage
by: Young, Richard J., et al.
Published: (2026)
by: Young, Richard J., et al.
Published: (2026)
Do LLMs have a Gender (Entropy) Bias?
by: Prabhune, Sonal, et al.
Published: (2025)
by: Prabhune, Sonal, et al.
Published: (2025)
Intersecting Liminality: Acquiring a Smartphone as a Blind or Low Vision Older Adult
by: Figueira, Isabela, et al.
Published: (2024)
by: Figueira, Isabela, et al.
Published: (2024)
Communicative Agents for Slideshow Storytelling Video Generation based on LLMs
by: Fan, Jingxing, et al.
Published: (2025)
by: Fan, Jingxing, et al.
Published: (2025)
HEARTS: A Holistic Framework for Explainable, Sustainable and Robust Text Stereotype Detection
by: King, Theo, et al.
Published: (2024)
by: King, Theo, et al.
Published: (2024)
Who's Asking? Investigating Bias Through the Lens of Disability Framed Queries in LLMs
by: Hari, Vishnu, et al.
Published: (2025)
by: Hari, Vishnu, et al.
Published: (2025)
A Multi-Agent Framework for Medical AI: Leveraging Fine-Tuned GPT, LLaMA, and DeepSeek R1 for Evidence-Based and Bias-Aware Clinical Query Processing
by: Nourmohammadi, Naeimeh, et al.
Published: (2026)
by: Nourmohammadi, Naeimeh, et al.
Published: (2026)
Universal Conditional Logic: A Formal Language for Prompt Engineering
by: Mikinka, Anthony
Published: (2025)
by: Mikinka, Anthony
Published: (2025)
Control Reinforcement Learning: Interpretable Token-Level Steering of LLMs via Sparse Autoencoder Features
by: Cho, Seonglae, et al.
Published: (2026)
by: Cho, Seonglae, et al.
Published: (2026)
CorrSteer: Generation-Time LLM Steering via Correlated Sparse Autoencoder Features
by: Cho, Seonglae, et al.
Published: (2025)
by: Cho, Seonglae, et al.
Published: (2025)
On Matrices Whose Distinct Eigenvalues Are Fully Captured by Quotient Matrices
by: Rather, Bilal Ahmad
Published: (2026)
by: Rather, Bilal Ahmad
Published: (2026)
Robot-Assisted Social Dining as a White Glove Service
by: Kashyap, Atharva S, et al.
Published: (2026)
by: Kashyap, Atharva S, et al.
Published: (2026)
GPT is Not an Annotator: The Necessity of Human Annotation in Fairness Benchmark Construction
by: Felkner, Virginia K., et al.
Published: (2024)
by: Felkner, Virginia K., et al.
Published: (2024)
Compositionality of Rewriting Rules with Conditions
by: Behr, Nicolas, et al.
Published: (2019)
by: Behr, Nicolas, et al.
Published: (2019)
Textual Data Bias Detection and Mitigation -- An Extensible Pipeline with Experimental Evaluation
by: Görge, Rebekka, et al.
Published: (2025)
by: Görge, Rebekka, et al.
Published: (2025)
Particulate organic carbon and nitrogen in bottom water at stations M42/2_362-2 to M42/2_434#2
by: Turnewitsch, Robert
Published: (2001)
by: Turnewitsch, Robert
Published: (2001)
Exact Bias of Linear TRNG Correctors -- Spectral Approach
by: Skorski, Maciej, et al.
Published: (2025)
by: Skorski, Maciej, et al.
Published: (2025)
A Three Steps Methodological Approach to Legal Governance Validation
by: Casanovas, Pompeu, et al.
Published: (2024)
by: Casanovas, Pompeu, et al.
Published: (2024)
Black-box Context-free Grammar Inference for Readable & Natural Grammars
by: Arefin, Mohammad Rifat, et al.
Published: (2025)
by: Arefin, Mohammad Rifat, et al.
Published: (2025)
Ichthyoplankton composition from PC-BELAP cruise 2, Southern Brazilian Shelf, in 1982
by: Muelbert, Jose H, et al.
Published: (2018)
by: Muelbert, Jose H, et al.
Published: (2018)
Fairness in Socio-technical Systems: a Case Study of Wikipedia
by: Damadi, Mir Saeed, et al.
Published: (2023)
by: Damadi, Mir Saeed, et al.
Published: (2023)
AgentBreeder: Mitigating the AI Safety Risks of Multi-Agent Scaffolds via Self-Improvement
by: Rosser, J, et al.
Published: (2025)
by: Rosser, J, et al.
Published: (2025)
Fairness Is Not Enough: Auditing Competence and Intersectional Bias in AI-powered Resume Screening
by: Webster, Kevin T
Published: (2025)
by: Webster, Kevin T
Published: (2025)
Unsolvability Ceiling in Multi-LLM Routing: An Empirical Study of Evaluation Artifacts
by: Garg, Saloni, et al.
Published: (2026)
by: Garg, Saloni, et al.
Published: (2026)
Bias in Surface Electromyography Features across a Demographically Diverse Cohort
by: Agrawal, Aditi, et al.
Published: (2026)
by: Agrawal, Aditi, et al.
Published: (2026)
Does Machine Bring in Extra Bias in Learning? Approximating Fairness in Models Promptly
by: Bian, Yijun, et al.
Published: (2024)
by: Bian, Yijun, et al.
Published: (2024)
Bias Amplification: Large Language Models as Increasingly Biased Media
by: Wang, Ze, et al.
Published: (2024)
by: Wang, Ze, et al.
Published: (2024)
Thorium 234 in bottom water on a profile between stations 13200-048 to M42/2_434-1
by: Turnewitsch, Robert
Published: (2001)
by: Turnewitsch, Robert
Published: (2001)
Tool Receipts, Not Zero-Knowledge Proofs: Practical Hallucination Detection for AI Agents
by: Basu, Abhinaba
Published: (2026)
by: Basu, Abhinaba
Published: (2026)
The Strategic Foresight of LLMs: Evidence from a Fully Prospective Venture Tournament
by: Csaszar, Felipe A., et al.
Published: (2026)
by: Csaszar, Felipe A., et al.
Published: (2026)
Drama Engine: A Framework for Narrative Agents
by: Pichlmair, Martin, et al.
Published: (2024)
by: Pichlmair, Martin, et al.
Published: (2024)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
by: Huo, Dongjie, et al.
Published: (2026)
by: Huo, Dongjie, et al.
Published: (2026)
Benchmarking Bengali Dialectal Bias: A Multi-Stage Framework Integrating RAG-Based Translation and Human-Augmented RLAIF
by: Sami, K. M. Jubair, et al.
Published: (2026)
by: Sami, K. M. Jubair, et al.
Published: (2026)
Geist in the Machine: Simulating Recognition and Inner Dialogue in AI-Mediated Teaching and Research
by: Magee, Liam
Published: (2026)
by: Magee, Liam
Published: (2026)
Pure Data Spaces
by: Youssef, Saul
Published: (2025)
by: Youssef, Saul
Published: (2025)
Textual Entailment is not a Better Bias Metric than Token Probability
by: Felkner, Virginia K., et al.
Published: (2025)
by: Felkner, Virginia K., et al.
Published: (2025)
Detection of Edges in Spectral Data II. Nonlinear Enhancement
by: Gelb, Anne, et al.
Published: (2001)
by: Gelb, Anne, et al.
Published: (2001)
Similar Items
-
Assessing Bias in Metric Models for LLM Open-Ended Generation Bias Benchmarks
by: Demchak, Nathaniel, et al.
Published: (2024) -
JobFair: A Framework for Benchmarking Gender Hiring Bias in Large Language Models
by: Wang, Ze, et al.
Published: (2024) -
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
by: Jia, Xiao
Published: (2026) -
EQUITRIAGE: A Fairness Audit of Gender Bias in LLM-Based Emergency Department Triage
by: Young, Richard J., et al.
Published: (2026) -
Do LLMs have a Gender (Entropy) Bias?
by: Prabhune, Sonal, et al.
Published: (2025)