Saved in:
| Main Authors: | Bommasani, Rishi, Liang, Percy |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2212.11672 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Ecosystem Graphs: The Social Footprint of Foundation Models
by: Bommasani, Rishi, et al.
Published: (2023)
by: Bommasani, Rishi, et al.
Published: (2023)
The Societal Impact of Foundation Models: Advancing Evidence-based AI Policy
by: Bommasani, Rishi
Published: (2025)
by: Bommasani, Rishi
Published: (2025)
Algorithmic Monocultures in Hiring
by: Bommasani, Rishi, et al.
Published: (2026)
by: Bommasani, Rishi, et al.
Published: (2026)
NeurIPS should lead scientific consensus on AI policy
by: Bommasani, Rishi
Published: (2025)
by: Bommasani, Rishi
Published: (2025)
Ecosystem-level Analysis of Deployed Machine Learning Reveals Homogeneous Outcomes
by: Toups, Connor, et al.
Published: (2023)
by: Toups, Connor, et al.
Published: (2023)
The 2024 Foundation Model Transparency Index
by: Bommasani, Rishi, et al.
Published: (2024)
by: Bommasani, Rishi, et al.
Published: (2024)
Replaying pre-training data improves fine-tuning
by: Kotha, Suhas, et al.
Published: (2026)
by: Kotha, Suhas, et al.
Published: (2026)
Self-Verified Distillation: Your Language Model Is Secretly Its Own Synthetic Data Pipeline
by: Lee, Tony, et al.
Published: (2026)
by: Lee, Tony, et al.
Published: (2026)
Language model developers should report train-test overlap
by: Zhang, Andy K, et al.
Published: (2024)
by: Zhang, Andy K, et al.
Published: (2024)
Evaluating Human-Language Model Interaction
by: Lee, Mina, et al.
Published: (2022)
by: Lee, Mina, et al.
Published: (2022)
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences
by: Shrivastava, Vaishnavi, et al.
Published: (2025)
by: Shrivastava, Vaishnavi, et al.
Published: (2025)
Foundation Model Transparency Reports
by: Bommasani, Rishi, et al.
Published: (2024)
by: Bommasani, Rishi, et al.
Published: (2024)
The 2025 Foundation Model Transparency Index
by: Wan, Alexander, et al.
Published: (2025)
by: Wan, Alexander, et al.
Published: (2025)
Trustworthy AI: Safety, Bias, and Privacy -- A Survey
by: Fang, Xingli, et al.
Published: (2025)
by: Fang, Xingli, et al.
Published: (2025)
VisBias: Measuring Explicit and Implicit Social Biases in Vision Language Models
by: Huang, Jen-tse, et al.
Published: (2025)
by: Huang, Jen-tse, et al.
Published: (2025)
Relative Scaling Laws for LLMs
by: Held, William, et al.
Published: (2025)
by: Held, William, et al.
Published: (2025)
IndiBias: A Benchmark Dataset to Measure Social Biases in Language Models for Indian Context
by: Sahoo, Nihar Ranjan, et al.
Published: (2024)
by: Sahoo, Nihar Ranjan, et al.
Published: (2024)
Measuring and Enhancing Trustworthiness of LLMs in RAG through Grounded Attributions and Learning to Refuse
by: Song, Maojia, et al.
Published: (2024)
by: Song, Maojia, et al.
Published: (2024)
What's Not Said Still Hurts: A Description-Based Evaluation Framework for Measuring Social Bias in LLMs
by: Pan, Jinhao, et al.
Published: (2025)
by: Pan, Jinhao, et al.
Published: (2025)
On the Entropy Calibration of Language Models
by: Cao, Steven, et al.
Published: (2025)
by: Cao, Steven, et al.
Published: (2025)
Do AI Companies Make Good on Voluntary Commitments to the White House?
by: Wang, Jennifer, et al.
Published: (2025)
by: Wang, Jennifer, et al.
Published: (2025)
Independence Tests for Language Models
by: Zhu, Sally, et al.
Published: (2025)
by: Zhu, Sally, et al.
Published: (2025)
BiasAlert: A Plug-and-play Tool for Social Bias Detection in LLMs
by: Fan, Zhiting, et al.
Published: (2024)
by: Fan, Zhiting, et al.
Published: (2024)
Bias in News Summarization: Measures, Pitfalls and Corpora
by: Steen, Julius, et al.
Published: (2023)
by: Steen, Julius, et al.
Published: (2023)
SpecEval: Evaluating Model Adherence to Behavior Specifications
by: Ahmed, Ahmed, et al.
Published: (2025)
by: Ahmed, Ahmed, et al.
Published: (2025)
Instruction Following without Instruction Tuning
by: Hewitt, John, et al.
Published: (2024)
by: Hewitt, John, et al.
Published: (2024)
Ask LLMs Directly, "What shapes your bias?": Measuring Social Bias in Large Language Models
by: Shin, Jisu, et al.
Published: (2024)
by: Shin, Jisu, et al.
Published: (2024)
Social Bias in Large Language Models For Bangla: An Empirical Study on Gender and Religious Bias
by: Sadhu, Jayanta, et al.
Published: (2024)
by: Sadhu, Jayanta, et al.
Published: (2024)
Measuring Bias or Measuring the Task: Understanding the Brittle Nature of LLM Gender Biases
by: Gao, Bufan, et al.
Published: (2025)
by: Gao, Bufan, et al.
Published: (2025)
Measuring Social Bias in Vision-Language Models with Face-Only Counterfactuals from Real Photos
by: Chen, Haodong, et al.
Published: (2026)
by: Chen, Haodong, et al.
Published: (2026)
GKnow: Measuring the Entanglement of Gender Bias and Factual Gender
by: Veloso, Leonor, et al.
Published: (2026)
by: Veloso, Leonor, et al.
Published: (2026)
Measuring Spiritual Values and Bias of Large Language Models
by: Liu, Songyuan, et al.
Published: (2024)
by: Liu, Songyuan, et al.
Published: (2024)
New Job, New Gender? Measuring the Social Bias in Image Generation Models
by: Wang, Wenxuan, et al.
Published: (2024)
by: Wang, Wenxuan, et al.
Published: (2024)
Identifying High-Confidence Social Biases in LLMs for Trustworthy Conversational Tutoring Agents
by: Alvarez, Aitor Arronte, et al.
Published: (2026)
by: Alvarez, Aitor Arronte, et al.
Published: (2026)
Bias Beyond English: Evaluating Social Bias and Debiasing Methods in a Low-Resource Setting
by: Zhou, Ej, et al.
Published: (2025)
by: Zhou, Ej, et al.
Published: (2025)
Probing Social Identity Bias in Chinese LLMs with Gendered Pronouns and Social Groups
by: Liu, Geng, et al.
Published: (2025)
by: Liu, Geng, et al.
Published: (2025)
Mitigating Social Desirability Bias in Random Silicon Sampling
by: Chapala, Sashank, et al.
Published: (2025)
by: Chapala, Sashank, et al.
Published: (2025)
Social Bias Probing: Fairness Benchmarking for Language Models
by: Manerba, Marta Marchiori, et al.
Published: (2023)
by: Manerba, Marta Marchiori, et al.
Published: (2023)
Social Bias in Multilingual Language Models: A Survey
by: Gamboa, Lance Calvin Lim, et al.
Published: (2025)
by: Gamboa, Lance Calvin Lim, et al.
Published: (2025)
The Responsible Foundation Model Development Cheatsheet: A Review of Tools & Resources
by: Longpre, Shayne, et al.
Published: (2024)
by: Longpre, Shayne, et al.
Published: (2024)
Similar Items
-
Ecosystem Graphs: The Social Footprint of Foundation Models
by: Bommasani, Rishi, et al.
Published: (2023) -
The Societal Impact of Foundation Models: Advancing Evidence-based AI Policy
by: Bommasani, Rishi
Published: (2025) -
Algorithmic Monocultures in Hiring
by: Bommasani, Rishi, et al.
Published: (2026) -
NeurIPS should lead scientific consensus on AI policy
by: Bommasani, Rishi
Published: (2025) -
Ecosystem-level Analysis of Deployed Machine Learning Reveals Homogeneous Outcomes
by: Toups, Connor, et al.
Published: (2023)