Measuring Social Biases in Masked Language Models by Proxy of Prediction Quality
Fuente:
arXiv
Salvato in:
| Autori principali: | Zalkikar, Rahul, Chandra, Kanchan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Limitation Learning: Catching Adverse Dialog with GAIL
di: Kasmanoff, Noah, et al.
Pubblicazione: (2025)
di: Kasmanoff, Noah, et al.
Pubblicazione: (2025)
Robust Evaluation Measures for Evaluating Social Biases in Masked Language Models
di: Liu, Yang
Pubblicazione: (2024)
di: Liu, Yang
Pubblicazione: (2024)
Evaluating Short-Term Temporal Fluctuations of Social Biases in Social Media Data and Masked Language Models
di: Zhou, Yi, et al.
Pubblicazione: (2024)
di: Zhou, Yi, et al.
Pubblicazione: (2024)
ProxyLM: Predicting Language Model Performance on Multilingual Tasks via Proxy Models
di: Anugraha, David, et al.
Pubblicazione: (2024)
di: Anugraha, David, et al.
Pubblicazione: (2024)
Estranged Predictions: Measuring Semantic Category Disruption with Masked Language Modelling
di: Liu, Yuxuan, et al.
Pubblicazione: (2025)
di: Liu, Yuxuan, et al.
Pubblicazione: (2025)
VisBias: Measuring Explicit and Implicit Social Biases in Vision Language Models
di: Huang, Jen-tse, et al.
Pubblicazione: (2025)
di: Huang, Jen-tse, et al.
Pubblicazione: (2025)
Measuring Stereotype and Deviation Biases in Large Language Models
di: Wang, Daniel, et al.
Pubblicazione: (2025)
di: Wang, Daniel, et al.
Pubblicazione: (2025)
IndiBias: A Benchmark Dataset to Measure Social Biases in Language Models for Indian Context
di: Sahoo, Nihar Ranjan, et al.
Pubblicazione: (2024)
di: Sahoo, Nihar Ranjan, et al.
Pubblicazione: (2024)
Dutch CrowS-Pairs: Adapting a Challenge Dataset for Measuring Social Biases in Language Models for Dutch
di: Strazda, Elza, et al.
Pubblicazione: (2025)
di: Strazda, Elza, et al.
Pubblicazione: (2025)
Tuning Language Models by Proxy
di: Liu, Alisa, et al.
Pubblicazione: (2024)
di: Liu, Alisa, et al.
Pubblicazione: (2024)
Generative Language Models Exhibit Social Identity Biases
di: Hu, Tiancheng, et al.
Pubblicazione: (2023)
di: Hu, Tiancheng, et al.
Pubblicazione: (2023)
Faithfulness Measurable Masked Language Models
di: Madsen, Andreas, et al.
Pubblicazione: (2023)
di: Madsen, Andreas, et al.
Pubblicazione: (2023)
Empirical Analysis of Decoding Biases in Masked Diffusion Models
di: Huang, Pengcheng, et al.
Pubblicazione: (2025)
di: Huang, Pengcheng, et al.
Pubblicazione: (2025)
Mitigating Social Biases in Language Models through Unlearning
di: Dige, Omkar, et al.
Pubblicazione: (2024)
di: Dige, Omkar, et al.
Pubblicazione: (2024)
UnMASKed: Quantifying Gender Biases in Masked Language Models through Linguistically Informed Job Market Prompts
di: Parra, Iñigo
Pubblicazione: (2024)
di: Parra, Iñigo
Pubblicazione: (2024)
Proxy Compression for Language Modeling
di: Zheng, Lin, et al.
Pubblicazione: (2026)
di: Zheng, Lin, et al.
Pubblicazione: (2026)
The Proxy Presumption: From Semantic Embeddings to Valid Social Measures
di: Li, Baishi, et al.
Pubblicazione: (2026)
di: Li, Baishi, et al.
Pubblicazione: (2026)
Diversity Measures: Domain-Independent Proxies for Failure in Language Model Queries
di: Ngu, Noel, et al.
Pubblicazione: (2023)
di: Ngu, Noel, et al.
Pubblicazione: (2023)
The Devil is in the Neurons: Interpreting and Mitigating Social Biases in Pre-trained Language Models
di: Liu, Yan, et al.
Pubblicazione: (2024)
di: Liu, Yan, et al.
Pubblicazione: (2024)
JBBQ: Japanese Bias Benchmark for Analyzing Social Biases in Large Language Models
di: Yanaka, Hitomi, et al.
Pubblicazione: (2024)
di: Yanaka, Hitomi, et al.
Pubblicazione: (2024)
BiasCause: Evaluate Socially Biased Causal Reasoning of Large Language Models
di: Xie, Tian, et al.
Pubblicazione: (2025)
di: Xie, Tian, et al.
Pubblicazione: (2025)
Identifying Implicit Social Biases in Vision-Language Models
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
Natural Language Interfaces for Spatial and Temporal Databases: A Comprehensive Overview of Methods, Taxonomy, and Future Directions
di: Acharja, Samya, et al.
Pubblicazione: (2026)
di: Acharja, Samya, et al.
Pubblicazione: (2026)
Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models
di: Khandelwal, Khyati, et al.
Pubblicazione: (2023)
di: Khandelwal, Khyati, et al.
Pubblicazione: (2023)
Proxy-RLHF: Decoupling Generation and Alignment in Large Language Model with Proxy
di: Zhu, Yu, et al.
Pubblicazione: (2024)
di: Zhu, Yu, et al.
Pubblicazione: (2024)
BanStereoSet: A Dataset to Measure Stereotypical Social Biases in LLMs for Bangla
di: Kamruzzaman, Mahammed, et al.
Pubblicazione: (2024)
di: Kamruzzaman, Mahammed, et al.
Pubblicazione: (2024)
Evaluating LLMs Robustness in Less Resourced Languages with Proxy Models
di: Chrabąszcz, Maciej, et al.
Pubblicazione: (2025)
di: Chrabąszcz, Maciej, et al.
Pubblicazione: (2025)
Large Language Models as Proxies for Theories of Human Linguistic Cognition
di: Ziv, Imry, et al.
Pubblicazione: (2025)
di: Ziv, Imry, et al.
Pubblicazione: (2025)
Large Language Models Develop Novel Social Biases Through Adaptive Exploration
di: Wu, Addison J., et al.
Pubblicazione: (2025)
di: Wu, Addison J., et al.
Pubblicazione: (2025)
Intent-Aware Self-Correction for Mitigating Social Biases in Large Language Models
di: Anantaprayoon, Panatchakorn, et al.
Pubblicazione: (2025)
di: Anantaprayoon, Panatchakorn, et al.
Pubblicazione: (2025)
Robust Infidelity: When Faithfulness Measures on Masked Language Models Are Misleading
di: Crothers, Evan, et al.
Pubblicazione: (2023)
di: Crothers, Evan, et al.
Pubblicazione: (2023)
Third-Party Language Model Performance Prediction from Instruction
di: Nadkarni, Rahul, et al.
Pubblicazione: (2024)
di: Nadkarni, Rahul, et al.
Pubblicazione: (2024)
LPZero: Language Model Zero-cost Proxy Search from Zero
di: Dong, Peijie, et al.
Pubblicazione: (2024)
di: Dong, Peijie, et al.
Pubblicazione: (2024)
Do LLMs Align Human Values Regarding Social Biases? Judging and Explaining Social Biases with LLMs
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
Language Complexity Measurement as a Noisy Zero-Shot Proxy for Evaluating LLM Performance
di: Moell, Birger, et al.
Pubblicazione: (2025)
di: Moell, Birger, et al.
Pubblicazione: (2025)
Leveraging Large Language Models for Predictive Analysis of Human Misery
di: Seal, Bishanka, et al.
Pubblicazione: (2025)
di: Seal, Bishanka, et al.
Pubblicazione: (2025)
Mitigating Biases in Language Models via Bias Unlearning
di: Liu, Dianqing, et al.
Pubblicazione: (2025)
di: Liu, Dianqing, et al.
Pubblicazione: (2025)
Entropy-aware Masking for Masked Language Modeling
di: Srinivasagan, Gokul, et al.
Pubblicazione: (2026)
di: Srinivasagan, Gokul, et al.
Pubblicazione: (2026)
Persona Setting Pitfall: Persistent Outgroup Biases in Large Language Models Arising from Social Identity Adoption
di: Dong, Wenchao, et al.
Pubblicazione: (2024)
di: Dong, Wenchao, et al.
Pubblicazione: (2024)
Large Language Models are Biased Because They Are Large Language Models
di: Resnik, Philip
Pubblicazione: (2024)
di: Resnik, Philip
Pubblicazione: (2024)
Documenti analoghi
-
Limitation Learning: Catching Adverse Dialog with GAIL
di: Kasmanoff, Noah, et al.
Pubblicazione: (2025) -
Robust Evaluation Measures for Evaluating Social Biases in Masked Language Models
di: Liu, Yang
Pubblicazione: (2024) -
Evaluating Short-Term Temporal Fluctuations of Social Biases in Social Media Data and Masked Language Models
di: Zhou, Yi, et al.
Pubblicazione: (2024) -
ProxyLM: Predicting Language Model Performance on Multilingual Tasks via Proxy Models
di: Anugraha, David, et al.
Pubblicazione: (2024) -
Estranged Predictions: Measuring Semantic Category Disruption with Masked Language Modelling
di: Liu, Yuxuan, et al.
Pubblicazione: (2025)