Biases Propagate in Encoder-based Vision-Language Models: A Systematic Analysis From Intrinsic Measures to Zero-shot Retrieval Outcomes
Fuente:
arXiv
Salvato in:
| Autori principali: | Ghate, Kshitish, Charlesworth, Tessa, Diab, Mona, Caliskan, Aylin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Intrinsic Bias is Predicted by Pretraining Data and Correlates with Downstream Performance in Vision-Language Encoders
di: Ghate, Kshitish, et al.
Pubblicazione: (2025)
di: Ghate, Kshitish, et al.
Pubblicazione: (2025)
Personal Information Parroting in Language Models
di: Subramani, Nishant, et al.
Pubblicazione: (2026)
di: Subramani, Nishant, et al.
Pubblicazione: (2026)
EVALUESTEER: Measuring Reward Model Steerability Towards Values and Preferences
di: Ghate, Kshitish, et al.
Pubblicazione: (2025)
di: Ghate, Kshitish, et al.
Pubblicazione: (2025)
Beyond Cooperative Simulators: Generating Realistic User Personas for Robust Evaluation of LLM Agents
di: Chopra, Harshita, et al.
Pubblicazione: (2026)
di: Chopra, Harshita, et al.
Pubblicazione: (2026)
Pre-Calc: Learning to Use the Calculator Improves Numeracy in Language Models
di: Veerendranath, Vishruth, et al.
Pubblicazione: (2024)
di: Veerendranath, Vishruth, et al.
Pubblicazione: (2024)
BiasDora: Exploring Hidden Biased Associations in Vision-Language Models
di: Raj, Chahat, et al.
Pubblicazione: (2024)
di: Raj, Chahat, et al.
Pubblicazione: (2024)
Generative Value Conflicts Reveal LLM Priorities
di: Liu, Andy, et al.
Pubblicazione: (2025)
di: Liu, Andy, et al.
Pubblicazione: (2025)
Evaluating Large Language Model Biases in Persona-Steered Generation
di: Liu, Andy, et al.
Pubblicazione: (2024)
di: Liu, Andy, et al.
Pubblicazione: (2024)
Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval
di: Wilson, Kyra, et al.
Pubblicazione: (2024)
di: Wilson, Kyra, et al.
Pubblicazione: (2024)
A Taxonomy of Stereotype Content in Large Language Models
di: Nicolas, Gandalf, et al.
Pubblicazione: (2024)
di: Nicolas, Gandalf, et al.
Pubblicazione: (2024)
VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models
di: Raj, Chahat, et al.
Pubblicazione: (2025)
di: Raj, Chahat, et al.
Pubblicazione: (2025)
Identifying Features Associated with Bias Against 93 Stigmatized Groups in Language Models and Guardrail Model Safety Mitigation
di: Gueorguieva, Anna-Maria, et al.
Pubblicazione: (2025)
di: Gueorguieva, Anna-Maria, et al.
Pubblicazione: (2025)
Deep Reasoning in General Purpose Agents via Structured Meta-Cognition
di: Light, Dean, et al.
Pubblicazione: (2026)
di: Light, Dean, et al.
Pubblicazione: (2026)
ChatGPT Perpetuates Gender Bias in Machine Translation and Ignores Non-Gendered Pronouns: Findings across Bengali and Five other Low-Resource Languages
di: Ghosh, Sourojit, et al.
Pubblicazione: (2023)
di: Ghosh, Sourojit, et al.
Pubblicazione: (2023)
Breaking Bias, Building Bridges: Evaluation and Mitigation of Social Biases in LLMs via Contact Hypothesis
di: Raj, Chahat, et al.
Pubblicazione: (2024)
di: Raj, Chahat, et al.
Pubblicazione: (2024)
Inductive Biases for Zero-shot Systematic Generalization in Language-informed Reinforcement Learning
di: Dijujin, Negin Hashemi, et al.
Pubblicazione: (2025)
di: Dijujin, Negin Hashemi, et al.
Pubblicazione: (2025)
Persona-Assigned Large Language Models Exhibit Human-Like Motivated Reasoning
di: Dash, Saloni, et al.
Pubblicazione: (2025)
di: Dash, Saloni, et al.
Pubblicazione: (2025)
No Thoughts Just AI: Biased LLM Hiring Recommendations Alter Human Decision Making and Limit Human Autonomy
di: Wilson, Kyra, et al.
Pubblicazione: (2025)
di: Wilson, Kyra, et al.
Pubblicazione: (2025)
CoRAG: Collaborative Retrieval-Augmented Generation
di: Muhamed, Aashiq, et al.
Pubblicazione: (2025)
di: Muhamed, Aashiq, et al.
Pubblicazione: (2025)
A Note on Bias to Complete
di: Xu, Jia, et al.
Pubblicazione: (2024)
di: Xu, Jia, et al.
Pubblicazione: (2024)
Talent or Luck? Evaluating Attribution Bias in Large Language Models
di: Raj, Chahat, et al.
Pubblicazione: (2025)
di: Raj, Chahat, et al.
Pubblicazione: (2025)
Emotion Classification in Low and Moderate Resource Languages
di: Tafreshi, Shabnam, et al.
Pubblicazione: (2024)
di: Tafreshi, Shabnam, et al.
Pubblicazione: (2024)
Shaping Shared Languages: Human and Large Language Models' Inductive Biases in Emergent Communication
di: Kouwenhoven, Tom, et al.
Pubblicazione: (2025)
di: Kouwenhoven, Tom, et al.
Pubblicazione: (2025)
SimBA: Simplifying Benchmark Analysis Using Performance Matrices Alone
di: Subramani, Nishant, et al.
Pubblicazione: (2025)
di: Subramani, Nishant, et al.
Pubblicazione: (2025)
BEIR-NL: Zero-shot Information Retrieval Benchmark for the Dutch Language
di: Banar, Nikolay, et al.
Pubblicazione: (2024)
di: Banar, Nikolay, et al.
Pubblicazione: (2024)
A Systematic Analysis of Biases in Large Language Models
di: Zhang, Xulang, et al.
Pubblicazione: (2025)
di: Zhang, Xulang, et al.
Pubblicazione: (2025)
Zero-shot Generative Large Language Models for Systematic Review Screening Automation
di: Wang, Shuai, et al.
Pubblicazione: (2024)
di: Wang, Shuai, et al.
Pubblicazione: (2024)
Zero-shot Context Biasing with Trie-based Decoding using Synthetic Multi-Pronunciation
di: Liu, Changsong, et al.
Pubblicazione: (2025)
di: Liu, Changsong, et al.
Pubblicazione: (2025)
Understanding Intrinsic Socioeconomic Biases in Large Language Models
di: Arzaghi, Mina, et al.
Pubblicazione: (2024)
di: Arzaghi, Mina, et al.
Pubblicazione: (2024)
REALM: A Dataset of Real-World LLM Use Cases
di: Cheng, Jingwen, et al.
Pubblicazione: (2025)
di: Cheng, Jingwen, et al.
Pubblicazione: (2025)
Decoding Dark Matter: Specialized Sparse Autoencoders for Interpreting Rare Concepts in Foundation Models
di: Muhamed, Aashiq, et al.
Pubblicazione: (2024)
di: Muhamed, Aashiq, et al.
Pubblicazione: (2024)
Combining Discrete Wavelet and Cosine Transforms for Efficient Sentence Embedding
di: Salama, Rana, et al.
Pubblicazione: (2025)
di: Salama, Rana, et al.
Pubblicazione: (2025)
Investigating Cultural Alignment of Large Language Models
di: AlKhamissi, Badr, et al.
Pubblicazione: (2024)
di: AlKhamissi, Badr, et al.
Pubblicazione: (2024)
Automatic Generation of Model and Data Cards: A Step Towards Responsible AI
di: Liu, Jiarui, et al.
Pubblicazione: (2024)
di: Liu, Jiarui, et al.
Pubblicazione: (2024)
LLM Microscope: What Model Internals Reveal About Answer Correctness and Context Utilization
di: Liu, Jiarui, et al.
Pubblicazione: (2025)
di: Liu, Jiarui, et al.
Pubblicazione: (2025)
VisBias: Measuring Explicit and Implicit Social Biases in Vision Language Models
di: Huang, Jen-tse, et al.
Pubblicazione: (2025)
di: Huang, Jen-tse, et al.
Pubblicazione: (2025)
StressRoBERTa: Cross-Condition Transfer Learning from Depression, Anxiety, and PTSD to Stress Detection
di: Alqahtani, Amal, et al.
Pubblicazione: (2025)
di: Alqahtani, Amal, et al.
Pubblicazione: (2025)
DWTSumm: Discrete Wavelet Transform for Document Summarization
di: Salama, Rana, et al.
Pubblicazione: (2026)
di: Salama, Rana, et al.
Pubblicazione: (2026)
Beyond Understanding: Evaluating the Pragmatic Gap in LLMs' Cultural Processing of Figurative Language
di: Attia, Mena, et al.
Pubblicazione: (2025)
di: Attia, Mena, et al.
Pubblicazione: (2025)
Semantic Compression for Word and Sentence Embeddings using Discrete Wavelet Transform
di: Salama, Rana Aref, et al.
Pubblicazione: (2025)
di: Salama, Rana Aref, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Intrinsic Bias is Predicted by Pretraining Data and Correlates with Downstream Performance in Vision-Language Encoders
di: Ghate, Kshitish, et al.
Pubblicazione: (2025) -
Personal Information Parroting in Language Models
di: Subramani, Nishant, et al.
Pubblicazione: (2026) -
EVALUESTEER: Measuring Reward Model Steerability Towards Values and Preferences
di: Ghate, Kshitish, et al.
Pubblicazione: (2025) -
Beyond Cooperative Simulators: Generating Realistic User Personas for Robust Evaluation of LLM Agents
di: Chopra, Harshita, et al.
Pubblicazione: (2026) -
Pre-Calc: Learning to Use the Calculator Improves Numeracy in Language Models
di: Veerendranath, Vishruth, et al.
Pubblicazione: (2024)