More Distinctively Black and Feminine Faces Lead to Increased Stereotyping in Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Messi H. J., Montgomery, Jacob M., Lai, Calvin K. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Visual Cues of Gender and Race are Associated with Stereotyping in Vision-Language Models
von: Lee, Messi H. J., et al.
Veröffentlicht: (2025)
von: Lee, Messi H. J., et al.
Veröffentlicht: (2025)
Vision-Language Models Generate More Homogeneous Stories for Phenotypically Black Individuals
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
Token-Level Entropy Reveals Demographic Disparities in Language Models
von: Lee, Messi H. J.
Veröffentlicht: (2025)
von: Lee, Messi H. J.
Veröffentlicht: (2025)
Large Language Models Portray Socially Subordinate Groups as More Homogeneous, Consistent with a Bias Observed in Humans
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024)
Examining the Robustness of Homogeneity Bias to Hyperparameter Adjustments in GPT-4
von: Lee, Messi H. J.
Veröffentlicht: (2025)
von: Lee, Messi H. J.
Veröffentlicht: (2025)
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
von: Woo, Sangmin, et al.
Veröffentlicht: (2025)
von: Woo, Sangmin, et al.
Veröffentlicht: (2025)
A Multimodal Recaptioning Framework to Account for Perceptual Diversity Across Languages in Vision-Language Modeling
von: Buettner, Kyle, et al.
Veröffentlicht: (2025)
von: Buettner, Kyle, et al.
Veröffentlicht: (2025)
Measuring Social Bias in Vision-Language Models with Face-Only Counterfactuals from Real Photos
von: Chen, Haodong, et al.
Veröffentlicht: (2026)
von: Chen, Haodong, et al.
Veröffentlicht: (2026)
Do Visual Imaginations Improve Vision-and-Language Navigation Agents?
von: Perincherry, Akhil, et al.
Veröffentlicht: (2025)
von: Perincherry, Akhil, et al.
Veröffentlicht: (2025)
FaceLLM: A Multimodal Large Language Model for Face Understanding
von: Shahreza, Hatef Otroshi, et al.
Veröffentlicht: (2025)
von: Shahreza, Hatef Otroshi, et al.
Veröffentlicht: (2025)
Aesthetic Assessment of Chinese Handwritings Based on Vision Language Models
von: Zheng, Chen, et al.
Veröffentlicht: (2026)
von: Zheng, Chen, et al.
Veröffentlicht: (2026)
Gender Stereotypes in Professional Roles Among Saudis: An Analytical Study of AI-Generated Images Using Language Models
von: AlKhalifah, Khaloud S., et al.
Veröffentlicht: (2025)
von: AlKhalifah, Khaloud S., et al.
Veröffentlicht: (2025)
VLind-Bench: Measuring Language Priors in Large Vision-Language Models
von: Lee, Kang-il, et al.
Veröffentlicht: (2024)
von: Lee, Kang-il, et al.
Veröffentlicht: (2024)
Benchmarking Multimodal Large Language Models for Face Recognition
von: Shahreza, Hatef Otroshi, et al.
Veröffentlicht: (2025)
von: Shahreza, Hatef Otroshi, et al.
Veröffentlicht: (2025)
Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding
von: Min, Kyungmin, et al.
Veröffentlicht: (2024)
von: Min, Kyungmin, et al.
Veröffentlicht: (2024)
Toward More Reliable Artificial Intelligence: Reducing Hallucinations in Vision-Language Models
von: Sanogo, Kassoum, et al.
Veröffentlicht: (2025)
von: Sanogo, Kassoum, et al.
Veröffentlicht: (2025)
Delve into Visual Contrastive Decoding for Hallucination Mitigation of Large Vision-Language Models
von: Lee, Yi-Lun, et al.
Veröffentlicht: (2024)
von: Lee, Yi-Lun, et al.
Veröffentlicht: (2024)
Better Safe Than Sorry? Overreaction Problem of Vision Language Models in Visual Emergency Recognition
von: Choi, Dasol, et al.
Veröffentlicht: (2025)
von: Choi, Dasol, et al.
Veröffentlicht: (2025)
Bridging the Missing-Modality Gap: Improving Text-Only Calibration of Vision Language Models
von: Kim, Mingyeong, et al.
Veröffentlicht: (2026)
von: Kim, Mingyeong, et al.
Veröffentlicht: (2026)
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning
von: Yang, Senqiao, et al.
Veröffentlicht: (2025)
von: Yang, Senqiao, et al.
Veröffentlicht: (2025)
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Survey on Vision-Language-Action Models
von: Adilkhanov, Adilzhan, et al.
Veröffentlicht: (2025)
von: Adilkhanov, Adilzhan, et al.
Veröffentlicht: (2025)
Revisiting the Role of Language Priors in Vision-Language Models
von: Lin, Zhiqiu, et al.
Veröffentlicht: (2023)
von: Lin, Zhiqiu, et al.
Veröffentlicht: (2023)
Do Vision-Language Models Have Internal World Models? Towards an Atomic Evaluation
von: Gao, Qiyue, et al.
Veröffentlicht: (2025)
von: Gao, Qiyue, et al.
Veröffentlicht: (2025)
WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences
von: Lu, Yujie, et al.
Veröffentlicht: (2024)
von: Lu, Yujie, et al.
Veröffentlicht: (2024)
EmbodiedMidtrain: Bridging the Gap between Vision-Language Models and Vision-Language-Action Models via Mid-training
von: Du, Yiyang, et al.
Veröffentlicht: (2026)
von: Du, Yiyang, et al.
Veröffentlicht: (2026)
Multi-Object Hallucination in Vision-Language Models
von: Chen, Xuweiyi, et al.
Veröffentlicht: (2024)
von: Chen, Xuweiyi, et al.
Veröffentlicht: (2024)
Benchmarking Vision Language Models for Cultural Understanding
von: Nayak, Shravan, et al.
Veröffentlicht: (2024)
von: Nayak, Shravan, et al.
Veröffentlicht: (2024)
VTCBench: Can Vision-Language Models Understand Long Context with Vision-Text Compression?
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
On the Cultural Anachronism and Temporal Reasoning in Vision Language Models
von: Ranjan, Mukul, et al.
Veröffentlicht: (2026)
von: Ranjan, Mukul, et al.
Veröffentlicht: (2026)
Can Vision Language Models Understand Mimed Actions?
von: Cho, Hyundong, et al.
Veröffentlicht: (2025)
von: Cho, Hyundong, et al.
Veröffentlicht: (2025)
An Examination of the Compositionality of Large Generative Vision-Language Models
von: Ma, Teli, et al.
Veröffentlicht: (2023)
von: Ma, Teli, et al.
Veröffentlicht: (2023)
Mitigating Multilingual Hallucination in Large Vision-Language Models
von: Qu, Xiaoye, et al.
Veröffentlicht: (2024)
von: Qu, Xiaoye, et al.
Veröffentlicht: (2024)
Benchmarking Deflection and Hallucination in Large Vision-Language Models
von: Moratelli, Nicholas, et al.
Veröffentlicht: (2026)
von: Moratelli, Nicholas, et al.
Veröffentlicht: (2026)
Are Large Vision Language Models Good Game Players?
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
Probing and Inducing Combinational Creativity in Vision-Language Models
von: Peng, Yongqian, et al.
Veröffentlicht: (2025)
von: Peng, Yongqian, et al.
Veröffentlicht: (2025)
Anthropogenic Regional Adaptation in Multimodal Vision-Language Model
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2026)
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2026)
Reward Design for Physical Reasoning in Vision-Language Models
von: Lilienthal, Derek, et al.
Veröffentlicht: (2026)
von: Lilienthal, Derek, et al.
Veröffentlicht: (2026)
PatientVLM Meets DocVLM: Pre-Consultation Dialogue Between Vision-Language Models for Efficient Diagnosis
von: Lokesh, K, et al.
Veröffentlicht: (2026)
von: Lokesh, K, et al.
Veröffentlicht: (2026)
Referring Expressions as a Lens into Spatial Language Grounding in Vision-Language Models
von: Tumu, Akshar, et al.
Veröffentlicht: (2025)
von: Tumu, Akshar, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Visual Cues of Gender and Race are Associated with Stereotyping in Vision-Language Models
von: Lee, Messi H. J., et al.
Veröffentlicht: (2025) -
Vision-Language Models Generate More Homogeneous Stories for Phenotypically Black Individuals
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024) -
Token-Level Entropy Reveals Demographic Disparities in Language Models
von: Lee, Messi H. J.
Veröffentlicht: (2025) -
Large Language Models Portray Socially Subordinate Groups as More Homogeneous, Consistent with a Bias Observed in Humans
von: Lee, Messi H. J., et al.
Veröffentlicht: (2024) -
Examining the Robustness of Homogeneity Bias to Hyperparameter Adjustments in GPT-4
von: Lee, Messi H. J.
Veröffentlicht: (2025)