Vision-Language Models Suppress Female Representations Under Ambiguous Input
Fuente:
arXiv
Guardado en:
| Autores principales: | Marin-Llobet, Arnau, Henniger, Simon, Banaji, Mahzarin R. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Investigating Disability Representations in Text-to-Image Models
por: Tian, Yang, et al.
Publicado: (2026)
por: Tian, Yang, et al.
Publicado: (2026)
Automated Interpretability and Feature Discovery in Language Models with Agents
por: Marin-Llobet, Arnau, et al.
Publicado: (2026)
por: Marin-Llobet, Arnau, et al.
Publicado: (2026)
Semantic and Expressive Variation in Image Captions Across Languages
por: Ye, Andre, et al.
Publicado: (2023)
por: Ye, Andre, et al.
Publicado: (2023)
CHART-6: Human-Centered Evaluation of Data Visualization Understanding in Vision-Language Models
por: Verma, Arnav, et al.
Publicado: (2025)
por: Verma, Arnav, et al.
Publicado: (2025)
Learning Multimodal Cues of Children's Uncertainty
por: Cheng, Qi, et al.
Publicado: (2024)
por: Cheng, Qi, et al.
Publicado: (2024)
GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents
por: Luo, Run, et al.
Publicado: (2025)
por: Luo, Run, et al.
Publicado: (2025)
Refusal as Silence: Gendered Disparities in Vision-Language Model Responses
por: Luo, Sha, et al.
Publicado: (2024)
por: Luo, Sha, et al.
Publicado: (2024)
Signformer is all you need: Towards Edge AI for Sign Language
por: Yang, Eta
Publicado: (2024)
por: Yang, Eta
Publicado: (2024)
A Review on Large Language Models for Visual Analytics
por: Agarwal, Navya Sonal, et al.
Publicado: (2025)
por: Agarwal, Navya Sonal, et al.
Publicado: (2025)
Human-Centred Evaluation of Text-to-Image Generation Models for Self-expression of Mental Distress: A Dataset Based on GPT-4o
por: He, Sui, et al.
Publicado: (2025)
por: He, Sui, et al.
Publicado: (2025)
GUICourse: From General Vision Language Models to Versatile GUI Agents
por: Chen, Wentong, et al.
Publicado: (2024)
por: Chen, Wentong, et al.
Publicado: (2024)
Towards Geographic Inclusion in the Evaluation of Text-to-Image Models
por: Hall, Melissa, et al.
Publicado: (2024)
por: Hall, Melissa, et al.
Publicado: (2024)
Navi-plus: Managing Ambiguous GUI Navigation Tasks with Follow-up Questions
por: Cheng, Ziming, et al.
Publicado: (2025)
por: Cheng, Ziming, et al.
Publicado: (2025)
VideoFDB: Evaluating Full-Duplex Vision-Speech Capabilities in Conversational Agents
por: Mazumdar, Amrita, et al.
Publicado: (2026)
por: Mazumdar, Amrita, et al.
Publicado: (2026)
ShowUI: One Vision-Language-Action Model for GUI Visual Agent
por: Lin, Kevin Qinghong, et al.
Publicado: (2024)
por: Lin, Kevin Qinghong, et al.
Publicado: (2024)
Mapping User Trust in Vision Language Models: Research Landscape, Challenges, and Prospects
por: Chiatti, Agnese, et al.
Publicado: (2025)
por: Chiatti, Agnese, et al.
Publicado: (2025)
Assessing Intersectional Bias in Representations of Pre-Trained Image Recognition Models
por: Krug, Valerie, et al.
Publicado: (2025)
por: Krug, Valerie, et al.
Publicado: (2025)
Shape vs. Context: Examining Human--AI Gaps in Ambiguous Japanese Character Recognition
por: Haraguchi, Daichi
Publicado: (2026)
por: Haraguchi, Daichi
Publicado: (2026)
Beyond Questionnaires: Video Analysis for Social Anxiety Detection
por: Sahu, Nilesh Kumar, et al.
Publicado: (2024)
por: Sahu, Nilesh Kumar, et al.
Publicado: (2024)
A Picture is Worth a Thousand (Correct) Captions: A Vision-Guided Judge-Corrector System for Multimodal Machine Translation
por: Betala, Siddharth, et al.
Publicado: (2025)
por: Betala, Siddharth, et al.
Publicado: (2025)
Vision Language Models as Values Detectors
por: Abbo, Giulio Antonio, et al.
Publicado: (2025)
por: Abbo, Giulio Antonio, et al.
Publicado: (2025)
Identifying & Interactively Refining Ambiguous User Goals for Data Visualization Code Generation
por: İnan, Mert, et al.
Publicado: (2025)
por: İnan, Mert, et al.
Publicado: (2025)
MAPWise: Evaluating Vision-Language Models for Advanced Map Queries
por: Mukhopadhyay, Srija, et al.
Publicado: (2024)
por: Mukhopadhyay, Srija, et al.
Publicado: (2024)
Forest-Chat: Adapting Vision-Language Agents for Interactive Forest Change Analysis
por: Brock, James, et al.
Publicado: (2026)
por: Brock, James, et al.
Publicado: (2026)
A Call to Arms: AI Should be Critical for Social Media Analysis of Conflict Zones
por: Abedin, Afia, et al.
Publicado: (2023)
por: Abedin, Afia, et al.
Publicado: (2023)
Improved Digital Therapy for Developmental Pediatrics Using Domain-Specific Artificial Intelligence: Machine Learning Study
por: Washington, Peter, et al.
Publicado: (2020)
por: Washington, Peter, et al.
Publicado: (2020)
A Comparison of Human and Machine Learning Errors in Face Recognition
por: Estévez-Almenzar, Marina, et al.
Publicado: (2025)
por: Estévez-Almenzar, Marina, et al.
Publicado: (2025)
Classification of the lunar surface pattern by AI architectures: Does AI see a rabbit in the Moon?
por: Shoji, Daigo
Publicado: (2023)
por: Shoji, Daigo
Publicado: (2023)
AIDEN: Design and Pilot Study of an AI Assistant for the Visually Impaired
por: Marquez-Carpintero, Luis, et al.
Publicado: (2025)
por: Marquez-Carpintero, Luis, et al.
Publicado: (2025)
DepMamba: Progressive Fusion Mamba for Multimodal Depression Detection
por: Ye, Jiaxin, et al.
Publicado: (2024)
por: Ye, Jiaxin, et al.
Publicado: (2024)
The Cadaver in the Machine: The Social Practices of Measurement and Validation in Motion Capture Technology
por: Harvey, Emma, et al.
Publicado: (2024)
por: Harvey, Emma, et al.
Publicado: (2024)
CAF-Mamba: Mamba-Based Cross-Modal Adaptive Attention Fusion for Multimodal Depression Detection
por: Zhou, Bowen, et al.
Publicado: (2026)
por: Zhou, Bowen, et al.
Publicado: (2026)
Intuitions of Machine Learning Researchers about Transfer Learning for Medical Image Classification
por: Lu, Yucheng, et al.
Publicado: (2025)
por: Lu, Yucheng, et al.
Publicado: (2025)
A Pilot Study on Curator-Guided Multilingual Art Description for Blind and Low-Vision Audiences with Small Vision-Language Models
por: Tsangko, Iosif, et al.
Publicado: (2026)
por: Tsangko, Iosif, et al.
Publicado: (2026)
UIClip: A Data-driven Model for Assessing User Interface Design
por: Wu, Jason, et al.
Publicado: (2024)
por: Wu, Jason, et al.
Publicado: (2024)
OS-ATLAS: A Foundation Action Model for Generalist GUI Agents
por: Wu, Zhiyong, et al.
Publicado: (2024)
por: Wu, Zhiyong, et al.
Publicado: (2024)
GPT-5 Model Corrected GPT-4V's Chart Reading Errors, Not Prompting
por: Yang, Kaichun, et al.
Publicado: (2025)
por: Yang, Kaichun, et al.
Publicado: (2025)
How Can Large Language Models Enable Better Socially Assistive Human-Robot Interaction: A Brief Survey
por: Shi, Zhonghao, et al.
Publicado: (2024)
por: Shi, Zhonghao, et al.
Publicado: (2024)
ImaGGen: Zero-Shot Generation of Co-Speech Semantic Gestures Grounded in Language and Image Input
por: Voss, Hendric, et al.
Publicado: (2025)
por: Voss, Hendric, et al.
Publicado: (2025)
An Egocentric Vision-Language Model based Portable Real-time Smart Assistant
por: Huang, Yifei, et al.
Publicado: (2025)
por: Huang, Yifei, et al.
Publicado: (2025)
Ejemplares similares
-
Investigating Disability Representations in Text-to-Image Models
por: Tian, Yang, et al.
Publicado: (2026) -
Automated Interpretability and Feature Discovery in Language Models with Agents
por: Marin-Llobet, Arnau, et al.
Publicado: (2026) -
Semantic and Expressive Variation in Image Captions Across Languages
por: Ye, Andre, et al.
Publicado: (2023) -
CHART-6: Human-Centered Evaluation of Data Visualization Understanding in Vision-Language Models
por: Verma, Arnav, et al.
Publicado: (2025) -
Learning Multimodal Cues of Children's Uncertainty
por: Cheng, Qi, et al.
Publicado: (2024)