Investigating Disability Representations in Text-to-Image Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tian, Yang, Fan, Yu, Zavolokina, Liudmila, Ebling, Sarah |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Geographic Inclusion in the Evaluation of Text-to-Image Models
von: Hall, Melissa, et al.
Veröffentlicht: (2024)
von: Hall, Melissa, et al.
Veröffentlicht: (2024)
Semantic and Expressive Variation in Image Captions Across Languages
von: Ye, Andre, et al.
Veröffentlicht: (2023)
von: Ye, Andre, et al.
Veröffentlicht: (2023)
SwissADT: An Audio Description Translation System for Swiss Languages
von: Fischer, Lukas, et al.
Veröffentlicht: (2024)
von: Fischer, Lukas, et al.
Veröffentlicht: (2024)
Human-Centred Evaluation of Text-to-Image Generation Models for Self-expression of Mental Distress: A Dataset Based on GPT-4o
von: He, Sui, et al.
Veröffentlicht: (2025)
von: He, Sui, et al.
Veröffentlicht: (2025)
Vision-Language Models Suppress Female Representations Under Ambiguous Input
von: Marin-Llobet, Arnau, et al.
Veröffentlicht: (2026)
von: Marin-Llobet, Arnau, et al.
Veröffentlicht: (2026)
Learning Multimodal Cues of Children's Uncertainty
von: Cheng, Qi, et al.
Veröffentlicht: (2024)
von: Cheng, Qi, et al.
Veröffentlicht: (2024)
Assessing Intersectional Bias in Representations of Pre-Trained Image Recognition Models
von: Krug, Valerie, et al.
Veröffentlicht: (2025)
von: Krug, Valerie, et al.
Veröffentlicht: (2025)
Signformer is all you need: Towards Edge AI for Sign Language
von: Yang, Eta
Veröffentlicht: (2024)
von: Yang, Eta
Veröffentlicht: (2024)
Intuitions of Machine Learning Researchers about Transfer Learning for Medical Image Classification
von: Lu, Yucheng, et al.
Veröffentlicht: (2025)
von: Lu, Yucheng, et al.
Veröffentlicht: (2025)
CHART-6: Human-Centered Evaluation of Data Visualization Understanding in Vision-Language Models
von: Verma, Arnav, et al.
Veröffentlicht: (2025)
von: Verma, Arnav, et al.
Veröffentlicht: (2025)
GesGPT: Speech Gesture Synthesis With Text Parsing from ChatGPT
von: Gao, Nan, et al.
Veröffentlicht: (2023)
von: Gao, Nan, et al.
Veröffentlicht: (2023)
GPT-5 Model Corrected GPT-4V's Chart Reading Errors, Not Prompting
von: Yang, Kaichun, et al.
Veröffentlicht: (2025)
von: Yang, Kaichun, et al.
Veröffentlicht: (2025)
Digital Comprehensibility Assessment of Simplified Texts among Persons with Intellectual Disabilities
von: Säuberli, Andreas, et al.
Veröffentlicht: (2024)
von: Säuberli, Andreas, et al.
Veröffentlicht: (2024)
Semantic Draw Engineering for Text-to-Image Creation
von: Li, Yang, et al.
Veröffentlicht: (2023)
von: Li, Yang, et al.
Veröffentlicht: (2023)
OS-ATLAS: A Foundation Action Model for Generalist GUI Agents
von: Wu, Zhiyong, et al.
Veröffentlicht: (2024)
von: Wu, Zhiyong, et al.
Veröffentlicht: (2024)
Mask-up: Investigating Biases in Face Re-identification for Masked Faces
von: Jaiswal, Siddharth D, et al.
Veröffentlicht: (2024)
von: Jaiswal, Siddharth D, et al.
Veröffentlicht: (2024)
CAF-Mamba: Mamba-Based Cross-Modal Adaptive Attention Fusion for Multimodal Depression Detection
von: Zhou, Bowen, et al.
Veröffentlicht: (2026)
von: Zhou, Bowen, et al.
Veröffentlicht: (2026)
SpatialViz-Bench: A Cognitively-Grounded Benchmark for Diagnosing Spatial Visualization in MLLMs
von: Wang, Siting, et al.
Veröffentlicht: (2025)
von: Wang, Siting, et al.
Veröffentlicht: (2025)
A Call to Arms: AI Should be Critical for Social Media Analysis of Conflict Zones
von: Abedin, Afia, et al.
Veröffentlicht: (2023)
von: Abedin, Afia, et al.
Veröffentlicht: (2023)
Improved Digital Therapy for Developmental Pediatrics Using Domain-Specific Artificial Intelligence: Machine Learning Study
von: Washington, Peter, et al.
Veröffentlicht: (2020)
von: Washington, Peter, et al.
Veröffentlicht: (2020)
A Comparison of Human and Machine Learning Errors in Face Recognition
von: Estévez-Almenzar, Marina, et al.
Veröffentlicht: (2025)
von: Estévez-Almenzar, Marina, et al.
Veröffentlicht: (2025)
Classification of the lunar surface pattern by AI architectures: Does AI see a rabbit in the Moon?
von: Shoji, Daigo
Veröffentlicht: (2023)
von: Shoji, Daigo
Veröffentlicht: (2023)
AIDEN: Design and Pilot Study of an AI Assistant for the Visually Impaired
von: Marquez-Carpintero, Luis, et al.
Veröffentlicht: (2025)
von: Marquez-Carpintero, Luis, et al.
Veröffentlicht: (2025)
Beyond Questionnaires: Video Analysis for Social Anxiety Detection
von: Sahu, Nilesh Kumar, et al.
Veröffentlicht: (2024)
von: Sahu, Nilesh Kumar, et al.
Veröffentlicht: (2024)
DepMamba: Progressive Fusion Mamba for Multimodal Depression Detection
von: Ye, Jiaxin, et al.
Veröffentlicht: (2024)
von: Ye, Jiaxin, et al.
Veröffentlicht: (2024)
The Cadaver in the Machine: The Social Practices of Measurement and Validation in Motion Capture Technology
von: Harvey, Emma, et al.
Veröffentlicht: (2024)
von: Harvey, Emma, et al.
Veröffentlicht: (2024)
A Review on Large Language Models for Visual Analytics
von: Agarwal, Navya Sonal, et al.
Veröffentlicht: (2025)
von: Agarwal, Navya Sonal, et al.
Veröffentlicht: (2025)
UIClip: A Data-driven Model for Assessing User Interface Design
von: Wu, Jason, et al.
Veröffentlicht: (2024)
von: Wu, Jason, et al.
Veröffentlicht: (2024)
GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents
von: Luo, Run, et al.
Veröffentlicht: (2025)
von: Luo, Run, et al.
Veröffentlicht: (2025)
Text-to-Image Representativity Fairness Evaluation Framework
von: Yamani, Asma, et al.
Veröffentlicht: (2024)
von: Yamani, Asma, et al.
Veröffentlicht: (2024)
ScreenQA: Large-Scale Question-Answer Pairs over Mobile App Screenshots
von: Hsiao, Yu-Chung, et al.
Veröffentlicht: (2022)
von: Hsiao, Yu-Chung, et al.
Veröffentlicht: (2022)
Computer-Use Agents as Judges for Generative User Interface
von: Lin, Kevin Qinghong, et al.
Veröffentlicht: (2025)
von: Lin, Kevin Qinghong, et al.
Veröffentlicht: (2025)
Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMs
von: You, Keen, et al.
Veröffentlicht: (2024)
von: You, Keen, et al.
Veröffentlicht: (2024)
SiMing-Bench: Evaluating Procedural Correctness from Continuous Interactions in Clinical Skill Videos
von: Huang, Xiyang, et al.
Veröffentlicht: (2026)
von: Huang, Xiyang, et al.
Veröffentlicht: (2026)
Bridging Text and Image for Artist Style Transfer via Contrastive Learning
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2024)
von: Liu, Zhi-Song, et al.
Veröffentlicht: (2024)
From Image Generation to Infrastructure Design: a Multi-agent Pipeline for Street Design Generation
von: Wang, Chenguang, et al.
Veröffentlicht: (2025)
von: Wang, Chenguang, et al.
Veröffentlicht: (2025)
A Survey on Trustworthiness in Foundation Models for Medical Image Analysis
von: Shi, Congzhen, et al.
Veröffentlicht: (2024)
von: Shi, Congzhen, et al.
Veröffentlicht: (2024)
POET: Supporting Prompting Creativity and Personalization with Automated Expansion of Text-to-Image Generation
von: Han, Evans Xu, et al.
Veröffentlicht: (2025)
von: Han, Evans Xu, et al.
Veröffentlicht: (2025)
True (VIS) Lies: Analyzing How Generative AI Recognizes Intentionality, Rhetoric, and Misleadingness in Visualization Lies
von: Blasilli, Graziano, et al.
Veröffentlicht: (2026)
von: Blasilli, Graziano, et al.
Veröffentlicht: (2026)
What They Saw, Not Just Where They Looked: Semantic Scanpath Similarity via VLMs and NLP metric
von: Kerkouri, Mohamed Amine, et al.
Veröffentlicht: (2026)
von: Kerkouri, Mohamed Amine, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards Geographic Inclusion in the Evaluation of Text-to-Image Models
von: Hall, Melissa, et al.
Veröffentlicht: (2024) -
Semantic and Expressive Variation in Image Captions Across Languages
von: Ye, Andre, et al.
Veröffentlicht: (2023) -
SwissADT: An Audio Description Translation System for Swiss Languages
von: Fischer, Lukas, et al.
Veröffentlicht: (2024) -
Human-Centred Evaluation of Text-to-Image Generation Models for Self-expression of Mental Distress: A Dataset Based on GPT-4o
von: He, Sui, et al.
Veröffentlicht: (2025) -
Vision-Language Models Suppress Female Representations Under Ambiguous Input
von: Marin-Llobet, Arnau, et al.
Veröffentlicht: (2026)