Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lyu, Yunbo, Yang, Zhou, Niu, Yuqing, Jiang, Jing, Lo, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
New Job, New Gender? Measuring the Social Bias in Image Generation Models
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
What You Trust Is Insecure: Demystifying How Developers (Mis)Use Trusted Execution Environments in Practice
von: Niu, Yuqing, et al.
Veröffentlicht: (2025)
von: Niu, Yuqing, et al.
Veröffentlicht: (2025)
Industry Practitioners Perspectives on AI Model Quality: Perceptions, Challenges, and Solutions
von: Wang, Chenyu, et al.
Veröffentlicht: (2024)
von: Wang, Chenyu, et al.
Veröffentlicht: (2024)
AgentSZZ: Teaching the LLM Agent to Play Detective with Bug-Inducing Commits
von: Lyu, Yunbo, et al.
Veröffentlicht: (2026)
von: Lyu, Yunbo, et al.
Veröffentlicht: (2026)
Benchmarking Image Perturbations for Testing Automated Driving Assistance Systems
von: Lambertenghi, Stefano Carlo, et al.
Veröffentlicht: (2025)
von: Lambertenghi, Stefano Carlo, et al.
Veröffentlicht: (2025)
A Functional Software Reference Architecture for LLM-Integrated Systems
von: Bucaioni, Alessio, et al.
Veröffentlicht: (2025)
von: Bucaioni, Alessio, et al.
Veröffentlicht: (2025)
Artificial Intelligence for Software Architecture: Literature Review and the Road Ahead
von: Bucaioni, Alessio, et al.
Veröffentlicht: (2025)
von: Bucaioni, Alessio, et al.
Veröffentlicht: (2025)
Evaluating and Enhancing Segmentation Model Robustness with Metamorphic Testing
von: Mzoughi, Seif, et al.
Veröffentlicht: (2025)
von: Mzoughi, Seif, et al.
Veröffentlicht: (2025)
Evaluating SZZ Implementations: An Empirical Study on the Linux Kernel
von: Lyu, Yunbo, et al.
Veröffentlicht: (2023)
von: Lyu, Yunbo, et al.
Veröffentlicht: (2023)
Interleaving Large Language Models for Compiler Testing
von: Ni, Yunbo, et al.
Veröffentlicht: (2025)
von: Ni, Yunbo, et al.
Veröffentlicht: (2025)
SLICEMATE: Accurate and Scalable Static Program Slicing via LLM-Powered Agents
von: Chang, Jianming, et al.
Veröffentlicht: (2025)
von: Chang, Jianming, et al.
Veröffentlicht: (2025)
APIDocBooster: An Extract-Then-Abstract Framework Leveraging Large Language Models for Augmenting API Documentation
von: Yang, Chengran, et al.
Veröffentlicht: (2023)
von: Yang, Chengran, et al.
Veröffentlicht: (2023)
Computer Vision Intelligence Test Modeling and Generation: A Case Study on Smart OCR
von: Shu, Jing, et al.
Veröffentlicht: (2024)
von: Shu, Jing, et al.
Veröffentlicht: (2024)
Bias Behind the Wheel: Fairness Testing of Autonomous Driving Systems
von: Li, Xinyue, et al.
Veröffentlicht: (2023)
von: Li, Xinyue, et al.
Veröffentlicht: (2023)
DiffGAN: A Test Generation Approach for Differential Testing of Deep Neural Networks for Image Analysis
von: Aghababaeyan, Zohreh, et al.
Veröffentlicht: (2024)
von: Aghababaeyan, Zohreh, et al.
Veröffentlicht: (2024)
What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs
von: Aich, Abhishek, et al.
Veröffentlicht: (2026)
von: Aich, Abhishek, et al.
Veröffentlicht: (2026)
VEglue: Testing Visual Entailment Systems via Object-Aligned Joint Erasing
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2024)
Hotfixing Large Language Models for Code
von: Yang, Zhou, et al.
Veröffentlicht: (2024)
von: Yang, Zhou, et al.
Veröffentlicht: (2024)
ROMAN: Reward-Orchestrated Multi-Head Attention Network for Autonomous Driving System Testing
von: Chi, Jianlei, et al.
Veröffentlicht: (2026)
von: Chi, Jianlei, et al.
Veröffentlicht: (2026)
TigAug: Data Augmentation for Testing Traffic Light Detection in Autonomous Driving Systems
von: Lu, You, et al.
Veröffentlicht: (2025)
von: Lu, You, et al.
Veröffentlicht: (2025)
Effort-Optimized, Accuracy-Driven Labelling and Validation of Test Inputs for DL Systems: A Mixed-Integer Linear Programming Approach
von: Amini, Mohammad Hossein, et al.
Veröffentlicht: (2025)
von: Amini, Mohammad Hossein, et al.
Veröffentlicht: (2025)
ITKIT: Feasible CT Image Analysis based on SimpleITK and MMEngine
von: Zhang, Yiqin, et al.
Veröffentlicht: (2026)
von: Zhang, Yiqin, et al.
Veröffentlicht: (2026)
Towards Fair Machine Learning Software: Understanding and Addressing Model Bias Through Counterfactual Thinking
von: Wang, Zichong, et al.
Veröffentlicht: (2023)
von: Wang, Zichong, et al.
Veröffentlicht: (2023)
SWAN -- Enabling Fast and Mobile Histopathology Image Annotation through Swipeable Interfaces
von: Banerjee, Sweta, et al.
Veröffentlicht: (2025)
von: Banerjee, Sweta, et al.
Veröffentlicht: (2025)
How Agentic AI Coding Assistants Become the Attacker's Shell
von: Liu, Yue, et al.
Veröffentlicht: (2026)
von: Liu, Yue, et al.
Veröffentlicht: (2026)
"Your AI, My Shell": Demystifying Prompt Injection Attacks on Agentic AI Coding Editors
von: Liu, Yue, et al.
Veröffentlicht: (2025)
von: Liu, Yue, et al.
Veröffentlicht: (2025)
Do AI Agents Really Improve Code Readability?
von: Horikawa, Kyogo, et al.
Veröffentlicht: (2026)
von: Horikawa, Kyogo, et al.
Veröffentlicht: (2026)
"My productivity is boosted, but ..." Demystifying Users' Perception on AI Coding Assistants
von: Lyu, Yunbo, et al.
Veröffentlicht: (2025)
von: Lyu, Yunbo, et al.
Veröffentlicht: (2025)
ARI3D: A Software for Interactive Quantification of Regions in X-Ray CT 3D Images
von: Albrecht, Jan Phillipp, et al.
Veröffentlicht: (2025)
von: Albrecht, Jan Phillipp, et al.
Veröffentlicht: (2025)
ClawMark: A Living-World Benchmark for Multi-Turn, Multi-Day, Multimodal Coworker Agents
von: Meng, Fanqing, et al.
Veröffentlicht: (2026)
von: Meng, Fanqing, et al.
Veröffentlicht: (2026)
Metamorphic Testing for Pose Estimation Systems
von: Duran, Matias, et al.
Veröffentlicht: (2025)
von: Duran, Matias, et al.
Veröffentlicht: (2025)
Distribution-aware Fairness Test Generation
von: Rajan, Sai Sathiesh, et al.
Veröffentlicht: (2023)
von: Rajan, Sai Sathiesh, et al.
Veröffentlicht: (2023)
Foundation Models in Remote Sensing: Evolving from Unimodality to Multimodality
von: Hong, Danfeng, et al.
Veröffentlicht: (2026)
von: Hong, Danfeng, et al.
Veröffentlicht: (2026)
Investigating Traffic Accident Detection Using Multimodal Large Language Models
von: Skender, Ilhan, et al.
Veröffentlicht: (2025)
von: Skender, Ilhan, et al.
Veröffentlicht: (2025)
Large Language Models for Code Analysis: Do LLMs Really Do Their Job?
von: Fang, Chongzhou, et al.
Veröffentlicht: (2023)
von: Fang, Chongzhou, et al.
Veröffentlicht: (2023)
DD-CAM: Minimal Sufficient Explanations for Vision Models Using Delta Debugging
von: Khadka, Krishna, et al.
Veröffentlicht: (2026)
von: Khadka, Krishna, et al.
Veröffentlicht: (2026)
CUARewardBench: A Benchmark for Evaluating Reward Models on Computer-using Agent
von: Lin, Haojia, et al.
Veröffentlicht: (2025)
von: Lin, Haojia, et al.
Veröffentlicht: (2025)
Efficient and Green Large Language Models for Software Engineering: Literature Review, Vision, and the Road Ahead
von: Shi, Jieke, et al.
Veröffentlicht: (2024)
von: Shi, Jieke, et al.
Veröffentlicht: (2024)
Distinguishability-guided Test Program Generation for WebAssembly Runtime Performance Testing
von: Jiang, Shuyao, et al.
Veröffentlicht: (2024)
von: Jiang, Shuyao, et al.
Veröffentlicht: (2024)
GUing: A Mobile GUI Search Engine using a Vision-Language Model
von: Wei, Jialiang, et al.
Veröffentlicht: (2024)
von: Wei, Jialiang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
New Job, New Gender? Measuring the Social Bias in Image Generation Models
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024) -
What You Trust Is Insecure: Demystifying How Developers (Mis)Use Trusted Execution Environments in Practice
von: Niu, Yuqing, et al.
Veröffentlicht: (2025) -
Industry Practitioners Perspectives on AI Model Quality: Perceptions, Challenges, and Solutions
von: Wang, Chenyu, et al.
Veröffentlicht: (2024) -
AgentSZZ: Teaching the LLM Agent to Play Detective with Bug-Inducing Commits
von: Lyu, Yunbo, et al.
Veröffentlicht: (2026) -
Benchmarking Image Perturbations for Testing Automated Driving Assistance Systems
von: Lambertenghi, Stefano Carlo, et al.
Veröffentlicht: (2025)