You Never Know: Quantization Induces Inconsistent Biases in Vision-Language Foundation Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Slyman, Eric, Kanneganti, Anirudh, Hong, Sanghyun, Lee, Stefan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VLSlice: Interactive Vision-and-Language Slice Discovery
di: Slyman, Eric, et al.
Pubblicazione: (2023)
di: Slyman, Eric, et al.
Pubblicazione: (2023)
Hijacking Vision-and-Language Navigation Agents with Adversarial Environmental Attacks
di: Yang, Zijiao, et al.
Pubblicazione: (2024)
di: Yang, Zijiao, et al.
Pubblicazione: (2024)
Statistical Challenges with Dataset Construction: Why You Will Never Have Enough Images
di: Goldman, Josh, et al.
Pubblicazione: (2024)
di: Goldman, Josh, et al.
Pubblicazione: (2024)
No One Knows the State of the Art in Geospatial Foundation Models
di: Corley, Isaac, et al.
Pubblicazione: (2026)
di: Corley, Isaac, et al.
Pubblicazione: (2026)
Know What You do Not Know: Verbalized Uncertainty Estimation Robustness on Corrupted Images in Vision-Language Models
di: Borszukovszki, Mirko, et al.
Pubblicazione: (2025)
di: Borszukovszki, Mirko, et al.
Pubblicazione: (2025)
Demographic Bias of Expert-Level Vision-Language Foundation Models in Medical Imaging
di: Yang, Yuzhe, et al.
Pubblicazione: (2024)
di: Yang, Yuzhe, et al.
Pubblicazione: (2024)
Calibrating MLLM-as-a-judge via Multimodal Bayesian Prompt Ensembles
di: Slyman, Eric, et al.
Pubblicazione: (2025)
di: Slyman, Eric, et al.
Pubblicazione: (2025)
Identifying Implicit Social Biases in Vision-Language Models
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
Vision Language Models are Biased
di: Vo, An, et al.
Pubblicazione: (2025)
di: Vo, An, et al.
Pubblicazione: (2025)
Are Models Biased on Text without Gender-related Language?
di: Belém, Catarina G, et al.
Pubblicazione: (2024)
di: Belém, Catarina G, et al.
Pubblicazione: (2024)
Advanced Knowledge Transfer: Refined Feature Distillation for Zero-Shot Quantization in Edge Computing
di: Hong, Inpyo, et al.
Pubblicazione: (2024)
di: Hong, Inpyo, et al.
Pubblicazione: (2024)
How Does Fine-Tuning Impact Out-of-Distribution Detection for Vision-Language Models?
di: Ming, Yifei, et al.
Pubblicazione: (2023)
di: Ming, Yifei, et al.
Pubblicazione: (2023)
Fantastic Biases (What are They) and Where to Find Them
di: Barriere, Valentin
Pubblicazione: (2024)
di: Barriere, Valentin
Pubblicazione: (2024)
CARES: A Comprehensive Benchmark of Trustworthiness in Medical Vision Language Models
di: Xia, Peng, et al.
Pubblicazione: (2024)
di: Xia, Peng, et al.
Pubblicazione: (2024)
FairDeDup: Detecting and Mitigating Vision-Language Fairness Disparities in Semantic Dataset Deduplication
di: Slyman, Eric, et al.
Pubblicazione: (2024)
di: Slyman, Eric, et al.
Pubblicazione: (2024)
Surgeons Are Indian Males and Speech Therapists Are White Females: Auditing Biases in Vision-Language Models for Healthcare Professionals
di: Siddiqui, Zohaib Hasan, et al.
Pubblicazione: (2025)
di: Siddiqui, Zohaib Hasan, et al.
Pubblicazione: (2025)
BLEnD-Vis: Benchmarking Multimodal Cultural Understanding in Vision Language Models
di: Tan, Bryan Chen Zhengyu, et al.
Pubblicazione: (2025)
di: Tan, Bryan Chen Zhengyu, et al.
Pubblicazione: (2025)
Social Perception of Faces in a Vision-Language Model
di: Hausladen, Carina I., et al.
Pubblicazione: (2024)
di: Hausladen, Carina I., et al.
Pubblicazione: (2024)
TIBET: Identifying and Evaluating Biases in Text-to-Image Generative Models
di: Chinchure, Aditya, et al.
Pubblicazione: (2023)
di: Chinchure, Aditya, et al.
Pubblicazione: (2023)
Emotional Images: Assessing Emotions in Images and Potential Biases in Generative Models
di: Mehta, Maneet, et al.
Pubblicazione: (2024)
di: Mehta, Maneet, et al.
Pubblicazione: (2024)
Bias Begets Bias: The Impact of Biased Embeddings on Diffusion Models
di: Kuchlous, Sahil, et al.
Pubblicazione: (2024)
di: Kuchlous, Sahil, et al.
Pubblicazione: (2024)
Specifying What You Know or Not for Multi-Label Class-Incremental Learning
di: Zhang, Aoting, et al.
Pubblicazione: (2025)
di: Zhang, Aoting, et al.
Pubblicazione: (2025)
A Large Scale Analysis of Gender Biases in Text-to-Image Generative Models
di: Girrbach, Leander, et al.
Pubblicazione: (2025)
di: Girrbach, Leander, et al.
Pubblicazione: (2025)
Exploiting Cultural Biases via Homoglyphs in Text-to-Image Synthesis
di: Struppek, Lukas, et al.
Pubblicazione: (2022)
di: Struppek, Lukas, et al.
Pubblicazione: (2022)
Magnet: We Never Know How Text-to-Image Diffusion Models Work, Until We Learn How Vision-Language Models Function
di: Zhuang, Chenyi, et al.
Pubblicazione: (2024)
di: Zhuang, Chenyi, et al.
Pubblicazione: (2024)
NegVQA: Can Vision Language Models Understand Negation?
di: Zhang, Yuhui, et al.
Pubblicazione: (2025)
di: Zhang, Yuhui, et al.
Pubblicazione: (2025)
Exposing and Addressing Cross-Task Inconsistency in Unified Vision-Language Models
di: Maharana, Adyasha, et al.
Pubblicazione: (2023)
di: Maharana, Adyasha, et al.
Pubblicazione: (2023)
RULE: Reliable Multimodal RAG for Factuality in Medical Vision Language Models
di: Xia, Peng, et al.
Pubblicazione: (2024)
di: Xia, Peng, et al.
Pubblicazione: (2024)
Language-Pretraining-Induced Bias: A Strong Foundation for General Vision Tasks
di: Luo, Yaxin, et al.
Pubblicazione: (2026)
di: Luo, Yaxin, et al.
Pubblicazione: (2026)
Harnessing Input-Adaptive Inference for Efficient VLN
di: Kang, Dongwoo, et al.
Pubblicazione: (2025)
di: Kang, Dongwoo, et al.
Pubblicazione: (2025)
Visually Dehallucinative Instruction Generation: Know What You Don't Know
di: Cha, Sungguk, et al.
Pubblicazione: (2024)
di: Cha, Sungguk, et al.
Pubblicazione: (2024)
Multimodal Language Models Cannot Spot Spatial Inconsistencies
di: Khangaonkar, Om, et al.
Pubblicazione: (2026)
di: Khangaonkar, Om, et al.
Pubblicazione: (2026)
Automated Generation of Challenging Multiple-Choice Questions for Vision Language Model Evaluation
di: Zhang, Yuhui, et al.
Pubblicazione: (2025)
di: Zhang, Yuhui, et al.
Pubblicazione: (2025)
Debiasing Methods for Fairer Neural Models in Vision and Language Research: A Survey
di: Parraga, Otávio, et al.
Pubblicazione: (2022)
di: Parraga, Otávio, et al.
Pubblicazione: (2022)
EaqVLA: Encoding-aligned Quantization for Vision-Language-Action Models
di: Jiang, Feng, et al.
Pubblicazione: (2025)
di: Jiang, Feng, et al.
Pubblicazione: (2025)
Reliable and Responsible Foundation Models: A Comprehensive Survey
di: Yang, Xinyu, et al.
Pubblicazione: (2026)
di: Yang, Xinyu, et al.
Pubblicazione: (2026)
PRISM: Reducing Spurious Implicit Biases in Vision-Language Models with LLM-Guided Embedding Projection
di: Molahasani, Mahdiyar, et al.
Pubblicazione: (2025)
di: Molahasani, Mahdiyar, et al.
Pubblicazione: (2025)
Are Vision Language Models Cross-Cultural Theory of Mind Reasoners?
di: Nazi, Zabir Al, et al.
Pubblicazione: (2025)
di: Nazi, Zabir Al, et al.
Pubblicazione: (2025)
Automatic Teaching Platform on Vision Language Retrieval Augmented Generation
di: Gokhman, Ruslan, et al.
Pubblicazione: (2025)
di: Gokhman, Ruslan, et al.
Pubblicazione: (2025)
Do VLMs Have a Moral Backbone? A Study on the Fragile Morality of Vision-Language Models
di: Liu, Zhining, et al.
Pubblicazione: (2026)
di: Liu, Zhining, et al.
Pubblicazione: (2026)
Documenti analoghi
-
VLSlice: Interactive Vision-and-Language Slice Discovery
di: Slyman, Eric, et al.
Pubblicazione: (2023) -
Hijacking Vision-and-Language Navigation Agents with Adversarial Environmental Attacks
di: Yang, Zijiao, et al.
Pubblicazione: (2024) -
Statistical Challenges with Dataset Construction: Why You Will Never Have Enough Images
di: Goldman, Josh, et al.
Pubblicazione: (2024) -
No One Knows the State of the Art in Geospatial Foundation Models
di: Corley, Isaac, et al.
Pubblicazione: (2026) -
Know What You do Not Know: Verbalized Uncertainty Estimation Robustness on Corrupted Images in Vision-Language Models
di: Borszukovszki, Mirko, et al.
Pubblicazione: (2025)