When Text and Images Don't Mix: Bias-Correcting Language-Image Similarity Scores for Anomaly Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Goodge, Adam, Hooi, Bryan, Ng, Wee Siong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Spatio-Temporal Foundation Models: Vision, Challenges, and Opportunities
di: Goodge, Adam, et al.
Pubblicazione: (2025)
di: Goodge, Adam, et al.
Pubblicazione: (2025)
SODA: Out-of-Distribution Detection in Domain-Shifted Point Clouds via Neighborhood Propagation
di: Goodge, Adam, et al.
Pubblicazione: (2025)
di: Goodge, Adam, et al.
Pubblicazione: (2025)
Are Anomaly Scores Telling the Whole Story? A Benchmark for Multilevel Anomaly Detection
di: Cao, Tri, et al.
Pubblicazione: (2024)
di: Cao, Tri, et al.
Pubblicazione: (2024)
Re-Scoring Using Image-Language Similarity for Few-Shot Object Detection
di: Jung, Min Jae, et al.
Pubblicazione: (2023)
di: Jung, Min Jae, et al.
Pubblicazione: (2023)
You Don't Need All That Attention: Surgical Memorization Mitigation in Text-to-Image Diffusion Models
di: Zhao, Kairan, et al.
Pubblicazione: (2026)
di: Zhao, Kairan, et al.
Pubblicazione: (2026)
I Detect What I Don't Know: Incremental Anomaly Learning with Stochastic Weight Averaging-Gaussian for Oracle-Free Medical Imaging
di: Yadav, Nand Kumar, et al.
Pubblicazione: (2025)
di: Yadav, Nand Kumar, et al.
Pubblicazione: (2025)
OpenBias: Open-set Bias Detection in Text-to-Image Generative Models
di: D'Incà, Moreno, et al.
Pubblicazione: (2024)
di: D'Incà, Moreno, et al.
Pubblicazione: (2024)
Focus, Don't Prune: Identifying Instruction-Relevant Regions for Information-Rich Image Understanding
di: Kwon, Mincheol, et al.
Pubblicazione: (2026)
di: Kwon, Mincheol, et al.
Pubblicazione: (2026)
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
di: Kalluri, Tarun, et al.
Pubblicazione: (2024)
di: Kalluri, Tarun, et al.
Pubblicazione: (2024)
When Cars Have Stereotypes: Auditing Demographic Bias in Objects from Text-to-Image Models
di: Choi, Dasol, et al.
Pubblicazione: (2025)
di: Choi, Dasol, et al.
Pubblicazione: (2025)
Text-Guided Variational Image Generation for Industrial Anomaly Detection and Segmentation
di: Lee, Mingyu, et al.
Pubblicazione: (2024)
di: Lee, Mingyu, et al.
Pubblicazione: (2024)
Deep Unsupervised Anomaly Detection in Brain Imaging: Large-Scale Benchmarking and Bias Analysis
di: Frotscher, Alexander, et al.
Pubblicazione: (2025)
di: Frotscher, Alexander, et al.
Pubblicazione: (2025)
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
di: Mei, Zhiting, et al.
Pubblicazione: (2025)
di: Mei, Zhiting, et al.
Pubblicazione: (2025)
Text Speaks Louder than Vision: ASCII Art Reveals Textual Biases in Vision-Language Models
di: Wang, Zhaochen, et al.
Pubblicazione: (2025)
di: Wang, Zhaochen, et al.
Pubblicazione: (2025)
Vision Transformers Don't Need Trained Registers
di: Jiang, Nick, et al.
Pubblicazione: (2025)
di: Jiang, Nick, et al.
Pubblicazione: (2025)
Diversify, Don't Fine-Tune: Scaling Up Visual Recognition Training with Synthetic Images
di: Yu, Zhuoran, et al.
Pubblicazione: (2023)
di: Yu, Zhuoran, et al.
Pubblicazione: (2023)
Don't Blame the Annotator: Bias Already Starts in the Annotation Instructions
di: Parmar, Mihir, et al.
Pubblicazione: (2022)
di: Parmar, Mihir, et al.
Pubblicazione: (2022)
When Cultures Meet: Multicultural Text-to-Image Generation
di: Bhalerao, Parth, et al.
Pubblicazione: (2025)
di: Bhalerao, Parth, et al.
Pubblicazione: (2025)
TypeScore: A Text Fidelity Metric for Text-to-Image Generative Models
di: Sampaio, Georgia Gabriela, et al.
Pubblicazione: (2024)
di: Sampaio, Georgia Gabriela, et al.
Pubblicazione: (2024)
Don't Fear Peculiar Activation Functions: EUAF and Beyond
di: Wang, Qianchao, et al.
Pubblicazione: (2024)
di: Wang, Qianchao, et al.
Pubblicazione: (2024)
Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models
di: Woo, Sangmin, et al.
Pubblicazione: (2024)
di: Woo, Sangmin, et al.
Pubblicazione: (2024)
How Bias Binds: Measuring Hidden Associations for Bias Control in Text-to-Image Compositions
di: Li, Jeng-Lin, et al.
Pubblicazione: (2025)
di: Li, Jeng-Lin, et al.
Pubblicazione: (2025)
Exploring Bias in over 100 Text-to-Image Generative Models
di: Vice, Jordan, et al.
Pubblicazione: (2025)
di: Vice, Jordan, et al.
Pubblicazione: (2025)
Anomaly Detection by Effectively Leveraging Synthetic Images
di: Kang, Sungho, et al.
Pubblicazione: (2025)
di: Kang, Sungho, et al.
Pubblicazione: (2025)
A Structured Benchmark for Text-Guided Anomaly Detection: When Language Stops Conditioning the Decision
di: Samele, Stefano, et al.
Pubblicazione: (2026)
di: Samele, Stefano, et al.
Pubblicazione: (2026)
Bias Detection and Rotation-Robustness Mitigation in Vision-Language Models and Generative Image Models
di: Mithila, Tarannum
Pubblicazione: (2026)
di: Mithila, Tarannum
Pubblicazione: (2026)
Implicit Bias Injection Attacks against Text-to-Image Diffusion Models
di: Huang, Huayang, et al.
Pubblicazione: (2025)
di: Huang, Huayang, et al.
Pubblicazione: (2025)
Don't Fight Hallucinations, Use Them: Estimating Image Realism using NLI over Atomic Facts
di: Rykov, Elisei, et al.
Pubblicazione: (2025)
di: Rykov, Elisei, et al.
Pubblicazione: (2025)
Words or Vision: Do Vision-Language Models Have Blind Faith in Text?
di: Deng, Ailin, et al.
Pubblicazione: (2025)
di: Deng, Ailin, et al.
Pubblicazione: (2025)
Don't Deceive Me: Mitigating Gaslighting through Attention Reallocation in LMMs
di: Jiao, Pengkun, et al.
Pubblicazione: (2025)
di: Jiao, Pengkun, et al.
Pubblicazione: (2025)
DCMM-Transformer: Degree-Corrected Mixed-Membership Attention for Medical Imaging
di: Cheng, Huimin, et al.
Pubblicazione: (2025)
di: Cheng, Huimin, et al.
Pubblicazione: (2025)
FAIntbench: A Holistic and Precise Benchmark for Bias Evaluation in Text-to-Image Models
di: Luo, Hanjun, et al.
Pubblicazione: (2024)
di: Luo, Hanjun, et al.
Pubblicazione: (2024)
IM-IAD: Industrial Image Anomaly Detection Benchmark in Manufacturing
di: Xie, Guoyang, et al.
Pubblicazione: (2023)
di: Xie, Guoyang, et al.
Pubblicazione: (2023)
Leveraging Hierarchical Image-Text Misalignment for Universal Fake Image Detection
di: Zhang, Daichi, et al.
Pubblicazione: (2025)
di: Zhang, Daichi, et al.
Pubblicazione: (2025)
Semantic Similarity Score for Measuring Visual Similarity at Semantic Level
di: Fan, Senran, et al.
Pubblicazione: (2024)
di: Fan, Senran, et al.
Pubblicazione: (2024)
Render, Don't Decode: Weight-Space World Models with Latent Structural Disentanglement
di: Nzoyem, Roussel Desmond, et al.
Pubblicazione: (2026)
di: Nzoyem, Roussel Desmond, et al.
Pubblicazione: (2026)
When Robots Should Say "I Don't Know": Benchmarking Abstention in Embodied Question Answering
di: Wu, Tao, et al.
Pubblicazione: (2025)
di: Wu, Tao, et al.
Pubblicazione: (2025)
Foundations for Unfairness in Anomaly Detection -- Case Studies in Facial Imaging Data
di: Livanos, Michael, et al.
Pubblicazione: (2024)
di: Livanos, Michael, et al.
Pubblicazione: (2024)
Large Pre-Training Datasets Don't Always Guarantee Robustness after Fine-Tuning
di: Hwang, Jaedong, et al.
Pubblicazione: (2024)
di: Hwang, Jaedong, et al.
Pubblicazione: (2024)
Invisible Relevance Bias: Text-Image Retrieval Models Prefer AI-Generated Images
di: Xu, Shicheng, et al.
Pubblicazione: (2023)
di: Xu, Shicheng, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Spatio-Temporal Foundation Models: Vision, Challenges, and Opportunities
di: Goodge, Adam, et al.
Pubblicazione: (2025) -
SODA: Out-of-Distribution Detection in Domain-Shifted Point Clouds via Neighborhood Propagation
di: Goodge, Adam, et al.
Pubblicazione: (2025) -
Are Anomaly Scores Telling the Whole Story? A Benchmark for Multilevel Anomaly Detection
di: Cao, Tri, et al.
Pubblicazione: (2024) -
Re-Scoring Using Image-Language Similarity for Few-Shot Object Detection
di: Jung, Min Jae, et al.
Pubblicazione: (2023) -
You Don't Need All That Attention: Surgical Memorization Mitigation in Text-to-Image Diffusion Models
di: Zhao, Kairan, et al.
Pubblicazione: (2026)