GS-Bias: Global-Spatial Bias Learner for Single-Image Test-Time Adaptation of Vision-Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Huang, Zhaohong, Zhang, Yuxin, Xie, Jingjing, Chao, Fei, Ji, Rongrong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Prototype-Based Test-Time Adaptation of Vision-Language Models
di: Huang, Zhaohong, et al.
Pubblicazione: (2026)
di: Huang, Zhaohong, et al.
Pubblicazione: (2026)
ID-Selection: Importance-Diversity Based Visual Token Selection for Efficient LVLM Inference
di: Huang, Zhaohong, et al.
Pubblicazione: (2026)
di: Huang, Zhaohong, et al.
Pubblicazione: (2026)
TextRefiner: Internal Visual Feature as Efficient Refiner for Vision-Language Models Prompt Tuning
di: Xie, Jingjing, et al.
Pubblicazione: (2024)
di: Xie, Jingjing, et al.
Pubblicazione: (2024)
Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation
di: Xie, Jingjing, et al.
Pubblicazione: (2024)
di: Xie, Jingjing, et al.
Pubblicazione: (2024)
DS$^2$Net: Detail-Semantic Deep Supervision Network for Medical Image Segmentation
di: Huang, Zhaohong, et al.
Pubblicazione: (2025)
di: Huang, Zhaohong, et al.
Pubblicazione: (2025)
Mitigating the Bias in the Model for Continual Test-Time Adaptation
di: Chung, Inseop, et al.
Pubblicazione: (2024)
di: Chung, Inseop, et al.
Pubblicazione: (2024)
Efficient Open Set Single Image Test Time Adaptation of Vision Language Models
di: Sreenivas, Manogna, et al.
Pubblicazione: (2024)
di: Sreenivas, Manogna, et al.
Pubblicazione: (2024)
Few-Shot Image Quality Assessment via Adaptation of Vision-Language Models
di: Li, Xudong, et al.
Pubblicazione: (2024)
di: Li, Xudong, et al.
Pubblicazione: (2024)
Investigating Spatial Attention Bias in Vision-Language Models
di: Chaudhary, Aryan, et al.
Pubblicazione: (2025)
di: Chaudhary, Aryan, et al.
Pubblicazione: (2025)
Egocentric Bias in Vision-Language Models
di: Wang, Maijunxian, et al.
Pubblicazione: (2026)
di: Wang, Maijunxian, et al.
Pubblicazione: (2026)
Learning Image Demoireing from Unpaired Real Data
di: Zhong, Yunshan, et al.
Pubblicazione: (2024)
di: Zhong, Yunshan, et al.
Pubblicazione: (2024)
Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models
di: Luo, Gen, et al.
Pubblicazione: (2024)
di: Luo, Gen, et al.
Pubblicazione: (2024)
Pathological Truth Bias in Vision-Language Models
di: Thube, Yash
Pubblicazione: (2025)
di: Thube, Yash
Pubblicazione: (2025)
Realistic Test-Time Adaptation of Vision-Language Models
di: Zanella, Maxime, et al.
Pubblicazione: (2025)
di: Zanella, Maxime, et al.
Pubblicazione: (2025)
Bayesian Test-Time Adaptation for Vision-Language Models
di: Zhou, Lihua, et al.
Pubblicazione: (2025)
di: Zhou, Lihua, et al.
Pubblicazione: (2025)
Efficient Test-Time Adaptation of Vision-Language Models
di: Karmanov, Adilbek, et al.
Pubblicazione: (2024)
di: Karmanov, Adilbek, et al.
Pubblicazione: (2024)
Beyond the Vision Encoder: Identifying and Mitigating Spatial Bias in Large Vision-Language Models
di: Zhu, Yingjie, et al.
Pubblicazione: (2025)
di: Zhu, Yingjie, et al.
Pubblicazione: (2025)
Single Image Test-Time Adaptation for Segmentation
di: Janouskova, Klara, et al.
Pubblicazione: (2023)
di: Janouskova, Klara, et al.
Pubblicazione: (2023)
The Bias of Harmful Label Associations in Vision-Language Models
di: Hazirbas, Caner, et al.
Pubblicazione: (2024)
di: Hazirbas, Caner, et al.
Pubblicazione: (2024)
Ultra-Light Test-Time Adaptation for Vision--Language Models
di: Kim, Byunghyun
Pubblicazione: (2025)
di: Kim, Byunghyun
Pubblicazione: (2025)
Online Gaussian Test-Time Adaptation of Vision-Language Models
di: Fuchs, Clément, et al.
Pubblicazione: (2025)
di: Fuchs, Clément, et al.
Pubblicazione: (2025)
Negation-Aware Test-Time Adaptation for Vision-Language Models
di: Han, Haochen, et al.
Pubblicazione: (2025)
di: Han, Haochen, et al.
Pubblicazione: (2025)
Flatness Guided Test-Time Adaptation for Vision-Language Models
di: Li, Aodi, et al.
Pubblicazione: (2025)
di: Li, Aodi, et al.
Pubblicazione: (2025)
Mechanistic Diagnostics of Spatial Lexical Bias in Multimodal Large Language Model Spatial Reasoning
di: Ma, Chuang, et al.
Pubblicazione: (2026)
di: Ma, Chuang, et al.
Pubblicazione: (2026)
Adaptive Cache Enhancement for Test-Time Adaptation of Vision-Language Models
di: Nguyen, Khanh-Binh, et al.
Pubblicazione: (2025)
di: Nguyen, Khanh-Binh, et al.
Pubblicazione: (2025)
Uncovering Bias in Large Vision-Language Models at Scale with Counterfactuals
di: Howard, Phillip, et al.
Pubblicazione: (2024)
di: Howard, Phillip, et al.
Pubblicazione: (2024)
Semantic Alignment and Reinforcement for Data-Free Quantization of Vision Transformers
di: Zhong, Yunshan, et al.
Pubblicazione: (2024)
di: Zhong, Yunshan, et al.
Pubblicazione: (2024)
Test-Time Computing for Referring Multimodal Large Language Models
di: Wu, Mingrui, et al.
Pubblicazione: (2026)
di: Wu, Mingrui, et al.
Pubblicazione: (2026)
Bidirectional Prototype-Reward co-Evolution for Test-Time Adaptation of Vision-Language Models
di: Qiao, Xiaozhen, et al.
Pubblicazione: (2025)
di: Qiao, Xiaozhen, et al.
Pubblicazione: (2025)
Towards Accurate Post-Training Quantization of Vision Transformers via Error Reduction
di: Zhong, Yunshan, et al.
Pubblicazione: (2024)
di: Zhong, Yunshan, et al.
Pubblicazione: (2024)
Prompting Medical Vision-Language Models to Mitigate Diagnosis Bias by Generating Realistic Dermoscopic Images
di: Munia, Nusrat, et al.
Pubblicazione: (2025)
di: Munia, Nusrat, et al.
Pubblicazione: (2025)
Interpretable Vision-Language Survival Analysis with Ordinal Inductive Bias for Computational Pathology
di: Liu, Pei, et al.
Pubblicazione: (2024)
di: Liu, Pei, et al.
Pubblicazione: (2024)
Interactive Test-Time Adaptation with Reliable Spatial-Temporal Voxels for Multi-Modal Segmentation
di: Cao, Haozhi, et al.
Pubblicazione: (2024)
di: Cao, Haozhi, et al.
Pubblicazione: (2024)
MEBench: A Novel Benchmark for Understanding Mutual Exclusivity Bias in Vision-Language Models
di: Thai, Anh, et al.
Pubblicazione: (2025)
di: Thai, Anh, et al.
Pubblicazione: (2025)
Spatial Re-parameterization for N:M Sparsity
di: Zhang, Yuxin, et al.
Pubblicazione: (2023)
di: Zhang, Yuxin, et al.
Pubblicazione: (2023)
Test-Time Adaptation of Vision-Language Models for Open-Vocabulary Semantic Segmentation
di: Noori, Mehrdad, et al.
Pubblicazione: (2025)
di: Noori, Mehrdad, et al.
Pubblicazione: (2025)
Mitigating Cache Noise in Test-Time Adaptation for Large Vision-Language Models
di: Zhai, Haotian, et al.
Pubblicazione: (2025)
di: Zhai, Haotian, et al.
Pubblicazione: (2025)
Semantic Anchor Transport: Robust Test-Time Adaptation for Vision-Language Models
di: Mishra, Shambhavi, et al.
Pubblicazione: (2024)
di: Mishra, Shambhavi, et al.
Pubblicazione: (2024)
Bias Detection and Rotation-Robustness Mitigation in Vision-Language Models and Generative Image Models
di: Mithila, Tarannum
Pubblicazione: (2026)
di: Mithila, Tarannum
Pubblicazione: (2026)
Benchmarking and Mitigating MCQA Selection Bias of Large Vision-Language Models
di: Atabuzzaman, Md., et al.
Pubblicazione: (2025)
di: Atabuzzaman, Md., et al.
Pubblicazione: (2025)
Documenti analoghi
-
Prototype-Based Test-Time Adaptation of Vision-Language Models
di: Huang, Zhaohong, et al.
Pubblicazione: (2026) -
ID-Selection: Importance-Diversity Based Visual Token Selection for Efficient LVLM Inference
di: Huang, Zhaohong, et al.
Pubblicazione: (2026) -
TextRefiner: Internal Visual Feature as Efficient Refiner for Vision-Language Models Prompt Tuning
di: Xie, Jingjing, et al.
Pubblicazione: (2024) -
Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation
di: Xie, Jingjing, et al.
Pubblicazione: (2024) -
DS$^2$Net: Detail-Semantic Deep Supervision Network for Medical Image Segmentation
di: Huang, Zhaohong, et al.
Pubblicazione: (2025)