Debiasing Vison-Language Models with Text-Only Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Yunfan, Jiang, Chaoquan, Lin, Zhiyu, Xiao, Jinlin, Zhang, Jiaming, Sang, Jitao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Debiased Prompt Tuning in Vision-Language Model without Annotations
von: Jiang, Chaoquan, et al.
Veröffentlicht: (2025)
von: Jiang, Chaoquan, et al.
Veröffentlicht: (2025)
Language-assisted Vision Model Debugger: A Sample-Free Approach to Finding and Fixing Bugs
von: Jiang, Chaoquan, et al.
Veröffentlicht: (2023)
von: Jiang, Chaoquan, et al.
Veröffentlicht: (2023)
Investigating and Enhancing Vision-Audio Capability in Omnimodal Large Language Models
von: Hu, Rui, et al.
Veröffentlicht: (2025)
von: Hu, Rui, et al.
Veröffentlicht: (2025)
Not Only Text: Exploring Compositionality of Visual Representations in Vision-Language Models
von: Berasi, Davide, et al.
Veröffentlicht: (2025)
von: Berasi, Davide, et al.
Veröffentlicht: (2025)
PARIC: Probabilistic Attention Regularization for Language Guided Image Classification from Pre-trained Vison Language Models
von: Nautiyal, Mayank, et al.
Veröffentlicht: (2025)
von: Nautiyal, Mayank, et al.
Veröffentlicht: (2025)
AIGCs Confuse AI Too: Investigating and Explaining Synthetic Image-induced Hallucinations in Large Vision-Language Models
von: Gao, Yifei, et al.
Veröffentlicht: (2024)
von: Gao, Yifei, et al.
Veröffentlicht: (2024)
An Experimental Study of Semantic Continuity for Deep Learning Models
von: Wu, Shangxi, et al.
Veröffentlicht: (2020)
von: Wu, Shangxi, et al.
Veröffentlicht: (2020)
You Only Train Once
von: Sakaridis, Christos
Veröffentlicht: (2025)
von: Sakaridis, Christos
Veröffentlicht: (2025)
Robust Emotion Recognition in Context Debiasing
von: Yang, Dingkang, et al.
Veröffentlicht: (2024)
von: Yang, Dingkang, et al.
Veröffentlicht: (2024)
Debiased Negative Mining Improves Out-of-distribution Detection with Pre-trained Vision-Language Models
von: Peng, Bo, et al.
Veröffentlicht: (2026)
von: Peng, Bo, et al.
Veröffentlicht: (2026)
BendVLM: Test-Time Debiasing of Vision-Language Embeddings
von: Gerych, Walter, et al.
Veröffentlicht: (2024)
von: Gerych, Walter, et al.
Veröffentlicht: (2024)
Model Debiasing by Learnable Data Augmentation
von: Morerio, Pietro, et al.
Veröffentlicht: (2024)
von: Morerio, Pietro, et al.
Veröffentlicht: (2024)
Catch-Up Distillation: You Only Need to Train Once for Accelerating Sampling
von: Shao, Shitong, et al.
Veröffentlicht: (2023)
von: Shao, Shitong, et al.
Veröffentlicht: (2023)
NanoVDR: Distilling a 2B Vision-Language Retriever into a 70M Text-Only Encoder for Visual Document Retrieval
von: Liu, Zhuchenyang, et al.
Veröffentlicht: (2026)
von: Liu, Zhuchenyang, et al.
Veröffentlicht: (2026)
Debiasing Large Vision-Language Models by Ablating Protected Attribute Representations
von: Ratzlaff, Neale, et al.
Veröffentlicht: (2024)
von: Ratzlaff, Neale, et al.
Veröffentlicht: (2024)
Navigate Beyond Shortcuts: Debiased Learning through the Lens of Neural Collapse
von: Wang, Yining, et al.
Veröffentlicht: (2024)
von: Wang, Yining, et al.
Veröffentlicht: (2024)
Growing Visual Generative Capacity for Pre-Trained MLLMs
von: Wang, Hanyu, et al.
Veröffentlicht: (2025)
von: Wang, Hanyu, et al.
Veröffentlicht: (2025)
SharpZO: Hybrid Sharpness-Aware Vision Language Model Prompt Tuning via Forward-Only Passes
von: Yang, Yifan, et al.
Veröffentlicht: (2025)
von: Yang, Yifan, et al.
Veröffentlicht: (2025)
Training-free Editioning of Text-to-Image Models
von: Wang, Jinqi, et al.
Veröffentlicht: (2024)
von: Wang, Jinqi, et al.
Veröffentlicht: (2024)
SurrogateSHAP: Training-Free Contributor Attribution for Text-to-Image (T2I) Models
von: Lu, Mingyu, et al.
Veröffentlicht: (2026)
von: Lu, Mingyu, et al.
Veröffentlicht: (2026)
Gradient Extrapolation for Debiased Representation Learning
von: Asaad, Ihab, et al.
Veröffentlicht: (2025)
von: Asaad, Ihab, et al.
Veröffentlicht: (2025)
Training-Only Heterogeneous Image-Patch-Text Graph Supervision for Advancing Few-Shot Learning Adapters
von: Mohammad, Mohammed Rahman Sherif Khan, et al.
Veröffentlicht: (2026)
von: Mohammad, Mohammed Rahman Sherif Khan, et al.
Veröffentlicht: (2026)
DiFiC: Your Diffusion Model Holds the Secret to Fine-Grained Clustering
von: Yang, Ruohong, et al.
Veröffentlicht: (2024)
von: Yang, Ruohong, et al.
Veröffentlicht: (2024)
Text-to-CAD Generation Through Infusing Visual Feedback in Large Language Models
von: Wang, Ruiyu, et al.
Veröffentlicht: (2025)
von: Wang, Ruiyu, et al.
Veröffentlicht: (2025)
Debiasing Diffusion Model: Enhancing Fairness through Latent Representation Learning in Stable Diffusion Model
von: Huang, Lin-Chun, et al.
Veröffentlicht: (2025)
von: Huang, Lin-Chun, et al.
Veröffentlicht: (2025)
Doubly Debiased Test-Time Prompt Tuning for Vision-Language Models
von: Song, Fei, et al.
Veröffentlicht: (2025)
von: Song, Fei, et al.
Veröffentlicht: (2025)
Transitive Vision-Language Prompt Learning for Domain Generalization
von: Wang, Liyuan, et al.
Veröffentlicht: (2024)
von: Wang, Liyuan, et al.
Veröffentlicht: (2024)
Fast-Slow Efficient Training for Multimodal Large Language Models via Visual Token Pruning
von: Zhang, Dingkun, et al.
Veröffentlicht: (2026)
von: Zhang, Dingkun, et al.
Veröffentlicht: (2026)
You Only Look One Step: Accelerating Backpropagation in Diffusion Sampling with Gradient Shortcuts
von: Dou, Hongkun, et al.
Veröffentlicht: (2025)
von: Dou, Hongkun, et al.
Veröffentlicht: (2025)
Learning Decomposable and Debiased Representations via Attribute-Centric Information Bottlenecks
von: Hong, Jinyung, et al.
Veröffentlicht: (2024)
von: Hong, Jinyung, et al.
Veröffentlicht: (2024)
SEM: Sparse Embedding Modulation for Post-Hoc Debiasing of Vision-Language Models
von: Guimard, Quentin, et al.
Veröffentlicht: (2026)
von: Guimard, Quentin, et al.
Veröffentlicht: (2026)
DUDE: Diffusion-Based Unsupervised Cross-Domain Image Retrieval
von: Yang, Ruohong, et al.
Veröffentlicht: (2025)
von: Yang, Ruohong, et al.
Veröffentlicht: (2025)
Personalized Generative Models for Contextual Debiasing
von: Liang, Xinran, et al.
Veröffentlicht: (2026)
von: Liang, Xinran, et al.
Veröffentlicht: (2026)
DeNetDM: Debiasing by Network Depth Modulation
von: Sreelatha, Silpa Vadakkeeveetil, et al.
Veröffentlicht: (2024)
von: Sreelatha, Silpa Vadakkeeveetil, et al.
Veröffentlicht: (2024)
Mirage in the Eyes: Hallucination Attack on Multi-modal Large Language Models with Only Attention Sink
von: Wang, Yining, et al.
Veröffentlicht: (2025)
von: Wang, Yining, et al.
Veröffentlicht: (2025)
High-Fidelity Text-to-Image Generation from Pre-Trained Vision-Language Models via Distribution-Conditioned Diffusion Decoding
von: Hong, Ji Woo, et al.
Veröffentlicht: (2026)
von: Hong, Ji Woo, et al.
Veröffentlicht: (2026)
A Survey on Deep Clustering: From the Prior Perspective
von: Lu, Yiding, et al.
Veröffentlicht: (2024)
von: Lu, Yiding, et al.
Veröffentlicht: (2024)
BiPrompt: Bilateral Prompt Optimization for Visual and Textual Debiasing in Vision-Language Models
von: Gupta, Sunny, et al.
Veröffentlicht: (2026)
von: Gupta, Sunny, et al.
Veröffentlicht: (2026)
Transferring Visual Explainability of Self-Explaining Models to Prediction-Only Models without Additional Training
von: Yoshikawa, Yuya, et al.
Veröffentlicht: (2025)
von: Yoshikawa, Yuya, et al.
Veröffentlicht: (2025)
Forward-Only Continual Learning
von: Chen, Jiao, et al.
Veröffentlicht: (2025)
von: Chen, Jiao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Debiased Prompt Tuning in Vision-Language Model without Annotations
von: Jiang, Chaoquan, et al.
Veröffentlicht: (2025) -
Language-assisted Vision Model Debugger: A Sample-Free Approach to Finding and Fixing Bugs
von: Jiang, Chaoquan, et al.
Veröffentlicht: (2023) -
Investigating and Enhancing Vision-Audio Capability in Omnimodal Large Language Models
von: Hu, Rui, et al.
Veröffentlicht: (2025) -
Not Only Text: Exploring Compositionality of Visual Representations in Vision-Language Models
von: Berasi, Davide, et al.
Veröffentlicht: (2025) -
PARIC: Probabilistic Attention Regularization for Language Guided Image Classification from Pre-trained Vison Language Models
von: Nautiyal, Mayank, et al.
Veröffentlicht: (2025)