Decoupling Augmentation Bias in Prompt Learning for Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Gahyeon, Kim, Sohee, Lee, Seokju |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AAPL: Adding Attributes to Prompt Learning for Vision-Language Models
von: Kim, Gahyeon, et al.
Veröffentlicht: (2024)
von: Kim, Gahyeon, et al.
Veröffentlicht: (2024)
CLIP Can Understand Depth
von: Kim, Sohee, et al.
Veröffentlicht: (2024)
von: Kim, Sohee, et al.
Veröffentlicht: (2024)
FFNet: MetaMixer-based Efficient Convolutional Mixer Design
von: Yun, Seokju, et al.
Veröffentlicht: (2024)
von: Yun, Seokju, et al.
Veröffentlicht: (2024)
Tree of Attributes Prompt Learning for Vision-Language Models
von: Ding, Tong, et al.
Veröffentlicht: (2024)
von: Ding, Tong, et al.
Veröffentlicht: (2024)
RainbowPrompt: Diversity-Enhanced Prompt-Evolving for Continual Learning
von: Hong, Kiseong, et al.
Veröffentlicht: (2025)
von: Hong, Kiseong, et al.
Veröffentlicht: (2025)
Configuring Data Augmentations to Reduce Variance Shift in Positional Embedding of Vision Transformers
von: Kim, Bum Jun, et al.
Veröffentlicht: (2024)
von: Kim, Bum Jun, et al.
Veröffentlicht: (2024)
Differentiable Prompt Learning for Vision Language Models
von: Huang, Zhenhan, et al.
Veröffentlicht: (2024)
von: Huang, Zhenhan, et al.
Veröffentlicht: (2024)
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
von: Park, Jihwan, et al.
Veröffentlicht: (2025)
von: Park, Jihwan, et al.
Veröffentlicht: (2025)
Rethinking Post-Unlearning Behavior of Large Vision-Language Models
von: Kim, Minsung, et al.
Veröffentlicht: (2025)
von: Kim, Minsung, et al.
Veröffentlicht: (2025)
Candidate Pseudolabel Learning: Enhancing Vision-Language Models by Prompt Tuning with Unlabeled Data
von: Zhang, Jiahan, et al.
Veröffentlicht: (2024)
von: Zhang, Jiahan, et al.
Veröffentlicht: (2024)
Diffusion-Based Offline RL for Improved Decision-Making in Augmented ARC Task
von: Kim, Yunho, et al.
Veröffentlicht: (2024)
von: Kim, Yunho, et al.
Veröffentlicht: (2024)
BiPrompt: Bilateral Prompt Optimization for Visual and Textual Debiasing in Vision-Language Models
von: Gupta, Sunny, et al.
Veröffentlicht: (2026)
von: Gupta, Sunny, et al.
Veröffentlicht: (2026)
Interpreting Attention Heads for Image-to-Text Information Flow in Large Vision-Language Models
von: Kim, Jinyeong, et al.
Veröffentlicht: (2025)
von: Kim, Jinyeong, et al.
Veröffentlicht: (2025)
GABInsight: Exploring Gender-Activity Binding Bias in Vision-Language Models
von: Abdollahi, Ali, et al.
Veröffentlicht: (2024)
von: Abdollahi, Ali, et al.
Veröffentlicht: (2024)
Self-Predictive Dynamics for Generalization of Vision-based Reinforcement Learning
von: Kim, Kyungsoo, et al.
Veröffentlicht: (2025)
von: Kim, Kyungsoo, et al.
Veröffentlicht: (2025)
ERGO: Efficient High-Resolution Visual Understanding for Vision-Language Models
von: Lee, Jewon, et al.
Veröffentlicht: (2025)
von: Lee, Jewon, et al.
Veröffentlicht: (2025)
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
von: Kim, Donghu, et al.
Veröffentlicht: (2024)
von: Kim, Donghu, et al.
Veröffentlicht: (2024)
Unveiling the Response of Large Vision-Language Models to Visually Absent Tokens
von: Kim, Sohee, et al.
Veröffentlicht: (2025)
von: Kim, Sohee, et al.
Veröffentlicht: (2025)
VLM in a flash: I/O-Efficient Sparsification of Vision-Language Model via Neuron Chunking
von: Yang, Kichang, et al.
Veröffentlicht: (2025)
von: Yang, Kichang, et al.
Veröffentlicht: (2025)
Cooperative Meta-Learning with Gradient Augmentation
von: Shin, Jongyun, et al.
Veröffentlicht: (2024)
von: Shin, Jongyun, et al.
Veröffentlicht: (2024)
Understanding Retrieval-Augmented Task Adaptation for Vision-Language Models
von: Ming, Yifei, et al.
Veröffentlicht: (2024)
von: Ming, Yifei, et al.
Veröffentlicht: (2024)
Doubly Debiased Test-Time Prompt Tuning for Vision-Language Models
von: Song, Fei, et al.
Veröffentlicht: (2025)
von: Song, Fei, et al.
Veröffentlicht: (2025)
Unified Supervision For Vision-Language Modeling in 3D Computed Tomography
von: Lee, Hao-Chih, et al.
Veröffentlicht: (2025)
von: Lee, Hao-Chih, et al.
Veröffentlicht: (2025)
Improving Consistency Models with Generator-Augmented Flows
von: Issenhuth, Thibaut, et al.
Veröffentlicht: (2024)
von: Issenhuth, Thibaut, et al.
Veröffentlicht: (2024)
ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
REP: Resource-Efficient Prompting for Rehearsal-Free Continual Learning
von: Jeon, Sungho, et al.
Veröffentlicht: (2024)
von: Jeon, Sungho, et al.
Veröffentlicht: (2024)
LAPT: Label-driven Automated Prompt Tuning for OOD Detection with Vision-Language Models
von: Zhang, Yabin, et al.
Veröffentlicht: (2024)
von: Zhang, Yabin, et al.
Veröffentlicht: (2024)
Visual Modality Prompt for Adapting Vision-Language Object Detectors
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
Extreme Blind Image Restoration via Prompt-Conditioned Information Bottleneck
von: Kim, Hongeun, et al.
Veröffentlicht: (2025)
von: Kim, Hongeun, et al.
Veröffentlicht: (2025)
Latent Schrodinger Bridge: Prompting Latent Diffusion for Fast Unpaired Image-to-Image Translation
von: Kim, Jeongsol, et al.
Veröffentlicht: (2024)
von: Kim, Jeongsol, et al.
Veröffentlicht: (2024)
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
von: Kim, Moo Jin, et al.
Veröffentlicht: (2025)
von: Kim, Moo Jin, et al.
Veröffentlicht: (2025)
One Prompt Word is Enough to Boost Adversarial Robustness for Pre-trained Vision-Language Models
von: Li, Lin, et al.
Veröffentlicht: (2024)
von: Li, Lin, et al.
Veröffentlicht: (2024)
Tiled Prompts: Overcoming Prompt Misguidance in Image and Video Super-Resolution
von: Kim, Bryan Sangwoo, et al.
Veröffentlicht: (2026)
von: Kim, Bryan Sangwoo, et al.
Veröffentlicht: (2026)
MedCLM: Learning to Localize and Reason via a CoT-Curriculum in Medical Vision-Language Models
von: Kim, Soo Yong, et al.
Veröffentlicht: (2025)
von: Kim, Soo Yong, et al.
Veröffentlicht: (2025)
Weighted Multi-Prompt Learning with Description-free Large Language Model Distillation
von: Lee, Sua, et al.
Veröffentlicht: (2025)
von: Lee, Sua, et al.
Veröffentlicht: (2025)
Learning to Explore for Stochastic Gradient MCMC
von: Kim, SeungHyun, et al.
Veröffentlicht: (2024)
von: Kim, SeungHyun, et al.
Veröffentlicht: (2024)
Towards Adversarially Robust Vision-Language Models: Insights from Design Choices and Prompt Formatting Techniques
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2024)
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2024)
Demographic Bias of Expert-Level Vision-Language Foundation Models in Medical Imaging
von: Yang, Yuzhe, et al.
Veröffentlicht: (2024)
von: Yang, Yuzhe, et al.
Veröffentlicht: (2024)
Compositional Entailment Learning for Hyperbolic Vision-Language Models
von: Pal, Avik, et al.
Veröffentlicht: (2024)
von: Pal, Avik, et al.
Veröffentlicht: (2024)
Parallel In-context Learning for Large Vision Language Models
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2026)
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
AAPL: Adding Attributes to Prompt Learning for Vision-Language Models
von: Kim, Gahyeon, et al.
Veröffentlicht: (2024) -
CLIP Can Understand Depth
von: Kim, Sohee, et al.
Veröffentlicht: (2024) -
FFNet: MetaMixer-based Efficient Convolutional Mixer Design
von: Yun, Seokju, et al.
Veröffentlicht: (2024) -
Tree of Attributes Prompt Learning for Vision-Language Models
von: Ding, Tong, et al.
Veröffentlicht: (2024) -
RainbowPrompt: Diversity-Enhanced Prompt-Evolving for Continual Learning
von: Hong, Kiseong, et al.
Veröffentlicht: (2025)