AAPL: Adding Attributes to Prompt Learning for Vision-Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Gahyeon, Kim, Sohee, Lee, Seokju |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Decoupling Augmentation Bias in Prompt Learning for Vision-Language Models
di: Kim, Gahyeon, et al.
Pubblicazione: (2025)
di: Kim, Gahyeon, et al.
Pubblicazione: (2025)
CLIP Can Understand Depth
di: Kim, Sohee, et al.
Pubblicazione: (2024)
di: Kim, Sohee, et al.
Pubblicazione: (2024)
Tree of Attributes Prompt Learning for Vision-Language Models
di: Ding, Tong, et al.
Pubblicazione: (2024)
di: Ding, Tong, et al.
Pubblicazione: (2024)
FFNet: MetaMixer-based Efficient Convolutional Mixer Design
di: Yun, Seokju, et al.
Pubblicazione: (2024)
di: Yun, Seokju, et al.
Pubblicazione: (2024)
RainbowPrompt: Diversity-Enhanced Prompt-Evolving for Continual Learning
di: Hong, Kiseong, et al.
Pubblicazione: (2025)
di: Hong, Kiseong, et al.
Pubblicazione: (2025)
Differentiable Prompt Learning for Vision Language Models
di: Huang, Zhenhan, et al.
Pubblicazione: (2024)
di: Huang, Zhenhan, et al.
Pubblicazione: (2024)
Private Attribute Inference from Images with Vision-Language Models
di: Tömekçe, Batuhan, et al.
Pubblicazione: (2024)
di: Tömekçe, Batuhan, et al.
Pubblicazione: (2024)
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
di: Park, Jihwan, et al.
Pubblicazione: (2025)
di: Park, Jihwan, et al.
Pubblicazione: (2025)
Rethinking Post-Unlearning Behavior of Large Vision-Language Models
di: Kim, Minsung, et al.
Pubblicazione: (2025)
di: Kim, Minsung, et al.
Pubblicazione: (2025)
Candidate Pseudolabel Learning: Enhancing Vision-Language Models by Prompt Tuning with Unlabeled Data
di: Zhang, Jiahan, et al.
Pubblicazione: (2024)
di: Zhang, Jiahan, et al.
Pubblicazione: (2024)
BiPrompt: Bilateral Prompt Optimization for Visual and Textual Debiasing in Vision-Language Models
di: Gupta, Sunny, et al.
Pubblicazione: (2026)
di: Gupta, Sunny, et al.
Pubblicazione: (2026)
Interpreting Attention Heads for Image-to-Text Information Flow in Large Vision-Language Models
di: Kim, Jinyeong, et al.
Pubblicazione: (2025)
di: Kim, Jinyeong, et al.
Pubblicazione: (2025)
Self-Predictive Dynamics for Generalization of Vision-based Reinforcement Learning
di: Kim, Kyungsoo, et al.
Pubblicazione: (2025)
di: Kim, Kyungsoo, et al.
Pubblicazione: (2025)
ERGO: Efficient High-Resolution Visual Understanding for Vision-Language Models
di: Lee, Jewon, et al.
Pubblicazione: (2025)
di: Lee, Jewon, et al.
Pubblicazione: (2025)
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
di: Kim, Donghu, et al.
Pubblicazione: (2024)
di: Kim, Donghu, et al.
Pubblicazione: (2024)
Unveiling the Response of Large Vision-Language Models to Visually Absent Tokens
di: Kim, Sohee, et al.
Pubblicazione: (2025)
di: Kim, Sohee, et al.
Pubblicazione: (2025)
VLM in a flash: I/O-Efficient Sparsification of Vision-Language Model via Neuron Chunking
di: Yang, Kichang, et al.
Pubblicazione: (2025)
di: Yang, Kichang, et al.
Pubblicazione: (2025)
Doubly Debiased Test-Time Prompt Tuning for Vision-Language Models
di: Song, Fei, et al.
Pubblicazione: (2025)
di: Song, Fei, et al.
Pubblicazione: (2025)
Unified Supervision For Vision-Language Modeling in 3D Computed Tomography
di: Lee, Hao-Chih, et al.
Pubblicazione: (2025)
di: Lee, Hao-Chih, et al.
Pubblicazione: (2025)
ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning
di: Kim, Taewhan, et al.
Pubblicazione: (2024)
di: Kim, Taewhan, et al.
Pubblicazione: (2024)
REP: Resource-Efficient Prompting for Rehearsal-Free Continual Learning
di: Jeon, Sungho, et al.
Pubblicazione: (2024)
di: Jeon, Sungho, et al.
Pubblicazione: (2024)
LAPT: Label-driven Automated Prompt Tuning for OOD Detection with Vision-Language Models
di: Zhang, Yabin, et al.
Pubblicazione: (2024)
di: Zhang, Yabin, et al.
Pubblicazione: (2024)
Visual Modality Prompt for Adapting Vision-Language Object Detectors
di: Medeiros, Heitor R., et al.
Pubblicazione: (2024)
di: Medeiros, Heitor R., et al.
Pubblicazione: (2024)
Extreme Blind Image Restoration via Prompt-Conditioned Information Bottleneck
di: Kim, Hongeun, et al.
Pubblicazione: (2025)
di: Kim, Hongeun, et al.
Pubblicazione: (2025)
Latent Schrodinger Bridge: Prompting Latent Diffusion for Fast Unpaired Image-to-Image Translation
di: Kim, Jeongsol, et al.
Pubblicazione: (2024)
di: Kim, Jeongsol, et al.
Pubblicazione: (2024)
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
di: Kim, Moo Jin, et al.
Pubblicazione: (2025)
di: Kim, Moo Jin, et al.
Pubblicazione: (2025)
One Prompt Word is Enough to Boost Adversarial Robustness for Pre-trained Vision-Language Models
di: Li, Lin, et al.
Pubblicazione: (2024)
di: Li, Lin, et al.
Pubblicazione: (2024)
Tiled Prompts: Overcoming Prompt Misguidance in Image and Video Super-Resolution
di: Kim, Bryan Sangwoo, et al.
Pubblicazione: (2026)
di: Kim, Bryan Sangwoo, et al.
Pubblicazione: (2026)
PFM-VEPAR: Prompting Foundation Models for RGB-Event Camera based Pedestrian Attribute Recognition
di: Xu, Minghe, et al.
Pubblicazione: (2026)
di: Xu, Minghe, et al.
Pubblicazione: (2026)
Learning to Explore for Stochastic Gradient MCMC
di: Kim, SeungHyun, et al.
Pubblicazione: (2024)
di: Kim, SeungHyun, et al.
Pubblicazione: (2024)
MedCLM: Learning to Localize and Reason via a CoT-Curriculum in Medical Vision-Language Models
di: Kim, Soo Yong, et al.
Pubblicazione: (2025)
di: Kim, Soo Yong, et al.
Pubblicazione: (2025)
Weighted Multi-Prompt Learning with Description-free Large Language Model Distillation
di: Lee, Sua, et al.
Pubblicazione: (2025)
di: Lee, Sua, et al.
Pubblicazione: (2025)
Towards Adversarially Robust Vision-Language Models: Insights from Design Choices and Prompt Formatting Techniques
di: Bhagwatkar, Rishika, et al.
Pubblicazione: (2024)
di: Bhagwatkar, Rishika, et al.
Pubblicazione: (2024)
Configuring Data Augmentations to Reduce Variance Shift in Positional Embedding of Vision Transformers
di: Kim, Bum Jun, et al.
Pubblicazione: (2024)
di: Kim, Bum Jun, et al.
Pubblicazione: (2024)
Compositional Entailment Learning for Hyperbolic Vision-Language Models
di: Pal, Avik, et al.
Pubblicazione: (2024)
di: Pal, Avik, et al.
Pubblicazione: (2024)
Parallel In-context Learning for Large Vision Language Models
di: Yamaguchi, Shin'ya, et al.
Pubblicazione: (2026)
di: Yamaguchi, Shin'ya, et al.
Pubblicazione: (2026)
FedCAR: Cross-client Adaptive Re-weighting for Generative Models in Federated Learning
di: Kim, Minjun, et al.
Pubblicazione: (2024)
di: Kim, Minjun, et al.
Pubblicazione: (2024)
FedWSQ: Efficient Federated Learning with Weight Standardization and Distribution-Aware Non-Uniform Quantization
di: Kim, Seung-Wook, et al.
Pubblicazione: (2025)
di: Kim, Seung-Wook, et al.
Pubblicazione: (2025)
CLEFT: Language-Image Contrastive Learning with Efficient Large Language Model and Prompt Fine-Tuning
di: Du, Yuexi, et al.
Pubblicazione: (2024)
di: Du, Yuexi, et al.
Pubblicazione: (2024)
Masking Teacher and Reinforcing Student for Distilling Vision-Language Models
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2025)
di: Lee, Byung-Kwan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Decoupling Augmentation Bias in Prompt Learning for Vision-Language Models
di: Kim, Gahyeon, et al.
Pubblicazione: (2025) -
CLIP Can Understand Depth
di: Kim, Sohee, et al.
Pubblicazione: (2024) -
Tree of Attributes Prompt Learning for Vision-Language Models
di: Ding, Tong, et al.
Pubblicazione: (2024) -
FFNet: MetaMixer-based Efficient Convolutional Mixer Design
di: Yun, Seokju, et al.
Pubblicazione: (2024) -
RainbowPrompt: Diversity-Enhanced Prompt-Evolving for Continual Learning
di: Hong, Kiseong, et al.
Pubblicazione: (2025)