IntCoOp: Interpretability-Aware Vision-Language Prompt Tuning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ghosal, Soumya Suvra, Basu, Samyadeep, Feizi, Soheil, Manocha, Dinesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2024)
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2024)
Rethinking Artistic Copyright Infringements in the Era of Text-to-Image Generative Models
von: Moayeri, Mazda, et al.
Veröffentlicht: (2024)
von: Moayeri, Mazda, et al.
Veröffentlicht: (2024)
Towards a Systematic Evaluation of Hallucinations in Large-Vision Language Models
von: Seth, Ashish, et al.
Veröffentlicht: (2024)
von: Seth, Ashish, et al.
Veröffentlicht: (2024)
PromptRefine: Enhancing Few-Shot Performance on Low-Resource Indic Languages with Example Selection from Related Example Banks
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2024)
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2024)
IConMark: Robust Interpretable Concept-Based Watermark For AI Images
von: Sadasivan, Vinu Sankar, et al.
Veröffentlicht: (2025)
von: Sadasivan, Vinu Sankar, et al.
Veröffentlicht: (2025)
A Closer Look at Bias and Chain-of-Thought Faithfulness of Large (Vision) Language Models
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2025)
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2025)
Understanding Information Storage and Transfer in Multi-modal Large Language Models
von: Basu, Samyadeep, et al.
Veröffentlicht: (2024)
von: Basu, Samyadeep, et al.
Veröffentlicht: (2024)
Distilling Knowledge from Text-to-Image Generative Models Improves Visio-Linguistic Reasoning in CLIP
von: Basu, Samyadeep, et al.
Veröffentlicht: (2023)
von: Basu, Samyadeep, et al.
Veröffentlicht: (2023)
CAPT: Confusion-Aware Prompt Tuning for Reducing Vision-Language Misalignment
von: Shao, Maoyuan, et al.
Veröffentlicht: (2026)
von: Shao, Maoyuan, et al.
Veröffentlicht: (2026)
GroupCoOp: Group-robust Fine-tuning via Group Prompt Learning
von: Kim, Nayeong, et al.
Veröffentlicht: (2025)
von: Kim, Nayeong, et al.
Veröffentlicht: (2025)
How Learnable Grids Recover Fine Detail in Low Dimensions: A Neural Tangent Kernel Analysis of Multigrid Parametric Encodings
von: Audia, Samuel, et al.
Veröffentlicht: (2025)
von: Audia, Samuel, et al.
Veröffentlicht: (2025)
SliderEdit: Continuous Image Editing with Fine-Grained Instruction Control
von: Zarei, Arman, et al.
Veröffentlicht: (2025)
von: Zarei, Arman, et al.
Veröffentlicht: (2025)
Localizing Knowledge in Diffusion Transformers
von: Zarei, Arman, et al.
Veröffentlicht: (2025)
von: Zarei, Arman, et al.
Veröffentlicht: (2025)
VisRef: Visual Refocusing while Thinking Improves Test-Time Scaling in Multi-Modal Large Reasoning Models
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2026)
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2026)
Adversarial Prompt Tuning for Vision-Language Models
von: Zhang, Jiaming, et al.
Veröffentlicht: (2023)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2023)
LPT: Less-overfitting Prompt Tuning for Vision-Language Model
von: Ding, Chenhao, et al.
Veröffentlicht: (2024)
von: Ding, Chenhao, et al.
Veröffentlicht: (2024)
Prompt Tuning with Soft Context Sharing for Vision-Language Models
von: Ding, Kun, et al.
Veröffentlicht: (2022)
von: Ding, Kun, et al.
Veröffentlicht: (2022)
Tuning Vision-Language Models with Candidate Labels by Prompt Alignment
von: Zhang, Zhifang, et al.
Veröffentlicht: (2024)
von: Zhang, Zhifang, et al.
Veröffentlicht: (2024)
NAP-Tuning: Neural Augmented Prompt Tuning for Adversarially Robust Vision-Language Models
von: Zhang, Jiaming, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2025)
Hierarchy-Aware Fine-Tuning of Vision-Language Models
von: Li, Jiayu, et al.
Veröffentlicht: (2025)
von: Li, Jiayu, et al.
Veröffentlicht: (2025)
Improving Compositional Attribute Binding in Text-to-Image Generative Models via Enhanced Text Embeddings
von: Zarei, Arman, et al.
Veröffentlicht: (2024)
von: Zarei, Arman, et al.
Veröffentlicht: (2024)
Biomed-DPT: Dual Modality Prompt Tuning for Biomedical Vision-Language Models
von: Peng, Wei, et al.
Veröffentlicht: (2025)
von: Peng, Wei, et al.
Veröffentlicht: (2025)
XCoOp: Explainable Prompt Learning for Computer-Aided Diagnosis via Concept-guided Context Optimization
von: Bie, Yequan, et al.
Veröffentlicht: (2024)
von: Bie, Yequan, et al.
Veröffentlicht: (2024)
Efficient Prompt Tuning of Large Vision-Language Model for Fine-Grained Ship Classification
von: Lan, Long, et al.
Veröffentlicht: (2024)
von: Lan, Long, et al.
Veröffentlicht: (2024)
ZSPAPrune: Zero-Shot Prompt-Aware Token Pruning for Vision-Language Models
von: Zhang, Pu, et al.
Veröffentlicht: (2025)
von: Zhang, Pu, et al.
Veröffentlicht: (2025)
3D-Aware Vision-Language Models Fine-Tuning with Geometric Distillation
von: Lee, Seonho, et al.
Veröffentlicht: (2025)
von: Lee, Seonho, et al.
Veröffentlicht: (2025)
BiomedCoOp: Learning to Prompt for Biomedical Vision-Language Models
von: Koleilat, Taha, et al.
Veröffentlicht: (2024)
von: Koleilat, Taha, et al.
Veröffentlicht: (2024)
On Mechanistic Knowledge Localization in Text-to-Image Generative Models
von: Basu, Samyadeep, et al.
Veröffentlicht: (2024)
von: Basu, Samyadeep, et al.
Veröffentlicht: (2024)
Prompt-CAM: Making Vision Transformers Interpretable for Fine-Grained Analysis
von: Chowdhury, Arpita, et al.
Veröffentlicht: (2025)
von: Chowdhury, Arpita, et al.
Veröffentlicht: (2025)
FDBPL: Faster Distillation-Based Prompt Learning for Region-Aware Vision-Language Models Adaptation
von: Zhang, Zherui, et al.
Veröffentlicht: (2025)
von: Zhang, Zherui, et al.
Veröffentlicht: (2025)
Doubly Debiased Test-Time Prompt Tuning for Vision-Language Models
von: Song, Fei, et al.
Veröffentlicht: (2025)
von: Song, Fei, et al.
Veröffentlicht: (2025)
ClinCoT: Clinical-Aware Visual Chain-of-Thought for Medical Vision Language Models
von: Liu, Xiwei, et al.
Veröffentlicht: (2026)
von: Liu, Xiwei, et al.
Veröffentlicht: (2026)
Revealing the Ancient Beauty: Digital Reconstruction of Temple Tiles using Computer Vision
von: Basu, Arkaprabha
Veröffentlicht: (2025)
von: Basu, Arkaprabha
Veröffentlicht: (2025)
FairLLaVA: Fairness-Aware Parameter-Efficient Fine-Tuning for Large Vision-Language Assistants
von: Bhosale, Mahesh, et al.
Veröffentlicht: (2026)
von: Bhosale, Mahesh, et al.
Veröffentlicht: (2026)
Adversarial Prompt Distillation for Vision-Language Models
von: Luo, Lin, et al.
Veröffentlicht: (2024)
von: Luo, Lin, et al.
Veröffentlicht: (2024)
Evolving Prompt Adaptation for Vision-Language Models
von: Zhang, Enming, et al.
Veröffentlicht: (2026)
von: Zhang, Enming, et al.
Veröffentlicht: (2026)
Grounding DINO-US-SAM: Text-Prompted Multi-Organ Segmentation in Ultrasound with LoRA-Tuned Vision-Language Models
von: Rasaee, Hamza, et al.
Veröffentlicht: (2025)
von: Rasaee, Hamza, et al.
Veröffentlicht: (2025)
Historical Test-time Prompt Tuning for Vision Foundation Models
von: Zhang, Jingyi, et al.
Veröffentlicht: (2024)
von: Zhang, Jingyi, et al.
Veröffentlicht: (2024)
ADAPT to Robustify Prompt Tuning Vision Transformers
von: Eskandar, Masih, et al.
Veröffentlicht: (2024)
von: Eskandar, Masih, et al.
Veröffentlicht: (2024)
Target Prompting for Information Extraction with Vision Language Model
von: Medhi, Dipankar
Veröffentlicht: (2024)
von: Medhi, Dipankar
Veröffentlicht: (2024)
Ähnliche Einträge
-
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2024) -
Rethinking Artistic Copyright Infringements in the Era of Text-to-Image Generative Models
von: Moayeri, Mazda, et al.
Veröffentlicht: (2024) -
Towards a Systematic Evaluation of Hallucinations in Large-Vision Language Models
von: Seth, Ashish, et al.
Veröffentlicht: (2024) -
PromptRefine: Enhancing Few-Shot Performance on Low-Resource Indic Languages with Example Selection from Related Example Banks
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2024) -
IConMark: Robust Interpretable Concept-Based Watermark For AI Images
von: Sadasivan, Vinu Sankar, et al.
Veröffentlicht: (2025)