DPA: Dual Prototypes Alignment for Unsupervised Adaptation of Vision-Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Ali, Eman, Silva, Sathira, Khan, Muhammad Haris |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Fine-Grained Adaptation of CLIP via a Self-Trained Alignment Score
di: Ali, Eman, et al.
Pubblicazione: (2025)
di: Ali, Eman, et al.
Pubblicazione: (2025)
microCLIP: Unsupervised CLIP Adaptation via Coarse-Fine Token Fusion for Fine-Grained Image Classification
di: Silva, Sathira, et al.
Pubblicazione: (2025)
di: Silva, Sathira, et al.
Pubblicazione: (2025)
Noise-Tolerant Few-Shot Unsupervised Adapter for Vision-Language Models
di: Ali, Eman, et al.
Pubblicazione: (2023)
di: Ali, Eman, et al.
Pubblicazione: (2023)
A Spatiotemporal Approach to Tri-Perspective Representation for 3D Semantic Occupancy Prediction
di: Silva, Sathira, et al.
Pubblicazione: (2024)
di: Silva, Sathira, et al.
Pubblicazione: (2024)
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
di: Sharifdeen, Ashshak, et al.
Pubblicazione: (2025)
di: Sharifdeen, Ashshak, et al.
Pubblicazione: (2025)
Calibration-Aware Prompt Learning for Medical Vision-Language Models
di: Basu, Abhishek, et al.
Pubblicazione: (2025)
di: Basu, Abhishek, et al.
Pubblicazione: (2025)
DATR: Unsupervised Domain Adaptive Detection Transformer with Dataset-Level Adaptation and Prototypical Alignment
di: Han, Jianhong, et al.
Pubblicazione: (2024)
di: Han, Jianhong, et al.
Pubblicazione: (2024)
Prototype-Based Test-Time Adaptation of Vision-Language Models
di: Huang, Zhaohong, et al.
Pubblicazione: (2026)
di: Huang, Zhaohong, et al.
Pubblicazione: (2026)
Improving Single Domain-Generalized Object Detection: A Focus on Diversification and Alignment
di: Danish, Muhammad Sohail, et al.
Pubblicazione: (2024)
di: Danish, Muhammad Sohail, et al.
Pubblicazione: (2024)
Pose-Guided Self-Training with Two-Stage Clustering for Unsupervised Landmark Discovery
di: Tourani, Siddharth, et al.
Pubblicazione: (2024)
di: Tourani, Siddharth, et al.
Pubblicazione: (2024)
Towards Calibrating Prompt Tuning of Vision-Language Models
di: Sharifdeen, Ashshak, et al.
Pubblicazione: (2026)
di: Sharifdeen, Ashshak, et al.
Pubblicazione: (2026)
A-TPT: Angular Diversity Calibration Properties for Test-Time Prompt Tuning of Vision-Language Models
di: Ahamed, Shihab Aaqil, et al.
Pubblicazione: (2025)
di: Ahamed, Shihab Aaqil, et al.
Pubblicazione: (2025)
Robust and Label-Efficient Deep Waste Detection
di: Abid, Hassan, et al.
Pubblicazione: (2025)
di: Abid, Hassan, et al.
Pubblicazione: (2025)
Dual-Foundation Models for Unsupervised Domain Adaptation
di: Cheon, Yerin, et al.
Pubblicazione: (2026)
di: Cheon, Yerin, et al.
Pubblicazione: (2026)
Unsupervised Deep Graph Matching Based on Cycle Consistency
di: Tourani, Siddharth, et al.
Pubblicazione: (2023)
di: Tourani, Siddharth, et al.
Pubblicazione: (2023)
Gradually Vanishing Gap in Prototypical Network for Unsupervised Domain Adaptation
di: Wang, Shanshan, et al.
Pubblicazione: (2024)
di: Wang, Shanshan, et al.
Pubblicazione: (2024)
Dual Prototype Attention for Unsupervised Video Object Segmentation
di: Cho, Suhwan, et al.
Pubblicazione: (2022)
di: Cho, Suhwan, et al.
Pubblicazione: (2022)
Bidirectional Prototype-Reward co-Evolution for Test-Time Adaptation of Vision-Language Models
di: Qiao, Xiaozhen, et al.
Pubblicazione: (2025)
di: Qiao, Xiaozhen, et al.
Pubblicazione: (2025)
CountZES: Counting via Zero-Shot Exemplar Selection
di: Siddiqui, Muhammad Ibraheem, et al.
Pubblicazione: (2025)
di: Siddiqui, Muhammad Ibraheem, et al.
Pubblicazione: (2025)
Class-Aware Prototype Learning with Negative Contrast for Test-Time Adaptation of Vision-Language Models
di: Qiao, Xiaozhen, et al.
Pubblicazione: (2025)
di: Qiao, Xiaozhen, et al.
Pubblicazione: (2025)
Exploring the Benefits of Vision Foundation Models for Unsupervised Domain Adaptation
di: Englert, Brunó B., et al.
Pubblicazione: (2024)
di: Englert, Brunó B., et al.
Pubblicazione: (2024)
Dual Prototype Evolving for Test-Time Generalization of Vision-Language Models
di: Zhang, Ce, et al.
Pubblicazione: (2024)
di: Zhang, Ce, et al.
Pubblicazione: (2024)
Prompt-based Distribution Alignment for Unsupervised Domain Adaptation
di: Bai, Shuanghao, et al.
Pubblicazione: (2023)
di: Bai, Shuanghao, et al.
Pubblicazione: (2023)
GEOBench-VLM: Benchmarking Vision-Language Models for Geospatial Tasks
di: Danish, Muhammad Sohail, et al.
Pubblicazione: (2024)
di: Danish, Muhammad Sohail, et al.
Pubblicazione: (2024)
Improving Pseudo-labelling and Enhancing Robustness for Semi-Supervised Domain Generalization
di: Khan, Adnan, et al.
Pubblicazione: (2024)
di: Khan, Adnan, et al.
Pubblicazione: (2024)
TerraFM: A Scalable Foundation Model for Unified Multisensor Earth Observation
di: Danish, Muhammad Sohail, et al.
Pubblicazione: (2025)
di: Danish, Muhammad Sohail, et al.
Pubblicazione: (2025)
TLAC: Two-stage LMM Augmented CLIP for Zero-Shot Classification
di: Munir, Ans, et al.
Pubblicazione: (2025)
di: Munir, Ans, et al.
Pubblicazione: (2025)
Compositional Zero-Shot Learning: A Survey
di: Munir, Ans, et al.
Pubblicazione: (2025)
di: Munir, Ans, et al.
Pubblicazione: (2025)
Multi-Prompt Alignment for Multi-Source Unsupervised Domain Adaptation
di: Chen, Haoran, et al.
Pubblicazione: (2022)
di: Chen, Haoran, et al.
Pubblicazione: (2022)
Unsupervised Domain Adaptation via Content Alignment for Hippocampus Segmentation
di: Kalabizadeh, Hoda, et al.
Pubblicazione: (2025)
di: Kalabizadeh, Hoda, et al.
Pubblicazione: (2025)
Dynamic Multimodal Prototype Learning in Vision-Language Models
di: Zhu, Xingyu, et al.
Pubblicazione: (2025)
di: Zhu, Xingyu, et al.
Pubblicazione: (2025)
Subspace Alignment for Vision-Language Model Test-time Adaptation
di: Zeng, Zhichen, et al.
Pubblicazione: (2026)
di: Zeng, Zhichen, et al.
Pubblicazione: (2026)
Supervised Classification Heads as Semantic Prototypes: Unlocking Vision-Language Alignment via Weight Recycling
di: Méndez, David, et al.
Pubblicazione: (2026)
di: Méndez, David, et al.
Pubblicazione: (2026)
Unsupervised Part Discovery via Dual Representation Alignment
di: Xia, Jiahao, et al.
Pubblicazione: (2024)
di: Xia, Jiahao, et al.
Pubblicazione: (2024)
FrogDogNet: Fourier frequency Retained visual prompt Output Guidance for Domain Generalization of CLIP in Remote Sensing
di: Gunduboina, Hariseetharam, et al.
Pubblicazione: (2025)
di: Gunduboina, Hariseetharam, et al.
Pubblicazione: (2025)
Integrating Frequency-Domain Representations with Low-Rank Adaptation in Vision-Language Models
di: Khan, Md Azim, et al.
Pubblicazione: (2025)
di: Khan, Md Azim, et al.
Pubblicazione: (2025)
Semi-Supervised Few-Shot Adaptation of Vision-Language Models
di: Silva-Rodríguez, Julio, et al.
Pubblicazione: (2026)
di: Silva-Rodríguez, Julio, et al.
Pubblicazione: (2026)
Temporal Prototyping and Hierarchical Alignment for Unsupervised Video-based Visible-Infrared Person Re-Identification
di: Li, Zhiyong, et al.
Pubblicazione: (2026)
di: Li, Zhiyong, et al.
Pubblicazione: (2026)
Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
di: Maaz, Muhammad, et al.
Pubblicazione: (2023)
di: Maaz, Muhammad, et al.
Pubblicazione: (2023)
Prototypical Distillation and Debiased Tuning for Black-box Unsupervised Domain Adaptation
di: Liang, Jian, et al.
Pubblicazione: (2024)
di: Liang, Jian, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Towards Fine-Grained Adaptation of CLIP via a Self-Trained Alignment Score
di: Ali, Eman, et al.
Pubblicazione: (2025) -
microCLIP: Unsupervised CLIP Adaptation via Coarse-Fine Token Fusion for Fine-Grained Image Classification
di: Silva, Sathira, et al.
Pubblicazione: (2025) -
Noise-Tolerant Few-Shot Unsupervised Adapter for Vision-Language Models
di: Ali, Eman, et al.
Pubblicazione: (2023) -
A Spatiotemporal Approach to Tri-Perspective Representation for 3D Semantic Occupancy Prediction
di: Silva, Sathira, et al.
Pubblicazione: (2024) -
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
di: Sharifdeen, Ashshak, et al.
Pubblicazione: (2025)