Optimizing CLIP Models for Image Retrieval with Maintained Joint-Embedding Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Schall, Konstantin, Barthel, Kai Uwe, Hezel, Nico, Jung, Klaus |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Creating Sorted Grid Layouts with Gradient-based Optimization
von: Barthel, Kai Uwe, et al.
Veröffentlicht: (2025)
von: Barthel, Kai Uwe, et al.
Veröffentlicht: (2025)
Permutation Learning with Only N Parameters: From SoftSort to Self-Organizing Gaussians
von: Barthel, Kai Uwe, et al.
Veröffentlicht: (2025)
von: Barthel, Kai Uwe, et al.
Veröffentlicht: (2025)
IsoCLIP: Decomposing CLIP Projectors for Efficient Intra-modal Alignment
von: Magistri, Simone, et al.
Veröffentlicht: (2026)
von: Magistri, Simone, et al.
Veröffentlicht: (2026)
Contrast-Aware Calibration for Fine-Tuned CLIP: Leveraging Image-Text Alignment
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
SHED: Style-Homogenized Embedding Alignment for Domain Generalization
von: Gan, Kai, et al.
Veröffentlicht: (2026)
von: Gan, Kai, et al.
Veröffentlicht: (2026)
SatCLIP: Global, General-Purpose Location Embeddings with Satellite Imagery
von: Klemmer, Konstantin, et al.
Veröffentlicht: (2023)
von: Klemmer, Konstantin, et al.
Veröffentlicht: (2023)
Isotropic3D: Image-to-3D Generation Based on a Single CLIP Embedding
von: Liu, Pengkun, et al.
Veröffentlicht: (2024)
von: Liu, Pengkun, et al.
Veröffentlicht: (2024)
NeuCLIP: Efficient Large-Scale CLIP Training with Neural Normalizer Optimization
von: Wei, Xiyuan, et al.
Veröffentlicht: (2025)
von: Wei, Xiyuan, et al.
Veröffentlicht: (2025)
FastCLIP: A Suite of Optimization Techniques to Accelerate CLIP Training with Limited Resources
von: Wei, Xiyuan, et al.
Veröffentlicht: (2024)
von: Wei, Xiyuan, et al.
Veröffentlicht: (2024)
Interpreting CLIP with Sparse Linear Concept Embeddings (SpLiCE)
von: Bhalla, Usha, et al.
Veröffentlicht: (2024)
von: Bhalla, Usha, et al.
Veröffentlicht: (2024)
Breaking the Limits of Open-Weight CLIP: An Optimization Framework for Self-supervised Fine-tuning of CLIP
von: Mehta, Anant, et al.
Veröffentlicht: (2026)
von: Mehta, Anant, et al.
Veröffentlicht: (2026)
Anomaly Detection by Clustering DINO Embeddings using a Dirichlet Process Mixture
von: Schulthess, Nico, et al.
Veröffentlicht: (2025)
von: Schulthess, Nico, et al.
Veröffentlicht: (2025)
Topological Alignment of Shared Vision-Language Embedding Space
von: You, Junwon, et al.
Veröffentlicht: (2025)
von: You, Junwon, et al.
Veröffentlicht: (2025)
Detecting AI-Generated Images via CLIP
von: Moskowitz, A. G., et al.
Veröffentlicht: (2024)
von: Moskowitz, A. G., et al.
Veröffentlicht: (2024)
Semantic Token Reweighting for Interpretable and Controllable Text Embeddings in CLIP
von: Kim, Eunji, et al.
Veröffentlicht: (2024)
von: Kim, Eunji, et al.
Veröffentlicht: (2024)
Denoising with a Joint-Embedding Predictive Architecture
von: Chen, Dengsheng, et al.
Veröffentlicht: (2024)
von: Chen, Dengsheng, et al.
Veröffentlicht: (2024)
Realistic Unsupervised CLIP Fine-tuning with Universal Entropy Optimization
von: Liang, Jian, et al.
Veröffentlicht: (2023)
von: Liang, Jian, et al.
Veröffentlicht: (2023)
TiC-CLIP: Continual Training of CLIP Models
von: Garg, Saurabh, et al.
Veröffentlicht: (2023)
von: Garg, Saurabh, et al.
Veröffentlicht: (2023)
Maintaining Performance with Less Data
von: Sanderson, Dominic, et al.
Veröffentlicht: (2022)
von: Sanderson, Dominic, et al.
Veröffentlicht: (2022)
MoP-CLIP: A Mixture of Prompt-Tuned CLIP Models for Domain Incremental Learning
von: Nicolas, Julien, et al.
Veröffentlicht: (2023)
von: Nicolas, Julien, et al.
Veröffentlicht: (2023)
MIP: CLIP-based Image Reconstruction from PEFT Gradients
von: Zhou, Peiheng, et al.
Veröffentlicht: (2024)
von: Zhou, Peiheng, et al.
Veröffentlicht: (2024)
kNN-CLIP: Retrieval Enables Training-Free Segmentation on Continually Expanding Large Vocabularies
von: Gui, Zhongrui, et al.
Veröffentlicht: (2024)
von: Gui, Zhongrui, et al.
Veröffentlicht: (2024)
DeCLIP: Decoding CLIP representations for deepfake localization
von: Smeu, Stefan, et al.
Veröffentlicht: (2024)
von: Smeu, Stefan, et al.
Veröffentlicht: (2024)
Highlighting What Matters: Promptable Embeddings for Attribute-Focused Image Retrieval
von: Li, Siting, et al.
Veröffentlicht: (2025)
von: Li, Siting, et al.
Veröffentlicht: (2025)
MedCLIP-SAM: Bridging Text and Image Towards Universal Medical Image Segmentation
von: Koleilat, Taha, et al.
Veröffentlicht: (2024)
von: Koleilat, Taha, et al.
Veröffentlicht: (2024)
Memorization In Stable Diffusion Is Unexpectedly Driven by CLIP Embeddings
von: Kim, Bumjun, et al.
Veröffentlicht: (2026)
von: Kim, Bumjun, et al.
Veröffentlicht: (2026)
Backdoor Attack on Unpaired Medical Image-Text Foundation Models: A Pilot Study on MedCLIP
von: Jin, Ruinan, et al.
Veröffentlicht: (2024)
von: Jin, Ruinan, et al.
Veröffentlicht: (2024)
FROSTER: Frozen CLIP Is A Strong Teacher for Open-Vocabulary Action Recognition
von: Huang, Xiaohu, et al.
Veröffentlicht: (2024)
von: Huang, Xiaohu, et al.
Veröffentlicht: (2024)
Divergence Minimization Preference Optimization for Diffusion Model Alignment
von: Li, Binxu, et al.
Veröffentlicht: (2025)
von: Li, Binxu, et al.
Veröffentlicht: (2025)
Embedding-perturbed Exploration Preference Optimization for Flow Models
von: Hu, Sujie, et al.
Veröffentlicht: (2026)
von: Hu, Sujie, et al.
Veröffentlicht: (2026)
FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
Deep Spectral Clustering via Joint Spectral Embedding and Kmeans
von: Guo, Wengang, et al.
Veröffentlicht: (2024)
von: Guo, Wengang, et al.
Veröffentlicht: (2024)
Learning Generalizable Prompt for CLIP with Class Similarity Knowledge
von: Jung, Sehun, et al.
Veröffentlicht: (2025)
von: Jung, Sehun, et al.
Veröffentlicht: (2025)
Text Embedding is Not All You Need: Attention Control for Text-to-Image Semantic Alignment with Text Self-Attention Maps
von: Kim, Jeeyung, et al.
Veröffentlicht: (2024)
von: Kim, Jeeyung, et al.
Veröffentlicht: (2024)
Enhancing CLIP with CLIP: Exploring Pseudolabeling for Limited-Label Prompt Tuning
von: Menghini, Cristina, et al.
Veröffentlicht: (2023)
von: Menghini, Cristina, et al.
Veröffentlicht: (2023)
External Knowledge Injection for CLIP-Based Class-Incremental Learning
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2025)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2025)
A Closer Look at the Robustness of Contrastive Language-Image Pre-Training (CLIP)
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
Efficiency without Compromise: CLIP-aided Text-to-Image GANs with Increased Diversity
von: Kobayashi, Yuya, et al.
Veröffentlicht: (2025)
von: Kobayashi, Yuya, et al.
Veröffentlicht: (2025)
Test-Time Alignment of Text-to-Image Diffusion Models via Null-Text Embedding Optimisation
von: Kim, Taehoon, et al.
Veröffentlicht: (2025)
von: Kim, Taehoon, et al.
Veröffentlicht: (2025)
FairerCLIP: Debiasing CLIP's Zero-Shot Predictions using Functions in RKHSs
von: Dehdashtian, Sepehr, et al.
Veröffentlicht: (2024)
von: Dehdashtian, Sepehr, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Creating Sorted Grid Layouts with Gradient-based Optimization
von: Barthel, Kai Uwe, et al.
Veröffentlicht: (2025) -
Permutation Learning with Only N Parameters: From SoftSort to Self-Organizing Gaussians
von: Barthel, Kai Uwe, et al.
Veröffentlicht: (2025) -
IsoCLIP: Decomposing CLIP Projectors for Efficient Intra-modal Alignment
von: Magistri, Simone, et al.
Veröffentlicht: (2026) -
Contrast-Aware Calibration for Fine-Tuned CLIP: Leveraging Image-Text Alignment
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025) -
SHED: Style-Homogenized Embedding Alignment for Domain Generalization
von: Gan, Kai, et al.
Veröffentlicht: (2026)