CLIP-VQDiffusion : Langauge Free Training of Text To Image generation using CLIP and vector quantized diffusion model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Han, Seungdae, Kim, Joohee |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ComCLIP: Training-Free Compositional Image and Text Matching
von: Jiang, Kenan, et al.
Veröffentlicht: (2022)
von: Jiang, Kenan, et al.
Veröffentlicht: (2022)
Text and Image Are Mutually Beneficial: Enhancing Training-Free Few-Shot Classification with CLIP
von: Li, Yayuan, et al.
Veröffentlicht: (2024)
von: Li, Yayuan, et al.
Veröffentlicht: (2024)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
Long-CLIP: Unlocking the Long-Text Capability of CLIP
von: Zhang, Beichen, et al.
Veröffentlicht: (2024)
von: Zhang, Beichen, et al.
Veröffentlicht: (2024)
Self-Calibrated CLIP for Training-Free Open-Vocabulary Segmentation
von: Bai, Sule, et al.
Veröffentlicht: (2024)
von: Bai, Sule, et al.
Veröffentlicht: (2024)
ABE-CLIP: Training-Free Attribute Binding Enhancement for Compositional Image-Text Matching
von: Zhang, Qi, et al.
Veröffentlicht: (2025)
von: Zhang, Qi, et al.
Veröffentlicht: (2025)
VTD-CLIP: Video-to-Text Discretization via Prompting CLIP
von: Zhu, Wencheng, et al.
Veröffentlicht: (2025)
von: Zhu, Wencheng, et al.
Veröffentlicht: (2025)
Raising the Bar of AI-generated Image Detection with CLIP
von: Cozzolino, Davide, et al.
Veröffentlicht: (2023)
von: Cozzolino, Davide, et al.
Veröffentlicht: (2023)
CLIP-IT: CLIP-based Pairing for Histology Images Classification
von: Karimian, Banafsheh, et al.
Veröffentlicht: (2025)
von: Karimian, Banafsheh, et al.
Veröffentlicht: (2025)
Text-to-Image Generation Via Energy-Based CLIP
von: Ganz, Roy, et al.
Veröffentlicht: (2024)
von: Ganz, Roy, et al.
Veröffentlicht: (2024)
DetailCLIP: Injecting Image Details into CLIP's Feature Space
von: Zhang, Zilun, et al.
Veröffentlicht: (2022)
von: Zhang, Zilun, et al.
Veröffentlicht: (2022)
Exploring the Adversarial Robustness of CLIP for AI-generated Image Detection
von: De Rosa, Vincenzo, et al.
Veröffentlicht: (2024)
von: De Rosa, Vincenzo, et al.
Veröffentlicht: (2024)
CLIP-KD: An Empirical Study of CLIP Model Distillation
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
HQ-CLIP: Leveraging Large Vision-Language Models to Create High-Quality Image-Text Datasets and CLIP Models
von: Wei, Zhixiang, et al.
Veröffentlicht: (2025)
von: Wei, Zhixiang, et al.
Veröffentlicht: (2025)
IPAD-CLIP: Teaching CLIP to Detect Image Local Perceptual Artifacts
von: Wang, Juan, et al.
Veröffentlicht: (2026)
von: Wang, Juan, et al.
Veröffentlicht: (2026)
TTD: Text-Tag Self-Distillation Enhancing Image-Text Alignment in CLIP to Alleviate Single Tag Bias
von: Jo, Sanghyun, et al.
Veröffentlicht: (2024)
von: Jo, Sanghyun, et al.
Veröffentlicht: (2024)
A Training-Free Framework for Open-Vocabulary Image Segmentation and Recognition with EfficientNet and CLIP
von: Dai, Ying, et al.
Veröffentlicht: (2025)
von: Dai, Ying, et al.
Veröffentlicht: (2025)
Explore the Potential of CLIP for Training-Free Open Vocabulary Semantic Segmentation
von: Shao, Tong, et al.
Veröffentlicht: (2024)
von: Shao, Tong, et al.
Veröffentlicht: (2024)
CLIP-Guided Source-Free Object Detection in Aerial Images
von: Liu, Nanqing, et al.
Veröffentlicht: (2024)
von: Liu, Nanqing, et al.
Veröffentlicht: (2024)
CLIP in Medical Imaging: A Survey
von: Zhao, Zihao, et al.
Veröffentlicht: (2023)
von: Zhao, Zihao, et al.
Veröffentlicht: (2023)
TiC-CLIP: Continual Training of CLIP Models
von: Garg, Saurabh, et al.
Veröffentlicht: (2023)
von: Garg, Saurabh, et al.
Veröffentlicht: (2023)
PureCLIP-Depth: Prompt-Free and Decoder-Free Monocular Depth Estimation within CLIP Embedding Space
von: Miya, Ryutaro, et al.
Veröffentlicht: (2026)
von: Miya, Ryutaro, et al.
Veröffentlicht: (2026)
MediCLIP: Adapting CLIP for Few-shot Medical Image Anomaly Detection
von: Zhang, Ximiao, et al.
Veröffentlicht: (2024)
von: Zhang, Ximiao, et al.
Veröffentlicht: (2024)
CLIP-AGIQA: Boosting the Performance of AI-Generated Image Quality Assessment with CLIP
von: Tang, Zhenchen, et al.
Veröffentlicht: (2024)
von: Tang, Zhenchen, et al.
Veröffentlicht: (2024)
VAR-CLIP: Text-to-Image Generator with Visual Auto-Regressive Modeling
von: Zhang, Qian, et al.
Veröffentlicht: (2024)
von: Zhang, Qian, et al.
Veröffentlicht: (2024)
GenCLIP: Generalizing CLIP Prompts for Zero-shot Anomaly Detection
von: Kim, Donghyeong, et al.
Veröffentlicht: (2025)
von: Kim, Donghyeong, et al.
Veröffentlicht: (2025)
DouC: Dual-Branch CLIP for Training-Free Open-Vocabulary Segmentation
von: Zamini, Mohamad, et al.
Veröffentlicht: (2026)
von: Zamini, Mohamad, et al.
Veröffentlicht: (2026)
Improving Visual Discriminability of CLIP for Training-Free Open-Vocabulary Semantic Segmentation
von: Zhou, Jinxin, et al.
Veröffentlicht: (2025)
von: Zhou, Jinxin, et al.
Veröffentlicht: (2025)
SuperCLIP: CLIP with Simple Classification Supervision
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025)
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025)
Extract Free Dense Misalignment from CLIP
von: Nam, JeongYeon, et al.
Veröffentlicht: (2024)
von: Nam, JeongYeon, et al.
Veröffentlicht: (2024)
SPACE-CLIP: Spatial Perception via Adaptive CLIP Embeddings for Monocular Depth Estimation
von: Cho, Taewan, et al.
Veröffentlicht: (2026)
von: Cho, Taewan, et al.
Veröffentlicht: (2026)
Erasing CLIP Memories: Non-Destructive, Data-Free Zero-Shot class Unlearning in CLIP Models
von: Mishra, Ashish, et al.
Veröffentlicht: (2025)
von: Mishra, Ashish, et al.
Veröffentlicht: (2025)
CLIP-Decoder : ZeroShot Multilabel Classification using Multimodal CLIP Aligned Representation
von: Ali, Muhammad, et al.
Veröffentlicht: (2024)
von: Ali, Muhammad, et al.
Veröffentlicht: (2024)
NeuroCLIP: Neuromorphic Data Understanding by CLIP and SNN
von: Guo, Yufei, et al.
Veröffentlicht: (2023)
von: Guo, Yufei, et al.
Veröffentlicht: (2023)
CalibCLIP: Contextual Calibration of Dominant Semantics for Text-Driven Image Retrieval
von: Kang, Bin, et al.
Veröffentlicht: (2025)
von: Kang, Bin, et al.
Veröffentlicht: (2025)
Detecting Deepfakes with Multivariate Soft Blending and CLIP-based Image-Text Alignment
von: Li, Jingwei, et al.
Veröffentlicht: (2026)
von: Li, Jingwei, et al.
Veröffentlicht: (2026)
Exploring Open-Vocabulary Object Recognition in Images using CLIP
von: Chen, Wei Yu, et al.
Veröffentlicht: (2026)
von: Chen, Wei Yu, et al.
Veröffentlicht: (2026)
MobileCLIP: Fast Image-Text Models through Multi-Modal Reinforced Training
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2023)
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2023)
CLIP-Free, Label Free, Unsupervised Concept Bottleneck Models
von: Sammani, Fawaz, et al.
Veröffentlicht: (2025)
von: Sammani, Fawaz, et al.
Veröffentlicht: (2025)
TNG-CLIP:Training-Time Negation Data Generation for Negation Awareness of CLIP
von: Cai, Yuliang, et al.
Veröffentlicht: (2025)
von: Cai, Yuliang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ComCLIP: Training-Free Compositional Image and Text Matching
von: Jiang, Kenan, et al.
Veröffentlicht: (2022) -
Text and Image Are Mutually Beneficial: Enhancing Training-Free Few-Shot Classification with CLIP
von: Li, Yayuan, et al.
Veröffentlicht: (2024) -
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023) -
Long-CLIP: Unlocking the Long-Text Capability of CLIP
von: Zhang, Beichen, et al.
Veröffentlicht: (2024) -
Self-Calibrated CLIP for Training-Free Open-Vocabulary Segmentation
von: Bai, Sule, et al.
Veröffentlicht: (2024)