Style Ambiguity Loss Using CLIP
Fuente:
arXiv
Guardado en:
| Autor principal: | Baker, James |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Using Multimodal Foundation Models and Clustering for Improved Style Ambiguity Loss
por: Baker, James
Publicado: (2024)
por: Baker, James
Publicado: (2024)
Control-CLIP: Decoupling Category and Style Guidance in CLIP for Specific-Domain Generation
por: Jia, Zexi, et al.
Publicado: (2025)
por: Jia, Zexi, et al.
Publicado: (2025)
StyleHumanCLIP: Text-guided Garment Manipulation for StyleGAN-Human
por: Yoshikawa, Takato, et al.
Publicado: (2023)
por: Yoshikawa, Takato, et al.
Publicado: (2023)
TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles
por: Ma, Yifeng, et al.
Publicado: (2023)
por: Ma, Yifeng, et al.
Publicado: (2023)
AD-CLIP: Adapting Domains in Prompt Space Using CLIP
por: Singha, Mainak, et al.
Publicado: (2023)
por: Singha, Mainak, et al.
Publicado: (2023)
Infrared and Visible Image Fusion with Language-Driven Loss in CLIP Embedding Space
por: Wang, Yuhao, et al.
Publicado: (2024)
por: Wang, Yuhao, et al.
Publicado: (2024)
A Picture is Worth More Than 77 Text Tokens: Evaluating CLIP-Style Models on Dense Captions
por: Urbanek, Jack, et al.
Publicado: (2023)
por: Urbanek, Jack, et al.
Publicado: (2023)
BRAT: Bonus oRthogonAl Token for Architecture Agnostic Textual Inversion
por: Baker, James
Publicado: (2024)
por: Baker, James
Publicado: (2024)
MONKEY: Masking ON KEY-Value Activation Adapter for Personalization
por: Baker, James
Publicado: (2025)
por: Baker, James
Publicado: (2025)
IRConStyle: Image Restoration Framework Using Contrastive Learning and Style Transfer
por: Fan, Dongqi, et al.
Publicado: (2024)
por: Fan, Dongqi, et al.
Publicado: (2024)
Recognizing Artistic Style of Archaeological Image Fragments Using Deep Style Extrapolation
por: Elkin, Gur, et al.
Publicado: (2025)
por: Elkin, Gur, et al.
Publicado: (2025)
SuperCLIP: CLIP with Simple Classification Supervision
por: Zhao, Weiheng, et al.
Publicado: (2025)
por: Zhao, Weiheng, et al.
Publicado: (2025)
Long-CLIP: Unlocking the Long-Text Capability of CLIP
por: Zhang, Beichen, et al.
Publicado: (2024)
por: Zhang, Beichen, et al.
Publicado: (2024)
CLIP-KD: An Empirical Study of CLIP Model Distillation
por: Yang, Chuanguang, et al.
Publicado: (2023)
por: Yang, Chuanguang, et al.
Publicado: (2023)
CLIP-IT: CLIP-based Pairing for Histology Images Classification
por: Karimian, Banafsheh, et al.
Publicado: (2025)
por: Karimian, Banafsheh, et al.
Publicado: (2025)
NeuroCLIP: Neuromorphic Data Understanding by CLIP and SNN
por: Guo, Yufei, et al.
Publicado: (2023)
por: Guo, Yufei, et al.
Publicado: (2023)
CLIP the Landscape: Automated Tagging of Crowdsourced Landscape Images
por: Ilyankou, Ilya, et al.
Publicado: (2025)
por: Ilyankou, Ilya, et al.
Publicado: (2025)
un$^2$CLIP: Improving CLIP's Visual Detail Capturing Ability via Inverting unCLIP
por: Li, Yinqi, et al.
Publicado: (2025)
por: Li, Yinqi, et al.
Publicado: (2025)
LocalStyleFool: Regional Video Style Transfer Attack Using Segment Anything Model
por: Cao, Yuxin, et al.
Publicado: (2024)
por: Cao, Yuxin, et al.
Publicado: (2024)
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks
por: Monsefi, Amin Karimi, et al.
Publicado: (2024)
por: Monsefi, Amin Karimi, et al.
Publicado: (2024)
CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation
por: Chung, Jeannie, et al.
Publicado: (2026)
por: Chung, Jeannie, et al.
Publicado: (2026)
VTD-CLIP: Video-to-Text Discretization via Prompting CLIP
por: Zhu, Wencheng, et al.
Publicado: (2025)
por: Zhu, Wencheng, et al.
Publicado: (2025)
MadCLIP: Few-shot Medical Anomaly Detection with CLIP
por: Shiri, Mahshid, et al.
Publicado: (2025)
por: Shiri, Mahshid, et al.
Publicado: (2025)
DetailCLIP: Injecting Image Details into CLIP's Feature Space
por: Zhang, Zilun, et al.
Publicado: (2022)
por: Zhang, Zilun, et al.
Publicado: (2022)
Acknowledging Focus Ambiguity in Visual Questions
por: Chen, Chongyan, et al.
Publicado: (2025)
por: Chen, Chongyan, et al.
Publicado: (2025)
Exploring Weak-to-Strong Generalization for CLIP-based Classification
por: Li, Jinhao, et al.
Publicado: (2025)
por: Li, Jinhao, et al.
Publicado: (2025)
ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation
por: Wang, Jingyun, et al.
Publicado: (2024)
por: Wang, Jingyun, et al.
Publicado: (2024)
CLIP-Mamba: CLIP Pretrained Mamba Models with OOD and Hessian Evaluation
por: Huang, Weiquan, et al.
Publicado: (2024)
por: Huang, Weiquan, et al.
Publicado: (2024)
CLIP-CID: Efficient CLIP Distillation via Cluster-Instance Discrimination
por: Yang, Kaicheng, et al.
Publicado: (2024)
por: Yang, Kaicheng, et al.
Publicado: (2024)
EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
por: Sun, Quan, et al.
Publicado: (2024)
por: Sun, Quan, et al.
Publicado: (2024)
ProxyCLIP: Proxy Attention Improves CLIP for Open-Vocabulary Segmentation
por: Lan, Mengcheng, et al.
Publicado: (2024)
por: Lan, Mengcheng, et al.
Publicado: (2024)
ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference
por: Lan, Mengcheng, et al.
Publicado: (2024)
por: Lan, Mengcheng, et al.
Publicado: (2024)
CLIP-VIS: Adapting CLIP for Open-Vocabulary Video Instance Segmentation
por: Zhu, Wenqi, et al.
Publicado: (2024)
por: Zhu, Wenqi, et al.
Publicado: (2024)
CLIP-Map: Structured Matrix Mapping for Parameter-Efficient CLIP Compression
por: Zhang, Kangjie, et al.
Publicado: (2026)
por: Zhang, Kangjie, et al.
Publicado: (2026)
MedP-CLIP: Medical CLIP with Region-Aware Prompt Integration
por: Peng, Jiahui, et al.
Publicado: (2026)
por: Peng, Jiahui, et al.
Publicado: (2026)
BadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIP
por: Bai, Jiawang, et al.
Publicado: (2023)
por: Bai, Jiawang, et al.
Publicado: (2023)
WP-CLIP: Leveraging CLIP to Predict Wölfflin's Principles in Visual Art
por: Ghildyal, Abhijay, et al.
Publicado: (2025)
por: Ghildyal, Abhijay, et al.
Publicado: (2025)
GenCLIP: Generalizing CLIP Prompts for Zero-shot Anomaly Detection
por: Kim, Donghyeong, et al.
Publicado: (2025)
por: Kim, Donghyeong, et al.
Publicado: (2025)
CLIP-VG: Self-paced Curriculum Adapting of CLIP for Visual Grounding
por: Xiao, Linhui, et al.
Publicado: (2023)
por: Xiao, Linhui, et al.
Publicado: (2023)
AnimalMotionCLIP: Embedding motion in CLIP for Animal Behavior Analysis
por: Zhong, Enmin, et al.
Publicado: (2025)
por: Zhong, Enmin, et al.
Publicado: (2025)
Ejemplares similares
-
Using Multimodal Foundation Models and Clustering for Improved Style Ambiguity Loss
por: Baker, James
Publicado: (2024) -
Control-CLIP: Decoupling Category and Style Guidance in CLIP for Specific-Domain Generation
por: Jia, Zexi, et al.
Publicado: (2025) -
StyleHumanCLIP: Text-guided Garment Manipulation for StyleGAN-Human
por: Yoshikawa, Takato, et al.
Publicado: (2023) -
TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles
por: Ma, Yifeng, et al.
Publicado: (2023) -
AD-CLIP: Adapting Domains in Prompt Space Using CLIP
por: Singha, Mainak, et al.
Publicado: (2023)