CLAP: Isolating Content from Style through Contrastive Learning with Augmented Prompts
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Yichao, Liu, Yuhang, Zhang, Zhen, Shi, Javen Qinfeng |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Causal Disentanglement and Cross-Modal Alignment for Enhanced Few-Shot Learning
by: Jiang, Tianjiao, et al.
Published: (2025)
by: Jiang, Tianjiao, et al.
Published: (2025)
On the Value of Cross-Modal Misalignment in Multimodal Representation Learning
by: Cai, Yichao, et al.
Published: (2025)
by: Cai, Yichao, et al.
Published: (2025)
A Survey on Deep Neural Network Pruning-Taxonomy, Comparison, Analysis, and Recommendations
by: Cheng, Hongrong, et al.
Published: (2023)
by: Cheng, Hongrong, et al.
Published: (2023)
A Simple-but-effective Baseline for Training-free Class-Agnostic Counting
by: Lin, Yuhao, et al.
Published: (2024)
by: Lin, Yuhao, et al.
Published: (2024)
Augmented Commonsense Knowledge for Remote Object Grounding
by: Mohammadi, Bahram, et al.
Published: (2024)
by: Mohammadi, Bahram, et al.
Published: (2024)
Beyond DAGs: A Latent Partial Causal Model for Multimodal Learning
by: Liu, Yuhang, et al.
Published: (2024)
by: Liu, Yuhang, et al.
Published: (2024)
The Devil is in the Distributions: Explicit Modeling of Scene Content is Key in Zero-Shot Video Captioning
by: Tian, Mingkai, et al.
Published: (2025)
by: Tian, Mingkai, et al.
Published: (2025)
Learning to Reason and Navigate: Parameter Efficient Action Planning with Large Language Models
by: Mohammadi, Bahram, et al.
Published: (2025)
by: Mohammadi, Bahram, et al.
Published: (2025)
CLAP: Contrastive Latent Action Pretraining for Learning Vision-Language-Action Models from Human Videos
by: Zhang, Chubin, et al.
Published: (2026)
by: Zhang, Chubin, et al.
Published: (2026)
The Geometric Mechanics of Contrastive Representation Learning: Alignment Potentials, Entropic Dispersion, and Cross-modal Divergence
by: Cai, Yichao, et al.
Published: (2026)
by: Cai, Yichao, et al.
Published: (2026)
Categorical Keypoint Positional Embedding for Robust Animal Re-Identification
by: Lin, Yuhao, et al.
Published: (2024)
by: Lin, Yuhao, et al.
Published: (2024)
CLAP: Contrastive Latent-space Prompt Optimization for End-to-end Autonomous Driving
by: Zhu, Ruiyang, et al.
Published: (2026)
by: Zhu, Ruiyang, et al.
Published: (2026)
SPG: Style-Prompting Guidance for Style-Specific Content Creation
by: Liang, Qian, et al.
Published: (2025)
by: Liang, Qian, et al.
Published: (2025)
Advancements in Point Cloud Data Augmentation for Deep Learning: A Survey
by: Zhu, Qinfeng, et al.
Published: (2023)
by: Zhu, Qinfeng, et al.
Published: (2023)
DICE: Disentangling Artist Style from Content via Contrastive Subspace Decomposition in Diffusion Models
by: Zhang, Tong, et al.
Published: (2026)
by: Zhang, Tong, et al.
Published: (2026)
Improving Data Augmentation for Robust Visual Question Answering with Effective Curriculum Learning
by: Zheng, Yuhang, et al.
Published: (2024)
by: Zheng, Yuhang, et al.
Published: (2024)
StyleTextGen: Style-Conditioned Multilingual Scene Text Generation
by: Chen, Zeyu, et al.
Published: (2026)
by: Chen, Zeyu, et al.
Published: (2026)
TeleStyle: Content-Preserving Style Transfer in Images and Videos
by: Zhang, Shiwen, et al.
Published: (2026)
by: Zhang, Shiwen, et al.
Published: (2026)
QwenStyle: Content-Preserving Style Transfer with Qwen-Image-Edit
by: Zhang, Shiwen, et al.
Published: (2026)
by: Zhang, Shiwen, et al.
Published: (2026)
IRConStyle: Image Restoration Framework Using Contrastive Learning and Style Transfer
by: Fan, Dongqi, et al.
Published: (2024)
by: Fan, Dongqi, et al.
Published: (2024)
ConsisLoRA: Enhancing Content and Style Consistency for LoRA-based Style Transfer
by: Chen, Bolin, et al.
Published: (2025)
by: Chen, Bolin, et al.
Published: (2025)
Style-Pro: Style-Guided Prompt Learning for Generalizable Vision-Language Models
by: Talemi, Niloufar Alipour, et al.
Published: (2024)
by: Talemi, Niloufar Alipour, et al.
Published: (2024)
SplitFlux: Learning to Decouple Content and Style from a Single Image
by: Yang, Yitong, et al.
Published: (2025)
by: Yang, Yitong, et al.
Published: (2025)
WikiStyle+: A Multimodal Approach to Content-Style Representation Disentanglement for Artistic Image Stylization
by: Zhuoqi, Ma, et al.
Published: (2024)
by: Zhuoqi, Ma, et al.
Published: (2024)
CRAFT-LoRA: Content-Style Personalization via Rank-Constrained Adaptation and Training-Free Fusion
by: Li, Yu, et al.
Published: (2026)
by: Li, Yu, et al.
Published: (2026)
SCAdapter: Content-Style Disentanglement for Diffusion Style Transfer
by: Trinh, Luan Thanh, et al.
Published: (2025)
by: Trinh, Luan Thanh, et al.
Published: (2025)
Masked Language Prompting for Generative Data Augmentation in Few-shot Fashion Style Recognition
by: Hirakawa, Yuki, et al.
Published: (2025)
by: Hirakawa, Yuki, et al.
Published: (2025)
CLAP4CLIP: Continual Learning with Probabilistic Finetuning for Vision-Language Models
by: Jha, Saurav, et al.
Published: (2024)
by: Jha, Saurav, et al.
Published: (2024)
Character-Adapter: Prompt-Guided Region Control for High-Fidelity Character Customization
by: Ma, Yuhang, et al.
Published: (2024)
by: Ma, Yuhang, et al.
Published: (2024)
CLAP Convolutional Lightweight Autoencoder for Plant Disease Classification
by: Bera, Asish, et al.
Published: (2026)
by: Bera, Asish, et al.
Published: (2026)
CLAP: Concave Linear APproximation for Quadratic Graph Matching
by: Liang, Yongqing, et al.
Published: (2024)
by: Liang, Yongqing, et al.
Published: (2024)
DRG-Font: Dynamic Reference-Guided Few-shot Font Generation via Contrastive Style-Content Disentanglement
by: Chakraborty, Rejoy, et al.
Published: (2026)
by: Chakraborty, Rejoy, et al.
Published: (2026)
CLAP: Unsupervised 3D Representation Learning for Fusion 3D Perception via Curvature Sampling and Prototype Learning
by: Chen, Runjian, et al.
Published: (2024)
by: Chen, Runjian, et al.
Published: (2024)
Rethink Arbitrary Style Transfer with Transformer and Contrastive Learning
by: Zhang, Zhanjie, et al.
Published: (2024)
by: Zhang, Zhanjie, et al.
Published: (2024)
ViTCN: Vision Transformer Contrastive Network For Reasoning
by: Song, Bo, et al.
Published: (2024)
by: Song, Bo, et al.
Published: (2024)
Expanding the Content-Style Frontier: a Balanced Subspace Blending Approach for Content-Style LoRA Fusion
by: Huang, Linhao
Published: (2025)
by: Huang, Linhao
Published: (2025)
Pluggable Style Representation Learning for Multi-Style Transfer
by: Liu, Hongda, et al.
Published: (2025)
by: Liu, Hongda, et al.
Published: (2025)
Attack-Augmentation Mixing-Contrastive Skeletal Representation Learning
by: Xu, Binqian, et al.
Published: (2023)
by: Xu, Binqian, et al.
Published: (2023)
StyleBooth: Image Style Editing with Multimodal Instruction
by: Han, Zhen, et al.
Published: (2024)
by: Han, Zhen, et al.
Published: (2024)
SO3UFormer: Learning Intrinsic Spherical Features for Rotation-Robust Panoramic Segmentation
by: Zhu, Qinfeng, et al.
Published: (2026)
by: Zhu, Qinfeng, et al.
Published: (2026)
Similar Items
-
Causal Disentanglement and Cross-Modal Alignment for Enhanced Few-Shot Learning
by: Jiang, Tianjiao, et al.
Published: (2025) -
On the Value of Cross-Modal Misalignment in Multimodal Representation Learning
by: Cai, Yichao, et al.
Published: (2025) -
A Survey on Deep Neural Network Pruning-Taxonomy, Comparison, Analysis, and Recommendations
by: Cheng, Hongrong, et al.
Published: (2023) -
A Simple-but-effective Baseline for Training-free Class-Agnostic Counting
by: Lin, Yuhao, et al.
Published: (2024) -
Augmented Commonsense Knowledge for Remote Object Grounding
by: Mohammadi, Bahram, et al.
Published: (2024)