Image-to-Image Translation with Diffusion Transformers and CLIP-Based Image Conditioning
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhu, Qiang, Lu, Kuan, Huo, Menghao, Li, Yuxiao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CLIP in Medical Imaging: A Survey
di: Zhao, Zihao, et al.
Pubblicazione: (2023)
di: Zhao, Zihao, et al.
Pubblicazione: (2023)
Image Synthesis with Graph Conditioning: CLIP-Guided Diffusion Models for Scene Graphs
di: Mishra, Rameshwar, et al.
Pubblicazione: (2024)
di: Mishra, Rameshwar, et al.
Pubblicazione: (2024)
RATLIP: Generative Adversarial CLIP Text-to-Image Synthesis Based on Recurrent Affine Transformations
di: Lin, Chengde, et al.
Pubblicazione: (2024)
di: Lin, Chengde, et al.
Pubblicazione: (2024)
A Diffusion Model Translator for Efficient Image-to-Image Translation
di: Xia, Mengfei, et al.
Pubblicazione: (2025)
di: Xia, Mengfei, et al.
Pubblicazione: (2025)
UltraImage: Rethinking Resolution Extrapolation in Image Diffusion Transformers
di: Zhao, Min, et al.
Pubblicazione: (2025)
di: Zhao, Min, et al.
Pubblicazione: (2025)
Image-Conditional Diffusion Transformer for Underwater Image Enhancement
di: Nie, Xingyang, et al.
Pubblicazione: (2024)
di: Nie, Xingyang, et al.
Pubblicazione: (2024)
Diffusion Model Compression for Image-to-Image Translation
di: Kim, Geonung, et al.
Pubblicazione: (2024)
di: Kim, Geonung, et al.
Pubblicazione: (2024)
Physics-Informed Conditional Diffusion for Motion-Robust Retinal Temporal Laser Speckle Contrast Imaging
di: Chen, Qian, et al.
Pubblicazione: (2026)
di: Chen, Qian, et al.
Pubblicazione: (2026)
Underwater Image Enhancement by Diffusion Model with Customized CLIP-Classifier
di: Liu, Shuaixin, et al.
Pubblicazione: (2024)
di: Liu, Shuaixin, et al.
Pubblicazione: (2024)
Diffusion-Based Image-to-Image Translation by Noise Correction via Prompt Interpolation
di: Lee, Junsung, et al.
Pubblicazione: (2024)
di: Lee, Junsung, et al.
Pubblicazione: (2024)
CLIP-based Point Cloud Classification via Point Cloud to Image Translation
di: Ghose, Shuvozit, et al.
Pubblicazione: (2024)
di: Ghose, Shuvozit, et al.
Pubblicazione: (2024)
Identifying and Solving Conditional Image Leakage in Image-to-Video Diffusion Model
di: Zhao, Min, et al.
Pubblicazione: (2024)
di: Zhao, Min, et al.
Pubblicazione: (2024)
Enhancing Multimodal Understanding with CLIP-Based Image-to-Text Transformation
di: Che, Chang, et al.
Pubblicazione: (2024)
di: Che, Chang, et al.
Pubblicazione: (2024)
Image Translation as Diffusion Visual Programmers
di: Han, Cheng, et al.
Pubblicazione: (2024)
di: Han, Cheng, et al.
Pubblicazione: (2024)
Inf-DiT: Upsampling Any-Resolution Image with Memory-Efficient Diffusion Transformer
di: Yang, Zhuoyi, et al.
Pubblicazione: (2024)
di: Yang, Zhuoyi, et al.
Pubblicazione: (2024)
Enhancing Image Restoration Transformer via Adaptive Translation Equivariance
di: Hu, JiaKui, et al.
Pubblicazione: (2025)
di: Hu, JiaKui, et al.
Pubblicazione: (2025)
GeoDiT: Point-Conditioned Diffusion Transformer for Satellite Image Synthesis
di: Sastry, Srikumar, et al.
Pubblicazione: (2026)
di: Sastry, Srikumar, et al.
Pubblicazione: (2026)
CLIP-IT: CLIP-based Pairing for Histology Images Classification
di: Karimian, Banafsheh, et al.
Pubblicazione: (2025)
di: Karimian, Banafsheh, et al.
Pubblicazione: (2025)
Heterogeneous Image GNN: Graph-Conditioned Diffusion for Image Synthesis
di: Menneer, Rupert, et al.
Pubblicazione: (2025)
di: Menneer, Rupert, et al.
Pubblicazione: (2025)
Text-to-Image Generation Via Energy-Based CLIP
di: Ganz, Roy, et al.
Pubblicazione: (2024)
di: Ganz, Roy, et al.
Pubblicazione: (2024)
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation
di: Csizmadia, Daniel, et al.
Pubblicazione: (2025)
di: Csizmadia, Daniel, et al.
Pubblicazione: (2025)
Image-to-Brain Signal Generation for Visual Prosthesis with CLIP Guided Multimodal Diffusion Models
di: Xu, Ganxi, et al.
Pubblicazione: (2025)
di: Xu, Ganxi, et al.
Pubblicazione: (2025)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
di: Kim, Seoyeon, et al.
Pubblicazione: (2023)
di: Kim, Seoyeon, et al.
Pubblicazione: (2023)
DCI: Dual-Conditional Inversion for Boosting Diffusion-Based Image Editing
di: Li, Zixiang, et al.
Pubblicazione: (2025)
di: Li, Zixiang, et al.
Pubblicazione: (2025)
IPAD-CLIP: Teaching CLIP to Detect Image Local Perceptual Artifacts
di: Wang, Juan, et al.
Pubblicazione: (2026)
di: Wang, Juan, et al.
Pubblicazione: (2026)
Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation
di: Gao, Xiang, et al.
Pubblicazione: (2024)
di: Gao, Xiang, et al.
Pubblicazione: (2024)
Benchmarking PathCLIP for Pathology Image Analysis
di: Zheng, Sunyi, et al.
Pubblicazione: (2024)
di: Zheng, Sunyi, et al.
Pubblicazione: (2024)
DetailCLIP: Injecting Image Details into CLIP's Feature Space
di: Zhang, Zilun, et al.
Pubblicazione: (2022)
di: Zhang, Zilun, et al.
Pubblicazione: (2022)
Cross-conditioned Diffusion Model for Medical Image to Image Translation
di: Xing, Zhaohu, et al.
Pubblicazione: (2024)
di: Xing, Zhaohu, et al.
Pubblicazione: (2024)
CycleDiff: Cycle Diffusion Models for Unpaired Image-to-image Translation
di: Zou, Shilong, et al.
Pubblicazione: (2025)
di: Zou, Shilong, et al.
Pubblicazione: (2025)
Are Conditional Latent Diffusion Models Effective for Image Restoration?
di: Yuan, Yunchen, et al.
Pubblicazione: (2024)
di: Yuan, Yunchen, et al.
Pubblicazione: (2024)
Wavelet-Based Image Tokenizer for Vision Transformers
di: Zhu, Zhenhai, et al.
Pubblicazione: (2024)
di: Zhu, Zhenhai, et al.
Pubblicazione: (2024)
Rethinking the Paradigm of Content Constraints in Unpaired Image-to-Image Translation
di: Cai, Xiuding, et al.
Pubblicazione: (2022)
di: Cai, Xiuding, et al.
Pubblicazione: (2022)
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers
di: Xie, Enze, et al.
Pubblicazione: (2024)
di: Xie, Enze, et al.
Pubblicazione: (2024)
ThermalDiffusion: Visual-to-Thermal Image-to-Image Translation for Autonomous Navigation
di: Bansal, Shruti, et al.
Pubblicazione: (2025)
di: Bansal, Shruti, et al.
Pubblicazione: (2025)
SPECIAL: Zero-shot Hyperspectral Image Classification With CLIP
di: Pang, Li, et al.
Pubblicazione: (2025)
di: Pang, Li, et al.
Pubblicazione: (2025)
Low-light Image Enhancement via CLIP-Fourier Guided Wavelet Diffusion
di: Xue, Minglong, et al.
Pubblicazione: (2024)
di: Xue, Minglong, et al.
Pubblicazione: (2024)
Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translation
di: Lin, Ziyue, et al.
Pubblicazione: (2026)
di: Lin, Ziyue, et al.
Pubblicazione: (2026)
Translating Imaging to Genomics: Leveraging Transformers for Predictive Modeling
di: Farooq, Aiman, et al.
Pubblicazione: (2024)
di: Farooq, Aiman, et al.
Pubblicazione: (2024)
Seed-to-Seed: Image Translation in Diffusion Seed Space
di: Greenberg, Or, et al.
Pubblicazione: (2024)
di: Greenberg, Or, et al.
Pubblicazione: (2024)
Documenti analoghi
-
CLIP in Medical Imaging: A Survey
di: Zhao, Zihao, et al.
Pubblicazione: (2023) -
Image Synthesis with Graph Conditioning: CLIP-Guided Diffusion Models for Scene Graphs
di: Mishra, Rameshwar, et al.
Pubblicazione: (2024) -
RATLIP: Generative Adversarial CLIP Text-to-Image Synthesis Based on Recurrent Affine Transformations
di: Lin, Chengde, et al.
Pubblicazione: (2024) -
A Diffusion Model Translator for Efficient Image-to-Image Translation
di: Xia, Mengfei, et al.
Pubblicazione: (2025) -
UltraImage: Rethinking Resolution Extrapolation in Image Diffusion Transformers
di: Zhao, Min, et al.
Pubblicazione: (2025)