RATLIP: Generative Adversarial CLIP Text-to-Image Synthesis Based on Recurrent Affine Transformations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Chengde, Lu, Xijun, Chen, Guangxi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DAFT-GAN: Dual Affine Transformation Generative Adversarial Network for Text-Guided Image Inpainting
von: Lee, Jihoon, et al.
Veröffentlicht: (2024)
von: Lee, Jihoon, et al.
Veröffentlicht: (2024)
Enhancing Multimodal Understanding with CLIP-Based Image-to-Text Transformation
von: Che, Chang, et al.
Veröffentlicht: (2024)
von: Che, Chang, et al.
Veröffentlicht: (2024)
Text-to-Image Generation Via Energy-Based CLIP
von: Ganz, Roy, et al.
Veröffentlicht: (2024)
von: Ganz, Roy, et al.
Veröffentlicht: (2024)
Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis
von: Yuan, Yu, et al.
Veröffentlicht: (2024)
von: Yuan, Yu, et al.
Veröffentlicht: (2024)
Image-to-Image Translation with Diffusion Transformers and CLIP-Based Image Conditioning
von: Zhu, Qiang, et al.
Veröffentlicht: (2025)
von: Zhu, Qiang, et al.
Veröffentlicht: (2025)
StegOT: Trade-offs in Steganography via Optimal Transport
von: Lin, Chengde, et al.
Veröffentlicht: (2025)
von: Lin, Chengde, et al.
Veröffentlicht: (2025)
Progressive Image Restoration via Text-Conditioned Video Generation
von: Kang, Peng, et al.
Veröffentlicht: (2025)
von: Kang, Peng, et al.
Veröffentlicht: (2025)
SwinTextUNet: Integrating CLIP-Based Text Guidance into Swin Transformer U-Nets for Medical Image Segmentation
von: Yeafi, Ashfak, et al.
Veröffentlicht: (2026)
von: Yeafi, Ashfak, et al.
Veröffentlicht: (2026)
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation
von: Csizmadia, Daniel, et al.
Veröffentlicht: (2025)
von: Csizmadia, Daniel, et al.
Veröffentlicht: (2025)
VAR-CLIP: Text-to-Image Generator with Visual Auto-Regressive Modeling
von: Zhang, Qian, et al.
Veröffentlicht: (2024)
von: Zhang, Qian, et al.
Veröffentlicht: (2024)
CLIP-Guided Generative Networks for Transferable Targeted Adversarial Attacks
von: Fang, Hao, et al.
Veröffentlicht: (2024)
von: Fang, Hao, et al.
Veröffentlicht: (2024)
E4C: Enhance Editability for Text-Based Image Editing by Harnessing Efficient CLIP Guidance
von: Huang, Tianrui, et al.
Veröffentlicht: (2024)
von: Huang, Tianrui, et al.
Veröffentlicht: (2024)
Exploring the Adversarial Robustness of CLIP for AI-generated Image Detection
von: De Rosa, Vincenzo, et al.
Veröffentlicht: (2024)
von: De Rosa, Vincenzo, et al.
Veröffentlicht: (2024)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
SAU: A Dual-Branch Network to Enhance Long-Tailed Recognition via Generative Models
von: Li, Guangxi, et al.
Veröffentlicht: (2024)
von: Li, Guangxi, et al.
Veröffentlicht: (2024)
Adversarial Backdoor Defense in CLIP
von: Kuang, Junhao, et al.
Veröffentlicht: (2024)
von: Kuang, Junhao, et al.
Veröffentlicht: (2024)
ATAC: Augmentation-Based Test-Time Adversarial Correction for CLIP
von: Su, Linxiang, et al.
Veröffentlicht: (2025)
von: Su, Linxiang, et al.
Veröffentlicht: (2025)
Lung Nodule Image Synthesis Driven by Two-Stage Generative Adversarial Networks
von: Cao, Lu, et al.
Veröffentlicht: (2026)
von: Cao, Lu, et al.
Veröffentlicht: (2026)
HQ-CLIP: Leveraging Large Vision-Language Models to Create High-Quality Image-Text Datasets and CLIP Models
von: Wei, Zhixiang, et al.
Veröffentlicht: (2025)
von: Wei, Zhixiang, et al.
Veröffentlicht: (2025)
Interpreting CLIP's Image Representation via Text-Based Decomposition
von: Gandelsman, Yossi, et al.
Veröffentlicht: (2023)
von: Gandelsman, Yossi, et al.
Veröffentlicht: (2023)
PixArt-$α$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis
von: Chen, Junsong, et al.
Veröffentlicht: (2023)
von: Chen, Junsong, et al.
Veröffentlicht: (2023)
Meissonic: Revitalizing Masked Generative Transformers for Efficient High-Resolution Text-to-Image Synthesis
von: Bai, Jinbin, et al.
Veröffentlicht: (2024)
von: Bai, Jinbin, et al.
Veröffentlicht: (2024)
Long-CLIP: Unlocking the Long-Text Capability of CLIP
von: Zhang, Beichen, et al.
Veröffentlicht: (2024)
von: Zhang, Beichen, et al.
Veröffentlicht: (2024)
CalibCLIP: Contextual Calibration of Dominant Semantics for Text-Driven Image Retrieval
von: Kang, Bin, et al.
Veröffentlicht: (2025)
von: Kang, Bin, et al.
Veröffentlicht: (2025)
Distilling Knowledge from Text-to-Image Generative Models Improves Visio-Linguistic Reasoning in CLIP
von: Basu, Samyadeep, et al.
Veröffentlicht: (2023)
von: Basu, Samyadeep, et al.
Veröffentlicht: (2023)
CLIP-AGIQA: Boosting the Performance of AI-Generated Image Quality Assessment with CLIP
von: Tang, Zhenchen, et al.
Veröffentlicht: (2024)
von: Tang, Zhenchen, et al.
Veröffentlicht: (2024)
CLIP-VQDiffusion : Langauge Free Training of Text To Image generation using CLIP and vector quantized diffusion model
von: Han, Seungdae, et al.
Veröffentlicht: (2024)
von: Han, Seungdae, et al.
Veröffentlicht: (2024)
Switti: Designing Scale-Wise Transformers for Text-to-Image Synthesis
von: Voronov, Anton, et al.
Veröffentlicht: (2024)
von: Voronov, Anton, et al.
Veröffentlicht: (2024)
VTD-CLIP: Video-to-Text Discretization via Prompting CLIP
von: Zhu, Wencheng, et al.
Veröffentlicht: (2025)
von: Zhu, Wencheng, et al.
Veröffentlicht: (2025)
Contrast-Aware Calibration for Fine-Tuned CLIP: Leveraging Image-Text Alignment
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
Semantic-aware Adversarial Fine-tuning for CLIP
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2026)
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2026)
DeeCLIP: A Robust and Generalizable Transformer-Based Framework for Detecting AI-Generated Images
von: Keita, Mamadou, et al.
Veröffentlicht: (2025)
von: Keita, Mamadou, et al.
Veröffentlicht: (2025)
DSE-GAN: Dynamic Semantic Evolution Generative Adversarial Network for Text-to-Image Generation
von: Huang, Mengqi, et al.
Veröffentlicht: (2022)
von: Huang, Mengqi, et al.
Veröffentlicht: (2022)
Benchmarking PathCLIP for Pathology Image Analysis
von: Zheng, Sunyi, et al.
Veröffentlicht: (2024)
von: Zheng, Sunyi, et al.
Veröffentlicht: (2024)
Detecting Deepfakes with Multivariate Soft Blending and CLIP-based Image-Text Alignment
von: Li, Jingwei, et al.
Veröffentlicht: (2026)
von: Li, Jingwei, et al.
Veröffentlicht: (2026)
AutoPrompt: Automated Red-Teaming of Text-to-Image Models via LLM-Driven Adversarial Prompts
von: Liu, Yufan, et al.
Veröffentlicht: (2025)
von: Liu, Yufan, et al.
Veröffentlicht: (2025)
ToxicTextCLIP: Text-Based Poisoning and Backdoor Attacks on CLIP Pre-training
von: Yao, Xin, et al.
Veröffentlicht: (2025)
von: Yao, Xin, et al.
Veröffentlicht: (2025)
Fine-grained Text to Image Synthesis
von: Ouyang, Xu, et al.
Veröffentlicht: (2024)
von: Ouyang, Xu, et al.
Veröffentlicht: (2024)
Adversarial Distribution Matching for Diffusion Distillation Towards Efficient Image and Video Synthesis
von: Lu, Yanzuo, et al.
Veröffentlicht: (2025)
von: Lu, Yanzuo, et al.
Veröffentlicht: (2025)
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers
von: Xie, Enze, et al.
Veröffentlicht: (2024)
von: Xie, Enze, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DAFT-GAN: Dual Affine Transformation Generative Adversarial Network for Text-Guided Image Inpainting
von: Lee, Jihoon, et al.
Veröffentlicht: (2024) -
Enhancing Multimodal Understanding with CLIP-Based Image-to-Text Transformation
von: Che, Chang, et al.
Veröffentlicht: (2024) -
Text-to-Image Generation Via Energy-Based CLIP
von: Ganz, Roy, et al.
Veröffentlicht: (2024) -
Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis
von: Yuan, Yu, et al.
Veröffentlicht: (2024) -
Image-to-Image Translation with Diffusion Transformers and CLIP-Based Image Conditioning
von: Zhu, Qiang, et al.
Veröffentlicht: (2025)