Text-to-Image GAN with Pretrained Representations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | You, Xiaozhou, Zhang, Jian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
E$^{2}$GAN: Efficient Training of Efficient GANs for Image-to-Image Translation
von: Gong, Yifan, et al.
Veröffentlicht: (2024)
von: Gong, Yifan, et al.
Veröffentlicht: (2024)
Pretrained Reversible Generation as Unsupervised Visual Representation Learning
von: Xue, Rongkun, et al.
Veröffentlicht: (2024)
von: Xue, Rongkun, et al.
Veröffentlicht: (2024)
Compositional Text-to-Image Generation with Dense Blob Representations
von: Nie, Weili, et al.
Veröffentlicht: (2024)
von: Nie, Weili, et al.
Veröffentlicht: (2024)
RankCLIP: Ranking-Consistent Language-Image Pretraining
von: Zhang, Yiming, et al.
Veröffentlicht: (2024)
von: Zhang, Yiming, et al.
Veröffentlicht: (2024)
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers
von: Roschmann, Simon, et al.
Veröffentlicht: (2025)
von: Roschmann, Simon, et al.
Veröffentlicht: (2025)
GeMix: Conditional GAN-Based Mixup for Improved Medical Image Augmentation
von: Carlesso, Hugo, et al.
Veröffentlicht: (2025)
von: Carlesso, Hugo, et al.
Veröffentlicht: (2025)
TextCraftor: Your Text Encoder Can be Image Quality Controller
von: Li, Yanyu, et al.
Veröffentlicht: (2024)
von: Li, Yanyu, et al.
Veröffentlicht: (2024)
HI-GAN: Hierarchical Inpainting GAN with Auxiliary Inputs for Combined RGB and Depth Inpainting
von: Dash, Ankan, et al.
Veröffentlicht: (2024)
von: Dash, Ankan, et al.
Veröffentlicht: (2024)
Visual Explanations of Image-Text Representations via Multi-Modal Information Bottleneck Attribution
von: Wang, Ying, et al.
Veröffentlicht: (2023)
von: Wang, Ying, et al.
Veröffentlicht: (2023)
Region-centric Image-Language Pretraining for Open-Vocabulary Detection
von: Kim, Dahun, et al.
Veröffentlicht: (2023)
von: Kim, Dahun, et al.
Veröffentlicht: (2023)
FastCLIPstyler: Optimisation-free Text-based Image Style Transfer Using Style Representations
von: Suresh, Ananda Padhmanabhan, et al.
Veröffentlicht: (2022)
von: Suresh, Ananda Padhmanabhan, et al.
Veröffentlicht: (2022)
Image Captions are Natural Prompts for Text-to-Image Models
von: Lei, Shiye, et al.
Veröffentlicht: (2023)
von: Lei, Shiye, et al.
Veröffentlicht: (2023)
Double InfoGAN for Contrastive Analysis
von: Carton, Florence, et al.
Veröffentlicht: (2024)
von: Carton, Florence, et al.
Veröffentlicht: (2024)
GoldiCLIP: The Goldilocks Approach for Balancing Explicit Supervision for Language-Image Pretraining
von: Mohan, Deen Dayal, et al.
Veröffentlicht: (2026)
von: Mohan, Deen Dayal, et al.
Veröffentlicht: (2026)
Improving GFlowNets for Text-to-Image Diffusion Alignment
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
Enhancing Generalization in Vision-Language-Action Models by Preserving Pretrained Representations
von: Grover, Shresth, et al.
Veröffentlicht: (2025)
von: Grover, Shresth, et al.
Veröffentlicht: (2025)
ProxT2I: Efficient Reward-Guided Text-to-Image Generation via Proximal Diffusion
von: Fang, Zhenghan, et al.
Veröffentlicht: (2025)
von: Fang, Zhenghan, et al.
Veröffentlicht: (2025)
LSAP: Rethinking Inversion Fidelity, Perception and Editability in GAN Latent Space
von: Zhao, Xuekun, et al.
Veröffentlicht: (2022)
von: Zhao, Xuekun, et al.
Veröffentlicht: (2022)
TULIP: Towards Unified Language-Image Pretraining
von: Tang, Zineng, et al.
Veröffentlicht: (2025)
von: Tang, Zineng, et al.
Veröffentlicht: (2025)
Text-Guided Multi-Scale Frequency Representation Adaptation
von: Yan, Weicai, et al.
Veröffentlicht: (2026)
von: Yan, Weicai, et al.
Veröffentlicht: (2026)
Symbolic Disentangled Representations for Images
von: Korchemnyi, Alexandr, et al.
Veröffentlicht: (2024)
von: Korchemnyi, Alexandr, et al.
Veröffentlicht: (2024)
GeMM-GAN: A Multimodal Generative Model Conditioned on Histopathology Images and Clinical Descriptions for Gene Expression Profile Generation
von: Panaccione, Francesca Pia, et al.
Veröffentlicht: (2026)
von: Panaccione, Francesca Pia, et al.
Veröffentlicht: (2026)
Self-Evaluation Unlocks Any-Step Text-to-Image Generation
von: Yu, Xin, et al.
Veröffentlicht: (2025)
von: Yu, Xin, et al.
Veröffentlicht: (2025)
CoDeGAN: Contrastive Disentanglement for Generative Adversarial Network
von: Zhao, Jiangwei, et al.
Veröffentlicht: (2021)
von: Zhao, Jiangwei, et al.
Veröffentlicht: (2021)
From Image to Video: An Empirical Study of Diffusion Representations
von: Vélez, Pedro, et al.
Veröffentlicht: (2025)
von: Vélez, Pedro, et al.
Veröffentlicht: (2025)
Semi-Supervised 3D Medical Segmentation from 2D Natural Images Pretrained Model
von: Yeung, Pak-Hei, et al.
Veröffentlicht: (2025)
von: Yeung, Pak-Hei, et al.
Veröffentlicht: (2025)
Text-To-Image with Generative Adversarial Networks
von: Momen-Tayefeh, Mehrshad
Veröffentlicht: (2024)
von: Momen-Tayefeh, Mehrshad
Veröffentlicht: (2024)
Contextualized Diffusion Models for Text-Guided Image and Video Generation
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
Collaborative Temporal Feature Generation via Critic-Free Reinforcement Learning for Cross-User Sensor-Based Activity Recognition
von: Ye, Xiaozhou, et al.
Veröffentlicht: (2026)
von: Ye, Xiaozhou, et al.
Veröffentlicht: (2026)
Pretrained Image-Text Models are Secretly Video Captioners
von: Zhang, Chunhui, et al.
Veröffentlicht: (2025)
von: Zhang, Chunhui, et al.
Veröffentlicht: (2025)
ReText: Text Boosts Generalization in Image-Based Person Re-identification
von: Mamedov, Timur, et al.
Veröffentlicht: (2026)
von: Mamedov, Timur, et al.
Veröffentlicht: (2026)
AlignGuard: Scalable Safety Alignment for Text-to-Image Generation
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models
von: Tewel, Yoad, et al.
Veröffentlicht: (2024)
von: Tewel, Yoad, et al.
Veröffentlicht: (2024)
On the Scalability of Diffusion-based Text-to-Image Generation
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
EdgeFusion: On-Device Text-to-Image Generation
von: Castells, Thibault, et al.
Veröffentlicht: (2024)
von: Castells, Thibault, et al.
Veröffentlicht: (2024)
Text-Aware Image Restoration with Diffusion Models
von: Min, Jaewon, et al.
Veröffentlicht: (2025)
von: Min, Jaewon, et al.
Veröffentlicht: (2025)
Conditioning GAN Without Training Dataset
von: Mekonnen, Kidist Amde
Veröffentlicht: (2024)
von: Mekonnen, Kidist Amde
Veröffentlicht: (2024)
Intervention-Aware Multiscale Representation Learning from Imaging Phenomics and Perturbation Transcriptomics
von: Chen, Jiayuan, et al.
Veröffentlicht: (2026)
von: Chen, Jiayuan, et al.
Veröffentlicht: (2026)
Diffuse and Disperse: Image Generation with Representation Regularization
von: Wang, Runqian, et al.
Veröffentlicht: (2025)
von: Wang, Runqian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
E$^{2}$GAN: Efficient Training of Efficient GANs for Image-to-Image Translation
von: Gong, Yifan, et al.
Veröffentlicht: (2024) -
Pretrained Reversible Generation as Unsupervised Visual Representation Learning
von: Xue, Rongkun, et al.
Veröffentlicht: (2024) -
Compositional Text-to-Image Generation with Dense Blob Representations
von: Nie, Weili, et al.
Veröffentlicht: (2024) -
RankCLIP: Ranking-Consistent Language-Image Pretraining
von: Zhang, Yiming, et al.
Veröffentlicht: (2024) -
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers
von: Roschmann, Simon, et al.
Veröffentlicht: (2025)