Saved in:
| Main Authors: | You, Xiaozhou, Zhang, Jian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.00116 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
E$^{2}$GAN: Efficient Training of Efficient GANs for Image-to-Image Translation
by: Gong, Yifan, et al.
Published: (2024)
by: Gong, Yifan, et al.
Published: (2024)
Pretrained Reversible Generation as Unsupervised Visual Representation Learning
by: Xue, Rongkun, et al.
Published: (2024)
by: Xue, Rongkun, et al.
Published: (2024)
Compositional Text-to-Image Generation with Dense Blob Representations
by: Nie, Weili, et al.
Published: (2024)
by: Nie, Weili, et al.
Published: (2024)
RankCLIP: Ranking-Consistent Language-Image Pretraining
by: Zhang, Yiming, et al.
Published: (2024)
by: Zhang, Yiming, et al.
Published: (2024)
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers
by: Roschmann, Simon, et al.
Published: (2025)
by: Roschmann, Simon, et al.
Published: (2025)
GeMix: Conditional GAN-Based Mixup for Improved Medical Image Augmentation
by: Carlesso, Hugo, et al.
Published: (2025)
by: Carlesso, Hugo, et al.
Published: (2025)
TextCraftor: Your Text Encoder Can be Image Quality Controller
by: Li, Yanyu, et al.
Published: (2024)
by: Li, Yanyu, et al.
Published: (2024)
HI-GAN: Hierarchical Inpainting GAN with Auxiliary Inputs for Combined RGB and Depth Inpainting
by: Dash, Ankan, et al.
Published: (2024)
by: Dash, Ankan, et al.
Published: (2024)
Visual Explanations of Image-Text Representations via Multi-Modal Information Bottleneck Attribution
by: Wang, Ying, et al.
Published: (2023)
by: Wang, Ying, et al.
Published: (2023)
Region-centric Image-Language Pretraining for Open-Vocabulary Detection
by: Kim, Dahun, et al.
Published: (2023)
by: Kim, Dahun, et al.
Published: (2023)
FastCLIPstyler: Optimisation-free Text-based Image Style Transfer Using Style Representations
by: Suresh, Ananda Padhmanabhan, et al.
Published: (2022)
by: Suresh, Ananda Padhmanabhan, et al.
Published: (2022)
Double InfoGAN for Contrastive Analysis
by: Carton, Florence, et al.
Published: (2024)
by: Carton, Florence, et al.
Published: (2024)
Image Captions are Natural Prompts for Text-to-Image Models
by: Lei, Shiye, et al.
Published: (2023)
by: Lei, Shiye, et al.
Published: (2023)
Enhancing Generalization in Vision-Language-Action Models by Preserving Pretrained Representations
by: Grover, Shresth, et al.
Published: (2025)
by: Grover, Shresth, et al.
Published: (2025)
TULIP: Towards Unified Language-Image Pretraining
by: Tang, Zineng, et al.
Published: (2025)
by: Tang, Zineng, et al.
Published: (2025)
GoldiCLIP: The Goldilocks Approach for Balancing Explicit Supervision for Language-Image Pretraining
by: Mohan, Deen Dayal, et al.
Published: (2026)
by: Mohan, Deen Dayal, et al.
Published: (2026)
Improving GFlowNets for Text-to-Image Diffusion Alignment
by: Zhang, Dinghuai, et al.
Published: (2024)
by: Zhang, Dinghuai, et al.
Published: (2024)
ProxT2I: Efficient Reward-Guided Text-to-Image Generation via Proximal Diffusion
by: Fang, Zhenghan, et al.
Published: (2025)
by: Fang, Zhenghan, et al.
Published: (2025)
LSAP: Rethinking Inversion Fidelity, Perception and Editability in GAN Latent Space
by: Zhao, Xuekun, et al.
Published: (2022)
by: Zhao, Xuekun, et al.
Published: (2022)
Collaborative Temporal Feature Generation via Critic-Free Reinforcement Learning for Cross-User Sensor-Based Activity Recognition
by: Ye, Xiaozhou, et al.
Published: (2026)
by: Ye, Xiaozhou, et al.
Published: (2026)
Text-Guided Multi-Scale Frequency Representation Adaptation
by: Yan, Weicai, et al.
Published: (2026)
by: Yan, Weicai, et al.
Published: (2026)
Deep Generative Domain Adaptation with Temporal Attention for Cross-User Activity Recognition
by: Ye, Xiaozhou, et al.
Published: (2024)
by: Ye, Xiaozhou, et al.
Published: (2024)
Deep Generative Domain Adaptation with Temporal Relation Knowledge for Cross-User Activity Recognition
by: Ye, Xiaozhou, et al.
Published: (2024)
by: Ye, Xiaozhou, et al.
Published: (2024)
Symbolic Disentangled Representations for Images
by: Korchemnyi, Alexandr, et al.
Published: (2024)
by: Korchemnyi, Alexandr, et al.
Published: (2024)
GeMM-GAN: A Multimodal Generative Model Conditioned on Histopathology Images and Clinical Descriptions for Gene Expression Profile Generation
by: Panaccione, Francesca Pia, et al.
Published: (2026)
by: Panaccione, Francesca Pia, et al.
Published: (2026)
CoDeGAN: Contrastive Disentanglement for Generative Adversarial Network
by: Zhao, Jiangwei, et al.
Published: (2021)
by: Zhao, Jiangwei, et al.
Published: (2021)
Self-Evaluation Unlocks Any-Step Text-to-Image Generation
by: Yu, Xin, et al.
Published: (2025)
by: Yu, Xin, et al.
Published: (2025)
On Unsupervised Image-to-image translation and GAN stability
by: AlAila, BahaaEddin, et al.
Published: (2023)
by: AlAila, BahaaEddin, et al.
Published: (2023)
From Image to Video: An Empirical Study of Diffusion Representations
by: Vélez, Pedro, et al.
Published: (2025)
by: Vélez, Pedro, et al.
Published: (2025)
Conditioning GAN Without Training Dataset
by: Mekonnen, Kidist Amde
Published: (2024)
by: Mekonnen, Kidist Amde
Published: (2024)
Semi-Supervised 3D Medical Segmentation from 2D Natural Images Pretrained Model
by: Yeung, Pak-Hei, et al.
Published: (2025)
by: Yeung, Pak-Hei, et al.
Published: (2025)
Text-To-Image with Generative Adversarial Networks
by: Momen-Tayefeh, Mehrshad
Published: (2024)
by: Momen-Tayefeh, Mehrshad
Published: (2024)
Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models
by: Tewel, Yoad, et al.
Published: (2024)
by: Tewel, Yoad, et al.
Published: (2024)
Contextualized Diffusion Models for Text-Guided Image and Video Generation
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
ReText: Text Boosts Generalization in Image-Based Person Re-identification
by: Mamedov, Timur, et al.
Published: (2026)
by: Mamedov, Timur, et al.
Published: (2026)
AlignGuard: Scalable Safety Alignment for Text-to-Image Generation
by: Liu, Runtao, et al.
Published: (2024)
by: Liu, Runtao, et al.
Published: (2024)
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
by: Oertell, Owen, et al.
Published: (2024)
by: Oertell, Owen, et al.
Published: (2024)
On the Scalability of Diffusion-based Text-to-Image Generation
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
EdgeFusion: On-Device Text-to-Image Generation
by: Castells, Thibault, et al.
Published: (2024)
by: Castells, Thibault, et al.
Published: (2024)
Text-Aware Image Restoration with Diffusion Models
by: Min, Jaewon, et al.
Published: (2025)
by: Min, Jaewon, et al.
Published: (2025)
Similar Items
-
E$^{2}$GAN: Efficient Training of Efficient GANs for Image-to-Image Translation
by: Gong, Yifan, et al.
Published: (2024) -
Pretrained Reversible Generation as Unsupervised Visual Representation Learning
by: Xue, Rongkun, et al.
Published: (2024) -
Compositional Text-to-Image Generation with Dense Blob Representations
by: Nie, Weili, et al.
Published: (2024) -
RankCLIP: Ranking-Consistent Language-Image Pretraining
by: Zhang, Yiming, et al.
Published: (2024) -
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers
by: Roschmann, Simon, et al.
Published: (2025)