Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Xiang, Xu, Zhengbo, Zhao, Junhan, Liu, Jiaying |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FBSDiff: Plug-and-Play Frequency Band Substitution of Diffusion Features for Highly Controllable Text-Driven Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
FBSDiff++: Improved Frequency Band Substitution of Diffusion Features for Efficient and Highly Controllable Text-Driven Image-to-Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2026)
von: Gao, Xiang, et al.
Veröffentlicht: (2026)
Diff-Tracker: Text-to-Image Diffusion Models are Unsupervised Trackers
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2024)
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2024)
Versatile Diffusion: Text, Images and Variations All in One Diffusion Model
von: Xu, Xingqian, et al.
Veröffentlicht: (2022)
von: Xu, Xingqian, et al.
Veröffentlicht: (2022)
Leveraging Text-to-Image Diffusion Models for Unsupervised Visual Object Tracking
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2026)
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2026)
Text Image Inpainting via Global Structure-Guided Diffusion Models
von: Zhu, Shipeng, et al.
Veröffentlicht: (2024)
von: Zhu, Shipeng, et al.
Veröffentlicht: (2024)
Contrastive Learning Guided Latent Diffusion Model for Image-to-Image Translation
von: Si, Qi, et al.
Veröffentlicht: (2025)
von: Si, Qi, et al.
Veröffentlicht: (2025)
Local Conditional Controlling for Text-to-Image Diffusion Models
von: Zhao, Yibo, et al.
Veröffentlicht: (2023)
von: Zhao, Yibo, et al.
Veröffentlicht: (2023)
One Stone with Two Birds: A Null-Text-Null Frequency-Aware Diffusion Models for Text-Guided Image Inpainting
von: Liu, Haipeng, et al.
Veröffentlicht: (2025)
von: Liu, Haipeng, et al.
Veröffentlicht: (2025)
Origin Identification for Text-Guided Image-to-Image Diffusion Models
von: Wang, Wenhao, et al.
Veröffentlicht: (2025)
von: Wang, Wenhao, et al.
Veröffentlicht: (2025)
A Diffusion Model Translator for Efficient Image-to-Image Translation
von: Xia, Mengfei, et al.
Veröffentlicht: (2025)
von: Xia, Mengfei, et al.
Veröffentlicht: (2025)
AnyControl: Create Your Artwork with Versatile Control on Text-to-Image Generation
von: Sun, Yanan, et al.
Veröffentlicht: (2024)
von: Sun, Yanan, et al.
Veröffentlicht: (2024)
PTDiffusion: Free Lunch for Generating Optical Illusion Hidden Pictures with Phase-Transferred Diffusion Model
von: Gao, Xiang, et al.
Veröffentlicht: (2025)
von: Gao, Xiang, et al.
Veröffentlicht: (2025)
FDS: Frequency-Aware Denoising Score for Text-Guided Latent Diffusion Image Editing
von: Ren, Yufan, et al.
Veröffentlicht: (2025)
von: Ren, Yufan, et al.
Veröffentlicht: (2025)
VideoElevator: Elevating Video Generation Quality with Versatile Text-to-Image Diffusion Models
von: Zhang, Yabo, et al.
Veröffentlicht: (2024)
von: Zhang, Yabo, et al.
Veröffentlicht: (2024)
Dynamic Frequency Modulation for Controllable Text-driven Image Generation
von: Shi, Tiandong, et al.
Veröffentlicht: (2026)
von: Shi, Tiandong, et al.
Veröffentlicht: (2026)
A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models
von: Shuai, Xincheng, et al.
Veröffentlicht: (2024)
von: Shuai, Xincheng, et al.
Veröffentlicht: (2024)
TextDiff: Mask-Guided Residual Diffusion Models for Scene Text Image Super-Resolution
von: Liu, Baolin, et al.
Veröffentlicht: (2023)
von: Liu, Baolin, et al.
Veröffentlicht: (2023)
Diffusion Model Compression for Image-to-Image Translation
von: Kim, Geonung, et al.
Veröffentlicht: (2024)
von: Kim, Geonung, et al.
Veröffentlicht: (2024)
VersaT2I: Improving Text-to-Image Models with Versatile Reward
von: Guo, Jianshu, et al.
Veröffentlicht: (2024)
von: Guo, Jianshu, et al.
Veröffentlicht: (2024)
GrOCE:Graph-Guided Online Concept Erasure for Text-to-Image Diffusion Models
von: Han, Ning, et al.
Veröffentlicht: (2025)
von: Han, Ning, et al.
Veröffentlicht: (2025)
Energy-Guided Optimization for Personalized Image Editing with Pretrained Text-to-Image Diffusion Models
von: Jiang, Rui, et al.
Veröffentlicht: (2025)
von: Jiang, Rui, et al.
Veröffentlicht: (2025)
ECNet: Effective Controllable Text-to-Image Diffusion Models
von: Li, Sicheng, et al.
Veröffentlicht: (2024)
von: Li, Sicheng, et al.
Veröffentlicht: (2024)
PICD: Versatile Perceptual Image Compression with Diffusion Rendering
von: Xu, Tongda, et al.
Veröffentlicht: (2025)
von: Xu, Tongda, et al.
Veröffentlicht: (2025)
Frequency-Enhanced Diffusion Models: Curriculum-Guided Semantic Alignment for Zero-Shot Skeleton Action Recognition
von: Zhou, Yuxi, et al.
Veröffentlicht: (2026)
von: Zhou, Yuxi, et al.
Veröffentlicht: (2026)
HDW-SR: High-Frequency Guided Diffusion Model based on Wavelet Decomposition for Image Super-Resolution
von: Yang, Chao, et al.
Veröffentlicht: (2025)
von: Yang, Chao, et al.
Veröffentlicht: (2025)
Pre-trained Text-to-Image Diffusion Models Are Versatile Representation Learners for Control
von: Gupta, Gunshi, et al.
Veröffentlicht: (2024)
von: Gupta, Gunshi, et al.
Veröffentlicht: (2024)
Scaling Down Text Encoders of Text-to-Image Diffusion Models
von: Wang, Lifu, et al.
Veröffentlicht: (2025)
von: Wang, Lifu, et al.
Veröffentlicht: (2025)
Plasticine: A Traceable Diffusion Model for Medical Image Translation
von: Zhang, Tianyang, et al.
Veröffentlicht: (2025)
von: Zhang, Tianyang, et al.
Veröffentlicht: (2025)
Versatile Transition Generation with Image-to-Video Diffusion
von: Yang, Zuhao, et al.
Veröffentlicht: (2025)
von: Yang, Zuhao, et al.
Veröffentlicht: (2025)
Controllable Generation with Text-to-Image Diffusion Models: A Survey
von: Cao, Pu, et al.
Veröffentlicht: (2024)
von: Cao, Pu, et al.
Veröffentlicht: (2024)
Controllable Longer Image Animation with Diffusion Models
von: Wang, Qiang, et al.
Veröffentlicht: (2024)
von: Wang, Qiang, et al.
Veröffentlicht: (2024)
TINA: Text-Free Inversion Attack for Unlearned Text-to-Image Diffusion Models
von: Xiang, Qianlong, et al.
Veröffentlicht: (2026)
von: Xiang, Qianlong, et al.
Veröffentlicht: (2026)
FIRE: Robust Detection of Diffusion-Generated Images via Frequency-Guided Reconstruction Error
von: Chu, Beilin, et al.
Veröffentlicht: (2024)
von: Chu, Beilin, et al.
Veröffentlicht: (2024)
VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control
von: Wu, Shaojin, et al.
Veröffentlicht: (2024)
von: Wu, Shaojin, et al.
Veröffentlicht: (2024)
Detecting Origin Attribution for Text-to-Image Diffusion Models
von: Xu, Katherine, et al.
Veröffentlicht: (2024)
von: Xu, Katherine, et al.
Veröffentlicht: (2024)
S$^2$Edit: Text-Guided Image Editing with Precise Semantic and Spatial Control
von: Liu, Xudong, et al.
Veröffentlicht: (2025)
von: Liu, Xudong, et al.
Veröffentlicht: (2025)
Instructing Text-to-Image Diffusion Models via Classifier-Guided Semantic Optimization
von: Chang, Yuanyuan, et al.
Veröffentlicht: (2025)
von: Chang, Yuanyuan, et al.
Veröffentlicht: (2025)
Translation of Text Embedding via Delta Vector to Suppress Strongly Entangled Content in Text-to-Image Diffusion Models
von: Koh, Eunseo, et al.
Veröffentlicht: (2025)
von: Koh, Eunseo, et al.
Veröffentlicht: (2025)
DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models
von: Kim, Sungnyun, et al.
Veröffentlicht: (2023)
von: Kim, Sungnyun, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
FBSDiff: Plug-and-Play Frequency Band Substitution of Diffusion Features for Highly Controllable Text-Driven Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2024) -
FBSDiff++: Improved Frequency Band Substitution of Diffusion Features for Efficient and Highly Controllable Text-Driven Image-to-Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2026) -
Diff-Tracker: Text-to-Image Diffusion Models are Unsupervised Trackers
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2024) -
Versatile Diffusion: Text, Images and Variations All in One Diffusion Model
von: Xu, Xingqian, et al.
Veröffentlicht: (2022) -
Leveraging Text-to-Image Diffusion Models for Unsupervised Visual Object Tracking
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2026)