FBSDiff++: Improved Frequency Band Substitution of Diffusion Features for Efficient and Highly Controllable Text-Driven Image-to-Image Translation
Fuente:
arXiv
Guardado en:
| Autores principales: | Gao, Xiang, Jia, Yunpeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
FBSDiff: Plug-and-Play Frequency Band Substitution of Diffusion Features for Highly Controllable Text-Driven Image Translation
por: Gao, Xiang, et al.
Publicado: (2024)
por: Gao, Xiang, et al.
Publicado: (2024)
Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation
por: Gao, Xiang, et al.
Publicado: (2024)
por: Gao, Xiang, et al.
Publicado: (2024)
A Diffusion Model Translator for Efficient Image-to-Image Translation
por: Xia, Mengfei, et al.
Publicado: (2025)
por: Xia, Mengfei, et al.
Publicado: (2025)
PiCo: Enhancing Text-Image Alignment with Improved Noise Selection and Precise Mask Control in Diffusion Models
por: Xie, Chang, et al.
Publicado: (2025)
por: Xie, Chang, et al.
Publicado: (2025)
VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control
por: Wu, Shaojin, et al.
Publicado: (2024)
por: Wu, Shaojin, et al.
Publicado: (2024)
Dynamic Frequency Modulation for Controllable Text-driven Image Generation
por: Shi, Tiandong, et al.
Publicado: (2026)
por: Shi, Tiandong, et al.
Publicado: (2026)
NSYNC: Negative Synthetic Image Generation for Contrastive Training to Improve Stylized Text-To-Image Translation
por: Ozturk, Serkan, et al.
Publicado: (2025)
por: Ozturk, Serkan, et al.
Publicado: (2025)
Reusing Computation in Text-to-Image Diffusion for Efficient Generation of Image Sets
por: Decatur, Dale, et al.
Publicado: (2025)
por: Decatur, Dale, et al.
Publicado: (2025)
IDAdapter: Learning Mixed Features for Tuning-Free Personalization of Text-to-Image Models
por: Cui, Siying, et al.
Publicado: (2024)
por: Cui, Siying, et al.
Publicado: (2024)
Text Modality Oriented Image Feature Extraction for Detecting Diffusion-based DeepFake
por: Yang, Di, et al.
Publicado: (2024)
por: Yang, Di, et al.
Publicado: (2024)
Diffusion Model Compression for Image-to-Image Translation
por: Kim, Geonung, et al.
Publicado: (2024)
por: Kim, Geonung, et al.
Publicado: (2024)
Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translation
por: Lin, Ziyue, et al.
Publicado: (2026)
por: Lin, Ziyue, et al.
Publicado: (2026)
Grounding Text-to-Image Diffusion Models for Controlled High-Quality Image Generation
por: Süleyman, Ahmad, et al.
Publicado: (2025)
por: Süleyman, Ahmad, et al.
Publicado: (2025)
Analyzing and Improving Fast Sampling of Text-to-Image Diffusion Models
por: Zhou, Zhenyu, et al.
Publicado: (2026)
por: Zhou, Zhenyu, et al.
Publicado: (2026)
Contrastive Prompts Improve Disentanglement in Text-to-Image Diffusion Models
por: Wu, Chen, et al.
Publicado: (2024)
por: Wu, Chen, et al.
Publicado: (2024)
ECNet: Effective Controllable Text-to-Image Diffusion Models
por: Li, Sicheng, et al.
Publicado: (2024)
por: Li, Sicheng, et al.
Publicado: (2024)
Local Conditional Controlling for Text-to-Image Diffusion Models
por: Zhao, Yibo, et al.
Publicado: (2023)
por: Zhao, Yibo, et al.
Publicado: (2023)
Customizing Text-to-Image Diffusion with Object Viewpoint Control
por: Kumari, Nupur, et al.
Publicado: (2024)
por: Kumari, Nupur, et al.
Publicado: (2024)
Image2Text2Image: A Novel Framework for Label-Free Evaluation of Image-to-Text Generation with Text-to-Image Diffusion Models
por: Huang, Jia-Hong, et al.
Publicado: (2024)
por: Huang, Jia-Hong, et al.
Publicado: (2024)
LASER: Tuning-Free LLM-Driven Attention Control for Efficient Text-conditioned Image-to-Animation
por: Zheng, Haoyu, et al.
Publicado: (2024)
por: Zheng, Haoyu, et al.
Publicado: (2024)
FreeGraftor: Training-Free Cross-Image Feature Grafting for Subject-Driven Text-to-Image Generation
por: Yao, Zebin, et al.
Publicado: (2025)
por: Yao, Zebin, et al.
Publicado: (2025)
Control-DINO: Feature Space Conditioning for Controllable Image-to-Video Diffusion
por: Dominici, Edoardo A., et al.
Publicado: (2026)
por: Dominici, Edoardo A., et al.
Publicado: (2026)
Image-to-Image Translation with Diffusion Transformers and CLIP-Based Image Conditioning
por: Zhu, Qiang, et al.
Publicado: (2025)
por: Zhu, Qiang, et al.
Publicado: (2025)
FIA-Edit: Frequency-Interactive Attention for Efficient and High-Fidelity Inversion-Free Text-Guided Image Editing
por: Yang, Kaixiang, et al.
Publicado: (2025)
por: Yang, Kaixiang, et al.
Publicado: (2025)
FDS: Frequency-Aware Denoising Score for Text-Guided Latent Diffusion Image Editing
por: Ren, Yufan, et al.
Publicado: (2025)
por: Ren, Yufan, et al.
Publicado: (2025)
Reliable and Efficient Concept Erasure of Text-to-Image Diffusion Models
por: Gong, Chao, et al.
Publicado: (2024)
por: Gong, Chao, et al.
Publicado: (2024)
Image Translation as Diffusion Visual Programmers
por: Han, Cheng, et al.
Publicado: (2024)
por: Han, Cheng, et al.
Publicado: (2024)
FAM Diffusion: Frequency and Attention Modulation for High-Resolution Image Generation with Stable Diffusion
por: Yang, Haosen, et al.
Publicado: (2024)
por: Yang, Haosen, et al.
Publicado: (2024)
Exposing Text-Image Inconsistency Using Diffusion Models
por: Huang, Mingzhen, et al.
Publicado: (2024)
por: Huang, Mingzhen, et al.
Publicado: (2024)
TINA: Text-Free Inversion Attack for Unlearned Text-to-Image Diffusion Models
por: Xiang, Qianlong, et al.
Publicado: (2026)
por: Xiang, Qianlong, et al.
Publicado: (2026)
An Intermediate Fusion ViT Enables Efficient Text-Image Alignment in Diffusion Models
por: Hu, Zizhao, et al.
Publicado: (2024)
por: Hu, Zizhao, et al.
Publicado: (2024)
Controllable Generation with Text-to-Image Diffusion Models: A Survey
por: Cao, Pu, et al.
Publicado: (2024)
por: Cao, Pu, et al.
Publicado: (2024)
Text-guided Controllable Diffusion for Realistic Camouflage Images Generation
por: Qian, Yuhang, et al.
Publicado: (2025)
por: Qian, Yuhang, et al.
Publicado: (2025)
Translation of Text Embedding via Delta Vector to Suppress Strongly Entangled Content in Text-to-Image Diffusion Models
por: Koh, Eunseo, et al.
Publicado: (2025)
por: Koh, Eunseo, et al.
Publicado: (2025)
Emphasizing Crucial Features for Efficient Image Restoration
por: Gao, Hu, et al.
Publicado: (2024)
por: Gao, Hu, et al.
Publicado: (2024)
DesignDiffusion: High-Quality Text-to-Design Image Generation with Diffusion Models
por: Wang, Zhendong, et al.
Publicado: (2025)
por: Wang, Zhendong, et al.
Publicado: (2025)
MobileDiffusion: Instant Text-to-Image Generation on Mobile Devices
por: Zhao, Yang, et al.
Publicado: (2023)
por: Zhao, Yang, et al.
Publicado: (2023)
TextDiffuser-RL: Efficient and Robust Text Layout Optimization for High-Fidelity Text-to-Image Synthesis
por: Rahman, Kazi Mahathir, et al.
Publicado: (2025)
por: Rahman, Kazi Mahathir, et al.
Publicado: (2025)
High-Frequency First: A Two-Stage Approach for Improving Image INR
por: Dam, Sumit Kumar, et al.
Publicado: (2025)
por: Dam, Sumit Kumar, et al.
Publicado: (2025)
StyleInject: Parameter Efficient Tuning of Text-to-Image Diffusion Models
por: Zhou, Mohan, et al.
Publicado: (2024)
por: Zhou, Mohan, et al.
Publicado: (2024)
Ejemplares similares
-
FBSDiff: Plug-and-Play Frequency Band Substitution of Diffusion Features for Highly Controllable Text-Driven Image Translation
por: Gao, Xiang, et al.
Publicado: (2024) -
Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation
por: Gao, Xiang, et al.
Publicado: (2024) -
A Diffusion Model Translator for Efficient Image-to-Image Translation
por: Xia, Mengfei, et al.
Publicado: (2025) -
PiCo: Enhancing Text-Image Alignment with Improved Noise Selection and Precise Mask Control in Diffusion Models
por: Xie, Chang, et al.
Publicado: (2025) -
VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control
por: Wu, Shaojin, et al.
Publicado: (2024)