Equilibrated Diffusion: Frequency-aware Textual Embedding for Equilibrated Image Customization
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Liyuan, Fang, Xueji, Qi, Guo-Jun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation
by: Fang, Xueji, et al.
Published: (2026)
by: Fang, Xueji, et al.
Published: (2026)
InfLVG: Reinforce Inference-Time Consistent Long Video Generation with GRPO
by: Fang, Xueji, et al.
Published: (2025)
by: Fang, Xueji, et al.
Published: (2025)
Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image Customization
by: Song, Yeji, et al.
Published: (2024)
by: Song, Yeji, et al.
Published: (2024)
CustomText: Customized Textual Image Generation using Diffusion Models
by: Paliwal, Shubham, et al.
Published: (2024)
by: Paliwal, Shubham, et al.
Published: (2024)
Self-Guidance: Boosting Flow and Diffusion Generation on Their Own
by: Li, Tiancheng, et al.
Published: (2024)
by: Li, Tiancheng, et al.
Published: (2024)
RealCustom++: Representing Images as Real Textual Word for Real-Time Customization
by: Mao, Zhendong, et al.
Published: (2024)
by: Mao, Zhendong, et al.
Published: (2024)
Underwater Image Enhancement by Diffusion Model with Customized CLIP-Classifier
by: Liu, Shuaixin, et al.
Published: (2024)
by: Liu, Shuaixin, et al.
Published: (2024)
Uncertainty-aware Spatial-Frequency Registration and Fusion for Infrared and Visible Images
by: Li, Xingyuan, et al.
Published: (2026)
by: Li, Xingyuan, et al.
Published: (2026)
Customizing Text-to-Image Diffusion with Object Viewpoint Control
by: Kumari, Nupur, et al.
Published: (2024)
by: Kumari, Nupur, et al.
Published: (2024)
Embedding Textual Information in Images Using Quinary Pixel Combinations
by: Kandala, A V Uday Kiran
Published: (2026)
by: Kandala, A V Uday Kiran
Published: (2026)
Frequency-aware Event Cloud Network
by: Ren, Hongwei, et al.
Published: (2024)
by: Ren, Hongwei, et al.
Published: (2024)
DECOR:Decomposition and Projection of Text Embeddings for Text-to-Image Customization
by: Jang, Geonhui, et al.
Published: (2024)
by: Jang, Geonhui, et al.
Published: (2024)
Disentangled Textual Priors for Diffusion-based Image Super-Resolution
by: Jiang, Lei, et al.
Published: (2026)
by: Jiang, Lei, et al.
Published: (2026)
Spatial-and-Frequency-aware Restoration method for Images based on Diffusion Models
by: Lee, Kyungsung, et al.
Published: (2024)
by: Lee, Kyungsung, et al.
Published: (2024)
Frequency-aware Neural Representation for Videos
by: Zhu, Jun, et al.
Published: (2026)
by: Zhu, Jun, et al.
Published: (2026)
Frequency Domain-Based Diffusion Model for Unpaired Image Dehazing
by: Liu, Chengxu, et al.
Published: (2025)
by: Liu, Chengxu, et al.
Published: (2025)
Event-Customized Image Generation
by: Wang, Zhen, et al.
Published: (2024)
by: Wang, Zhen, et al.
Published: (2024)
PIDiff: Image Customization for Personalized Identities with Diffusion Models
by: Gu, Jinyu, et al.
Published: (2025)
by: Gu, Jinyu, et al.
Published: (2025)
Learning to Customize Text-to-Image Diffusion In Diverse Context
by: Kim, Taewook, et al.
Published: (2024)
by: Kim, Taewook, et al.
Published: (2024)
Contact-aware Human Motion Generation from Textual Descriptions
by: Ma, Sihan, et al.
Published: (2024)
by: Ma, Sihan, et al.
Published: (2024)
LMHaze: Intensity-aware Image Dehazing with a Large-scale Multi-intensity Real Haze Dataset
by: Zhang, Ruikun, et al.
Published: (2024)
by: Zhang, Ruikun, et al.
Published: (2024)
Instance-aware Image Colorization with Controllable Textual Descriptions and Segmentation Masks
by: An, Yanru, et al.
Published: (2025)
by: An, Yanru, et al.
Published: (2025)
LatexBlend: Scaling Multi-concept Customized Generation with Latent Textual Blending
by: Jin, Jian, et al.
Published: (2025)
by: Jin, Jian, et al.
Published: (2025)
FrequencyBooster: Full-Frequency Modeling for High-Fidelity Pixel Diffusion
by: Ma, Lichen, et al.
Published: (2026)
by: Ma, Lichen, et al.
Published: (2026)
Infusion: Preventing Customized Text-to-Image Diffusion from Overfitting
by: Zeng, Weili, et al.
Published: (2024)
by: Zeng, Weili, et al.
Published: (2024)
DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation
by: Ma, Zehong, et al.
Published: (2025)
by: Ma, Zehong, et al.
Published: (2025)
EmoAttack: Emotion-to-Image Diffusion Models for Emotional Backdoor Generation
by: Wei, Tianyu, et al.
Published: (2024)
by: Wei, Tianyu, et al.
Published: (2024)
Diff-PC: Identity-preserving and 3D-aware Controllable Diffusion for Zero-shot Portrait Customization
by: Xu, Yifang, et al.
Published: (2026)
by: Xu, Yifang, et al.
Published: (2026)
CustomVideo: Customizing Text-to-Video Generation with Multiple Subjects
by: Wang, Zhao, et al.
Published: (2024)
by: Wang, Zhao, et al.
Published: (2024)
FAIR: Frequency-aware Image Restoration for Industrial Visual Anomaly Detection
by: Liu, Tongkun, et al.
Published: (2023)
by: Liu, Tongkun, et al.
Published: (2023)
Direct Consistency Optimization for Robust Customization of Text-to-Image Diffusion Models
by: Lee, Kyungmin, et al.
Published: (2024)
by: Lee, Kyungmin, et al.
Published: (2024)
T-LoRA: Single Image Diffusion Model Customization Without Overfitting
by: Soboleva, Vera, et al.
Published: (2025)
by: Soboleva, Vera, et al.
Published: (2025)
How to Continually Adapt Text-to-Image Diffusion Models for Flexible Customization?
by: Dong, Jiahua, et al.
Published: (2024)
by: Dong, Jiahua, et al.
Published: (2024)
Object-Driven One-Shot Fine-tuning of Text-to-Image Diffusion with Prototypical Embedding
by: Lu, Jianxiang, et al.
Published: (2024)
by: Lu, Jianxiang, et al.
Published: (2024)
Dif-Fusion: Towards High Color Fidelity in Infrared and Visible Image Fusion with Diffusion Models
by: Yue, Jun, et al.
Published: (2023)
by: Yue, Jun, et al.
Published: (2023)
Diffusion-Driven Inter-Outer Surface Separation for Point Clouds with Open Boundaries
by: Qin, Zhengyan, et al.
Published: (2026)
by: Qin, Zhengyan, et al.
Published: (2026)
Dynamic Frequency Modulation for Controllable Text-driven Image Generation
by: Shi, Tiandong, et al.
Published: (2026)
by: Shi, Tiandong, et al.
Published: (2026)
Residual Prior-driven Frequency-aware Network for Image Fusion
by: Zheng, Guan, et al.
Published: (2025)
by: Zheng, Guan, et al.
Published: (2025)
Toward Sufficient Spatial-Frequency Interaction for Gradient-aware Underwater Image Enhancement
by: Zhao, Chen, et al.
Published: (2023)
by: Zhao, Chen, et al.
Published: (2023)
When Images Speak Louder: Mitigating Language Bias-induced Hallucinations in VLMs through Cross-Modal Guidance
by: Cao, Jinjin, et al.
Published: (2025)
by: Cao, Jinjin, et al.
Published: (2025)
Similar Items
-
FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation
by: Fang, Xueji, et al.
Published: (2026) -
InfLVG: Reinforce Inference-Time Consistent Long Video Generation with GRPO
by: Fang, Xueji, et al.
Published: (2025) -
Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image Customization
by: Song, Yeji, et al.
Published: (2024) -
CustomText: Customized Textual Image Generation using Diffusion Models
by: Paliwal, Shubham, et al.
Published: (2024) -
Self-Guidance: Boosting Flow and Diffusion Generation on Their Own
by: Li, Tiancheng, et al.
Published: (2024)