Dynamic Frequency Modulation for Controllable Text-driven Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shi, Tiandong, Zhao, Ling, Qi, Ji, Ma, Jiayi, Peng, Chengli |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FICGen: Frequency-Inspired Contextual Disentanglement for Layout-driven Degraded Image Generation
von: Wang, Wenzhuang, et al.
Veröffentlicht: (2025)
von: Wang, Wenzhuang, et al.
Veröffentlicht: (2025)
Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
Seeing It Before It Happens: In-Generation NSFW Detection for Diffusion-Based Text-to-Image Models
von: Yang, Fan, et al.
Veröffentlicht: (2025)
von: Yang, Fan, et al.
Veröffentlicht: (2025)
DynamicControl: Adaptive Condition Selection for Improved Text-to-Image Generation
von: He, Qingdong, et al.
Veröffentlicht: (2024)
von: He, Qingdong, et al.
Veröffentlicht: (2024)
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
Text2Earth: Unlocking Text-driven Remote Sensing Image Generation with a Global-Scale Dataset and a Foundation Model
von: Liu, Chenyang, et al.
Veröffentlicht: (2025)
von: Liu, Chenyang, et al.
Veröffentlicht: (2025)
Text-DiFuse: An Interactive Multi-Modal Image Fusion Framework based on Text-modulated Diffusion Model
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
Image Captioning via Dynamic Path Customization
von: Ma, Yiwei, et al.
Veröffentlicht: (2024)
von: Ma, Yiwei, et al.
Veröffentlicht: (2024)
Exploring Phrase-Level Grounding with Text-to-Image Diffusion Model
von: Yang, Danni, et al.
Veröffentlicht: (2024)
von: Yang, Danni, et al.
Veröffentlicht: (2024)
X-Oscar: A Progressive Framework for High-quality Text-guided 3D Animatable Avatar Generation
von: Ma, Yiwei, et al.
Veröffentlicht: (2024)
von: Ma, Yiwei, et al.
Veröffentlicht: (2024)
InsightTok: Improving Text and Face Fidelity in Discrete Tokenization for Autoregressive Image Generation
von: Yue, Yang, et al.
Veröffentlicht: (2026)
von: Yue, Yang, et al.
Veröffentlicht: (2026)
Text-IF: Leveraging Semantic Text Guidance for Degradation-Aware and Interactive Image Fusion
von: Yi, Xunpeng, et al.
Veröffentlicht: (2024)
von: Yi, Xunpeng, et al.
Veröffentlicht: (2024)
DyCoRM: Dynamic Criterion-Aware Reward Modeling for Text-to-Image Generation
von: Qian, Jiaying, et al.
Veröffentlicht: (2026)
von: Qian, Jiaying, et al.
Veröffentlicht: (2026)
Visual Concept-driven Image Generation with Text-to-Image Diffusion Model
von: Rahman, Tanzila, et al.
Veröffentlicht: (2024)
von: Rahman, Tanzila, et al.
Veröffentlicht: (2024)
X-Dreamer: Creating High-quality 3D Content by Bridging the Domain Gap Between Text-to-2D and Text-to-3D Generation
von: Ma, Yiwei, et al.
Veröffentlicht: (2023)
von: Ma, Yiwei, et al.
Veröffentlicht: (2023)
MIGC: Multi-Instance Generation Controller for Text-to-Image Synthesis
von: Zhou, Dewei, et al.
Veröffentlicht: (2024)
von: Zhou, Dewei, et al.
Veröffentlicht: (2024)
MICON-Bench: Benchmarking and Enhancing Multi-Image Context Image Generation in Unified Multimodal Models
von: Wu, Mingrui, et al.
Veröffentlicht: (2026)
von: Wu, Mingrui, et al.
Veröffentlicht: (2026)
Residual Prior-driven Frequency-aware Network for Image Fusion
von: Zheng, Guan, et al.
Veröffentlicht: (2025)
von: Zheng, Guan, et al.
Veröffentlicht: (2025)
Identity-Preserving Text-to-Video Generation via Training-Free Prompt, Image, and Guidance Enhancement
von: Gao, Jiayi, et al.
Veröffentlicht: (2025)
von: Gao, Jiayi, et al.
Veröffentlicht: (2025)
Exploring Timeline Control for Facial Motion Generation
von: Ma, Yifeng, et al.
Veröffentlicht: (2025)
von: Ma, Yifeng, et al.
Veröffentlicht: (2025)
Dynamic Prompt Optimizing for Text-to-Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
Beat: Bi-directional One-to-Many Embedding Alignment for Text-based Person Retrieval
von: Ma, Yiwei, et al.
Veröffentlicht: (2024)
von: Ma, Yiwei, et al.
Veröffentlicht: (2024)
RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation
von: Pang, Lexi, et al.
Veröffentlicht: (2025)
von: Pang, Lexi, et al.
Veröffentlicht: (2025)
Local Conditional Controlling for Text-to-Image Diffusion Models
von: Zhao, Yibo, et al.
Veröffentlicht: (2023)
von: Zhao, Yibo, et al.
Veröffentlicht: (2023)
Mixed Degradation Image Restoration via Local Dynamic Optimization and Conditional Embedding
von: Gu, Yubin, et al.
Veröffentlicht: (2024)
von: Gu, Yubin, et al.
Veröffentlicht: (2024)
FlexEControl: Flexible and Efficient Multimodal Control for Text-to-Image Generation
von: He, Xuehai, et al.
Veröffentlicht: (2024)
von: He, Xuehai, et al.
Veröffentlicht: (2024)
EmotiCrafter: Text-to-Emotional-Image Generation based on Valence-Arousal Model
von: Dang, Shengqi, et al.
Veröffentlicht: (2025)
von: Dang, Shengqi, et al.
Veröffentlicht: (2025)
FAM Diffusion: Frequency and Attention Modulation for High-Resolution Image Generation with Stable Diffusion
von: Yang, Haosen, et al.
Veröffentlicht: (2024)
von: Yang, Haosen, et al.
Veröffentlicht: (2024)
Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining
von: Liu, Dongyang, et al.
Veröffentlicht: (2024)
von: Liu, Dongyang, et al.
Veröffentlicht: (2024)
CIR-CoT: Towards Interpretable Composed Image Retrieval via End-to-End Chain-of-Thought Reasoning
von: Lin, Weihuang, et al.
Veröffentlicht: (2025)
von: Lin, Weihuang, et al.
Veröffentlicht: (2025)
Equilibrated Diffusion: Frequency-aware Textual Embedding for Equilibrated Image Customization
von: Ma, Liyuan, et al.
Veröffentlicht: (2026)
von: Ma, Liyuan, et al.
Veröffentlicht: (2026)
PixelPonder: Dynamic Patch Adaptation for Enhanced Multi-Conditional Text-to-Image Generation
von: Pan, Yanjie, et al.
Veröffentlicht: (2025)
von: Pan, Yanjie, et al.
Veröffentlicht: (2025)
Any-to-3D Generation via Hybrid Diffusion Supervision
von: Fan, Yijun, et al.
Veröffentlicht: (2024)
von: Fan, Yijun, et al.
Veröffentlicht: (2024)
TextOVSR: Text-Guided Real-World Opera Video Super-Resolution
von: Chang, Hua, et al.
Veröffentlicht: (2026)
von: Chang, Hua, et al.
Veröffentlicht: (2026)
Text-Animator: Controllable Visual Text Video Generation
von: Liu, Lin, et al.
Veröffentlicht: (2024)
von: Liu, Lin, et al.
Veröffentlicht: (2024)
MLLM-Selector: Necessity and Diversity-driven High-Value Data Selection for Enhanced Visual Instruction Tuning
von: Ma, Yiwei, et al.
Veröffentlicht: (2025)
von: Ma, Yiwei, et al.
Veröffentlicht: (2025)
Learning to Sample Effective and Diverse Prompts for Text-to-Image Generation
von: Yun, Taeyoung, et al.
Veröffentlicht: (2025)
von: Yun, Taeyoung, et al.
Veröffentlicht: (2025)
ID-EA: Identity-driven Text Enhancement and Adaptation with Textual Inversion for Personalized Text-to-Image Generation
von: Jin, Hyun-Jun, et al.
Veröffentlicht: (2025)
von: Jin, Hyun-Jun, et al.
Veröffentlicht: (2025)
EvoIR: Towards All-in-One Image Restoration via Evolutionary Frequency Modulation
von: Ma, Jiaqi, et al.
Veröffentlicht: (2025)
von: Ma, Jiaqi, et al.
Veröffentlicht: (2025)
IFAdapter: Instance Feature Control for Grounded Text-to-Image Generation
von: Wu, Yinwei, et al.
Veröffentlicht: (2024)
von: Wu, Yinwei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FICGen: Frequency-Inspired Contextual Disentanglement for Layout-driven Degraded Image Generation
von: Wang, Wenzhuang, et al.
Veröffentlicht: (2025) -
Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2024) -
Seeing It Before It Happens: In-Generation NSFW Detection for Diffusion-Based Text-to-Image Models
von: Yang, Fan, et al.
Veröffentlicht: (2025) -
DynamicControl: Adaptive Condition Selection for Improved Text-to-Image Generation
von: He, Qingdong, et al.
Veröffentlicht: (2024) -
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
von: Zhao, Yu, et al.
Veröffentlicht: (2024)