Chain-of-Image Generation: Toward Monitorable and Controllable Image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Young Kyung, Schlesinger, Oded, Zhao, Yuzhou, Di Martino, J. Matias, Sapiro, Guillermo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SPOT: Sparsification with Attention Dynamics via Token Relevance in Vision Transformers
by: Schlesinger, Oded, et al.
Published: (2025)
by: Schlesinger, Oded, et al.
Published: (2025)
Vision Transformers with Natural Language Semantics
by: Kim, Young Kyung, et al.
Published: (2024)
by: Kim, Young Kyung, et al.
Published: (2024)
Markerless Head Tracking for Accurate and Accessible Neuronavigation
by: Xie, Ziye, et al.
Published: (2026)
by: Xie, Ziye, et al.
Published: (2026)
M2P: Improving Visual Foundation Models with Mask-to-Point Weakly-Supervised Learning for Dense Point Tracking
by: Wu, Qiangqiang, et al.
Published: (2026)
by: Wu, Qiangqiang, et al.
Published: (2026)
Order-Aware Test-Time Adaptation: Leveraging Temporal Dynamics for Robust Streaming Inference
by: Kim, Young Kyung, et al.
Published: (2026)
by: Kim, Young Kyung, et al.
Published: (2026)
Generative Visual Chain-of-Thought for Image Editing
by: Yin, Zijin, et al.
Published: (2026)
by: Yin, Zijin, et al.
Published: (2026)
Unlocking Compositional Control: Self-Supervision for LVLM-Based Image Generation
by: Garcia, Fernando Gabriela, et al.
Published: (2025)
by: Garcia, Fernando Gabriela, et al.
Published: (2025)
ChangeBridge: Spatiotemporal Image Generation with Multimodal Controls for Remote Sensing
by: Zhao, Zhenghui, et al.
Published: (2025)
by: Zhao, Zhenghui, et al.
Published: (2025)
Phy124: Fast Physics-Driven 4D Content Generation from a Single Image
by: Lin, Jiajing, et al.
Published: (2024)
by: Lin, Jiajing, et al.
Published: (2024)
Generating Storytelling Images with Rich Chains-of-Reasoning
by: Song, Xiujie, et al.
Published: (2025)
by: Song, Xiujie, et al.
Published: (2025)
SCU-CGAN: Enhancing Fire Detection through Synthetic Fire Image Generation and Dataset Augmentation
by: Kim, Ju-Young, et al.
Published: (2025)
by: Kim, Ju-Young, et al.
Published: (2025)
Canvas-to-Image: Compositional Image Generation with Multimodal Controls
by: Dalva, Yusuf, et al.
Published: (2025)
by: Dalva, Yusuf, et al.
Published: (2025)
Attack Deterministic Conditional Image Generative Models for Diverse and Controllable Generation
by: Chu, Tianyi, et al.
Published: (2024)
by: Chu, Tianyi, et al.
Published: (2024)
Towards Enhanced Image Generation Via Multi-modal Chain of Thought in Unified Generative Models
by: Wang, Yi, et al.
Published: (2025)
by: Wang, Yi, et al.
Published: (2025)
Improving Synthetic Image Detection Towards Generalization: An Image Transformation Perspective
by: Li, Ouxiang, et al.
Published: (2024)
by: Li, Ouxiang, et al.
Published: (2024)
Improving Text Generation on Images with Synthetic Captions
by: Koh, Jun Young, et al.
Published: (2024)
by: Koh, Jun Young, et al.
Published: (2024)
Dynamic Frequency Modulation for Controllable Text-driven Image Generation
by: Shi, Tiandong, et al.
Published: (2026)
by: Shi, Tiandong, et al.
Published: (2026)
Controlling the Latent Diffusion Model for Generative Image Shadow Removal via Residual Generation
by: Li, Xinjie, et al.
Published: (2024)
by: Li, Xinjie, et al.
Published: (2024)
ControlAR: Controllable Image Generation with Autoregressive Models
by: Li, Zongming, et al.
Published: (2024)
by: Li, Zongming, et al.
Published: (2024)
Adaptive Domain Shift in Diffusion Models for Cross-Modality Image Translation
by: Wang, Zihao, et al.
Published: (2026)
by: Wang, Zihao, et al.
Published: (2026)
Layout-Guided Controllable Pathology Image Generation with In-Context Diffusion Transformers
by: Shou, Yuntao, et al.
Published: (2026)
by: Shou, Yuntao, et al.
Published: (2026)
WithAnyone: Towards Controllable and ID Consistent Image Generation
by: Xu, Hengyuan, et al.
Published: (2025)
by: Xu, Hengyuan, et al.
Published: (2025)
Towards Better & Faster Autoregressive Image Generation: From the Perspective of Entropy
by: Ma, Xiaoxiao, et al.
Published: (2025)
by: Ma, Xiaoxiao, et al.
Published: (2025)
GPS as a Control Signal for Image Generation
by: Feng, Chao, et al.
Published: (2025)
by: Feng, Chao, et al.
Published: (2025)
Filter-Guided Diffusion for Controllable Image Generation
by: Gu, Zeqi, et al.
Published: (2023)
by: Gu, Zeqi, et al.
Published: (2023)
ID-EA: Identity-driven Text Enhancement and Adaptation with Textual Inversion for Personalized Text-to-Image Generation
by: Jin, Hyun-Jun, et al.
Published: (2025)
by: Jin, Hyun-Jun, et al.
Published: (2025)
Visual-CoG: Stage-Aware Reinforcement Learning with Chain of Guidance for Text-to-Image Generation
by: Li, Yaqi, et al.
Published: (2025)
by: Li, Yaqi, et al.
Published: (2025)
A Preliminary Exploration Towards General Image Restoration
by: Kong, Xiangtao, et al.
Published: (2024)
by: Kong, Xiangtao, et al.
Published: (2024)
ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image Generation
by: Qian, Guocheng Gordon, et al.
Published: (2025)
by: Qian, Guocheng Gordon, et al.
Published: (2025)
Towards Generalizable AI-Generated Image Detection via Image-Adaptive Prompt Learning
by: Li, Yiheng, et al.
Published: (2025)
by: Li, Yiheng, et al.
Published: (2025)
Test-time Controllable Image Generation by Explicit Spatial Constraint Enforcement
by: Zhang, Z., et al.
Published: (2025)
by: Zhang, Z., et al.
Published: (2025)
AttnDreamBooth: Towards Text-Aligned Personalized Text-to-Image Generation
by: Pang, Lianyu, et al.
Published: (2024)
by: Pang, Lianyu, et al.
Published: (2024)
PIXART-δ: Fast and Controllable Image Generation with Latent Consistency Models
by: Chen, Junsong, et al.
Published: (2024)
by: Chen, Junsong, et al.
Published: (2024)
Towards Consistent and Controllable Image Synthesis for Face Editing
by: Wei, Mengting, et al.
Published: (2025)
by: Wei, Mengting, et al.
Published: (2025)
OmniDrag: Enabling Motion Control for Omnidirectional Image-to-Video Generation
by: Li, Weiqi, et al.
Published: (2024)
by: Li, Weiqi, et al.
Published: (2024)
FlowTurbo: Towards Real-time Flow-Based Image Generation with Velocity Refiner
by: Zhao, Wenliang, et al.
Published: (2024)
by: Zhao, Wenliang, et al.
Published: (2024)
A Review on Generative AI For Text-To-Image and Image-To-Image Generation and Implications To Scientific Images
by: Sordo, Zineb, et al.
Published: (2025)
by: Sordo, Zineb, et al.
Published: (2025)
LayeringDiff: Layered Image Synthesis via Generation, then Disassembly with Generative Knowledge
by: Kang, Kyoungkook, et al.
Published: (2025)
by: Kang, Kyoungkook, et al.
Published: (2025)
Towards Controllable Image Generation through Representation-Conditioned Diffusion Models
by: Karthikeyan, Nithesh Chandher, et al.
Published: (2026)
by: Karthikeyan, Nithesh Chandher, et al.
Published: (2026)
Seedream 4.0: Toward Next-generation Multimodal Image Generation
by: Seedream, Team, et al.
Published: (2025)
by: Seedream, Team, et al.
Published: (2025)
Similar Items
-
SPOT: Sparsification with Attention Dynamics via Token Relevance in Vision Transformers
by: Schlesinger, Oded, et al.
Published: (2025) -
Vision Transformers with Natural Language Semantics
by: Kim, Young Kyung, et al.
Published: (2024) -
Markerless Head Tracking for Accurate and Accessible Neuronavigation
by: Xie, Ziye, et al.
Published: (2026) -
M2P: Improving Visual Foundation Models with Mask-to-Point Weakly-Supervised Learning for Dense Point Tracking
by: Wu, Qiangqiang, et al.
Published: (2026) -
Order-Aware Test-Time Adaptation: Leveraging Temporal Dynamics for Robust Streaming Inference
by: Kim, Young Kyung, et al.
Published: (2026)