Ctrl-X: Controlling Structure and Appearance for Text-To-Image Generation Without Guidance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Kuan Heng, Mo, Sicheng, Klingher, Ben, Mu, Fangzhou, Zhou, Bolei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SnAG: Scalable and Accurate Video Grounding
von: Mu, Fangzhou, et al.
Veröffentlicht: (2024)
von: Mu, Fangzhou, et al.
Veröffentlicht: (2024)
TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
CFG-Ctrl: Control-Based Classifier-Free Diffusion Guidance
von: Wang, Hanyang, et al.
Veröffentlicht: (2026)
von: Wang, Hanyang, et al.
Veröffentlicht: (2026)
Dreamland: Controllable World Creation with Simulator and Generative Models
von: Mo, Sicheng, et al.
Veröffentlicht: (2025)
von: Mo, Sicheng, et al.
Veröffentlicht: (2025)
CameraCtrl: Enabling Camera Control for Text-to-Video Generation
von: He, Hao, et al.
Veröffentlicht: (2024)
von: He, Hao, et al.
Veröffentlicht: (2024)
RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation
von: Pang, Lexi, et al.
Veröffentlicht: (2025)
von: Pang, Lexi, et al.
Veröffentlicht: (2025)
EmoCtrl: Controllable Emotional Image Content Generation
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
SimGen: Simulator-conditioned Driving Scene Generation
von: Zhou, Yunsong, et al.
Veröffentlicht: (2024)
von: Zhou, Yunsong, et al.
Veröffentlicht: (2024)
Ctrl-GenAug: Controllable Generative Augmentation for Medical Sequence Classification
von: Zhou, Xinrui, et al.
Veröffentlicht: (2024)
von: Zhou, Xinrui, et al.
Veröffentlicht: (2024)
Visual Generation Without Guidance
von: Chen, Huayu, et al.
Veröffentlicht: (2025)
von: Chen, Huayu, et al.
Veröffentlicht: (2025)
LumiCtrl : Learning Illuminant Prompts for Lighting Control in Personalized Text-to-Image Models
von: Butt, Muhammad Atif, et al.
Veröffentlicht: (2025)
von: Butt, Muhammad Atif, et al.
Veröffentlicht: (2025)
IPDreamer: Appearance-Controllable 3D Object Generation with Complex Image Prompts
von: Zeng, Bohan, et al.
Veröffentlicht: (2023)
von: Zeng, Bohan, et al.
Veröffentlicht: (2023)
Ctrl-Room: Controllable Text-to-3D Room Meshes Generation with Layout Constraints
von: Fang, Chuan, et al.
Veröffentlicht: (2023)
von: Fang, Chuan, et al.
Veröffentlicht: (2023)
DynamiCtrl: Rethinking the Basic Structure and the Role of Text for High-quality Human Image Animation
von: Zhao, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhao, Haoyu, et al.
Veröffentlicht: (2025)
SafeCtrl: Region-Based Safety Control for Text-to-Image Diffusion via Detect-Then-Suppress
von: Zhang, Lingyun, et al.
Veröffentlicht: (2025)
von: Zhang, Lingyun, et al.
Veröffentlicht: (2025)
SafeCtrl: Region-Aware Safety Control for Text-to-Image Diffusion via Detect-Then-Suppress
von: Zhang, Lingyun, et al.
Veröffentlicht: (2026)
von: Zhang, Lingyun, et al.
Veröffentlicht: (2026)
MotionCtrl: A Unified and Flexible Motion Controller for Video Generation
von: Wang, Zhouxia, et al.
Veröffentlicht: (2023)
von: Wang, Zhouxia, et al.
Veröffentlicht: (2023)
Orchid: Image Latent Diffusion for Joint Appearance and Geometry Generation
von: Krishnan, Akshay, et al.
Veröffentlicht: (2025)
von: Krishnan, Akshay, et al.
Veröffentlicht: (2025)
CtrlSynth: Controllable Image Text Synthesis for Data-Efficient Multimodal Learning
von: Cao, Qingqing, et al.
Veröffentlicht: (2024)
von: Cao, Qingqing, et al.
Veröffentlicht: (2024)
Ctrl-Adapter: An Efficient and Versatile Framework for Adapting Diverse Controls to Any Diffusion Model
von: Lin, Han, et al.
Veröffentlicht: (2024)
von: Lin, Han, et al.
Veröffentlicht: (2024)
Steering Guidance for Personalized Text-to-Image Diffusion Models
von: Park, Sunghyun, et al.
Veröffentlicht: (2025)
von: Park, Sunghyun, et al.
Veröffentlicht: (2025)
Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
Group Diffusion: Enhancing Image Generation by Unlocking Cross-Sample Collaboration
von: Mo, Sicheng, et al.
Veröffentlicht: (2025)
von: Mo, Sicheng, et al.
Veröffentlicht: (2025)
Street-View Image Generation from a Bird's-Eye View Layout
von: Swerdlow, Alexander, et al.
Veröffentlicht: (2023)
von: Swerdlow, Alexander, et al.
Veröffentlicht: (2023)
PhysCtrl: Generative Physics for Controllable and Physics-Grounded Video Generation
von: Wang, Chen, et al.
Veröffentlicht: (2025)
von: Wang, Chen, et al.
Veröffentlicht: (2025)
Text-Driven Weakly Supervised OCT Lesion Segmentation with Structural Guidance
von: Yang, Jiaqi, et al.
Veröffentlicht: (2024)
von: Yang, Jiaqi, et al.
Veröffentlicht: (2024)
High Fidelity Text to Image Generation with Contrastive Alignment and Structural Guidance
von: Gao, Danyi
Veröffentlicht: (2025)
von: Gao, Danyi
Veröffentlicht: (2025)
CADKnitter: Compositional CAD Generation from Text and Geometry Guidance
von: Le, Tri, et al.
Veröffentlicht: (2025)
von: Le, Tri, et al.
Veröffentlicht: (2025)
TokenPure: Watermark Removal through Tokenized Appearance and Structural Guidance
von: Yang, Pei, et al.
Veröffentlicht: (2025)
von: Yang, Pei, et al.
Veröffentlicht: (2025)
Ctrl-A: Control-Driven Online Data Augmentation
von: Christensen, Jesper B., et al.
Veröffentlicht: (2026)
von: Christensen, Jesper B., et al.
Veröffentlicht: (2026)
Joint Learning of Depth and Appearance for Portrait Image Animation
von: Ji, Xinya, et al.
Veröffentlicht: (2025)
von: Ji, Xinya, et al.
Veröffentlicht: (2025)
LumiX: Structured and Coherent Text-to-Intrinsic Generation
von: Han, Xu, et al.
Veröffentlicht: (2025)
von: Han, Xu, et al.
Veröffentlicht: (2025)
EditCtrl: Disentangled Local and Global Control for Real-Time Generative Video Editing
von: Litman, Yehonathan, et al.
Veröffentlicht: (2026)
von: Litman, Yehonathan, et al.
Veröffentlicht: (2026)
CtrlFuse: Mask-Prompt Guided Controllable Infrared and Visible Image Fusion
von: Sun, Yiming, et al.
Veröffentlicht: (2026)
von: Sun, Yiming, et al.
Veröffentlicht: (2026)
Conditional Text-to-Image Generation with Reference Guidance
von: Kim, Taewook, et al.
Veröffentlicht: (2024)
von: Kim, Taewook, et al.
Veröffentlicht: (2024)
Estimating Appearance Models for Image Segmentation via Tensor Factorization
von: Neto, Jeova Farias Sales Rocha
Veröffentlicht: (2022)
von: Neto, Jeova Farias Sales Rocha
Veröffentlicht: (2022)
DiffArtist: Towards Structure and Appearance Controllable Image Stylization
von: Jiang, Ruixiang, et al.
Veröffentlicht: (2024)
von: Jiang, Ruixiang, et al.
Veröffentlicht: (2024)
Performance Plateaus in Inference-Time Scaling for Text-to-Image Diffusion Without External Models
von: Choi, Changhyun, et al.
Veröffentlicht: (2025)
von: Choi, Changhyun, et al.
Veröffentlicht: (2025)
Zero-Shot Visual Concept Blending Without Text Guidance
von: Makino, Hiroya, et al.
Veröffentlicht: (2025)
von: Makino, Hiroya, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SnAG: Scalable and Accurate Video Grounding
von: Mu, Fangzhou, et al.
Veröffentlicht: (2024) -
TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
von: Zeng, Weichao, et al.
Veröffentlicht: (2024) -
CFG-Ctrl: Control-Based Classifier-Free Diffusion Guidance
von: Wang, Hanyang, et al.
Veröffentlicht: (2026) -
Dreamland: Controllable World Creation with Simulator and Generative Models
von: Mo, Sicheng, et al.
Veröffentlicht: (2025) -
CameraCtrl: Enabling Camera Control for Text-to-Video Generation
von: He, Hao, et al.
Veröffentlicht: (2024)