DC-ControlNet: Decoupling Inter- and Intra-Element Conditions in Image Generation with Diffusion Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Hongji, Han, Wencheng, Zhou, Yucheng, Shen, Jianbing |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Self-Rewarding Large Vision-Language Models for Optimizing Prompts in Text-to-Image Generation
por: Yang, Hongji, et al.
Publicado: (2025)
por: Yang, Hongji, et al.
Publicado: (2025)
HiCoGen: Hierarchical Compositional Text-to-Image Generation in Diffusion Models via Reinforcement Learning
por: Yang, Hongji, et al.
Publicado: (2025)
por: Yang, Hongji, et al.
Publicado: (2025)
HanMoVLM: Large Vision-Language Models for Professional Artistic Painting Evaluation
por: Yang, Hongji, et al.
Publicado: (2026)
por: Yang, Hongji, et al.
Publicado: (2026)
Decoupling Fine Detail and Global Geometry for Compressed Depth Map Super-Resolution
por: Zheng, Huan, et al.
Publicado: (2024)
por: Zheng, Huan, et al.
Publicado: (2024)
Condition Errors Refinement in Autoregressive Image Generation with Diffusion Loss
por: Zhou, Yucheng, et al.
Publicado: (2026)
por: Zhou, Yucheng, et al.
Publicado: (2026)
PG-ControlNet: A Physics-Guided ControlNet for Generative Spatially Varying Image Deblurring
por: Motorcu, Hakki, et al.
Publicado: (2025)
por: Motorcu, Hakki, et al.
Publicado: (2025)
High-Precision Self-Supervised Monocular Depth Estimation with Rich-Resource Prior
por: Han, Wencheng, et al.
Publicado: (2024)
por: Han, Wencheng, et al.
Publicado: (2024)
EasyControl: Transfer ControlNet to Video Diffusion for Controllable Generation and Interpolation
por: Wang, Cong, et al.
Publicado: (2024)
por: Wang, Cong, et al.
Publicado: (2024)
CogOmniControl: Reasoning-Driven Controllable Video Generation via Creative Intent Cognition
por: Yang, Hongji, et al.
Publicado: (2026)
por: Yang, Hongji, et al.
Publicado: (2026)
RepControlNet: ControlNet Reparameterization
por: Deng, Zhaoli, et al.
Publicado: (2024)
por: Deng, Zhaoli, et al.
Publicado: (2024)
Towards Better Cephalometric Landmark Detection with Diffusion Data Generation
por: Guo, Dongqian, et al.
Publicado: (2025)
por: Guo, Dongqian, et al.
Publicado: (2025)
ControlNet-XS: Rethinking the Control of Text-to-Image Diffusion Models as Feedback-Control Systems
por: Zavadski, Denis, et al.
Publicado: (2023)
por: Zavadski, Denis, et al.
Publicado: (2023)
SmartControl: Enhancing ControlNet for Handling Rough Visual Conditions
por: Liu, Xiaoyu, et al.
Publicado: (2024)
por: Liu, Xiaoyu, et al.
Publicado: (2024)
Towards High-Fidelity 3D Portrait Generation with Rich Details by Cross-View Prior-Aware Diffusion
por: Wei, Haoran, et al.
Publicado: (2024)
por: Wei, Haoran, et al.
Publicado: (2024)
Layout-to-Image Generation with Localized Descriptions using ControlNet with Cross-Attention Control
por: Lukovnikov, Denis, et al.
Publicado: (2024)
por: Lukovnikov, Denis, et al.
Publicado: (2024)
Mask-ControlNet: Higher-Quality Image Generation with An Additional Mask Prompt
por: Huang, Zhiqi, et al.
Publicado: (2024)
por: Huang, Zhiqi, et al.
Publicado: (2024)
ViscoNet: Bridging and Harmonizing Visual and Textual Conditioning for ControlNet
por: Cheong, Soon Yau, et al.
Publicado: (2023)
por: Cheong, Soon Yau, et al.
Publicado: (2023)
From Broad Exploration to Stable Synthesis: Entropy-Guided Optimization for Autoregressive Image Generation
por: Song, Han, et al.
Publicado: (2026)
por: Song, Han, et al.
Publicado: (2026)
ControlNet++: Improving Conditional Controls with Efficient Consistency Feedback
por: Li, Ming, et al.
Publicado: (2024)
por: Li, Ming, et al.
Publicado: (2024)
RAWMamba: Unified sRGB-to-RAW De-rendering With State Space Model
por: Chen, Hongjun, et al.
Publicado: (2024)
por: Chen, Hongjun, et al.
Publicado: (2024)
When ControlNet Meets Inexplicit Masks: A Case Study of ControlNet on its Contour-following Ability
por: Xuan, Wenjie, et al.
Publicado: (2024)
por: Xuan, Wenjie, et al.
Publicado: (2024)
Adaptively Distilled ControlNet: Accelerated Training and Superior Sampling for Medical Image Synthesis
por: Qiu, Kunpeng, et al.
Publicado: (2025)
por: Qiu, Kunpeng, et al.
Publicado: (2025)
AdaOcc: Adaptive Forward View Transformation and Flow Modeling for 3D Occupancy and Flow Prediction
por: Chen, Dubing, et al.
Publicado: (2024)
por: Chen, Dubing, et al.
Publicado: (2024)
Uncertainty-Aware ControlNet: Bridging Domain Gaps with Synthetic Image Generation
por: Niemeijer, Joshua, et al.
Publicado: (2025)
por: Niemeijer, Joshua, et al.
Publicado: (2025)
Towards Geometry-Aware and Motion-Guided Video Human Mesh Recovery
por: Chen, Hongjun, et al.
Publicado: (2026)
por: Chen, Hongjun, et al.
Publicado: (2026)
MambaControl: Anatomy Graph-Enhanced Mamba ControlNet with Fourier Refinement for Diffusion-Based Disease Trajectory Prediction
por: Yang, Hao, et al.
Publicado: (2025)
por: Yang, Hao, et al.
Publicado: (2025)
JieHua Paintings Style Feature Extracting Model using Stable Diffusion with ControlNet
por: Gu, Yujia, et al.
Publicado: (2024)
por: Gu, Yujia, et al.
Publicado: (2024)
Reducing CT Metal Artifacts by Learning Latent Space Alignment with Gemstone Spectral Imaging Data
por: Han, Wencheng, et al.
Publicado: (2025)
por: Han, Wencheng, et al.
Publicado: (2025)
SemanticControl: A Training-Free Approach for Handling Loosely Aligned Visual Conditions in ControlNet
por: Joung, Woosung, et al.
Publicado: (2025)
por: Joung, Woosung, et al.
Publicado: (2025)
RepVF: A Unified Vector Fields Representation for Multi-task 3D Perception
por: Li, Chunliang, et al.
Publicado: (2024)
por: Li, Chunliang, et al.
Publicado: (2024)
DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous Driving
por: Han, Wencheng, et al.
Publicado: (2024)
por: Han, Wencheng, et al.
Publicado: (2024)
Visual In-Context Learning for Large Vision-Language Models
por: Zhou, Yucheng, et al.
Publicado: (2024)
por: Zhou, Yucheng, et al.
Publicado: (2024)
Meta ControlNet: Enhancing Task Adaptation via Meta Learning
por: Yang, Junjie, et al.
Publicado: (2023)
por: Yang, Junjie, et al.
Publicado: (2023)
LOVECon: Text-driven Training-Free Long Video Editing with ControlNet
por: Liao, Zhenyi, et al.
Publicado: (2023)
por: Liao, Zhenyi, et al.
Publicado: (2023)
Active Learning Inspired ControlNet Guidance for Augmenting Semantic Segmentation Datasets
por: Kniesel, Hannah, et al.
Publicado: (2025)
por: Kniesel, Hannah, et al.
Publicado: (2025)
Adaptive Whole-Body PET Image Denoising Using 3D Diffusion Models with ControlNet
por: Yu, Boxiao, et al.
Publicado: (2024)
por: Yu, Boxiao, et al.
Publicado: (2024)
RLGF: Reinforcement Learning with Geometric Feedback for Autonomous Driving Video Generation
por: Yan, Tianyi, et al.
Publicado: (2025)
por: Yan, Tianyi, et al.
Publicado: (2025)
Rethinking Visual Dependency in Long-Context Reasoning for Large Vision-Language Models
por: Zhou, Yucheng, et al.
Publicado: (2024)
por: Zhou, Yucheng, et al.
Publicado: (2024)
ContRail: A Framework for Realistic Railway Image Synthesis using ControlNet
por: Alexandrescu, Andrei-Robert, et al.
Publicado: (2024)
por: Alexandrescu, Andrei-Robert, et al.
Publicado: (2024)
Abstract Art Interpretation Using ControlNet
por: Srivastava, Rishabh, et al.
Publicado: (2024)
por: Srivastava, Rishabh, et al.
Publicado: (2024)
Ejemplares similares
-
Self-Rewarding Large Vision-Language Models for Optimizing Prompts in Text-to-Image Generation
por: Yang, Hongji, et al.
Publicado: (2025) -
HiCoGen: Hierarchical Compositional Text-to-Image Generation in Diffusion Models via Reinforcement Learning
por: Yang, Hongji, et al.
Publicado: (2025) -
HanMoVLM: Large Vision-Language Models for Professional Artistic Painting Evaluation
por: Yang, Hongji, et al.
Publicado: (2026) -
Decoupling Fine Detail and Global Geometry for Compressed Depth Map Super-Resolution
por: Zheng, Huan, et al.
Publicado: (2024) -
Condition Errors Refinement in Autoregressive Image Generation with Diffusion Loss
por: Zhou, Yucheng, et al.
Publicado: (2026)