CyCLeGen: Cycle-Consistent Layout Prediction and Image Generation in Vision Foundation Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Shan, Xiaojun, Shen, Haoyu, Mao, Yucheng, Zhang, Xiang, Anand, Abhay, Li, Bingnan, Xu, Haiyang, Tu, Zhuowen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
OverLayBench: A Benchmark for Layout-to-Image Generation with Dense Overlaps
di: Li, Bingnan, et al.
Pubblicazione: (2025)
di: Li, Bingnan, et al.
Pubblicazione: (2025)
YOLO-Count: Differentiable Object Counting for Text-to-Image Generation
di: Zeng, Guanning, et al.
Pubblicazione: (2025)
di: Zeng, Guanning, et al.
Pubblicazione: (2025)
EchoGen: Cycle-Consistent Learning for Unified Layout-Image Generation and Understanding
di: Zou, Kai, et al.
Pubblicazione: (2026)
di: Zou, Kai, et al.
Pubblicazione: (2026)
Lay-Your-Scene: Natural Scene Layout Generation with Diffusion Transformers
di: Srivastava, Divyansh, et al.
Pubblicazione: (2025)
di: Srivastava, Divyansh, et al.
Pubblicazione: (2025)
CVP: Central-Peripheral Vision-Inspired Multimodal Model for Spatial Reasoning
di: Chen, Zeyuan, et al.
Pubblicazione: (2025)
di: Chen, Zeyuan, et al.
Pubblicazione: (2025)
OmniControlNet: Dual-stage Integration for Conditional Image Generation
di: Wang, Yilin, et al.
Pubblicazione: (2024)
di: Wang, Yilin, et al.
Pubblicazione: (2024)
Exploring the Equivalence of Closed-Set Generative and Real Data Augmentation in Image Classification
di: Wang, Haowen, et al.
Pubblicazione: (2025)
di: Wang, Haowen, et al.
Pubblicazione: (2025)
VideoNSA: Native Sparse Attention Scales Video Understanding
di: Song, Enxin, et al.
Pubblicazione: (2025)
di: Song, Enxin, et al.
Pubblicazione: (2025)
ContextGen: Contextual Layout Anchoring for Identity-Consistent Multi-Instance Generation
di: Xu, Ruihang, et al.
Pubblicazione: (2025)
di: Xu, Ruihang, et al.
Pubblicazione: (2025)
AWRaCLe: All-Weather Image Restoration using Visual In-Context Learning
di: Rajagopalan, Sudarshan, et al.
Pubblicazione: (2024)
di: Rajagopalan, Sudarshan, et al.
Pubblicazione: (2024)
PlanGen: Towards Unified Layout Planning and Image Generation in Auto-Regressive Vision Language Models
di: He, Runze, et al.
Pubblicazione: (2025)
di: He, Runze, et al.
Pubblicazione: (2025)
CLeAN: Continual Learning Adaptive Normalization in Dynamic Environments
di: Marasco, Isabella, et al.
Pubblicazione: (2026)
di: Marasco, Isabella, et al.
Pubblicazione: (2026)
R-C2: Cycle-Consistent Reinforcement Learning Improves Multimodal Reasoning
di: Zhang, Zirui, et al.
Pubblicazione: (2026)
di: Zhang, Zirui, et al.
Pubblicazione: (2026)
C3Editor: Achieving Controllable Consistency in 2D Model for 3D Editing
di: Tao, Zeng, et al.
Pubblicazione: (2025)
di: Tao, Zeng, et al.
Pubblicazione: (2025)
ConsID-Gen: View-Consistent and Identity-Preserving Image-to-Video Generation
di: Wu, Mingyang, et al.
Pubblicazione: (2026)
di: Wu, Mingyang, et al.
Pubblicazione: (2026)
IMAGHarmony: Controllable Image Editing with Consistent Object Quantity and Layout
di: Shen, Fei, et al.
Pubblicazione: (2025)
di: Shen, Fei, et al.
Pubblicazione: (2025)
UnCLe: Benchmarking Unsupervised Continual Learning for Depth Completion
di: Chen, Xien, et al.
Pubblicazione: (2024)
di: Chen, Xien, et al.
Pubblicazione: (2024)
Towards Geometric and Textural Consistency 3D Scene Generation via Single Image-guided Model Generation and Layout Optimization
di: Tang, Xiang, et al.
Pubblicazione: (2025)
di: Tang, Xiang, et al.
Pubblicazione: (2025)
Bayesian Diffusion Models for 3D Shape Reconstruction
di: Xu, Haiyang, et al.
Pubblicazione: (2024)
di: Xu, Haiyang, et al.
Pubblicazione: (2024)
DepR: Depth Guided Single-view Scene Reconstruction with Instance-level Diffusion
di: Zhao, Qingcheng, et al.
Pubblicazione: (2025)
di: Zhao, Qingcheng, et al.
Pubblicazione: (2025)
SpotActor: Training-Free Layout-Controlled Consistent Image Generation
di: Wang, Jiahao, et al.
Pubblicazione: (2024)
di: Wang, Jiahao, et al.
Pubblicazione: (2024)
Consistent Image Layout Editing with Diffusion Models
di: Xia, Tao, et al.
Pubblicazione: (2025)
di: Xia, Tao, et al.
Pubblicazione: (2025)
CityGen: Infinite and Controllable City Layout Generation
di: Deng, Jie, et al.
Pubblicazione: (2023)
di: Deng, Jie, et al.
Pubblicazione: (2023)
LLMs as Layout Designers: Enhanced Spatial Reasoning for Content-Aware Layout Generation
di: Li, Sha, et al.
Pubblicazione: (2025)
di: Li, Sha, et al.
Pubblicazione: (2025)
Restoration by Generation with Constrained Priors
di: Ding, Zheng, et al.
Pubblicazione: (2023)
di: Ding, Zheng, et al.
Pubblicazione: (2023)
SPGen: Spherical Projection as Consistent and Flexible Representation for Single Image 3D Shape Generation
di: Zhang, Jingdong, et al.
Pubblicazione: (2025)
di: Zhang, Jingdong, et al.
Pubblicazione: (2025)
SheetDesigner: MLLM-Powered Spreadsheet Layout Generation with Rule-Based and Vision-Based Reflection
di: Chen, Qin, et al.
Pubblicazione: (2025)
di: Chen, Qin, et al.
Pubblicazione: (2025)
Non-autoregressive Sequence-to-Sequence Vision-Language Models
di: Shi, Kunyu, et al.
Pubblicazione: (2024)
di: Shi, Kunyu, et al.
Pubblicazione: (2024)
Self-Rewarding Large Vision-Language Models for Optimizing Prompts in Text-to-Image Generation
di: Yang, Hongji, et al.
Pubblicazione: (2025)
di: Yang, Hongji, et al.
Pubblicazione: (2025)
ToLL: Topological Layout Learning with Asymmetric Cross-View Structural Distillation for 3D Scene Graph Generation Pretraining
di: Huang, Yucheng, et al.
Pubblicazione: (2026)
di: Huang, Yucheng, et al.
Pubblicazione: (2026)
UnCLe: Towards Scalable Dynamic Causal Discovery in Non-linear Temporal Systems
di: Bi, Tingzhu, et al.
Pubblicazione: (2025)
di: Bi, Tingzhu, et al.
Pubblicazione: (2025)
HiCoGen: Hierarchical Compositional Text-to-Image Generation in Diffusion Models via Reinforcement Learning
di: Yang, Hongji, et al.
Pubblicazione: (2025)
di: Yang, Hongji, et al.
Pubblicazione: (2025)
LayoutDiffusion: Controllable Diffusion Model for Layout-to-image Generation
di: Zheng, Guangcong, et al.
Pubblicazione: (2023)
di: Zheng, Guangcong, et al.
Pubblicazione: (2023)
ConsistCompose: Unified Multimodal Layout Control for Image Composition
di: Shi, Xuanke, et al.
Pubblicazione: (2025)
di: Shi, Xuanke, et al.
Pubblicazione: (2025)
ViTGAN: Training GANs with Vision Transformers
di: Lee, Kwonjoon, et al.
Pubblicazione: (2021)
di: Lee, Kwonjoon, et al.
Pubblicazione: (2021)
Deep-Learning-Based Pre-Layout Parasitic Capacitance Prediction on SRAM Designs
di: Shen, Shan, et al.
Pubblicazione: (2025)
di: Shen, Shan, et al.
Pubblicazione: (2025)
Visual In-Context Learning for Large Vision-Language Models
di: Zhou, Yucheng, et al.
Pubblicazione: (2024)
di: Zhou, Yucheng, et al.
Pubblicazione: (2024)
SemLayer: Semantic-aware Generative Segmentation and Layer Construction for Abstract Icons
di: Xu, Haiyang, et al.
Pubblicazione: (2026)
di: Xu, Haiyang, et al.
Pubblicazione: (2026)
Optimization of the Woodcock Particle Tracking Method Using Neural Network
di: Zhang, Bingnan
Pubblicazione: (2025)
di: Zhang, Bingnan
Pubblicazione: (2025)
CoGen: 3D Consistent Video Generation via Adaptive Conditioning for Autonomous Driving
di: Ji, Yishen, et al.
Pubblicazione: (2025)
di: Ji, Yishen, et al.
Pubblicazione: (2025)
Documenti analoghi
-
OverLayBench: A Benchmark for Layout-to-Image Generation with Dense Overlaps
di: Li, Bingnan, et al.
Pubblicazione: (2025) -
YOLO-Count: Differentiable Object Counting for Text-to-Image Generation
di: Zeng, Guanning, et al.
Pubblicazione: (2025) -
EchoGen: Cycle-Consistent Learning for Unified Layout-Image Generation and Understanding
di: Zou, Kai, et al.
Pubblicazione: (2026) -
Lay-Your-Scene: Natural Scene Layout Generation with Diffusion Transformers
di: Srivastava, Divyansh, et al.
Pubblicazione: (2025) -
CVP: Central-Peripheral Vision-Inspired Multimodal Model for Spatial Reasoning
di: Chen, Zeyuan, et al.
Pubblicazione: (2025)