VASCAR: Content-Aware Layout Generation via Visual-Aware Self-Correction
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Jiahao, Yoshihashi, Ryota, Kitada, Shunsuke, Osanai, Atsuki, Nakashima, Yuta |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Layout-Corrector: Alleviating Layout Sticking Phenomenon in Discrete Diffusion Model
by: Iwai, Shoma, et al.
Published: (2024)
by: Iwai, Shoma, et al.
Published: (2024)
SCAdapter: Content-Style Disentanglement for Diffusion Style Transfer
by: Trinh, Luan Thanh, et al.
Published: (2025)
by: Trinh, Luan Thanh, et al.
Published: (2025)
Retrieval-Augmented Layout Transformer for Content-Aware Layout Generation
by: Horita, Daichi, et al.
Published: (2023)
by: Horita, Daichi, et al.
Published: (2023)
ReLayout: Towards Real-World Document Understanding via Layout-enhanced Pre-training
by: Jiang, Zhouqiang, et al.
Published: (2024)
by: Jiang, Zhouqiang, et al.
Published: (2024)
What-Where Transformer: A Slot-Centric Visual Backbone for Concurrent Representation and Localization
by: Yoshihashi, Ryota, et al.
Published: (2026)
by: Yoshihashi, Ryota, et al.
Published: (2026)
Improving Prediction Performance and Model Interpretability through Attention Mechanisms from Basic and Applied Research Perspectives
by: Kitada, Shunsuke
Published: (2023)
by: Kitada, Shunsuke
Published: (2023)
EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors
by: Miyazato, Ryuhei, et al.
Published: (2026)
by: Miyazato, Ryuhei, et al.
Published: (2026)
PANICL: Mitigating Over-Reliance on Single Prompt in Visual In-Context Learning
by: Zhang, Jiahao, et al.
Published: (2025)
by: Zhang, Jiahao, et al.
Published: (2025)
Constant Rate Scheduling: A General Framework for Optimizing Diffusion Noise Schedule via Distributional Change
by: Okada, Shuntaro, et al.
Published: (2024)
by: Okada, Shuntaro, et al.
Published: (2024)
PosterLlama: Bridging Design Ability of Langauge Model to Contents-Aware Layout Generation
by: Seol, Jaejung, et al.
Published: (2024)
by: Seol, Jaejung, et al.
Published: (2024)
UniLayDiff: A Unified Diffusion Transformer for Content-Aware Layout Generation
by: Liu, Zeyang, et al.
Published: (2025)
by: Liu, Zeyang, et al.
Published: (2025)
SEGA: A Stepwise Evolution Paradigm for Content-Aware Layout Generation with Design Prior
by: Wang, Haoran, et al.
Published: (2025)
by: Wang, Haoran, et al.
Published: (2025)
Explainable Image Recognition via Enhanced Slot-attention Based Classifier
by: Wang, Bowen, et al.
Published: (2024)
by: Wang, Bowen, et al.
Published: (2024)
E-InMeMo: Enhanced Prompting for Visual In-Context Learning
by: Zhang, Jiahao, et al.
Published: (2025)
by: Zhang, Jiahao, et al.
Published: (2025)
InstanceAssemble: Layout-Aware Image Generation via Instance Assembling Attention
by: Xiang, Qiang, et al.
Published: (2025)
by: Xiang, Qiang, et al.
Published: (2025)
Exploring Limits of Diffusion-Synthetic Training with Weakly Supervised Semantic Segmentation
by: Yoshihashi, Ryota, et al.
Published: (2023)
by: Yoshihashi, Ryota, et al.
Published: (2023)
Teacher-Guided Routing for Sparse Vision Mixture-of-Experts
by: Kada, Masahiro, et al.
Published: (2026)
by: Kada, Masahiro, et al.
Published: (2026)
Enhancing Retinal Vessel Segmentation Generalization via Layout-Aware Generative Modelling
by: Fhima, Jonathan, et al.
Published: (2025)
by: Fhima, Jonathan, et al.
Published: (2025)
Relation-Aware Diffusion Model for Controllable Poster Layout Generation
by: Li, Fengheng, et al.
Published: (2023)
by: Li, Fengheng, et al.
Published: (2023)
Generating Synthetic Invoices via Layout-Preserving Content Replacement
by: V, Bevin, et al.
Published: (2025)
by: V, Bevin, et al.
Published: (2025)
ReLayout: Versatile and Structure-Preserving Design Layout Editing via Relation-Aware Design Reconstruction
by: Lin, Jiawei, et al.
Published: (2026)
by: Lin, Jiawei, et al.
Published: (2026)
LayoutDiT: Exploring Content-Graphic Balance in Layout Generation with Diffusion Transformer
by: Li, Yu, et al.
Published: (2024)
by: Li, Yu, et al.
Published: (2024)
LayoutRAG: Retrieval-Augmented Model for Content-agnostic Conditional Layout Generation
by: Wu, Yuxuan, et al.
Published: (2025)
by: Wu, Yuxuan, et al.
Published: (2025)
Content-Aware Ad Banner Layout Generation with Two-Stage Chain-of-Thought in Vision Language Models
by: Yoshitake, Kei, et al.
Published: (2025)
by: Yoshitake, Kei, et al.
Published: (2025)
Content-Aware Preserving Image Generation
by: Le, Giang H., et al.
Published: (2024)
by: Le, Giang H., et al.
Published: (2024)
LAMIC: Layout-Aware Multi-Image Composition via Scalability of Multimodal Diffusion Transformer
by: Chen, Yuzhuo, et al.
Published: (2025)
by: Chen, Yuzhuo, et al.
Published: (2025)
TableSeq: Unified Generation of Structure, Content, and Layout
by: Hamdi, Laziz, et al.
Published: (2026)
by: Hamdi, Laziz, et al.
Published: (2026)
PARL: Position-Aware Relation Learning Network for Document Layout Analysis
by: Liu, Fuyuan, et al.
Published: (2026)
by: Liu, Fuyuan, et al.
Published: (2026)
Efficient 3D Content Reconstruction and Generation
by: Li, Jiahao
Published: (2026)
by: Li, Jiahao
Published: (2026)
SciGA: A Comprehensive Dataset for Designing Graphical Abstracts in Academic Papers
by: Kawada, Takuro, et al.
Published: (2025)
by: Kawada, Takuro, et al.
Published: (2025)
SDesc3D: Towards Layout-Aware 3D Indoor Scene Generation from Short Descriptions
by: Feng, Jie, et al.
Published: (2026)
by: Feng, Jie, et al.
Published: (2026)
LAPDoc: Layout-Aware Prompting for Documents
by: Lamott, Marcel, et al.
Published: (2024)
by: Lamott, Marcel, et al.
Published: (2024)
MV-SAM3D: Adaptive Multi-View Fusion for Layout-Aware 3D Generation
by: Li, Baicheng, et al.
Published: (2026)
by: Li, Baicheng, et al.
Published: (2026)
Measure Twice, Cut Once: A Semantic-Oriented Approach to Video Temporal Localization with Video LLMs
by: Pang, Zongshang, et al.
Published: (2025)
by: Pang, Zongshang, et al.
Published: (2025)
Stable Diffusion Exposed: Gender Bias from Prompt to Image
by: Wu, Yankun, et al.
Published: (2023)
by: Wu, Yankun, et al.
Published: (2023)
EMMA: Concept Erasure Benchmark with Comprehensive Semantic Metrics and Diverse Categories
by: Wei, Lu, et al.
Published: (2025)
by: Wei, Lu, et al.
Published: (2025)
MotionAdapter: Video Motion Transfer via Content-Aware Attention Customization
by: Zhang, Zhexin, et al.
Published: (2026)
by: Zhang, Zhexin, et al.
Published: (2026)
Content-Aware Mamba for Learned Image Compression
by: Chen, Yunuo, et al.
Published: (2025)
by: Chen, Yunuo, et al.
Published: (2025)
Pool-Select-Refine: Allocation-Aware Generative Dataset Distillation with Soft-Label-Guided Latent Refinement
by: Li, Wenmin, et al.
Published: (2026)
by: Li, Wenmin, et al.
Published: (2026)
Reward-Aware Trajectory Shaping for Few-step Visual Generation
by: Li, Rui, et al.
Published: (2026)
by: Li, Rui, et al.
Published: (2026)
Similar Items
-
Layout-Corrector: Alleviating Layout Sticking Phenomenon in Discrete Diffusion Model
by: Iwai, Shoma, et al.
Published: (2024) -
SCAdapter: Content-Style Disentanglement for Diffusion Style Transfer
by: Trinh, Luan Thanh, et al.
Published: (2025) -
Retrieval-Augmented Layout Transformer for Content-Aware Layout Generation
by: Horita, Daichi, et al.
Published: (2023) -
ReLayout: Towards Real-World Document Understanding via Layout-enhanced Pre-training
by: Jiang, Zhouqiang, et al.
Published: (2024) -
What-Where Transformer: A Slot-Centric Visual Backbone for Concurrent Representation and Localization
by: Yoshihashi, Ryota, et al.
Published: (2026)