Benchmarking Layout-Guided Diffusion Models through Unified Semantic-Spatial Evaluation in Closed and Open Settings
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Parolari, Luca, Faccioli, Nicla, Ballan, Lamberto |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
7Bench: a Comprehensive Benchmark for Layout-guided Text-to-image Models
von: Izzo, Elena, et al.
Veröffentlicht: (2025)
von: Izzo, Elena, et al.
Veröffentlicht: (2025)
Harlequin: Color-driven Generation of Synthetic Data for Referring Expression Comprehension
von: Parolari, Luca, et al.
Veröffentlicht: (2024)
von: Parolari, Luca, et al.
Veröffentlicht: (2024)
Towards Polyp Counting In Full-Procedure Colonoscopy Videos
von: Parolari, Luca, et al.
Veröffentlicht: (2025)
von: Parolari, Luca, et al.
Veröffentlicht: (2025)
Temporally-Aware Supervised Contrastive Learning for Polyp Counting in Colonoscopy
von: Parolari, Luca, et al.
Veröffentlicht: (2025)
von: Parolari, Luca, et al.
Veröffentlicht: (2025)
Contrastive Learning under Noisy Temporal Self-Supervision for Colonoscopy Videos
von: Parolari, Luca, et al.
Veröffentlicht: (2026)
von: Parolari, Luca, et al.
Veröffentlicht: (2026)
Multiview Progress Prediction of Robot Activities
von: Zoppellari, Elena, et al.
Veröffentlicht: (2026)
von: Zoppellari, Elena, et al.
Veröffentlicht: (2026)
PersONAL: Towards a Comprehensive Benchmark for Personalized Embodied Agents
von: Ziliotto, Filippo, et al.
Veröffentlicht: (2025)
von: Ziliotto, Filippo, et al.
Veröffentlicht: (2025)
You Only Landmark Once: Lightweight U-Net Face Super Resolution with YOLO-World Landmark Heatmaps
von: Carraro, Riccardo, et al.
Veröffentlicht: (2026)
von: Carraro, Riccardo, et al.
Veröffentlicht: (2026)
MLFM: Multi-Layered Feature Maps for Richer Language Understanding in Zero-Shot Semantic Navigation
von: Raychaudhuri, Sonia, et al.
Veröffentlicht: (2025)
von: Raychaudhuri, Sonia, et al.
Veröffentlicht: (2025)
Assessing the Visual Enumeration Abilities of Specialized Counting Architectures and Vision-Language Models
von: Hou, Kuinan, et al.
Veröffentlicht: (2025)
von: Hou, Kuinan, et al.
Veröffentlicht: (2025)
SSMG: Spatial-Semantic Map Guided Diffusion Model for Free-form Layout-to-Image Generation
von: Jia, Chengyou, et al.
Veröffentlicht: (2023)
von: Jia, Chengyou, et al.
Veröffentlicht: (2023)
Following the Human Thread in Social Navigation
von: Scofano, Luca, et al.
Veröffentlicht: (2024)
von: Scofano, Luca, et al.
Veröffentlicht: (2024)
Distilling Knowledge for Short-to-Long Term Trajectory Prediction
von: Das, Sourav, et al.
Veröffentlicht: (2023)
von: Das, Sourav, et al.
Veröffentlicht: (2023)
Spatial Diffusion for Cell Layout Generation
von: Li, Chen, et al.
Veröffentlicht: (2024)
von: Li, Chen, et al.
Veröffentlicht: (2024)
Layout2Scene: 3D Semantic Layout Guided Scene Generation via Geometry and Appearance Diffusion Priors
von: Chen, Minglin, et al.
Veröffentlicht: (2025)
von: Chen, Minglin, et al.
Veröffentlicht: (2025)
Open-Set Biometrics: Beyond Good Closed-Set Models
von: Su, Yiyang, et al.
Veröffentlicht: (2024)
von: Su, Yiyang, et al.
Veröffentlicht: (2024)
SemLayoutDiff: Semantic Layout Generation with Diffusion Model for Indoor Scene Synthesis
von: Sun, Xiaohao, et al.
Veröffentlicht: (2025)
von: Sun, Xiaohao, et al.
Veröffentlicht: (2025)
LayoutDiffusion: Controllable Diffusion Model for Layout-to-image Generation
von: Zheng, Guangcong, et al.
Veröffentlicht: (2023)
von: Zheng, Guangcong, et al.
Veröffentlicht: (2023)
Bayesian Fields: Task-driven Open-Set Semantic Gaussian Splatting
von: Maggio, Dominic, et al.
Veröffentlicht: (2025)
von: Maggio, Dominic, et al.
Veröffentlicht: (2025)
Layout Stroke Imitation: A Layout Guided Handwriting Stroke Generation for Style Imitation with Diffusion Model
von: Hanif, Sidra, et al.
Veröffentlicht: (2025)
von: Hanif, Sidra, et al.
Veröffentlicht: (2025)
Spatial-DISE: A Unified Benchmark for Evaluating Spatial Reasoning in Vision-Language Models
von: Huang, Xinmiao, et al.
Veröffentlicht: (2025)
von: Huang, Xinmiao, et al.
Veröffentlicht: (2025)
STAY Diffusion: Styled Layout Diffusion Model for Diverse Layout-to-Image Generation
von: Wang, Ruyu, et al.
Veröffentlicht: (2025)
von: Wang, Ruyu, et al.
Veröffentlicht: (2025)
LayoutAgent: A Vision-Language Agent Guided Compositional Diffusion for Spatial Layout Planning
von: Fan, Zezhong, et al.
Veröffentlicht: (2025)
von: Fan, Zezhong, et al.
Veröffentlicht: (2025)
Box It to Bind It: Unified Layout Control and Attribute Binding in T2I Diffusion Models
von: Taghipour, Ashkan, et al.
Veröffentlicht: (2024)
von: Taghipour, Ashkan, et al.
Veröffentlicht: (2024)
MUSE: Multi-Subject Unified Synthesis via Explicit Layout Semantic Expansion
von: Peng, Fei, et al.
Veröffentlicht: (2025)
von: Peng, Fei, et al.
Veröffentlicht: (2025)
Layout Control and Semantic Guidance with Attention Loss Backward for T2I Diffusion Model
von: Li, Guandong
Veröffentlicht: (2024)
von: Li, Guandong
Veröffentlicht: (2024)
CoSMo3D: Open-World Promptable 3D Semantic Part Segmentation through LLM-Guided Canonical Spatial Modeling
von: Jin, Li, et al.
Veröffentlicht: (2026)
von: Jin, Li, et al.
Veröffentlicht: (2026)
Setting-Matched and Semantics-Scaled Benchmarking of One-Step Generative Models Against Multistep Diffusion and Flow Models
von: Ravishankar, Advaith, et al.
Veröffentlicht: (2026)
von: Ravishankar, Advaith, et al.
Veröffentlicht: (2026)
Semantic Foam: Unifying Spatial and Semantic Scene Decomposition
von: Sharafeldin, Amr, et al.
Veröffentlicht: (2026)
von: Sharafeldin, Amr, et al.
Veröffentlicht: (2026)
Layout-Guided Controllable Pathology Image Generation with In-Context Diffusion Transformers
von: Shou, Yuntao, et al.
Veröffentlicht: (2026)
von: Shou, Yuntao, et al.
Veröffentlicht: (2026)
uLayout: Unified Room Layout Estimation for Perspective and Panoramic Images
von: Lee, Jonathan, et al.
Veröffentlicht: (2025)
von: Lee, Jonathan, et al.
Veröffentlicht: (2025)
POCI-Diff: Position Objects Consistently and Interactively with 3D-Layout Guided Diffusion
von: Rigo, Andrea, et al.
Veröffentlicht: (2026)
von: Rigo, Andrea, et al.
Veröffentlicht: (2026)
Consistent Image Layout Editing with Diffusion Models
von: Xia, Tao, et al.
Veröffentlicht: (2025)
von: Xia, Tao, et al.
Veröffentlicht: (2025)
Guiding Diffusion Models with Semantically Degraded Conditions
von: Han, Shilong, et al.
Veröffentlicht: (2026)
von: Han, Shilong, et al.
Veröffentlicht: (2026)
UniLayDiff: A Unified Diffusion Transformer for Content-Aware Layout Generation
von: Liu, Zeyang, et al.
Veröffentlicht: (2025)
von: Liu, Zeyang, et al.
Veröffentlicht: (2025)
LED Benchmark: Diagnosing Structural Layout Errors for Document Layout Analysis
von: Heo, Inbum, et al.
Veröffentlicht: (2025)
von: Heo, Inbum, et al.
Veröffentlicht: (2025)
Multi-Scale Diffusion: Enhancing Spatial Layout in High-Resolution Panoramic Image Generation
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2024)
Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models
von: Monon, Mashrafi, et al.
Veröffentlicht: (2026)
von: Monon, Mashrafi, et al.
Veröffentlicht: (2026)
Enhancing Image Layout Control with Loss-Guided Diffusion Models
von: Patel, Zakaria, et al.
Veröffentlicht: (2024)
von: Patel, Zakaria, et al.
Veröffentlicht: (2024)
Open-Set Domain Generalization through Spectral-Spatial Uncertainty Disentanglement for Hyperspectral Image Classification
von: Khoshbakht, Amirreza, et al.
Veröffentlicht: (2025)
von: Khoshbakht, Amirreza, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
7Bench: a Comprehensive Benchmark for Layout-guided Text-to-image Models
von: Izzo, Elena, et al.
Veröffentlicht: (2025) -
Harlequin: Color-driven Generation of Synthetic Data for Referring Expression Comprehension
von: Parolari, Luca, et al.
Veröffentlicht: (2024) -
Towards Polyp Counting In Full-Procedure Colonoscopy Videos
von: Parolari, Luca, et al.
Veröffentlicht: (2025) -
Temporally-Aware Supervised Contrastive Learning for Polyp Counting in Colonoscopy
von: Parolari, Luca, et al.
Veröffentlicht: (2025) -
Contrastive Learning under Noisy Temporal Self-Supervision for Colonoscopy Videos
von: Parolari, Luca, et al.
Veröffentlicht: (2026)