LayoutAgent: A Vision-Language Agent Guided Compositional Diffusion for Spatial Layout Planning
Fuente:
arXiv
Guardado en:
| Autores principales: | Fan, Zezhong, Li, Xiaohan, Ma, Luyi, Zhao, Kai, Peng, Liang, Biswas, Topojoy, Korpeoglu, Evren, Nag, Kaushiki, Achan, Kannan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Segment and Matte Anything in a Unified Model
por: Fan, Zezhong, et al.
Publicado: (2026)
por: Fan, Zezhong, et al.
Publicado: (2026)
CAL-RAG: Retrieval-Augmented Multi-Agent Generation for Content-Aware Layout Design
por: Forouzandehmehr, Najmeh, et al.
Publicado: (2025)
por: Forouzandehmehr, Najmeh, et al.
Publicado: (2025)
Prompt Optimizer of Text-to-Image Diffusion Models for Abstract Concept Understanding
por: Fan, Zezhong, et al.
Publicado: (2024)
por: Fan, Zezhong, et al.
Publicado: (2024)
LLM-Ensemble: Optimal Large Language Model Ensemble Method for E-commerce Product Attribute Value Extraction
por: Fang, Chenhao, et al.
Publicado: (2024)
por: Fang, Chenhao, et al.
Publicado: (2024)
Spatial Reasoning in Foundation Models: Benchmarking Object-Centric Spatial Understanding
por: Mirjalili, Vahid, et al.
Publicado: (2025)
por: Mirjalili, Vahid, et al.
Publicado: (2025)
CASE: Cadence-Aware Set Encoding for Large-Scale Next Basket Repurchase Recommendation
por: Cao, Yanan, et al.
Publicado: (2026)
por: Cao, Yanan, et al.
Publicado: (2026)
CRAB: Codebook Rebalancing for Bias Mitigation in Generative Recommendation
por: Fan, Zezhong, et al.
Publicado: (2026)
por: Fan, Zezhong, et al.
Publicado: (2026)
Triple Modality Fusion: Aligning Visual, Textual, and Graph Data with Large Language Models for Multi-Behavior Recommendations
por: Ma, Luyi, et al.
Publicado: (2024)
por: Ma, Luyi, et al.
Publicado: (2024)
VL-CLIP: Enhancing Multimodal Recommendations via Visual Grounding and LLM-Augmented CLIP Embeddings
por: Giahi, Ramin, et al.
Publicado: (2025)
por: Giahi, Ramin, et al.
Publicado: (2025)
Decoding Style: Efficient Fine-Tuning of LLMs for Image-Guided Outfit Recommendation with Preference
por: Forouzandehmehr, Najmeh, et al.
Publicado: (2024)
por: Forouzandehmehr, Najmeh, et al.
Publicado: (2024)
Is More Context Always Better? Examining LLM Reasoning Capability for Time Interval Prediction
por: Cao, Yanan, et al.
Publicado: (2026)
por: Cao, Yanan, et al.
Publicado: (2026)
Character-based Outfit Generation with Vision-augmented Style Extraction via LLMs
por: Forouzandehmehr, Najmeh, et al.
Publicado: (2024)
por: Forouzandehmehr, Najmeh, et al.
Publicado: (2024)
CARTS: Collaborative Agents for Recommendation Textual Summarization
por: Chen, Jiao, et al.
Publicado: (2025)
por: Chen, Jiao, et al.
Publicado: (2025)
Leveraging User-Generated Reviews for Recommender Systems with Dynamic Headers
por: Vashishtha, Shanu, et al.
Publicado: (2024)
por: Vashishtha, Shanu, et al.
Publicado: (2024)
S2SRec2: Set-to-Set Recommendation for Basket Completion with Recipe
por: Cao, Yanan, et al.
Publicado: (2025)
por: Cao, Yanan, et al.
Publicado: (2025)
LLM-HYPER: Generative CTR Modeling for Cold-Start Ad Personalization via LLM-Based Hypernetworks
por: Ma, Luyi, et al.
Publicado: (2026)
por: Ma, Luyi, et al.
Publicado: (2026)
No-Human in the Loop: Agentic Evaluation at Scale for Recommendation
por: Zhang, Tao, et al.
Publicado: (2025)
por: Zhang, Tao, et al.
Publicado: (2025)
Improving Sequential Recommender Systems with Online and In-store User Behavior
por: Ma, Luyi, et al.
Publicado: (2024)
por: Ma, Luyi, et al.
Publicado: (2024)
GRACE: Generative Recommendation via Journey-Aware Sparse Attention on Chain-of-Thought Tokenization
por: Ma, Luyi, et al.
Publicado: (2025)
por: Ma, Luyi, et al.
Publicado: (2025)
Chaining text-to-image and large language model: A novel approach for generating personalized e-commerce banners
por: Vashishtha, Shanu, et al.
Publicado: (2024)
por: Vashishtha, Shanu, et al.
Publicado: (2024)
Grocery to General Merchandise: A Cross-Pollination Recommender using LLMs and Real-Time Cart Context
por: Kekuda, Akshay, et al.
Publicado: (2025)
por: Kekuda, Akshay, et al.
Publicado: (2025)
Spatial Diffusion for Cell Layout Generation
por: Li, Chen, et al.
Publicado: (2024)
por: Li, Chen, et al.
Publicado: (2024)
Latent Customer Segmentation and Value-Based Recommendation Leveraging a Two-Stage Model with Missing Labels
por: Gopalakrishnan, Keerthi, et al.
Publicado: (2026)
por: Gopalakrishnan, Keerthi, et al.
Publicado: (2026)
Layout2Scene: 3D Semantic Layout Guided Scene Generation via Geometry and Appearance Diffusion Priors
por: Chen, Minglin, et al.
Publicado: (2025)
por: Chen, Minglin, et al.
Publicado: (2025)
LayoutDiffusion: Controllable Diffusion Model for Layout-to-image Generation
por: Zheng, Guangcong, et al.
Publicado: (2023)
por: Zheng, Guangcong, et al.
Publicado: (2023)
Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation
por: Brioschi, Riccardo, et al.
Publicado: (2025)
por: Brioschi, Riccardo, et al.
Publicado: (2025)
CogniPlan: Uncertainty-Guided Path Planning with Conditional Generative Layout Prediction
por: Wang, Yizhuo, et al.
Publicado: (2025)
por: Wang, Yizhuo, et al.
Publicado: (2025)
Personalized Product Search Ranking: A Multi-Task Learning Approach with Tabular and Non-Tabular Data
por: Morishetti, Lalitesh, et al.
Publicado: (2025)
por: Morishetti, Lalitesh, et al.
Publicado: (2025)
LAMIC: Layout-Aware Multi-Image Composition via Scalability of Multimodal Diffusion Transformer
por: Chen, Yuzhuo, et al.
Publicado: (2025)
por: Chen, Yuzhuo, et al.
Publicado: (2025)
Layout Generation Agents with Large Language Models
por: Sasazawa, Yuichi, et al.
Publicado: (2024)
por: Sasazawa, Yuichi, et al.
Publicado: (2024)
Griffin: Generative Reference and Layout Guided Image Composition
por: Mikaeili, Aryan, et al.
Publicado: (2025)
por: Mikaeili, Aryan, et al.
Publicado: (2025)
Campaign-2-PT-RAG: LLM-Guided Semantic Product Type Attribution for Scalable Campaign Ranking
por: Che, Yiming, et al.
Publicado: (2026)
por: Che, Yiming, et al.
Publicado: (2026)
Layout Stroke Imitation: A Layout Guided Handwriting Stroke Generation for Style Imitation with Diffusion Model
por: Hanif, Sidra, et al.
Publicado: (2025)
por: Hanif, Sidra, et al.
Publicado: (2025)
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
por: Sun, Fan-Yun, et al.
Publicado: (2024)
por: Sun, Fan-Yun, et al.
Publicado: (2024)
To See or To Read: User Behavior Reasoning in Multimodal LLMs
por: Dong, Tianning, et al.
Publicado: (2025)
por: Dong, Tianning, et al.
Publicado: (2025)
STAY Diffusion: Styled Layout Diffusion Model for Diverse Layout-to-Image Generation
por: Wang, Ruyu, et al.
Publicado: (2025)
por: Wang, Ruyu, et al.
Publicado: (2025)
LocRef-Diffusion:Tuning-Free Layout and Appearance-Guided Generation
por: Deng, Fan, et al.
Publicado: (2024)
por: Deng, Fan, et al.
Publicado: (2024)
LLMs as Layout Designers: Enhanced Spatial Reasoning for Content-Aware Layout Generation
por: Li, Sha, et al.
Publicado: (2025)
por: Li, Sha, et al.
Publicado: (2025)
DogLayout: Denoising Diffusion GAN for Discrete and Continuous Layout Generation
por: Gan, Zhaoxing, et al.
Publicado: (2024)
por: Gan, Zhaoxing, et al.
Publicado: (2024)
Layout-Corrector: Alleviating Layout Sticking Phenomenon in Discrete Diffusion Model
por: Iwai, Shoma, et al.
Publicado: (2024)
por: Iwai, Shoma, et al.
Publicado: (2024)
Ejemplares similares
-
Segment and Matte Anything in a Unified Model
por: Fan, Zezhong, et al.
Publicado: (2026) -
CAL-RAG: Retrieval-Augmented Multi-Agent Generation for Content-Aware Layout Design
por: Forouzandehmehr, Najmeh, et al.
Publicado: (2025) -
Prompt Optimizer of Text-to-Image Diffusion Models for Abstract Concept Understanding
por: Fan, Zezhong, et al.
Publicado: (2024) -
LLM-Ensemble: Optimal Large Language Model Ensemble Method for E-commerce Product Attribute Value Extraction
por: Fang, Chenhao, et al.
Publicado: (2024) -
Spatial Reasoning in Foundation Models: Benchmarking Object-Centric Spatial Understanding
por: Mirjalili, Vahid, et al.
Publicado: (2025)