From Image Generation to Infrastructure Design: a Multi-agent Pipeline for Street Design Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Chenguang, Yan, Xiang, Dai, Yilong, Wang, Ziyi, Xu, Susu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
StreetDesignAI: Broadening Designer Perspectives Through Multi-Persona Evaluation of Cycling Infrastructure
by: Wang, Ziyi, et al.
Published: (2026)
by: Wang, Ziyi, et al.
Published: (2026)
AIDEN: Design and Pilot Study of an AI Assistant for the Visually Impaired
by: Marquez-Carpintero, Luis, et al.
Published: (2025)
by: Marquez-Carpintero, Luis, et al.
Published: (2025)
GenColor: Generative Color-Concept Association in Visual Design
by: Hou, Yihan, et al.
Published: (2025)
by: Hou, Yihan, et al.
Published: (2025)
Design First, Code Later: Aesthetically Pleasing Template-Free Slides Generation
by: Cui, Zhiyao, et al.
Published: (2026)
by: Cui, Zhiyao, et al.
Published: (2026)
Towards Geographic Inclusion in the Evaluation of Text-to-Image Models
by: Hall, Melissa, et al.
Published: (2024)
by: Hall, Melissa, et al.
Published: (2024)
Using Text-to-Image Generation for Architectural Design Ideation
by: Paananen, Ville, et al.
Published: (2023)
by: Paananen, Ville, et al.
Published: (2023)
Intuitions of Machine Learning Researchers about Transfer Learning for Medical Image Classification
by: Lu, Yucheng, et al.
Published: (2025)
by: Lu, Yucheng, et al.
Published: (2025)
Persona-aware and Explainable Bikeability Assessment: A Vision-Language Model Approach
by: Dai, Yilong, et al.
Published: (2026)
by: Dai, Yilong, et al.
Published: (2026)
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
by: Yang, Boyin, et al.
Published: (2025)
by: Yang, Boyin, et al.
Published: (2025)
Steering Generative Models for Accessibility: EasyRead Image Generation
by: Dickenmann, Nicolas, et al.
Published: (2026)
by: Dickenmann, Nicolas, et al.
Published: (2026)
Designing Multi-Robot Ground Video Sensemaking with Public Safety Professionals
by: Zhou, Puqi, et al.
Published: (2026)
by: Zhou, Puqi, et al.
Published: (2026)
Multi-scale structural complexity as a quantitative measure of visual complexity
by: Kravchenko, Anna, et al.
Published: (2024)
by: Kravchenko, Anna, et al.
Published: (2024)
DiffGaze: A Diffusion Model for Continuous Gaze Sequence Generation on 360° Images
by: Jiao, Chuhan, et al.
Published: (2024)
by: Jiao, Chuhan, et al.
Published: (2024)
Investigating Disability Representations in Text-to-Image Models
by: Tian, Yang, et al.
Published: (2026)
by: Tian, Yang, et al.
Published: (2026)
Semantic and Expressive Variation in Image Captions Across Languages
by: Ye, Andre, et al.
Published: (2023)
by: Ye, Andre, et al.
Published: (2023)
POET: Supporting Prompting Creativity and Personalization with Automated Expansion of Text-to-Image Generation
by: Han, Evans Xu, et al.
Published: (2025)
by: Han, Evans Xu, et al.
Published: (2025)
BLK-Assist: A Methodological Framework for Artist-Led Co-Creation with Generative AI Models
by: Grimes, Daniel, et al.
Published: (2026)
by: Grimes, Daniel, et al.
Published: (2026)
A Call to Arms: AI Should be Critical for Social Media Analysis of Conflict Zones
by: Abedin, Afia, et al.
Published: (2023)
by: Abedin, Afia, et al.
Published: (2023)
Improved Digital Therapy for Developmental Pediatrics Using Domain-Specific Artificial Intelligence: Machine Learning Study
by: Washington, Peter, et al.
Published: (2020)
by: Washington, Peter, et al.
Published: (2020)
A Comparison of Human and Machine Learning Errors in Face Recognition
by: Estévez-Almenzar, Marina, et al.
Published: (2025)
by: Estévez-Almenzar, Marina, et al.
Published: (2025)
Classification of the lunar surface pattern by AI architectures: Does AI see a rabbit in the Moon?
by: Shoji, Daigo
Published: (2023)
by: Shoji, Daigo
Published: (2023)
Beyond Questionnaires: Video Analysis for Social Anxiety Detection
by: Sahu, Nilesh Kumar, et al.
Published: (2024)
by: Sahu, Nilesh Kumar, et al.
Published: (2024)
DepMamba: Progressive Fusion Mamba for Multimodal Depression Detection
by: Ye, Jiaxin, et al.
Published: (2024)
by: Ye, Jiaxin, et al.
Published: (2024)
The Cadaver in the Machine: The Social Practices of Measurement and Validation in Motion Capture Technology
by: Harvey, Emma, et al.
Published: (2024)
by: Harvey, Emma, et al.
Published: (2024)
CAF-Mamba: Mamba-Based Cross-Modal Adaptive Attention Fusion for Multimodal Depression Detection
by: Zhou, Bowen, et al.
Published: (2026)
by: Zhou, Bowen, et al.
Published: (2026)
DesignPref: Capturing Personal Preferences in Visual Design Generation
by: Peng, Yi-Hao, et al.
Published: (2025)
by: Peng, Yi-Hao, et al.
Published: (2025)
Shifts in Doctors' Eye Movements Between Real and AI-Generated Medical Images
by: Wong, David C, et al.
Published: (2025)
by: Wong, David C, et al.
Published: (2025)
DIG In: Evaluating Disparities in Image Generations with Indicators for Geographic Diversity
by: Hall, Melissa, et al.
Published: (2023)
by: Hall, Melissa, et al.
Published: (2023)
Generating Robot Constitutions & Benchmarks for Semantic Safety
by: Sermanet, Pierre, et al.
Published: (2025)
by: Sermanet, Pierre, et al.
Published: (2025)
Human-Centred Evaluation of Text-to-Image Generation Models for Self-expression of Mental Distress: A Dataset Based on GPT-4o
by: He, Sui, et al.
Published: (2025)
by: He, Sui, et al.
Published: (2025)
Assessing Intersectional Bias in Representations of Pre-Trained Image Recognition Models
by: Krug, Valerie, et al.
Published: (2025)
by: Krug, Valerie, et al.
Published: (2025)
ViT-Explainer: An Interactive Walkthrough of the Vision Transformer Pipeline
by: Hernandez, Juan Manuel, et al.
Published: (2026)
by: Hernandez, Juan Manuel, et al.
Published: (2026)
Design and Evaluation of Camera-Centric Mobile Crowdsourcing Applications
by: Stylianou, Abby, et al.
Published: (2024)
by: Stylianou, Abby, et al.
Published: (2024)
SwipeGANSpace: Swipe-to-Compare Image Generation via Efficient Latent Space Exploration
by: Nakashima, Yuto, et al.
Published: (2024)
by: Nakashima, Yuto, et al.
Published: (2024)
Clinically Aware Synthetic Image Generation for Concept Coverage in Chest X-ray Models
by: Rafferty, Amy, et al.
Published: (2026)
by: Rafferty, Amy, et al.
Published: (2026)
Learning Multimodal Cues of Children's Uncertainty
by: Cheng, Qi, et al.
Published: (2024)
by: Cheng, Qi, et al.
Published: (2024)
SCHEMA for Gemini 3 Pro Image: A Structured Methodology for Controlled AI Image Generation on Google's Native Multimodal Model
by: Cazzaniga, Luca
Published: (2026)
by: Cazzaniga, Luca
Published: (2026)
Facial recognition technology and human raters can predict political orientation from images of expressionless faces even when controlling for demographics and self-presentation
by: Kosinski, Michal, et al.
Published: (2023)
by: Kosinski, Michal, et al.
Published: (2023)
Inkspire: Supporting Design Exploration with Generative AI through Analogical Sketching
by: Lin, David Chuan-En, et al.
Published: (2025)
by: Lin, David Chuan-En, et al.
Published: (2025)
Stratified Avatar Generation from Sparse Observations
by: Feng, Han, et al.
Published: (2024)
by: Feng, Han, et al.
Published: (2024)
Similar Items
-
StreetDesignAI: Broadening Designer Perspectives Through Multi-Persona Evaluation of Cycling Infrastructure
by: Wang, Ziyi, et al.
Published: (2026) -
AIDEN: Design and Pilot Study of an AI Assistant for the Visually Impaired
by: Marquez-Carpintero, Luis, et al.
Published: (2025) -
GenColor: Generative Color-Concept Association in Visual Design
by: Hou, Yihan, et al.
Published: (2025) -
Design First, Code Later: Aesthetically Pleasing Template-Free Slides Generation
by: Cui, Zhiyao, et al.
Published: (2026) -
Towards Geographic Inclusion in the Evaluation of Text-to-Image Models
by: Hall, Melissa, et al.
Published: (2024)