From Bird's-Eye to Street View: Crafting Diverse and Condition-Aligned Images with Latent Diffusion Model
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Xiaojie, Xu, Tianshuo, Ma, Fulong, Chen, Yingcong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bird Eye-View to Street-View: A Survey
by: Bajbaa, Khawlah, et al.
Published: (2024)
by: Bajbaa, Khawlah, et al.
Published: (2024)
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
by: Xu, Tianshuo, et al.
Published: (2024)
by: Xu, Tianshuo, et al.
Published: (2024)
Street-View Image Generation from a Bird's-Eye View Layout
by: Swerdlow, Alexander, et al.
Published: (2023)
by: Swerdlow, Alexander, et al.
Published: (2023)
MagicDrive: Street View Generation with Diverse 3D Geometry Control
by: Gao, Ruiyuan, et al.
Published: (2023)
by: Gao, Ruiyuan, et al.
Published: (2023)
SDGOCC: Semantic and Depth-Guided Bird's-Eye View Transformation for 3D Multimodal Occupancy Prediction
by: Duan, Zaipeng, et al.
Published: (2025)
by: Duan, Zaipeng, et al.
Published: (2025)
Improving Bird's Eye View Semantic Segmentation by Task Decomposition
by: Zhao, Tianhao, et al.
Published: (2024)
by: Zhao, Tianhao, et al.
Published: (2024)
DreamDrive: Generative 4D Scene Modeling from Street View Images
by: Mao, Jiageng, et al.
Published: (2024)
by: Mao, Jiageng, et al.
Published: (2024)
An Initial Study of Bird's-Eye View Generation for Autonomous Vehicles using Cross-View Transformers
by: Santos, Felipe Carlos dos, et al.
Published: (2025)
by: Santos, Felipe Carlos dos, et al.
Published: (2025)
CycleBEV: Regularizing View Transformation Networks via View Cycle Consistency for Bird's-Eye-View Semantic Segmentation
by: Hong, Jeongbin, et al.
Published: (2026)
by: Hong, Jeongbin, et al.
Published: (2026)
STANCE: Motion Coherent Video Generation Via Sparse-to-Dense Anchored Encoding
by: Chen, Zhifei, et al.
Published: (2025)
by: Chen, Zhifei, et al.
Published: (2025)
Aligning Diffusion Models with Noise-Conditioned Perception
by: Gambashidze, Alexander, et al.
Published: (2024)
by: Gambashidze, Alexander, et al.
Published: (2024)
Fuse Your Latents: Video Editing with Multi-source Latent Diffusion Models
by: Lu, Tianyi, et al.
Published: (2023)
by: Lu, Tianyi, et al.
Published: (2023)
PolarBEVDet: Exploring Polar Representation for Multi-View 3D Object Detection in Bird's-Eye-View
by: Yu, Zichen, et al.
Published: (2024)
by: Yu, Zichen, et al.
Published: (2024)
Training-free Dense-Aligned Diffusion Guidance for Modular Conditional Image Synthesis
by: Wang, Zixuan, et al.
Published: (2025)
by: Wang, Zixuan, et al.
Published: (2025)
Discovering Latent Graphs with GFlowNets for Diverse Conditional Image Generation
by: Trang, Bailey, et al.
Published: (2025)
by: Trang, Bailey, et al.
Published: (2025)
Post Fusion Bird's Eye View Feature Stabilization for Robust Multimodal 3D Detection
by: Dong, Trung Tien, et al.
Published: (2026)
by: Dong, Trung Tien, et al.
Published: (2026)
Taming Identity Consistency and Prompt Diversity in Diffusion Models via Latent Concatenation and Masked Conditional Flow Matching
by: Singhania, Aditi, et al.
Published: (2025)
by: Singhania, Aditi, et al.
Published: (2025)
Towards Full-scene Domain Generalization in Multi-agent Collaborative Bird's Eye View Segmentation for Connected and Autonomous Driving
by: Hu, Senkang, et al.
Published: (2023)
by: Hu, Senkang, et al.
Published: (2023)
Addressing Diverging Training Costs using BEVRestore for High-resolution Bird's Eye View Map Construction
by: Kim, Minsu, et al.
Published: (2024)
by: Kim, Minsu, et al.
Published: (2024)
BEVTrack: A Simple and Strong Baseline for 3D Single Object Tracking in Bird's-Eye View
by: Yang, Yuxiang, et al.
Published: (2023)
by: Yang, Yuxiang, et al.
Published: (2023)
REFNet++: Multi-Task Efficient Fusion of Camera and Radar Sensor Data in Bird's-Eye Polar View
by: Chandrasekaran, Kavin, et al.
Published: (2026)
by: Chandrasekaran, Kavin, et al.
Published: (2026)
Learning Content-Aware Multi-Modal Joint Input Pruning via Bird's-Eye-View Representation
by: Li, Yuxin, et al.
Published: (2024)
by: Li, Yuxin, et al.
Published: (2024)
MagicDrive3D: Controllable 3D Generation for Any-View Rendering in Street Scenes
by: Gao, Ruiyuan, et al.
Published: (2024)
by: Gao, Ruiyuan, et al.
Published: (2024)
A Mechanistic View on Video Generation as World Models: State and Dynamics
by: Wang, Luozhou, et al.
Published: (2026)
by: Wang, Luozhou, et al.
Published: (2026)
BEVLM: Distilling Semantic Knowledge from LLMs into Bird's-Eye View Representations
by: Monninger, Thomas, et al.
Published: (2026)
by: Monninger, Thomas, et al.
Published: (2026)
VQ-Map: Bird's-Eye-View Map Layout Estimation in Tokenized Discrete Space via Vector Quantization
by: Zhang, Yiwei, et al.
Published: (2024)
by: Zhang, Yiwei, et al.
Published: (2024)
Towards Efficient 3D Object Detection in Bird's-Eye-View Space for Autonomous Driving: A Convolutional-Only Approach
by: Li, Yuxin, et al.
Published: (2023)
by: Li, Yuxin, et al.
Published: (2023)
Learning Street View Representations with Spatiotemporal Contrast
by: Li, Yong, et al.
Published: (2025)
by: Li, Yong, et al.
Published: (2025)
Conditional Image Synthesis with Diffusion Models: A Survey
by: Zhan, Zheyuan, et al.
Published: (2024)
by: Zhan, Zheyuan, et al.
Published: (2024)
A Resource Efficient Fusion Network for Object Detection in Bird's-Eye View using Camera and Raw Radar Data
by: Chandrasekaran, Kavin, et al.
Published: (2024)
by: Chandrasekaran, Kavin, et al.
Published: (2024)
DocShaDiffusion: Diffusion Model in Latent Space for Document Image Shadow Removal
by: Liu, Wenjie, et al.
Published: (2025)
by: Liu, Wenjie, et al.
Published: (2025)
From Satellite to Street: A Hybrid Framework Integrating Stable Diffusion and PanoGAN for Consistent Cross-View Synthesis
by: Bajbaa, Khawlah, et al.
Published: (2025)
by: Bajbaa, Khawlah, et al.
Published: (2025)
nuCarla: A nuScenes-Style Bird's-Eye View Perception Dataset for CARLA Simulation
by: Qiao, Zhijie, et al.
Published: (2025)
by: Qiao, Zhijie, et al.
Published: (2025)
Examining the Commitments and Difficulties Inherent in Multimodal Foundation Models for Street View Imagery
by: Yang, Zhenyuan, et al.
Published: (2024)
by: Yang, Zhenyuan, et al.
Published: (2024)
Latent Diffusion Models for Attribute-Preserving Image Anonymization
by: Piano, Luca, et al.
Published: (2024)
by: Piano, Luca, et al.
Published: (2024)
Contrastive Learning Guided Latent Diffusion Model for Image-to-Image Translation
by: Si, Qi, et al.
Published: (2025)
by: Si, Qi, et al.
Published: (2025)
When Preferences Diverge: Aligning Diffusion Models with Minority-Aware Adaptive DPO
by: Zhang, Lingfan, et al.
Published: (2025)
by: Zhang, Lingfan, et al.
Published: (2025)
Gradient-Guided Conditional Diffusion Models for Private Image Reconstruction: Analyzing Adversarial Impacts of Differential Privacy and Denoising
by: Huang, Tao, et al.
Published: (2024)
by: Huang, Tao, et al.
Published: (2024)
From Street to Orbit: Training-Free Cross-View Retrieval via Location Semantics and LLM Guidance
by: Min, Jeongho, et al.
Published: (2025)
by: Min, Jeongho, et al.
Published: (2025)
OT-ALD: Aligning Latent Distributions with Optimal Transport for Accelerated Image-to-Image Translation
by: Wang, Zhanpeng, et al.
Published: (2025)
by: Wang, Zhanpeng, et al.
Published: (2025)
Similar Items
-
Bird Eye-View to Street-View: A Survey
by: Bajbaa, Khawlah, et al.
Published: (2024) -
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
by: Xu, Tianshuo, et al.
Published: (2024) -
Street-View Image Generation from a Bird's-Eye View Layout
by: Swerdlow, Alexander, et al.
Published: (2023) -
MagicDrive: Street View Generation with Diverse 3D Geometry Control
by: Gao, Ruiyuan, et al.
Published: (2023) -
SDGOCC: Semantic and Depth-Guided Bird's-Eye View Transformation for 3D Multimodal Occupancy Prediction
by: Duan, Zaipeng, et al.
Published: (2025)