StreetCrafter: Street View Synthesis with Controllable Video Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Yan, Yunzhi, Xu, Zhen, Lin, Haotong, Jin, Haian, Guo, Haoyu, Wang, Yida, Zhan, Kun, Lang, Xianpeng, Bao, Hujun, Zhou, Xiaowei, Peng, Sida |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting
by: Yan, Yunzhi, et al.
Published: (2024)
by: Yan, Yunzhi, et al.
Published: (2024)
StyledStreets: Multi-style Street Simulator with Spatial and Temporal Consistency
by: Chen, Yuyin, et al.
Published: (2025)
by: Chen, Yuyin, et al.
Published: (2025)
Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models
by: Jin, Yudong, et al.
Published: (2025)
by: Jin, Yudong, et al.
Published: (2025)
Multi-view Reconstruction via SfM-guided Monocular Depth Estimation
by: Guo, Haoyu, et al.
Published: (2025)
by: Guo, Haoyu, et al.
Published: (2025)
StreetForward: Perceiving Dynamic Street with Feedforward Causal Attention
by: Yu, Zhongrui, et al.
Published: (2026)
by: Yu, Zhongrui, et al.
Published: (2026)
UniVerse: Unleashing the Scene Prior of Video Diffusion Models for Robust Radiance Field Reconstruction
by: Cao, Jin, et al.
Published: (2025)
by: Cao, Jin, et al.
Published: (2025)
HiNeuS: High-fidelity Neural Surface Mitigating Low-texture and Reflective Ambiguity
by: Wang, Yida, et al.
Published: (2025)
by: Wang, Yida, et al.
Published: (2025)
Precise Action-to-Video Generation Through Visual Action Prompts
by: Wang, Yuang, et al.
Published: (2025)
by: Wang, Yuang, et al.
Published: (2025)
Representing Long Volumetric Video with Temporal Gaussian Hierarchy
by: Xu, Zhen, et al.
Published: (2024)
by: Xu, Zhen, et al.
Published: (2024)
SpatialCrafter: Unleashing the Imagination of Video Diffusion Models for Scene Reconstruction from Limited Observations
by: Zhang, Songchun, et al.
Published: (2025)
by: Zhang, Songchun, et al.
Published: (2025)
Text2Street: Controllable Text-to-image Generation for Street Views
by: Su, Jinming, et al.
Published: (2024)
by: Su, Jinming, et al.
Published: (2024)
EnvGS: Modeling View-Dependent Appearance with Environment Gaussian
by: Xie, Tao, et al.
Published: (2024)
by: Xie, Tao, et al.
Published: (2024)
SGD: Street View Synthesis with Gaussian Splatting and Diffusion Prior
by: Yu, Zhongrui, et al.
Published: (2024)
by: Yu, Zhongrui, et al.
Published: (2024)
Enhancing the Understanding of Urban Street Perception With LLM s and Street View Imagery
by: Xin Han, et al.
Published: (2026)
by: Xin Han, et al.
Published: (2026)
PerLDiff: Controllable Street View Synthesis Using Perspective-Layout Diffusion Models
by: Zhang, Jinhua, et al.
Published: (2024)
by: Zhang, Jinhua, et al.
Published: (2024)
ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis
by: Yu, Wangbo, et al.
Published: (2024)
by: Yu, Wangbo, et al.
Published: (2024)
CrossViewDiff: A Cross-View Diffusion Model for Satellite-to-Street View Synthesis
by: Li, Weijia, et al.
Published: (2024)
by: Li, Weijia, et al.
Published: (2024)
Split4D: Decomposed 4D Scene Reconstruction Without Video Segmentation
by: Hu, Yongzhen, et al.
Published: (2025)
by: Hu, Yongzhen, et al.
Published: (2025)
Streetscapes: Large-scale Consistent Street View Generation Using Autoregressive Video Diffusion
by: Deng, Boyang, et al.
Published: (2024)
by: Deng, Boyang, et al.
Published: (2024)
Controllable Satellite-to-Street-View Synthesis with Precise Pose Alignment and Zero-Shot Environmental Control
by: Ze, Xianghui, et al.
Published: (2025)
by: Ze, Xianghui, et al.
Published: (2025)
Painting 3D Nature in 2D: View Synthesis of Natural Scenes from a Single Semantic Mask
by: Zhang, Shangzan, et al.
Published: (2023)
by: Zhang, Shangzan, et al.
Published: (2023)
Prompting Depth Anything for 4K Resolution Accurate Metric Depth Estimation
by: Lin, Haotong, et al.
Published: (2024)
by: Lin, Haotong, et al.
Published: (2024)
Hallucinating 360°: Panoramic Street-View Generation via Local Scenes Diffusion and Probabilistic Prompting
by: Teng, Fei, et al.
Published: (2025)
by: Teng, Fei, et al.
Published: (2025)
World-Grounded Human Motion Recovery via Gravity-View Coordinates
by: Shen, Zehong, et al.
Published: (2024)
by: Shen, Zehong, et al.
Published: (2024)
Map2Video: Street View Imagery Driven AI Video Generation
by: Jo, Hye-Young, et al.
Published: (2025)
by: Jo, Hye-Young, et al.
Published: (2025)
Segmentation-Guided Neural Radiance Fields for Novel Street View Synthesis
by: Li, Yizhou, et al.
Published: (2025)
by: Li, Yizhou, et al.
Published: (2025)
Bird Eye-View to Street-View: A Survey
by: Bajbaa, Khawlah, et al.
Published: (2024)
by: Bajbaa, Khawlah, et al.
Published: (2024)
StreetReaderAI: Making Street View Accessible Using Context-Aware Multimodal AI
by: Froehlich, Jon E., et al.
Published: (2025)
by: Froehlich, Jon E., et al.
Published: (2025)
Seeing through Satellite Images at Street Views
by: Qian, Ming, et al.
Published: (2025)
by: Qian, Ming, et al.
Published: (2025)
Artifacts of Idiosyncracy in Global Street View Data
by: Alpherts, Tim, et al.
Published: (2025)
by: Alpherts, Tim, et al.
Published: (2025)
Learning Street View Representations with Spatiotemporal Contrast
by: Li, Yong, et al.
Published: (2025)
by: Li, Yong, et al.
Published: (2025)
StreetLens: Enabling Human-Centered AI Agents for Neighborhood Assessment from Street View Imagery
by: Kim, Jina, et al.
Published: (2025)
by: Kim, Jina, et al.
Published: (2025)
DiffPlace: Street View Generation via Place-Controllable Diffusion Model Enhancing Place Recognition
by: Li, Ji, et al.
Published: (2026)
by: Li, Ji, et al.
Published: (2026)
Semantic-Aware Label Placement for Augmented Reality in Street View
by: Jia, Jianqing, et al.
Published: (2019)
by: Jia, Jianqing, et al.
Published: (2019)
MatchAnything: Universal Cross-Modality Image Matching with Large-Scale Pre-Training
by: He, Xingyi, et al.
Published: (2025)
by: He, Xingyi, et al.
Published: (2025)
StreamingTalker: Audio-driven 3D Facial Animation with Autoregressive Diffusion Model
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
MagicDrive: Street View Generation with Diverse 3D Geometry Control
by: Gao, Ruiyuan, et al.
Published: (2023)
by: Gao, Ruiyuan, et al.
Published: (2023)
Visualizing Routes with AI-Discovered Street-View Patterns
by: Wu, Tsung Heng, et al.
Published: (2024)
by: Wu, Tsung Heng, et al.
Published: (2024)
Spatially-Weighted CLIP for Street-View Geo-localization
by: Han, Ting, et al.
Published: (2026)
by: Han, Ting, et al.
Published: (2026)
The Process of Permanence on the Streets. Street Children in Mexico City
by: Patricia Murrieta
Published: (2010)
by: Patricia Murrieta
Published: (2010)
Similar Items
-
Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting
by: Yan, Yunzhi, et al.
Published: (2024) -
StyledStreets: Multi-style Street Simulator with Spatial and Temporal Consistency
by: Chen, Yuyin, et al.
Published: (2025) -
Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models
by: Jin, Yudong, et al.
Published: (2025) -
Multi-view Reconstruction via SfM-guided Monocular Depth Estimation
by: Guo, Haoyu, et al.
Published: (2025) -
StreetForward: Perceiving Dynamic Street with Feedforward Causal Attention
by: Yu, Zhongrui, et al.
Published: (2026)