SyntheOcc: Synthesize Geometric-Controlled Street View Images through 3D Semantic MPIs
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Leheng, Qiu, Weichao, Cai, Yingjie, Yan, Xu, Lian, Qing, Liu, Bingbing, Chen, Ying-Cong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
OmniBooth: Learning Latent Control for Image Synthesis with Multi-modal Instruction
di: Li, Leheng, et al.
Pubblicazione: (2024)
di: Li, Leheng, et al.
Pubblicazione: (2024)
SPOT-Occ: Sparse Prototype-guided Transformer for Camera-based 3D Occupancy Prediction
di: Chen, Suzeyu, et al.
Pubblicazione: (2026)
di: Chen, Suzeyu, et al.
Pubblicazione: (2026)
Adv3D: Generating 3D Adversarial Examples for 3D Object Detection in Driving Scenarios with NeRF
di: Li, Leheng, et al.
Pubblicazione: (2023)
di: Li, Leheng, et al.
Pubblicazione: (2023)
Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models
di: Xu, Tianshuo, et al.
Pubblicazione: (2025)
di: Xu, Tianshuo, et al.
Pubblicazione: (2025)
DisEnvisioner: Disentangled and Enriched Visual Prompt for Customized Image Generation
di: He, Jing, et al.
Pubblicazione: (2024)
di: He, Jing, et al.
Pubblicazione: (2024)
S-VAM: Shortcut Video-Action Model by Self-Distilling Geometric and Semantic Foresight
di: Yan, Haodong, et al.
Pubblicazione: (2026)
di: Yan, Haodong, et al.
Pubblicazione: (2026)
RoboOcc: Enhancing the Geometric and Semantic Scene Understanding for Robots
di: Zhang, Zhang, et al.
Pubblicazione: (2025)
di: Zhang, Zhang, et al.
Pubblicazione: (2025)
Seeing through Satellite Images at Street Views
di: Qian, Ming, et al.
Pubblicazione: (2025)
di: Qian, Ming, et al.
Pubblicazione: (2025)
StreetCrafter: Street View Synthesis with Controllable Video Diffusion Models
di: Yan, Yunzhi, et al.
Pubblicazione: (2024)
di: Yan, Yunzhi, et al.
Pubblicazione: (2024)
Statistical inference for high-dimensional convoluted rank regression
di: Cai, Leheng, et al.
Pubblicazione: (2024)
di: Cai, Leheng, et al.
Pubblicazione: (2024)
OccLE: Label-Efficient 3D Semantic Occupancy Prediction
di: Fang, Naiyu, et al.
Pubblicazione: (2025)
di: Fang, Naiyu, et al.
Pubblicazione: (2025)
MagicDrive: Street View Generation with Diverse 3D Geometry Control
di: Gao, Ruiyuan, et al.
Pubblicazione: (2023)
di: Gao, Ruiyuan, et al.
Pubblicazione: (2023)
Efficient Depth-Guided Urban View Synthesis
di: Miao, Sheng, et al.
Pubblicazione: (2024)
di: Miao, Sheng, et al.
Pubblicazione: (2024)
An Efficient Occupancy World Model via Decoupled Dynamic Flow and Image-assisted Training
di: Zhang, Haiming, et al.
Pubblicazione: (2024)
di: Zhang, Haiming, et al.
Pubblicazione: (2024)
SynthForge: Synthesizing High-Quality Face Dataset with Controllable 3D Generative Models
di: Rawat, Abhay, et al.
Pubblicazione: (2024)
di: Rawat, Abhay, et al.
Pubblicazione: (2024)
DreamDrive: Generative 4D Scene Modeling from Street View Images
di: Mao, Jiageng, et al.
Pubblicazione: (2024)
di: Mao, Jiageng, et al.
Pubblicazione: (2024)
MegaSynth: Scaling Up 3D Scene Reconstruction with Synthesized Data
di: Jiang, Hanwen, et al.
Pubblicazione: (2024)
di: Jiang, Hanwen, et al.
Pubblicazione: (2024)
Street-View Image Generation from a Bird's-Eye View Layout
di: Swerdlow, Alexander, et al.
Pubblicazione: (2023)
di: Swerdlow, Alexander, et al.
Pubblicazione: (2023)
PerLDiff: Controllable Street View Synthesis Using Perspective-Layout Diffusion Models
di: Zhang, Jinhua, et al.
Pubblicazione: (2024)
di: Zhang, Jinhua, et al.
Pubblicazione: (2024)
Text2Street: Controllable Text-to-image Generation for Street Views
di: Su, Jinming, et al.
Pubblicazione: (2024)
di: Su, Jinming, et al.
Pubblicazione: (2024)
MS-Occ: Multi-Stage LiDAR-Camera Fusion for 3D Semantic Occupancy Prediction
di: Wei, Zhiqiang, et al.
Pubblicazione: (2025)
di: Wei, Zhiqiang, et al.
Pubblicazione: (2025)
From Sparse to Dense Functional Data: Phase Transitions from a Simultaneous Inference Perspective
di: Cai, Leheng, et al.
Pubblicazione: (2024)
di: Cai, Leheng, et al.
Pubblicazione: (2024)
From sparse to dense functional time series: phase transitions of detecting structural breaks and beyond
di: Cai, Leheng, et al.
Pubblicazione: (2024)
di: Cai, Leheng, et al.
Pubblicazione: (2024)
Uplifting Range-View-based 3D Semantic Segmentation in Real-Time with Multi-Sensor Fusion
di: Tan, Shiqi, et al.
Pubblicazione: (2024)
di: Tan, Shiqi, et al.
Pubblicazione: (2024)
MagicDrive3D: Controllable 3D Generation for Any-View Rendering in Street Scenes
di: Gao, Ruiyuan, et al.
Pubblicazione: (2024)
di: Gao, Ruiyuan, et al.
Pubblicazione: (2024)
SatSynth: Augmenting Image-Mask Pairs through Diffusion Models for Aerial Semantic Segmentation
di: Toker, Aysim, et al.
Pubblicazione: (2024)
di: Toker, Aysim, et al.
Pubblicazione: (2024)
Semantic-Aware Label Placement for Augmented Reality in Street View
di: Jia, Jianqing, et al.
Pubblicazione: (2019)
di: Jia, Jianqing, et al.
Pubblicazione: (2019)
Test and Measure for Partial Mean Dependence Based on Machine Learning Methods
di: Cai, Leheng, et al.
Pubblicazione: (2022)
di: Cai, Leheng, et al.
Pubblicazione: (2022)
3D-LMVIC: Learning-based Multi-View Image Coding with 3D Gaussian Geometric Priors
di: Huang, Yujun, et al.
Pubblicazione: (2024)
di: Huang, Yujun, et al.
Pubblicazione: (2024)
FastOcc: Accelerating 3D Occupancy Prediction by Fusing the 2D Bird's-Eye View and Perspective View
di: Hou, Jiawei, et al.
Pubblicazione: (2024)
di: Hou, Jiawei, et al.
Pubblicazione: (2024)
SynthCloner: Synthesizer-style Audio Transfer via Factorized Codec with ADSR Envelope Control
di: Liu, Jeng-Yue, et al.
Pubblicazione: (2025)
di: Liu, Jeng-Yue, et al.
Pubblicazione: (2025)
SynthTab: Leveraging Synthesized Data for Guitar Tablature Transcription
di: Zang, Yongyi, et al.
Pubblicazione: (2023)
di: Zang, Yongyi, et al.
Pubblicazione: (2023)
OSMLoc: Single Image-Based Visual Localization in OpenStreetMap with Fused Geometric and Semantic Guidance
di: Liao, Youqi, et al.
Pubblicazione: (2024)
di: Liao, Youqi, et al.
Pubblicazione: (2024)
QueryOcc: Query-based Self-Supervision for 3D Semantic Occupancy
di: Lilja, Adam, et al.
Pubblicazione: (2025)
di: Lilja, Adam, et al.
Pubblicazione: (2025)
SinoSynth: A Physics-based Domain Randomization Approach for Generalizable CBCT Image Enhancement
di: Pang, Yunkui, et al.
Pubblicazione: (2024)
di: Pang, Yunkui, et al.
Pubblicazione: (2024)
Satellite-to-Street: Synthesizing Post-Disaster Views from Satellite Imagery via Generative Vision Models
di: Yang, Yifan, et al.
Pubblicazione: (2026)
di: Yang, Yifan, et al.
Pubblicazione: (2026)
MonoOcc: Digging into Monocular Semantic Occupancy Prediction
di: Zheng, Yupeng, et al.
Pubblicazione: (2024)
di: Zheng, Yupeng, et al.
Pubblicazione: (2024)
ForecastOcc: Vision-based Semantic Occupancy Forecasting
di: Mohan, Riya, et al.
Pubblicazione: (2026)
di: Mohan, Riya, et al.
Pubblicazione: (2026)
VectorSynth: Fine-Grained Satellite Image Synthesis with Structured Semantics
di: Cher, Daniel, et al.
Pubblicazione: (2025)
di: Cher, Daniel, et al.
Pubblicazione: (2025)
Learning to Synthesize Graphics Programs for Geometric Artworks
di: Bing, Qi, et al.
Pubblicazione: (2024)
di: Bing, Qi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
OmniBooth: Learning Latent Control for Image Synthesis with Multi-modal Instruction
di: Li, Leheng, et al.
Pubblicazione: (2024) -
SPOT-Occ: Sparse Prototype-guided Transformer for Camera-based 3D Occupancy Prediction
di: Chen, Suzeyu, et al.
Pubblicazione: (2026) -
Adv3D: Generating 3D Adversarial Examples for 3D Object Detection in Driving Scenarios with NeRF
di: Li, Leheng, et al.
Pubblicazione: (2023) -
Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models
di: Xu, Tianshuo, et al.
Pubblicazione: (2025) -
DisEnvisioner: Disentangled and Enriched Visual Prompt for Customized Image Generation
di: He, Jing, et al.
Pubblicazione: (2024)