SV3D: Novel Multi-view Synthesis and 3D Generation from a Single Image using Latent Video Diffusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Voleti, Vikram, Yao, Chun-Han, Boss, Mark, Letts, Adam, Pankratz, David, Tochilkin, Dmitry, Laforte, Christian, Rombach, Robin, Jampani, Varun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TripoSR: Fast 3D Object Reconstruction from a Single Image
von: Tochilkin, Dmitry, et al.
Veröffentlicht: (2024)
von: Tochilkin, Dmitry, et al.
Veröffentlicht: (2024)
SViM3D: Stable Video Material Diffusion for Single Image 3D Generation
von: Engelhardt, Andreas, et al.
Veröffentlicht: (2025)
von: Engelhardt, Andreas, et al.
Veröffentlicht: (2025)
SV4D 2.0: Enhancing Spatio-Temporal Consistency in Multi-View Video Diffusion for High-Quality 4D Generation
von: Yao, Chun-Han, et al.
Veröffentlicht: (2025)
von: Yao, Chun-Han, et al.
Veröffentlicht: (2025)
SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency
von: Xie, Yiming, et al.
Veröffentlicht: (2024)
von: Xie, Yiming, et al.
Veröffentlicht: (2024)
HouseCrafter: Lifting Floorplans to 3D Scenes with 2D Diffusion Model
von: Nguyen, Hieu T., et al.
Veröffentlicht: (2024)
von: Nguyen, Hieu T., et al.
Veröffentlicht: (2024)
Stable Virtual Camera: Generative View Synthesis with Diffusion Models
von: Zhou, Jensen, et al.
Veröffentlicht: (2025)
von: Zhou, Jensen, et al.
Veröffentlicht: (2025)
SF3D: Stable Fast 3D Mesh Reconstruction with UV-unwrapping and Illumination Disentanglement
von: Boss, Mark, et al.
Veröffentlicht: (2024)
von: Boss, Mark, et al.
Veröffentlicht: (2024)
SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
von: Huang, Zixuan, et al.
Veröffentlicht: (2025)
von: Huang, Zixuan, et al.
Veröffentlicht: (2025)
MVD-Fusion: Single-view 3D via Depth-consistent Multi-view Generation
von: Hu, Hanzhe, et al.
Veröffentlicht: (2024)
von: Hu, Hanzhe, et al.
Veröffentlicht: (2024)
ReLi3D: Relightable Multi-view 3D Reconstruction with Disentangled Illumination
von: Dihlmann, Jan-Niklas, et al.
Veröffentlicht: (2026)
von: Dihlmann, Jan-Niklas, et al.
Veröffentlicht: (2026)
Stable Video-Driven Portraits
von: R., Mallikarjun B., et al.
Veröffentlicht: (2025)
von: R., Mallikarjun B., et al.
Veröffentlicht: (2025)
OCTOPUS: Optimized KV Cache for Transformers via Octahedral Parametrization Under optimal Squared error quantization
von: Boss, Mark, et al.
Veröffentlicht: (2026)
von: Boss, Mark, et al.
Veröffentlicht: (2026)
FaceCraft4D: Animated 3D Facial Avatar Generation from a Single Image
von: Yin, Fei, et al.
Veröffentlicht: (2025)
von: Yin, Fei, et al.
Veröffentlicht: (2025)
HumANDiff: Articulated Noise Diffusion for Motion-Consistent Human Video Generation
von: Hu, Tao, et al.
Veröffentlicht: (2026)
von: Hu, Tao, et al.
Veröffentlicht: (2026)
Stable Part Diffusion 4D: Multi-View RGB and Kinematic Parts Video Generation
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
Human Video Generation from a Single Image with 3D Pose and View Control
von: Wang, Tiantian, et al.
Veröffentlicht: (2026)
von: Wang, Tiantian, et al.
Veröffentlicht: (2026)
ReSWD: ReSTIR'd, not shaken. Combining Reservoir Sampling and Sliced Wasserstein Distance for Variance Reduction
von: Boss, Mark, et al.
Veröffentlicht: (2025)
von: Boss, Mark, et al.
Veröffentlicht: (2025)
HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models
von: Peng, Xiaogang, et al.
Veröffentlicht: (2023)
von: Peng, Xiaogang, et al.
Veröffentlicht: (2023)
MARBLE: Material Recomposition and Blending in CLIP-Space
von: Cheng, Ta-Ying, et al.
Veröffentlicht: (2025)
von: Cheng, Ta-Ying, et al.
Veröffentlicht: (2025)
Computational Tradeoffs in Image Synthesis: Diffusion, Masked-Token, and Next-Token Prediction
von: Kilian, Maciej, et al.
Veröffentlicht: (2024)
von: Kilian, Maciej, et al.
Veröffentlicht: (2024)
Fast High-Resolution Image Synthesis with Latent Adversarial Diffusion Distillation
von: Sauer, Axel, et al.
Veröffentlicht: (2024)
von: Sauer, Axel, et al.
Veröffentlicht: (2024)
SV3.3B: A Sports Video Understanding Model for Action Recognition
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025)
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025)
Shaping Realities: Enhancing 3D Generative AI with Fabrication Constraints
von: Faruqi, Faraz, et al.
Veröffentlicht: (2024)
von: Faruqi, Faraz, et al.
Veröffentlicht: (2024)
WordRobe: Text-Guided Generation of Textured 3D Garments
von: Srivastava, Astitva, et al.
Veröffentlicht: (2024)
von: Srivastava, Astitva, et al.
Veröffentlicht: (2024)
SV-DRR: High-Fidelity Novel View X-Ray Synthesis Using Diffusion Model
von: Xie, Chun, et al.
Veröffentlicht: (2025)
von: Xie, Chun, et al.
Veröffentlicht: (2025)
3D Congealing: 3D-Aware Image Alignment in the Wild
von: Zhang, Yunzhi, et al.
Veröffentlicht: (2024)
von: Zhang, Yunzhi, et al.
Veröffentlicht: (2024)
Foley Control: Aligning a Frozen Latent Text-to-Audio Model to Video
von: Rowles, Ciara, et al.
Veröffentlicht: (2025)
von: Rowles, Ciara, et al.
Veröffentlicht: (2025)
CausNVS: Autoregressive Multi-view Diffusion for Flexible 3D Novel View Synthesis
von: Kong, Xin, et al.
Veröffentlicht: (2025)
von: Kong, Xin, et al.
Veröffentlicht: (2025)
Diffusion$^2$: Dynamic 3D Content Generation via Score Composition of Video and Multi-view Diffusion Models
von: Yang, Zeyu, et al.
Veröffentlicht: (2024)
von: Yang, Zeyu, et al.
Veröffentlicht: (2024)
Unified Dense Prediction of Video Diffusion
von: Yang, Lehan, et al.
Veröffentlicht: (2025)
von: Yang, Lehan, et al.
Veröffentlicht: (2025)
DecompDreamer: A Composition-Aware Curriculum for Structured 3D Asset Generation
von: Nath, Utkarsh, et al.
Veröffentlicht: (2025)
von: Nath, Utkarsh, et al.
Veröffentlicht: (2025)
Dehaze-then-Splat: Generative Dehazing with Physics-Informed 3D Gaussian Splatting for Smoke-Free Novel View Synthesis
von: Chen, Boss, et al.
Veröffentlicht: (2026)
von: Chen, Boss, et al.
Veröffentlicht: (2026)
Animate3D: Animating Any 3D Model with Multi-view Video Diffusion
von: Jiang, Yanqin, et al.
Veröffentlicht: (2024)
von: Jiang, Yanqin, et al.
Veröffentlicht: (2024)
CompGS: Unleashing 2D Compositionality for Compositional Text-to-3D via Dynamically Optimizing 3D Gaussians
von: Ge, Chongjian, et al.
Veröffentlicht: (2024)
von: Ge, Chongjian, et al.
Veröffentlicht: (2024)
ICE-G: Image Conditional Editing of 3D Gaussian Splats
von: Jaganathan, Vishnu, et al.
Veröffentlicht: (2024)
von: Jaganathan, Vishnu, et al.
Veröffentlicht: (2024)
CFSynthesis: Controllable and Free-view 3D Human Video Synthesis
von: Cui, Liyuan, et al.
Veröffentlicht: (2024)
von: Cui, Liyuan, et al.
Veröffentlicht: (2024)
ContactArt: Learning 3D Interaction Priors for Category-level Articulated Object and Hand Poses Estimation
von: Zhu, Zehao, et al.
Veröffentlicht: (2023)
von: Zhu, Zehao, et al.
Veröffentlicht: (2023)
Dual3D: Efficient and Consistent Text-to-3D Generation with Dual-mode Multi-view Latent Diffusion
von: Li, Xinyang, et al.
Veröffentlicht: (2024)
von: Li, Xinyang, et al.
Veröffentlicht: (2024)
SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing
von: Li, Ruihuang, et al.
Veröffentlicht: (2024)
von: Li, Ruihuang, et al.
Veröffentlicht: (2024)
3D-LENS: A 3D Lifting-based Elevated Novel-view Synthesis method for Single-View Aerial-Ground Re-Identification
von: Grolleau, William, et al.
Veröffentlicht: (2026)
von: Grolleau, William, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
TripoSR: Fast 3D Object Reconstruction from a Single Image
von: Tochilkin, Dmitry, et al.
Veröffentlicht: (2024) -
SViM3D: Stable Video Material Diffusion for Single Image 3D Generation
von: Engelhardt, Andreas, et al.
Veröffentlicht: (2025) -
SV4D 2.0: Enhancing Spatio-Temporal Consistency in Multi-View Video Diffusion for High-Quality 4D Generation
von: Yao, Chun-Han, et al.
Veröffentlicht: (2025) -
SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency
von: Xie, Yiming, et al.
Veröffentlicht: (2024) -
HouseCrafter: Lifting Floorplans to 3D Scenes with 2D Diffusion Model
von: Nguyen, Hieu T., et al.
Veröffentlicht: (2024)