WildFusion: Learning 3D-Aware Latent Diffusion Models in View Space
Fuente:
arXiv
Saved in:
| Main Authors: | Schwarz, Katja, Kim, Seung Wook, Gao, Jun, Fidler, Sanja, Geiger, Andreas, Kreis, Karsten |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Align Your Gaussians: Text-to-4D with Dynamic 3D Gaussians and Composed Diffusion Models
by: Ling, Huan, et al.
Published: (2023)
by: Ling, Huan, et al.
Published: (2023)
Align Your Steps: Optimizing Sampling Schedules in Diffusion Models
by: Sabour, Amirmojtaba, et al.
Published: (2024)
by: Sabour, Amirmojtaba, et al.
Published: (2024)
Align Your Flow: Scaling Continuous-Time Flow Map Distillation
by: Sabour, Amirmojtaba, et al.
Published: (2025)
by: Sabour, Amirmojtaba, et al.
Published: (2025)
EmerDiff: Emerging Pixel-level Semantic Knowledge in Diffusion Models
by: Namekata, Koichi, et al.
Published: (2024)
by: Namekata, Koichi, et al.
Published: (2024)
WildFusion: Individual Animal Identification with Calibrated Similarity Fusion
by: Cermak, Vojtěch, et al.
Published: (2024)
by: Cermak, Vojtěch, et al.
Published: (2024)
RefFusion: Reference Adapted Diffusion Models for 3D Scene Inpainting
by: Mirzaei, Ashkan, et al.
Published: (2024)
by: Mirzaei, Ashkan, et al.
Published: (2024)
WildFusion: Multimodal Implicit 3D Reconstructions in the Wild
by: Liu, Yanbaihui, et al.
Published: (2024)
by: Liu, Yanbaihui, et al.
Published: (2024)
L4GM: Large 4D Gaussian Reconstruction Model
by: Ren, Jiawei, et al.
Published: (2024)
by: Ren, Jiawei, et al.
Published: (2024)
Outdoor Scene Extrapolation with Hierarchical Generative Cellular Automata
by: Zhang, Dongsu, et al.
Published: (2024)
by: Zhang, Dongsu, et al.
Published: (2024)
WildCAT3D: Appearance-Aware Multi-View Diffusion in the Wild
by: Alper, Morris, et al.
Published: (2025)
by: Alper, Morris, et al.
Published: (2025)
PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion
by: Lu, Yifan, et al.
Published: (2026)
by: Lu, Yifan, et al.
Published: (2026)
Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene Generation
by: Yang, Yuanbo, et al.
Published: (2024)
by: Yang, Yuanbo, et al.
Published: (2024)
3D-LATTE: Latent Space 3D Editing from Textual Instructions
by: Parelli, Maria, et al.
Published: (2025)
by: Parelli, Maria, et al.
Published: (2025)
PARTFIELD: Learning 3D Feature Fields for Part Segmentation and Beyond
by: Liu, Minghua, et al.
Published: (2025)
by: Liu, Minghua, et al.
Published: (2025)
DisCo-Diff: Enhancing Continuous Diffusion Models with Discrete Latents
by: Xu, Yilun, et al.
Published: (2024)
by: Xu, Yilun, et al.
Published: (2024)
Difix3D+: Improving 3D Reconstructions with Single-Step Diffusion Models
by: Wu, Jay Zhangjie, et al.
Published: (2025)
by: Wu, Jay Zhangjie, et al.
Published: (2025)
sshELF: Single-Shot Hierarchical Extrapolation of Latent Features for 3D Reconstruction from Sparse-Views
by: Najafli, Eyvaz, et al.
Published: (2025)
by: Najafli, Eyvaz, et al.
Published: (2025)
SpaceMesh: A Continuous Representation for Learning Manifold Surface Meshes
by: Shen, Tianchang, et al.
Published: (2024)
by: Shen, Tianchang, et al.
Published: (2024)
DistillNeRF: Perceiving 3D Scenes from Single-Glance Images by Distilling Neural Fields and Foundation Model Features
by: Wang, Letian, et al.
Published: (2024)
by: Wang, Letian, et al.
Published: (2024)
Generative Gaussian Splatting: Generating 3D Scenes with Video Diffusion Priors
by: Schwarz, Katja, et al.
Published: (2025)
by: Schwarz, Katja, et al.
Published: (2025)
GTA: A Geometry-Aware Attention Mechanism for Multi-View Transformers
by: Miyato, Takeru, et al.
Published: (2023)
by: Miyato, Takeru, et al.
Published: (2023)
MV-SAM3D: Adaptive Multi-View Fusion for Layout-Aware 3D Generation
by: Li, Baicheng, et al.
Published: (2026)
by: Li, Baicheng, et al.
Published: (2026)
CATSplat: Context-Aware Transformer with Spatial Guidance for Generalizable 3D Gaussian Splatting from A Single-View Image
by: Roh, Wonseok, et al.
Published: (2024)
by: Roh, Wonseok, et al.
Published: (2024)
Real-Time Fusion of Visual and Chart Data for Enhanced Maritime Vision
by: Kreis, Marten, et al.
Published: (2025)
by: Kreis, Marten, et al.
Published: (2025)
Lyra: Generative 3D Scene Reconstruction via Video Diffusion Model Self-Distillation
by: Bahmani, Sherwin, et al.
Published: (2025)
by: Bahmani, Sherwin, et al.
Published: (2025)
LATTE3D: Large-scale Amortized Text-To-Enhanced3D Synthesis
by: Xie, Kevin, et al.
Published: (2024)
by: Xie, Kevin, et al.
Published: (2024)
XCube: Large-Scale 3D Generative Modeling using Sparse Voxel Hierarchies
by: Ren, Xuanchi, et al.
Published: (2023)
by: Ren, Xuanchi, et al.
Published: (2023)
Multi-student Diffusion Distillation for Better One-step Generators
by: Song, Yanke, et al.
Published: (2024)
by: Song, Yanke, et al.
Published: (2024)
3D Congealing: 3D-Aware Image Alignment in the Wild
by: Zhang, Yunzhi, et al.
Published: (2024)
by: Zhang, Yunzhi, et al.
Published: (2024)
GEN3C: 3D-Informed World-Consistent Video Generation with Precise Camera Control
by: Ren, Xuanchi, et al.
Published: (2025)
by: Ren, Xuanchi, et al.
Published: (2025)
Can Large Vision-Language Models Correct Semantic Grounding Errors By Themselves?
by: Liao, Yuan-Hong, et al.
Published: (2024)
by: Liao, Yuan-Hong, et al.
Published: (2024)
DiffusionRenderer: Neural Inverse and Forward Rendering with Video Diffusion Models
by: Liang, Ruofan, et al.
Published: (2025)
by: Liang, Ruofan, et al.
Published: (2025)
FedWSQ: Efficient Federated Learning with Weight Standardization and Distribution-Aware Non-Uniform Quantization
by: Kim, Seung-Wook, et al.
Published: (2025)
by: Kim, Seung-Wook, et al.
Published: (2025)
Direct Diffusion Score Preference Optimization via Stepwise Contrastive Policy-Pair Supervision
by: Kim, Dohyun, et al.
Published: (2025)
by: Kim, Dohyun, et al.
Published: (2025)
Sparse-View 3D Gaussian Splatting in the Wild
by: Park, Wongi, et al.
Published: (2026)
by: Park, Wongi, et al.
Published: (2026)
ReMatching Dynamic Reconstruction Flow
by: Oblak, Sara, et al.
Published: (2024)
by: Oblak, Sara, et al.
Published: (2024)
NeRF-XL: Scaling NeRFs with Multiple GPUs
by: Li, Ruilong, et al.
Published: (2024)
by: Li, Ruilong, et al.
Published: (2024)
GeoFusion-CAD: Structure-Aware Diffusion with Geometric State Space for Parametric 3D Design
by: Zhou, Xiaolei, et al.
Published: (2026)
by: Zhou, Xiaolei, et al.
Published: (2026)
Reasoning Paths with Reference Objects Elicit Quantitative Spatial Reasoning in Large Vision-Language Models
by: Liao, Yuan-Hong, et al.
Published: (2024)
by: Liao, Yuan-Hong, et al.
Published: (2024)
Controllable Weather Synthesis and Removal with Video Diffusion Models
by: Lin, Chih-Hao, et al.
Published: (2025)
by: Lin, Chih-Hao, et al.
Published: (2025)
Similar Items
-
Align Your Gaussians: Text-to-4D with Dynamic 3D Gaussians and Composed Diffusion Models
by: Ling, Huan, et al.
Published: (2023) -
Align Your Steps: Optimizing Sampling Schedules in Diffusion Models
by: Sabour, Amirmojtaba, et al.
Published: (2024) -
Align Your Flow: Scaling Continuous-Time Flow Map Distillation
by: Sabour, Amirmojtaba, et al.
Published: (2025) -
EmerDiff: Emerging Pixel-level Semantic Knowledge in Diffusion Models
by: Namekata, Koichi, et al.
Published: (2024) -
WildFusion: Individual Animal Identification with Calibrated Similarity Fusion
by: Cermak, Vojtěch, et al.
Published: (2024)