3D Congealing: 3D-Aware Image Alignment in the Wild
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yunzhi, Li, Zizhang, Raj, Amit, Engelhardt, Andreas, Li, Yuanzhen, Hou, Tingbo, Wu, Jiajun, Jampani, Varun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Probing the 3D Awareness of Visual Foundation Models
by: Banani, Mohamed El, et al.
Published: (2024)
by: Banani, Mohamed El, et al.
Published: (2024)
WordRobe: Text-Guided Generation of Textured 3D Garments
by: Srivastava, Astitva, et al.
Published: (2024)
by: Srivastava, Astitva, et al.
Published: (2024)
Learning the 3D Fauna of the Web
by: Li, Zizhang, et al.
Published: (2024)
by: Li, Zizhang, et al.
Published: (2024)
SHINOBI: Shape and Illumination using Neural Object Decomposition via BRDF Optimization In-the-wild
by: Engelhardt, Andreas, et al.
Published: (2024)
by: Engelhardt, Andreas, et al.
Published: (2024)
SViM3D: Stable Video Material Diffusion for Single Image 3D Generation
by: Engelhardt, Andreas, et al.
Published: (2025)
by: Engelhardt, Andreas, et al.
Published: (2025)
CamCtrl3D: Single-Image Scene Exploration with Precise 3D Camera Control
by: Popov, Stefan, et al.
Published: (2025)
by: Popov, Stefan, et al.
Published: (2025)
ReLi3D: Relightable Multi-view 3D Reconstruction with Disentangled Illumination
by: Dihlmann, Jan-Niklas, et al.
Published: (2026)
by: Dihlmann, Jan-Niklas, et al.
Published: (2026)
SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
by: Huang, Zixuan, et al.
Published: (2025)
by: Huang, Zixuan, et al.
Published: (2025)
ICE-G: Image Conditional Editing of 3D Gaussian Splats
by: Jaganathan, Vishnu, et al.
Published: (2024)
by: Jaganathan, Vishnu, et al.
Published: (2024)
HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image Models
by: Ruiz, Nataniel, et al.
Published: (2023)
by: Ruiz, Nataniel, et al.
Published: (2023)
Dress-Me-Up: A Dataset & Method for Self-Supervised 3D Garment Retargeting
by: Naik, Shanthika, et al.
Published: (2024)
by: Naik, Shanthika, et al.
Published: (2024)
The Scene Language: Representing Scenes with Programs, Words, and Embeddings
by: Zhang, Yunzhi, et al.
Published: (2024)
by: Zhang, Yunzhi, et al.
Published: (2024)
ReSWD: ReSTIR'd, not shaken. Combining Reservoir Sampling and Sliced Wasserstein Distance for Variance Reduction
by: Boss, Mark, et al.
Published: (2025)
by: Boss, Mark, et al.
Published: (2025)
SF3D: Stable Fast 3D Mesh Reconstruction with UV-unwrapping and Illumination Disentanglement
by: Boss, Mark, et al.
Published: (2024)
by: Boss, Mark, et al.
Published: (2024)
Product of Experts for Visual Generation
by: Zhang, Yunzhi, et al.
Published: (2025)
by: Zhang, Yunzhi, et al.
Published: (2025)
PerpetualWonder: Long-Horizon Action-Conditioned 4D Scene Generation
by: Zhan, Jiahao, et al.
Published: (2026)
by: Zhan, Jiahao, et al.
Published: (2026)
FaceCraft4D: Animated 3D Facial Avatar Generation from a Single Image
by: Yin, Fei, et al.
Published: (2025)
by: Yin, Fei, et al.
Published: (2025)
HouseCrafter: Lifting Floorplans to 3D Scenes with 2D Diffusion Model
by: Nguyen, Hieu T., et al.
Published: (2024)
by: Nguyen, Hieu T., et al.
Published: (2024)
MVD-Fusion: Single-view 3D via Depth-consistent Multi-view Generation
by: Hu, Hanzhe, et al.
Published: (2024)
by: Hu, Hanzhe, et al.
Published: (2024)
Ponymation: Learning Articulated 3D Animal Motions from Unlabeled Online Videos
by: Sun, Keqiang, et al.
Published: (2023)
by: Sun, Keqiang, et al.
Published: (2023)
SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing
by: Li, Ruihuang, et al.
Published: (2024)
by: Li, Ruihuang, et al.
Published: (2024)
SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency
by: Xie, Yiming, et al.
Published: (2024)
by: Xie, Yiming, et al.
Published: (2024)
ConDense: Consistent 2D/3D Pre-training for Dense and Sparse Features from Multi-View Images
by: Zhang, Xiaoshuai, et al.
Published: (2024)
by: Zhang, Xiaoshuai, et al.
Published: (2024)
WonderPlay: Dynamic 3D Scene Generation from a Single Image and Actions
by: Li, Zizhang, et al.
Published: (2025)
by: Li, Zizhang, et al.
Published: (2025)
DecompDreamer: A Composition-Aware Curriculum for Structured 3D Asset Generation
by: Nath, Utkarsh, et al.
Published: (2025)
by: Nath, Utkarsh, et al.
Published: (2025)
HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models
by: Peng, Xiaogang, et al.
Published: (2023)
by: Peng, Xiaogang, et al.
Published: (2023)
HumANDiff: Articulated Noise Diffusion for Motion-Consistent Human Video Generation
by: Hu, Tao, et al.
Published: (2026)
by: Hu, Tao, et al.
Published: (2026)
Computational Tradeoffs in Image Synthesis: Diffusion, Masked-Token, and Next-Token Prediction
by: Kilian, Maciej, et al.
Published: (2024)
by: Kilian, Maciej, et al.
Published: (2024)
Human Video Generation from a Single Image with 3D Pose and View Control
by: Wang, Tiantian, et al.
Published: (2026)
by: Wang, Tiantian, et al.
Published: (2026)
SV3D: Novel Multi-view Synthesis and 3D Generation from a Single Image using Latent Video Diffusion
by: Voleti, Vikram, et al.
Published: (2024)
by: Voleti, Vikram, et al.
Published: (2024)
TripoSR: Fast 3D Object Reconstruction from a Single Image
by: Tochilkin, Dmitry, et al.
Published: (2024)
by: Tochilkin, Dmitry, et al.
Published: (2024)
LightHeadEd: Relightable & Editable Head Avatars from a Smartphone
by: Manu, Pranav, et al.
Published: (2025)
by: Manu, Pranav, et al.
Published: (2025)
CompGS: Unleashing 2D Compositionality for Compositional Text-to-3D via Dynamically Optimizing 3D Gaussians
by: Ge, Chongjian, et al.
Published: (2024)
by: Ge, Chongjian, et al.
Published: (2024)
Stanford-ORB: A Real-World 3D Object Inverse Rendering Benchmark
by: Kuang, Zhengfei, et al.
Published: (2023)
by: Kuang, Zhengfei, et al.
Published: (2023)
WildSeg3D: Segment Any 3D Objects in the Wild from 2D Images
by: Guo, Yansong, et al.
Published: (2025)
by: Guo, Yansong, et al.
Published: (2025)
WildDet3D: Scaling Promptable 3D Detection in the Wild
by: Huang, Weikai, et al.
Published: (2026)
by: Huang, Weikai, et al.
Published: (2026)
Gaussian in the Wild: 3D Gaussian Splatting for Unconstrained Image Collections
by: Zhang, Dongbin, et al.
Published: (2024)
by: Zhang, Dongbin, et al.
Published: (2024)
Detect Anything 3D in the Wild
by: Zhang, Hanxue, et al.
Published: (2025)
by: Zhang, Hanxue, et al.
Published: (2025)
Stable Part Diffusion 4D: Multi-View RGB and Kinematic Parts Video Generation
by: Zhang, Hao, et al.
Published: (2025)
by: Zhang, Hao, et al.
Published: (2025)
ZipLoRA: Any Subject in Any Style by Effectively Merging LoRAs
by: Shah, Viraj, et al.
Published: (2023)
by: Shah, Viraj, et al.
Published: (2023)
Similar Items
-
Probing the 3D Awareness of Visual Foundation Models
by: Banani, Mohamed El, et al.
Published: (2024) -
WordRobe: Text-Guided Generation of Textured 3D Garments
by: Srivastava, Astitva, et al.
Published: (2024) -
Learning the 3D Fauna of the Web
by: Li, Zizhang, et al.
Published: (2024) -
SHINOBI: Shape and Illumination using Neural Object Decomposition via BRDF Optimization In-the-wild
by: Engelhardt, Andreas, et al.
Published: (2024) -
SViM3D: Stable Video Material Diffusion for Single Image 3D Generation
by: Engelhardt, Andreas, et al.
Published: (2025)