CAT3D: Create Anything in 3D with Multi-View Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Ruiqi, Holynski, Aleksander, Henzler, Philipp, Brussee, Arthur, Martin-Brualla, Ricardo, Srinivasan, Pratul, Barron, Jonathan T., Poole, Ben |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bolt3D: Generating 3D Scenes in Seconds
by: Szymanowicz, Stanislaw, et al.
Published: (2025)
by: Szymanowicz, Stanislaw, et al.
Published: (2025)
CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models
by: Wu, Rundi, et al.
Published: (2024)
by: Wu, Rundi, et al.
Published: (2024)
IllumiNeRF: 3D Relighting Without Inverse Rendering
by: Zhao, Xiaoming, et al.
Published: (2024)
by: Zhao, Xiaoming, et al.
Published: (2024)
SimVS: Simulating World Inconsistencies for Robust View Synthesis
by: Trevithick, Alex, et al.
Published: (2024)
by: Trevithick, Alex, et al.
Published: (2024)
ROGR: Relightable 3D Objects using Generative Relighting
by: Tang, Jiapeng, et al.
Published: (2025)
by: Tang, Jiapeng, et al.
Published: (2025)
GR3EN: Generative Relighting for 3D Environments
by: Xing, Xiaoyan, et al.
Published: (2026)
by: Xing, Xiaoyan, et al.
Published: (2026)
Generative Multiview Relighting for 3D Reconstruction under Extreme Illumination Variation
by: Alzayer, Hadi, et al.
Published: (2024)
by: Alzayer, Hadi, et al.
Published: (2024)
Disentangled 3D Scene Generation with Layout Learning
by: Epstein, Dave, et al.
Published: (2024)
by: Epstein, Dave, et al.
Published: (2024)
ZipMap: Linear-Time Stateful 3D Reconstruction via Test-Time Training
by: Jin, Haian, et al.
Published: (2026)
by: Jin, Haian, et al.
Published: (2026)
VLIC: Vision-Language Models As Perceptual Judges for Human-Aligned Image Compression
by: Sargent, Kyle, et al.
Published: (2025)
by: Sargent, Kyle, et al.
Published: (2025)
Video Interpolation with Diffusion Models
by: Jain, Siddhant, et al.
Published: (2024)
by: Jain, Siddhant, et al.
Published: (2024)
NeRF-Casting: Improved View-Dependent Appearance with Consistent Reflections
by: Verbin, Dor, et al.
Published: (2024)
by: Verbin, Dor, et al.
Published: (2024)
WildCAT3D: Appearance-Aware Multi-View Diffusion in the Wild
by: Alper, Morris, et al.
Published: (2025)
by: Alper, Morris, et al.
Published: (2025)
Generative Powers of Ten
by: Wang, Xiaojuan, et al.
Published: (2023)
by: Wang, Xiaojuan, et al.
Published: (2023)
Can Generative Video Models Help Pose Estimation?
by: Cai, Ruojin, et al.
Published: (2024)
by: Cai, Ruojin, et al.
Published: (2024)
MVPaint: Synchronized Multi-View Diffusion for Painting Anything 3D
by: Cheng, Wei, et al.
Published: (2024)
by: Cheng, Wei, et al.
Published: (2024)
LFM-3D: Learnable Feature Matching Across Wide Baselines Using 3D Signals
by: Karpur, Arjun, et al.
Published: (2023)
by: Karpur, Arjun, et al.
Published: (2023)
Perturb-and-Revise: Flexible 3D Editing with Generative Trajectories
by: Hong, Susung, et al.
Published: (2024)
by: Hong, Susung, et al.
Published: (2024)
Continuous 3D Perception Model with Persistent State
by: Wang, Qianqian, et al.
Published: (2025)
by: Wang, Qianqian, et al.
Published: (2025)
Binary Opacity Grids: Capturing Fine Geometric Detail for Mesh-Based View Synthesis
by: Reiser, Christian, et al.
Published: (2024)
by: Reiser, Christian, et al.
Published: (2024)
Stereo4D: Learning How Things Move in 3D from Internet Stereo Videos
by: Jin, Linyi, et al.
Published: (2024)
by: Jin, Linyi, et al.
Published: (2024)
Revealing the 3D Cosmic Web through Gravitationally Constrained Neural Fields
by: Zhao, Brandon, et al.
Published: (2025)
by: Zhao, Brandon, et al.
Published: (2025)
Flash Cache: Reducing Bias in Radiance Cache Based Inverse Rendering
by: Attal, Benjamin, et al.
Published: (2024)
by: Attal, Benjamin, et al.
Published: (2024)
Dual-Process Image Generation
by: Luo, Grace, et al.
Published: (2025)
by: Luo, Grace, et al.
Published: (2025)
Fourier Feature Pyramids for Physics-Informed Neural Networks
by: Zhao, Brandon, et al.
Published: (2026)
by: Zhao, Brandon, et al.
Published: (2026)
Infinite Texture: Text-guided High Resolution Diffusion Texture Synthesis
by: Wang, Yifan, et al.
Published: (2024)
by: Wang, Yifan, et al.
Published: (2024)
Diffusion Models as Data Mining Tools
by: Siglidis, Ioannis, et al.
Published: (2024)
by: Siglidis, Ioannis, et al.
Published: (2024)
ExtraNeRF: Visibility-Aware View Extrapolation of Neural Radiance Fields with Diffusion Models
by: Shih, Meng-Li, et al.
Published: (2024)
by: Shih, Meng-Li, et al.
Published: (2024)
Readout Guidance: Learning Control from Diffusion Features
by: Luo, Grace, et al.
Published: (2023)
by: Luo, Grace, et al.
Published: (2023)
3DEnhancer: Consistent Multi-View Diffusion for 3D Enhancement
by: Luo, Yihang, et al.
Published: (2024)
by: Luo, Yihang, et al.
Published: (2024)
Diffusion Hyperfeatures: Searching Through Time and Space for Semantic Correspondence
by: Luo, Grace, et al.
Published: (2023)
by: Luo, Grace, et al.
Published: (2023)
CAP4D: Creating Animatable 4D Portrait Avatars with Morphable Multi-View Diffusion Models
by: Taubner, Felix, et al.
Published: (2024)
by: Taubner, Felix, et al.
Published: (2024)
DreamEdit3D: Personalization of Multi-View Diffusion Models for 3D Editing
by: Ai, Jinxin, et al.
Published: (2026)
by: Ai, Jinxin, et al.
Published: (2026)
The new era of Eurocapitalism
by: Henzler, Herbert
Published: (1992)
by: Henzler, Herbert
Published: (1992)
3D-Adapter: Geometry-Consistent Multi-View Diffusion for High-Quality 3D Generation
by: Chen, Hansheng, et al.
Published: (2024)
by: Chen, Hansheng, et al.
Published: (2024)
Detect Anything 3D in the Wild
by: Zhang, Hanxue, et al.
Published: (2025)
by: Zhang, Hanxue, et al.
Published: (2025)
Generative Image Dynamics
by: Li, Zhengqi, et al.
Published: (2023)
by: Li, Zhengqi, et al.
Published: (2023)
Gaussian Grouping: Segment and Edit Anything in 3D Scenes
by: Ye, Mingqiao, et al.
Published: (2023)
by: Ye, Mingqiao, et al.
Published: (2023)
Material Anything: Generating Materials for Any 3D Object via Diffusion
by: Huang, Xin, et al.
Published: (2024)
by: Huang, Xin, et al.
Published: (2024)
HiScene: Creating Hierarchical 3D Scenes with Isometric View Generation
by: Dong, Wenqi, et al.
Published: (2025)
by: Dong, Wenqi, et al.
Published: (2025)
Similar Items
-
Bolt3D: Generating 3D Scenes in Seconds
by: Szymanowicz, Stanislaw, et al.
Published: (2025) -
CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models
by: Wu, Rundi, et al.
Published: (2024) -
IllumiNeRF: 3D Relighting Without Inverse Rendering
by: Zhao, Xiaoming, et al.
Published: (2024) -
SimVS: Simulating World Inconsistencies for Robust View Synthesis
by: Trevithick, Alex, et al.
Published: (2024) -
ROGR: Relightable 3D Objects using Generative Relighting
by: Tang, Jiapeng, et al.
Published: (2025)