CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Rundi, Gao, Ruiqi, Poole, Ben, Trevithick, Alex, Zheng, Changxi, Barron, Jonathan T., Holynski, Aleksander |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CAT3D: Create Anything in 3D with Multi-View Diffusion Models
by: Gao, Ruiqi, et al.
Published: (2024)
by: Gao, Ruiqi, et al.
Published: (2024)
SimVS: Simulating World Inconsistencies for Robust View Synthesis
by: Trevithick, Alex, et al.
Published: (2024)
by: Trevithick, Alex, et al.
Published: (2024)
ZipMap: Linear-Time Stateful 3D Reconstruction via Test-Time Training
by: Jin, Haian, et al.
Published: (2026)
by: Jin, Haian, et al.
Published: (2026)
Video Interpolation with Diffusion Models
by: Jain, Siddhant, et al.
Published: (2024)
by: Jain, Siddhant, et al.
Published: (2024)
Spatiotemporally Consistent Indoor Lighting Estimation with Diffusion Priors
by: Tong, Mutian, et al.
Published: (2025)
by: Tong, Mutian, et al.
Published: (2025)
Sin3DM: Learning a Diffusion Model from a Single 3D Textured Shape
by: Wu, Rundi, et al.
Published: (2023)
by: Wu, Rundi, et al.
Published: (2023)
Disentangled 3D Scene Generation with Layout Learning
by: Epstein, Dave, et al.
Published: (2024)
by: Epstein, Dave, et al.
Published: (2024)
Bolt3D: Generating 3D Scenes in Seconds
by: Szymanowicz, Stanislaw, et al.
Published: (2025)
by: Szymanowicz, Stanislaw, et al.
Published: (2025)
Stereo4D: Learning How Things Move in 3D from Internet Stereo Videos
by: Jin, Linyi, et al.
Published: (2024)
by: Jin, Linyi, et al.
Published: (2024)
PhysDreamer: Physics-Based Interaction with 3D Objects via Video Generation
by: Zhang, Tianyuan, et al.
Published: (2024)
by: Zhang, Tianyuan, et al.
Published: (2024)
CAP4D: Creating Animatable 4D Portrait Avatars with Morphable Multi-View Diffusion Models
by: Taubner, Felix, et al.
Published: (2024)
by: Taubner, Felix, et al.
Published: (2024)
WildCAT3D: Appearance-Aware Multi-View Diffusion in the Wild
by: Alper, Morris, et al.
Published: (2025)
by: Alper, Morris, et al.
Published: (2025)
RealmDreamer: Text-Driven 3D Scene Generation with Inpainting and Depth Diffusion
by: Shriram, Jaidev, et al.
Published: (2024)
by: Shriram, Jaidev, et al.
Published: (2024)
MVPaint: Synchronized Multi-View Diffusion for Painting Anything 3D
by: Cheng, Wei, et al.
Published: (2024)
by: Cheng, Wei, et al.
Published: (2024)
DetAny4D: Detect Anything 4D Temporally in a Streaming RGB Video
by: Hou, Jiawei, et al.
Published: (2025)
by: Hou, Jiawei, et al.
Published: (2025)
MVP4D: Multi-View Portrait Video Diffusion for Animatable 4D Avatars
by: Taubner, Felix, et al.
Published: (2025)
by: Taubner, Felix, et al.
Published: (2025)
Portrait4D-v2: Pseudo Multi-View Data Creates Better 4D Head Synthesizer
by: Deng, Yu, et al.
Published: (2024)
by: Deng, Yu, et al.
Published: (2024)
Continuous 3D Perception Model with Persistent State
by: Wang, Qianqian, et al.
Published: (2025)
by: Wang, Qianqian, et al.
Published: (2025)
Stable Part Diffusion 4D: Multi-View RGB and Kinematic Parts Video Generation
by: Zhang, Hao, et al.
Published: (2025)
by: Zhang, Hao, et al.
Published: (2025)
Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models
by: Jin, Yudong, et al.
Published: (2025)
by: Jin, Yudong, et al.
Published: (2025)
4Diffusion: Multi-view Video Diffusion Model for 4D Generation
by: Zhang, Haiyu, et al.
Published: (2024)
by: Zhang, Haiyu, et al.
Published: (2024)
Infinite Texture: Text-guided High Resolution Diffusion Texture Synthesis
by: Wang, Yifan, et al.
Published: (2024)
by: Wang, Yifan, et al.
Published: (2024)
SV4D 2.0: Enhancing Spatio-Temporal Consistency in Multi-View Video Diffusion for High-Quality 4D Generation
by: Yao, Chun-Han, et al.
Published: (2025)
by: Yao, Chun-Han, et al.
Published: (2025)
ExtraNeRF: Visibility-Aware View Extrapolation of Neural Radiance Fields with Diffusion Models
by: Shih, Meng-Li, et al.
Published: (2024)
by: Shih, Meng-Li, et al.
Published: (2024)
ViPS: Video-informed Pose Spaces for Auto-Rigged Meshes
by: Chen, Honglin, et al.
Published: (2026)
by: Chen, Honglin, et al.
Published: (2026)
VLIC: Vision-Language Models As Perceptual Judges for Human-Aligned Image Compression
by: Sargent, Kyle, et al.
Published: (2025)
by: Sargent, Kyle, et al.
Published: (2025)
Readout Guidance: Learning Control from Diffusion Features
by: Luo, Grace, et al.
Published: (2023)
by: Luo, Grace, et al.
Published: (2023)
Generative Camera Dolly: Extreme Monocular Dynamic Novel View Synthesis
by: Van Hoorick, Basile, et al.
Published: (2024)
by: Van Hoorick, Basile, et al.
Published: (2024)
Dual-Process Image Generation
by: Luo, Grace, et al.
Published: (2025)
by: Luo, Grace, et al.
Published: (2025)
VLMaterial: Procedural Material Generation with Large Vision-Language Models
by: Li, Beichen, et al.
Published: (2025)
by: Li, Beichen, et al.
Published: (2025)
4D-CAT: Synthesis of 4D Coronary Artery Trees from Systole and Diastole
by: Hu, Daosong, et al.
Published: (2024)
by: Hu, Daosong, et al.
Published: (2024)
Diffusion Hyperfeatures: Searching Through Time and Space for Semantic Correspondence
by: Luo, Grace, et al.
Published: (2023)
by: Luo, Grace, et al.
Published: (2023)
Trace Anything: Representing Any Video in 4D via Trajectory Fields
by: Liu, Xinhang, et al.
Published: (2025)
by: Liu, Xinhang, et al.
Published: (2025)
Diffusion Models as Data Mining Tools
by: Siglidis, Ioannis, et al.
Published: (2024)
by: Siglidis, Ioannis, et al.
Published: (2024)
Place Anything into Any Video
by: Liu, Ziling, et al.
Published: (2024)
by: Liu, Ziling, et al.
Published: (2024)
Generative Image Dynamics
by: Li, Zhengqi, et al.
Published: (2023)
by: Li, Zhengqi, et al.
Published: (2023)
MultiEgo: A Multi-View Egocentric Video Dataset for 4D Scene Reconstruction
by: Li, Bate, et al.
Published: (2025)
by: Li, Bate, et al.
Published: (2025)
AdaViewPlanner: Adapting Video Diffusion Models for Viewpoint Planning in 4D Scenes
by: Li, Yu, et al.
Published: (2025)
by: Li, Yu, et al.
Published: (2025)
Create Anything Anywhere: Layout-Controllable Personalized Diffusion Model for Multiple Subjects
by: Li, Wei, et al.
Published: (2025)
by: Li, Wei, et al.
Published: (2025)
Can Video Diffusion Model Reconstruct 4D Geometry?
by: Mai, Jinjie, et al.
Published: (2025)
by: Mai, Jinjie, et al.
Published: (2025)
Similar Items
-
CAT3D: Create Anything in 3D with Multi-View Diffusion Models
by: Gao, Ruiqi, et al.
Published: (2024) -
SimVS: Simulating World Inconsistencies for Robust View Synthesis
by: Trevithick, Alex, et al.
Published: (2024) -
ZipMap: Linear-Time Stateful 3D Reconstruction via Test-Time Training
by: Jin, Haian, et al.
Published: (2026) -
Video Interpolation with Diffusion Models
by: Jain, Siddhant, et al.
Published: (2024) -
Spatiotemporally Consistent Indoor Lighting Estimation with Diffusion Priors
by: Tong, Mutian, et al.
Published: (2025)