Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Jin, Yudong, Peng, Sida, Wang, Xuan, Xie, Tao, Xu, Zhen, Yang, Yifan, Shen, Yujun, Bao, Hujun, Zhou, Xiaowei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EnvGS: Modeling View-Dependent Appearance with Environment Gaussian
by: Xie, Tao, et al.
Published: (2024)
by: Xie, Tao, et al.
Published: (2024)
StreetCrafter: Street View Synthesis with Controllable Video Diffusion Models
by: Yan, Yunzhi, et al.
Published: (2024)
by: Yan, Yunzhi, et al.
Published: (2024)
SV4D 2.0: Enhancing Spatio-Temporal Consistency in Multi-View Video Diffusion for High-Quality 4D Generation
by: Yao, Chun-Han, et al.
Published: (2025)
by: Yao, Chun-Han, et al.
Published: (2025)
Split4D: Decomposed 4D Scene Reconstruction Without Video Segmentation
by: Hu, Yongzhen, et al.
Published: (2025)
by: Hu, Yongzhen, et al.
Published: (2025)
World-Grounded Human Motion Recovery via Gravity-View Coordinates
by: Shen, Zehong, et al.
Published: (2024)
by: Shen, Zehong, et al.
Published: (2024)
Representing Long Volumetric Video with Temporal Gaussian Hierarchy
by: Xu, Zhen, et al.
Published: (2024)
by: Xu, Zhen, et al.
Published: (2024)
Binocular-Guided 3D Gaussian Splatting with View Consistency for Sparse View Synthesis
by: Han, Liang, et al.
Published: (2024)
by: Han, Liang, et al.
Published: (2024)
UniVerse: Unleashing the Scene Prior of Video Diffusion Models for Robust Radiance Field Reconstruction
by: Cao, Jin, et al.
Published: (2025)
by: Cao, Jin, et al.
Published: (2025)
GaussianPrediction: Dynamic 3D Gaussian Prediction for Motion Extrapolation and Free View Synthesis
by: Zhao, Boming, et al.
Published: (2024)
by: Zhao, Boming, et al.
Published: (2024)
SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse Cameras
by: Pan, Weihong, et al.
Published: (2026)
by: Pan, Weihong, et al.
Published: (2026)
Geometry-Consistent 4D Gaussian Splatting for Sparse-Input Dynamic View Synthesis
by: Li, Yiwei, et al.
Published: (2025)
by: Li, Yiwei, et al.
Published: (2025)
Generating Human Motion in 3D Scenes from Text Descriptions
by: Cen, Zhi, et al.
Published: (2024)
by: Cen, Zhi, et al.
Published: (2024)
Multi-View Video Diffusion Policy: A 3D Spatio-Temporal-Aware Video Action Model
by: Li, Peiyan, et al.
Published: (2026)
by: Li, Peiyan, et al.
Published: (2026)
MaPa: Text-driven Photorealistic Material Painting for 3D Shapes
by: Zhang, Shangzan, et al.
Published: (2024)
by: Zhang, Shangzan, et al.
Published: (2024)
Geometry-guided Online 3D Video Synthesis with Multi-View Temporal Consistency
by: Ha, Hyunho, et al.
Published: (2025)
by: Ha, Hyunho, et al.
Published: (2025)
StreamingTalker: Audio-driven 3D Facial Animation with Autoregressive Diffusion Model
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
View-Consistent Diffusion Representations for 3D-Consistent Video Generation
by: Danier, Duolikun, et al.
Published: (2025)
by: Danier, Duolikun, et al.
Published: (2025)
OmniView: An All-Seeing Diffusion Model for 3D and 4D View Synthesis
by: Fan, Xiang, et al.
Published: (2025)
by: Fan, Xiang, et al.
Published: (2025)
From Sparse to Dense: Spatio-Temporal Fusion for Multi-View 3D Human Pose Estimation with DenseWarper
by: Li, Ling, et al.
Published: (2026)
by: Li, Ling, et al.
Published: (2026)
Ready-to-React: Online Reaction Policy for Two-Character Interaction Generation
by: Cen, Zhi, et al.
Published: (2025)
by: Cen, Zhi, et al.
Published: (2025)
PointGS: Point Attention-Aware Sparse View Synthesis with Gaussian Splatting
by: Xiang, Lintao, et al.
Published: (2025)
by: Xiang, Lintao, et al.
Published: (2025)
CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models
by: Wu, Rundi, et al.
Published: (2024)
by: Wu, Rundi, et al.
Published: (2024)
MVP4D: Multi-View Portrait Video Diffusion for Animatable 4D Avatars
by: Taubner, Felix, et al.
Published: (2025)
by: Taubner, Felix, et al.
Published: (2025)
Painting 3D Nature in 2D: View Synthesis of Natural Scenes from a Single Semantic Mask
by: Zhang, Shangzan, et al.
Published: (2023)
by: Zhang, Shangzan, et al.
Published: (2023)
CFSynthesis: Controllable and Free-view 3D Human Video Synthesis
by: Cui, Liyuan, et al.
Published: (2024)
by: Cui, Liyuan, et al.
Published: (2024)
MoVieS: Motion-Aware 4D Dynamic View Synthesis in One Second
by: Lin, Chenguo, et al.
Published: (2025)
by: Lin, Chenguo, et al.
Published: (2025)
ViewFusion: Structured Spatial Thinking Chains for Multi-View Reasoning
by: Tao, Xingjian, et al.
Published: (2026)
by: Tao, Xingjian, et al.
Published: (2026)
HiScene: Creating Hierarchical 3D Scenes with Isometric View Generation
by: Dong, Wenqi, et al.
Published: (2025)
by: Dong, Wenqi, et al.
Published: (2025)
MetaOcc: Spatio-Temporal Fusion of Surround-View 4D Radar and Camera for 3D Occupancy Prediction with Dual Training Strategies
by: Yang, Long, et al.
Published: (2025)
by: Yang, Long, et al.
Published: (2025)
SparseGS-W: Sparse-View 3D Gaussian Splatting in the Wild with Generative Priors
by: Li, Yiqing, et al.
Published: (2025)
by: Li, Yiqing, et al.
Published: (2025)
ViewCraft3D: High-Fidelity and View-Consistent 3D Vector Graphics Synthesis
by: Wang, Chuang, et al.
Published: (2025)
by: Wang, Chuang, et al.
Published: (2025)
4D-GSW: Kinematic-Aware Spatio-Temporal Consistent Watermarking for 4D Gaussian Splatting
by: Zhou, Sifan, et al.
Published: (2026)
by: Zhou, Sifan, et al.
Published: (2026)
STP4D: Spatio-Temporal-Prompt Consistent Modeling for Text-to-4D Gaussian Splatting
by: Deng, Yunze, et al.
Published: (2025)
by: Deng, Yunze, et al.
Published: (2025)
SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency
by: Xie, Yiming, et al.
Published: (2024)
by: Xie, Yiming, et al.
Published: (2024)
STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
PSF-4D: A Progressive Sampling Framework for View Consistent 4D Editing
by: Iqbal, Hasan, et al.
Published: (2025)
by: Iqbal, Hasan, et al.
Published: (2025)
FLARE: Feed-forward Geometry, Appearance and Camera Estimation from Uncalibrated Sparse Views
by: Zhang, Shangzhan, et al.
Published: (2025)
by: Zhang, Shangzhan, et al.
Published: (2025)
SpatialTracker: Tracking Any 2D Pixels in 3D Space
by: Xiao, Yuxi, et al.
Published: (2024)
by: Xiao, Yuxi, et al.
Published: (2024)
Generalizable Human Gaussians for Sparse View Synthesis
by: Kwon, Youngjoong, et al.
Published: (2024)
by: Kwon, Youngjoong, et al.
Published: (2024)
Tinker: Diffusion's Gift to 3D--Multi-View Consistent Editing From Sparse Inputs without Per-Scene Optimization
by: Zhao, Canyu, et al.
Published: (2025)
by: Zhao, Canyu, et al.
Published: (2025)
Similar Items
-
EnvGS: Modeling View-Dependent Appearance with Environment Gaussian
by: Xie, Tao, et al.
Published: (2024) -
StreetCrafter: Street View Synthesis with Controllable Video Diffusion Models
by: Yan, Yunzhi, et al.
Published: (2024) -
SV4D 2.0: Enhancing Spatio-Temporal Consistency in Multi-View Video Diffusion for High-Quality 4D Generation
by: Yao, Chun-Han, et al.
Published: (2025) -
Split4D: Decomposed 4D Scene Reconstruction Without Video Segmentation
by: Hu, Yongzhen, et al.
Published: (2025) -
World-Grounded Human Motion Recovery via Gravity-View Coordinates
by: Shen, Zehong, et al.
Published: (2024)