Sculpt4D: Generating 4D Shapes via Sparse-Attention Diffusion Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Yin, Minghao, Hu, Wenbo, Xu, Jiale, Shan, Ying, Han, Kai |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Splat4D: Diffusion-Enhanced 4D Gaussian Splatting for Temporally and Spatially Consistent Content Creation
by: Yin, Minghao, et al.
Published: (2025)
by: Yin, Minghao, et al.
Published: (2025)
VerseCrafter: Dynamic Realistic Video World Model with 4D Geometric Control
by: Zheng, Sixiao, et al.
Published: (2026)
by: Zheng, Sixiao, et al.
Published: (2026)
FreeSplatter: Pose-free Gaussian Splatting for Sparse-view 3D Reconstruction
by: Xu, Jiale, et al.
Published: (2024)
by: Xu, Jiale, et al.
Published: (2024)
Wukong's 72 Transformations: High-fidelity Textured 3D Morphing via Flow Models
by: Yin, Minghao, et al.
Published: (2025)
by: Yin, Minghao, et al.
Published: (2025)
Rolling Forcing: Autoregressive Long Video Diffusion in Real Time
by: Liu, Kunhao, et al.
Published: (2025)
by: Liu, Kunhao, et al.
Published: (2025)
TopoSculpt: Betti-Steered Topological Sculpting of 3D Fine-grained Tubular Shapes
by: Zhang, Minghui, et al.
Published: (2025)
by: Zhang, Minghui, et al.
Published: (2025)
Assembler: Scalable 3D Part Assembly via Anchor Point Diffusion
by: Zhao, Wang, et al.
Published: (2025)
by: Zhao, Wang, et al.
Published: (2025)
Sculpt3D: Multi-View Consistent Text-to-3D Generation with Sparse 3D Prior
by: Chen, Cheng, et al.
Published: (2024)
by: Chen, Cheng, et al.
Published: (2024)
TextMesh4D: Text-to-4D Mesh Generation via Jacobian Deformation Field
by: Dai, Sisi, et al.
Published: (2025)
by: Dai, Sisi, et al.
Published: (2025)
InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models
by: Xu, Jiale, et al.
Published: (2024)
by: Xu, Jiale, et al.
Published: (2024)
Diffusion4D: Fast Spatial-temporal Consistent 4D Generation via Video Diffusion Models
by: Liang, Hanwen, et al.
Published: (2024)
by: Liang, Hanwen, et al.
Published: (2024)
Sparse4DGS: 4D Gaussian Splatting for Sparse-Frame Dynamic Scene Reconstruction
by: Shi, Changyue, et al.
Published: (2025)
by: Shi, Changyue, et al.
Published: (2025)
NVComposer: Boosting Generative Novel View Synthesis with Multiple Sparse and Unposed Images
by: Li, Lingen, et al.
Published: (2024)
by: Li, Lingen, et al.
Published: (2024)
Track4World: Feedforward World-centric Dense 3D Tracking of All Pixels
by: Lu, Jiahao, et al.
Published: (2026)
by: Lu, Jiahao, et al.
Published: (2026)
Veda: Scalable Video Diffusion via Distilled Sparse Attention
by: Han, Shihao, et al.
Published: (2026)
by: Han, Shihao, et al.
Published: (2026)
Part-aware Shape Generation with Latent 3D Diffusion of Neural Voxel Fields
by: Huang, Yuhang, et al.
Published: (2024)
by: Huang, Yuhang, et al.
Published: (2024)
ShapeGen4D: Towards High Quality 4D Shape Generation from Videos
by: Yenphraphai, Jiraphon, et al.
Published: (2025)
by: Yenphraphai, Jiraphon, et al.
Published: (2025)
InfoSculpt: Sculpting the Latent Space for Generalized Category Discovery
by: Liao, Wenwen, et al.
Published: (2026)
by: Liao, Wenwen, et al.
Published: (2026)
4DVGGT-D: 4D Visual Geometry Transformer with Improved Dynamic Depth Estimation
by: Zang, Ying, et al.
Published: (2026)
by: Zang, Ying, et al.
Published: (2026)
Learning Neural Deformation Representation for 4D Dynamic Shape Generation
by: Han, Gyojin, et al.
Published: (2026)
by: Han, Gyojin, et al.
Published: (2026)
4Real: Towards Photorealistic 4D Scene Generation via Video Diffusion Models
by: Yu, Heng, et al.
Published: (2024)
by: Yu, Heng, et al.
Published: (2024)
Sparse-Up: Learnable Sparse Upsampling for 3D Generation with High-Fidelity Textures
by: Xiao, Lu, et al.
Published: (2025)
by: Xiao, Lu, et al.
Published: (2025)
DI-PCG: Diffusion-based Efficient Inverse Procedural Content Generation for High-quality 3D Asset Creation
by: Zhao, Wang, et al.
Published: (2024)
by: Zhao, Wang, et al.
Published: (2024)
SC4D: Sparse-Controlled Video-to-4D Generation and Motion Transfer
by: Wu, Zijie, et al.
Published: (2024)
by: Wu, Zijie, et al.
Published: (2024)
Pixal3D: Pixel-Aligned 3D Generation from Images
by: Li, Dong-Yang, et al.
Published: (2026)
by: Li, Dong-Yang, et al.
Published: (2026)
Human4DiT: 360-degree Human Video Generation with 4D Diffusion Transformer
by: Shao, Ruizhi, et al.
Published: (2024)
by: Shao, Ruizhi, et al.
Published: (2024)
En3D: An Enhanced Generative Model for Sculpting 3D Humans from 2D Synthetic Data
by: Men, Yifang, et al.
Published: (2024)
by: Men, Yifang, et al.
Published: (2024)
MotionCrafter: Dense Geometry and Motion Reconstruction with a 4D VAE
by: Zhu, Ruijie, et al.
Published: (2026)
by: Zhu, Ruijie, et al.
Published: (2026)
Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models
by: Jin, Yudong, et al.
Published: (2025)
by: Jin, Yudong, et al.
Published: (2025)
Texture-GS: Disentangling the Geometry and Texture for 3D Gaussian Splatting Editing
by: Xu, Tian-Xing, et al.
Published: (2024)
by: Xu, Tian-Xing, et al.
Published: (2024)
Interspatial Attention for Efficient 4D Human Video Generation
by: Shao, Ruizhi, et al.
Published: (2025)
by: Shao, Ruizhi, et al.
Published: (2025)
SALAD: Achieve High-Sparsity Attention via Efficient Linear Attention Tuning for Video Diffusion Transformer
by: Fang, Tongcheng, et al.
Published: (2026)
by: Fang, Tongcheng, et al.
Published: (2026)
Efficient Camera-Controlled Video Generation of Static Scenes via Sparse Diffusion and 3D Rendering
by: Chen, Jieying, et al.
Published: (2026)
by: Chen, Jieying, et al.
Published: (2026)
TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models
by: YU, Mark, et al.
Published: (2025)
by: YU, Mark, et al.
Published: (2025)
ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices
by: Du, Kunpeng, et al.
Published: (2026)
by: Du, Kunpeng, et al.
Published: (2026)
SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse Cameras
by: Pan, Weihong, et al.
Published: (2026)
by: Pan, Weihong, et al.
Published: (2026)
AnchorHOI: Zero-shot Generation of 4D Human-Object Interaction via Anchor-based Prior Distillation
by: Dai, Sisi, et al.
Published: (2025)
by: Dai, Sisi, et al.
Published: (2025)
Robust 4D Visual Geometry Transformer with Uncertainty-Aware Priors
by: Zang, Ying, et al.
Published: (2026)
by: Zang, Ying, et al.
Published: (2026)
Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer
by: Wu, Shuang, et al.
Published: (2024)
by: Wu, Shuang, et al.
Published: (2024)
4Diffusion: Multi-view Video Diffusion Model for 4D Generation
by: Zhang, Haiyu, et al.
Published: (2024)
by: Zhang, Haiyu, et al.
Published: (2024)
Similar Items
-
Splat4D: Diffusion-Enhanced 4D Gaussian Splatting for Temporally and Spatially Consistent Content Creation
by: Yin, Minghao, et al.
Published: (2025) -
VerseCrafter: Dynamic Realistic Video World Model with 4D Geometric Control
by: Zheng, Sixiao, et al.
Published: (2026) -
FreeSplatter: Pose-free Gaussian Splatting for Sparse-view 3D Reconstruction
by: Xu, Jiale, et al.
Published: (2024) -
Wukong's 72 Transformations: High-fidelity Textured 3D Morphing via Flow Models
by: Yin, Minghao, et al.
Published: (2025) -
Rolling Forcing: Autoregressive Long Video Diffusion in Real Time
by: Liu, Kunhao, et al.
Published: (2025)