SS4D: Native 4D Generative Model via Structured Spacetime Latents
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Zhibing, Zhang, Mengchen, Wu, Tong, Tan, Jing, Wang, Jiaqi, Lin, Dahua |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
IDArb: Intrinsic Decomposition for Arbitrary Number of Input Views and Illuminations
by: Li, Zhibing, et al.
Published: (2024)
by: Li, Zhibing, et al.
Published: (2024)
LayerPano3D: Layered 3D Panorama for Hyper-Immersive Scene Generation
by: Yang, Shuai, et al.
Published: (2024)
by: Yang, Shuai, et al.
Published: (2024)
GPT-4V(ision) is a Human-Aligned Evaluator for Text-to-3D Generation
by: Wu, Tong, et al.
Published: (2024)
by: Wu, Tong, et al.
Published: (2024)
3DGen-Bench: Comprehensive Benchmark Suite for 3D Generative Models
by: Zhang, Yuhan, et al.
Published: (2025)
by: Zhang, Yuhan, et al.
Published: (2025)
GenDoP: Auto-regressive Camera Trajectory Generation as a Director of Photography
by: Zhang, Mengchen, et al.
Published: (2025)
by: Zhang, Mengchen, et al.
Published: (2025)
Tailor3D: Customized 3D Assets Editing and Generation with Dual-Side Images
by: Qi, Zhangyang, et al.
Published: (2024)
by: Qi, Zhangyang, et al.
Published: (2024)
Hi3DEval: Advancing 3D Generation Evaluation with Hierarchical Validity
by: Zhang, Yuhan, et al.
Published: (2025)
by: Zhang, Yuhan, et al.
Published: (2025)
Omni6D: Large-Vocabulary 3D Object Dataset for Category-Level 6D Object Pose Estimation
by: Zhang, Mengchen, et al.
Published: (2024)
by: Zhang, Mengchen, et al.
Published: (2024)
ViSAudio: End-to-End Video-Driven Binaural Spatial Audio Generation
by: Zhang, Mengchen, et al.
Published: (2025)
by: Zhang, Mengchen, et al.
Published: (2025)
Native and Compact Structured Latents for 3D Generation
by: Xiang, Jianfeng, et al.
Published: (2025)
by: Xiang, Jianfeng, et al.
Published: (2025)
FiVA: Fine-grained Visual Attribute Dataset for Text-to-Image Diffusion Models
by: Wu, Tong, et al.
Published: (2024)
by: Wu, Tong, et al.
Published: (2024)
GPT4Point: A Unified Framework for Point-Language Understanding and Generation
by: Qi, Zhangyang, et al.
Published: (2023)
by: Qi, Zhangyang, et al.
Published: (2023)
SepRep-Net: Multi-source Free Domain Adaptation via Model Separation And Reparameterization
by: Jin, Ying, et al.
Published: (2024)
by: Jin, Ying, et al.
Published: (2024)
Diff4Splat: Controllable 4D Scene Generation with Latent Dynamic Reconstruction Models
by: Pan, Panwang, et al.
Published: (2025)
by: Pan, Panwang, et al.
Published: (2025)
4D Gaussian Splatting: Modeling Dynamic Scenes with Native 4D Primitives
by: Yang, Zeyu, et al.
Published: (2024)
by: Yang, Zeyu, et al.
Published: (2024)
Structured 3D Latents for Scalable and Versatile 3D Generation
by: Xiang, Jianfeng, et al.
Published: (2024)
by: Xiang, Jianfeng, et al.
Published: (2024)
MORPHOS: Autoregressive 4D Generation with Temporal Structured Latents
by: Kwon, Minkyung, et al.
Published: (2026)
by: Kwon, Minkyung, et al.
Published: (2026)
Improving Viewpoint Consistency in 3D Generation via Structure Feature and CLIP Guidance
by: Zhang, Qing, et al.
Published: (2024)
by: Zhang, Qing, et al.
Published: (2024)
Hitem3D 2.0: Multi-View Guided Native 3D Texture Generation
by: He, Huiang, et al.
Published: (2026)
by: He, Huiang, et al.
Published: (2026)
Make-it-Real: Unleashing Large Multimodal Model for Painting 3D Objects with Realistic Materials
by: Fang, Ye, et al.
Published: (2024)
by: Fang, Ye, et al.
Published: (2024)
Imagine360: Immersive 360 Video Generation from Perspective Anchor
by: Tan, Jing, et al.
Published: (2024)
by: Tan, Jing, et al.
Published: (2024)
SK-Adapter: Skeleton-Based Structural Control for Native 3D Generation
by: Wang, Anbang, et al.
Published: (2026)
by: Wang, Anbang, et al.
Published: (2026)
GeoMotion: Rethinking Motion Segmentation via Latent 4D Geometry
by: He, Xiankang, et al.
Published: (2026)
by: He, Xiankang, et al.
Published: (2026)
TEXTRIX: Latent Attribute Grid for Native Texture Generation and Beyond
by: Zeng, Yifei, et al.
Published: (2025)
by: Zeng, Yifei, et al.
Published: (2025)
Structural Energy Guidance for View-Consistent Text-to-3D Generation
by: Zhang, Qing, et al.
Published: (2026)
by: Zhang, Qing, et al.
Published: (2026)
Learning Coherent Matrixized Representation in Latent Space for Volumetric 4D Generation
by: Yang, Qitong, et al.
Published: (2024)
by: Yang, Qitong, et al.
Published: (2024)
ByTheWay: Boost Your Text-to-Video Generation Model to Higher Quality in a Training-free Way
by: Bu, Jiazi, et al.
Published: (2024)
by: Bu, Jiazi, et al.
Published: (2024)
4D-RaDiff: Latent Diffusion for 4D Radar Point Cloud Generation
by: Kwok, Jimmie, et al.
Published: (2025)
by: Kwok, Jimmie, et al.
Published: (2025)
LaFiTe: A Generative Latent Field for 3D Native Texturing
by: Chen, Chia-Hao, et al.
Published: (2025)
by: Chen, Chia-Hao, et al.
Published: (2025)
Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer
by: Wu, Shuang, et al.
Published: (2024)
by: Wu, Shuang, et al.
Published: (2024)
HyperHuman: Hyper-Realistic Human Generation with Latent Structural Diffusion
by: Liu, Xian, et al.
Published: (2023)
by: Liu, Xian, et al.
Published: (2023)
PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion Transformers
by: Lin, Yuchen, et al.
Published: (2025)
by: Lin, Yuchen, et al.
Published: (2025)
3DTopia: Large Text-to-3D Generation Model with Hybrid Diffusion Priors
by: Hong, Fangzhou, et al.
Published: (2024)
by: Hong, Fangzhou, et al.
Published: (2024)
Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control
by: Song, Chenxi, et al.
Published: (2025)
by: Song, Chenxi, et al.
Published: (2025)
EG4D: Explicit Generation of 4D Object without Score Distillation
by: Sun, Qi, et al.
Published: (2024)
by: Sun, Qi, et al.
Published: (2024)
JADE: Joint-aware Latent Diffusion for 3D Human Generative Modeling
by: Ji, Haorui, et al.
Published: (2024)
by: Ji, Haorui, et al.
Published: (2024)
Bootstrap3D: Improving Multi-view Diffusion Model with Synthetic Data
by: Sun, Zeyi, et al.
Published: (2024)
by: Sun, Zeyi, et al.
Published: (2024)
Towards Native Generative Model for 3D Head Avatar
by: Zhuang, Yiyu, et al.
Published: (2024)
by: Zhuang, Yiyu, et al.
Published: (2024)
DualFocus: Integrating Macro and Micro Perspectives in Multi-modal Large Language Models
by: Cao, Yuhang, et al.
Published: (2024)
by: Cao, Yuhang, et al.
Published: (2024)
Optimizing 4D Wires for Sparse 3D Abstraction
by: Wu, Dong-Yi, et al.
Published: (2026)
by: Wu, Dong-Yi, et al.
Published: (2026)
Similar Items
-
IDArb: Intrinsic Decomposition for Arbitrary Number of Input Views and Illuminations
by: Li, Zhibing, et al.
Published: (2024) -
LayerPano3D: Layered 3D Panorama for Hyper-Immersive Scene Generation
by: Yang, Shuai, et al.
Published: (2024) -
GPT-4V(ision) is a Human-Aligned Evaluator for Text-to-3D Generation
by: Wu, Tong, et al.
Published: (2024) -
3DGen-Bench: Comprehensive Benchmark Suite for 3D Generative Models
by: Zhang, Yuhan, et al.
Published: (2025) -
GenDoP: Auto-regressive Camera Trajectory Generation as a Director of Photography
by: Zhang, Mengchen, et al.
Published: (2025)