Direct3D-S2: Gigascale 3D Generation Made Easy with Spatial Sparse Attention
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Shuang, Lin, Youtian, Zhang, Feihu, Zeng, Yifei, Yang, Yikang, Bao, Yajie, Qian, Jiachen, Zhu, Siyu, Cao, Xun, Torr, Philip, Yao, Yao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer
by: Wu, Shuang, et al.
Published: (2024)
by: Wu, Shuang, et al.
Published: (2024)
TEXTRIX: Latent Attribute Grid for Native Texture Generation and Beyond
by: Zeng, Yifei, et al.
Published: (2025)
by: Zeng, Yifei, et al.
Published: (2025)
STAG4D: Spatial-Temporal Anchored Generative 4D Gaussians
by: Zeng, Yifei, et al.
Published: (2024)
by: Zeng, Yifei, et al.
Published: (2024)
Flow Distillation Sampling: Regularizing 3D Gaussians with Pre-trained Matching Priors
by: Chen, Lin-Zhuo, et al.
Published: (2025)
by: Chen, Lin-Zhuo, et al.
Published: (2025)
SpatialTrackerV2: 3D Point Tracking Made Easy
by: Xiao, Yuxi, et al.
Published: (2025)
by: Xiao, Yuxi, et al.
Published: (2025)
Relightable 3D Gaussians: Realistic Point Cloud Relighting with BRDF Decomposition and Ray Tracing
by: Gao, Jian, et al.
Published: (2023)
by: Gao, Jian, et al.
Published: (2023)
DUSt3R: Geometric 3D Vision Made Easy
by: Wang, Shuzhe, et al.
Published: (2023)
by: Wang, Shuzhe, et al.
Published: (2023)
DragMesh: Interactive 3D Generation Made Easy
by: Zhang, Tianshan, et al.
Published: (2025)
by: Zhang, Tianshan, et al.
Published: (2025)
SpatialVID: A Large-Scale Video Dataset with Spatial Annotations
by: Wang, Jiahao, et al.
Published: (2025)
by: Wang, Jiahao, et al.
Published: (2025)
P3P Made Easy
by: Lee, Seong Hun, et al.
Published: (2025)
by: Lee, Seong Hun, et al.
Published: (2025)
Towards Native Generative Model for 3D Head Avatar
by: Zhuang, Yiyu, et al.
Published: (2024)
by: Zhuang, Yiyu, et al.
Published: (2024)
VFusion3D: Learning Scalable 3D Generative Models from Video Diffusion Models
by: Han, Junlin, et al.
Published: (2024)
by: Han, Junlin, et al.
Published: (2024)
ComGS: Efficient 3D Object-Scene Composition via Surface Octahedral Probes
by: Gao, Jian, et al.
Published: (2025)
by: Gao, Jian, et al.
Published: (2025)
Direct2.5: Diverse Text-to-3D Generation via Multi-view 2.5D Diffusion
by: Lu, Yuanxun, et al.
Published: (2023)
by: Lu, Yuanxun, et al.
Published: (2023)
MoDification: Mixture of Depths Made Easy
by: Zhang, Chen, et al.
Published: (2024)
by: Zhang, Chen, et al.
Published: (2024)
Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
by: Zhu, Shenhao, et al.
Published: (2024)
by: Zhu, Shenhao, et al.
Published: (2024)
4DNeX: Feed-Forward 4D Generative Modeling Made Easy
by: Chen, Zhaoxi, et al.
Published: (2025)
by: Chen, Zhaoxi, et al.
Published: (2025)
Head360: Learning a Parametric 3D Full-Head for Free-View Synthesis in 360°
by: He, Yuxiao, et al.
Published: (2024)
by: He, Yuxiao, et al.
Published: (2024)
HumDex: Humanoid Dexterous Manipulation Made Easy
by: Heng, Liang, et al.
Published: (2026)
by: Heng, Liang, et al.
Published: (2026)
High-Fidelity 3D Facial Avatar Synthesis with Controllable Fine-Grained Expressions
by: He, Yikang, et al.
Published: (2026)
by: He, Yikang, et al.
Published: (2026)
EasyTime: Time Series Forecasting Made Easy
by: Qiu, Xiangfei, et al.
Published: (2024)
by: Qiu, Xiangfei, et al.
Published: (2024)
EmoTalk3D: High-Fidelity Free-View Synthesis of Emotional 3D Talking Head
by: He, Qianyun, et al.
Published: (2024)
by: He, Qianyun, et al.
Published: (2024)
Differential Privacy Made Easy
by: Aitsam, Muhammad
Published: (2022)
by: Aitsam, Muhammad
Published: (2022)
Consistency Models Made Easy
by: Geng, Zhengyang, et al.
Published: (2024)
by: Geng, Zhengyang, et al.
Published: (2024)
Articulate3D: Zero-Shot Text-Driven 3D Object Posing
by: Deb, Oishi, et al.
Published: (2025)
by: Deb, Oishi, et al.
Published: (2025)
Hybrid 3D Human Pose Estimation with Monocular Video and Sparse IMUs
by: Bao, Yiming, et al.
Published: (2024)
by: Bao, Yiming, et al.
Published: (2024)
Scene-Conditional 3D Object Stylization and Composition
by: Zhou, Jinghao, et al.
Published: (2023)
by: Zhou, Jinghao, et al.
Published: (2023)
DreamPolisher: Towards High-Quality Text-to-3D Generation via Geometric Diffusion
by: Lin, Yuanze, et al.
Published: (2024)
by: Lin, Yuanze, et al.
Published: (2024)
Optimal Low-Thrust Orbit Transfers Made Easy: A Direct Approach
by: Leomanni, Mirko, et al.
Published: (2021)
by: Leomanni, Mirko, et al.
Published: (2021)
Flex3D: Feed-Forward 3D Generation with Flexible Reconstruction Model and Input View Curation
by: Han, Junlin, et al.
Published: (2024)
by: Han, Junlin, et al.
Published: (2024)
Urban Neural Surface Reconstruction from Constrained Sparse Aerial Imagery with 3D SAR Fusion
by: Li, Da, et al.
Published: (2026)
by: Li, Da, et al.
Published: (2026)
E3x: $\mathrm{E}(3)$-Equivariant Deep Learning Made Easy
by: Unke, Oliver T., et al.
Published: (2024)
by: Unke, Oliver T., et al.
Published: (2024)
TeRA: Rethinking Text-guided Realistic 3D Avatar Generation
by: Wang, Yanwen, et al.
Published: (2025)
by: Wang, Yanwen, et al.
Published: (2025)
JDT3D: Addressing the Gaps in LiDAR-Based Tracking-by-Attention
by: Cheong, Brian, et al.
Published: (2024)
by: Cheong, Brian, et al.
Published: (2024)
Semantic Score Distillation Sampling for Compositional Text-to-3D Generation
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
Matrix3D: Large Photogrammetry Model All-in-One
by: Lu, Yuanxun, et al.
Published: (2025)
by: Lu, Yuanxun, et al.
Published: (2025)
A Lightweight Sparse Focus Transformer for Remote Sensing Image Change Captioning
by: Sun, Dongwei, et al.
Published: (2024)
by: Sun, Dongwei, et al.
Published: (2024)
SmartSpatial: Enhancing the 3D Spatial Arrangement Capabilities of Stable Diffusion Models and Introducing a Novel 3D Spatial Evaluation Framework
by: Huang, Mao Xun, et al.
Published: (2025)
by: Huang, Mao Xun, et al.
Published: (2025)
Wilkins: HPC In Situ Workflows Made Easy
by: Yildiz, Orcun, et al.
Published: (2024)
by: Yildiz, Orcun, et al.
Published: (2024)
AnyDepth: Depth Estimation Made Easy
by: Ren, Zeyu, et al.
Published: (2026)
by: Ren, Zeyu, et al.
Published: (2026)
Similar Items
-
Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer
by: Wu, Shuang, et al.
Published: (2024) -
TEXTRIX: Latent Attribute Grid for Native Texture Generation and Beyond
by: Zeng, Yifei, et al.
Published: (2025) -
STAG4D: Spatial-Temporal Anchored Generative 4D Gaussians
by: Zeng, Yifei, et al.
Published: (2024) -
Flow Distillation Sampling: Regularizing 3D Gaussians with Pre-trained Matching Priors
by: Chen, Lin-Zhuo, et al.
Published: (2025) -
SpatialTrackerV2: 3D Point Tracking Made Easy
by: Xiao, Yuxi, et al.
Published: (2025)