Can3Tok: Canonical 3D Tokenization and Latent Modeling of Scene-Level 3D Gaussians
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Quankai, Georgiev, Iliyan, Wang, Tuanfeng Y., Singh, Krishna Kumar, Neumann, Ulrich, Yoon, Jae Shin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stochastic Ray Tracing for the Reconstruction of 3D Gaussian Splatting
by: Xu, Peiyu, et al.
Published: (2026)
by: Xu, Peiyu, et al.
Published: (2026)
SceneTok: A Compressed, Diffusable Token Space for 3D Scenes
by: Asim, Mohammad, et al.
Published: (2026)
by: Asim, Mohammad, et al.
Published: (2026)
GaussianFlow: Splatting Gaussian Dynamics for 4D Content Creation
by: Gao, Quankai, et al.
Published: (2024)
by: Gao, Quankai, et al.
Published: (2024)
3D-Fixup: Advancing Photo Editing with 3D Priors
by: Cheng, Yen-Chi, et al.
Published: (2025)
by: Cheng, Yen-Chi, et al.
Published: (2025)
Stochastic Ray Tracing of Transparent 3D Gaussians
by: Sun, Xin, et al.
Published: (2025)
by: Sun, Xin, et al.
Published: (2025)
SAMa: Material-aware 3D Selection and Segmentation
by: Fischer, Michael, et al.
Published: (2024)
by: Fischer, Michael, et al.
Published: (2024)
DriveTok: 3D Driving Scene Tokenization for Unified Multi-View Reconstruction and Understanding
by: Zhuo, Dong, et al.
Published: (2026)
by: Zhuo, Dong, et al.
Published: (2026)
Comprehensive Relighting: Generalizable and Consistent Monocular Human Relighting and Harmonization
by: Wang, Junying, et al.
Published: (2025)
by: Wang, Junying, et al.
Published: (2025)
Sampling 3D Gaussian Scenes in Seconds with Latent Diffusion Models
by: Henderson, Paul, et al.
Published: (2024)
by: Henderson, Paul, et al.
Published: (2024)
Synergy between 3DMM and 3D Landmarks for Accurate 3D Facial Geometry
by: Wu, Cho-Ying, et al.
Published: (2021)
by: Wu, Cho-Ying, et al.
Published: (2021)
Neural Product Importance Sampling via Warp Composition
by: Litalien, Joey, et al.
Published: (2024)
by: Litalien, Joey, et al.
Published: (2024)
Canonical correlation analysis of stochastic trends via functional approximation
by: Franchi, Massimo, et al.
Published: (2024)
by: Franchi, Massimo, et al.
Published: (2024)
SceneFactor: Factored Latent 3D Diffusion for Controllable 3D Scene Generation
by: Bokhovkin, Alexey, et al.
Published: (2024)
by: Bokhovkin, Alexey, et al.
Published: (2024)
Drag4D: Align Your Motion with Text-Driven 3D Scene Generation
by: Kang, Minjun, et al.
Published: (2025)
by: Kang, Minjun, et al.
Published: (2025)
IntrinsicEdit: Precise generative image manipulation in intrinsic space
by: Lyu, Linjie, et al.
Published: (2025)
by: Lyu, Linjie, et al.
Published: (2025)
Bolt3D: Generating 3D Scenes in Seconds
by: Szymanowicz, Stanislaw, et al.
Published: (2025)
by: Szymanowicz, Stanislaw, et al.
Published: (2025)
GaussianGraph: 3D Gaussian-based Scene Graph Generation for Open-world Scene Understanding
by: Wang, Xihan, et al.
Published: (2025)
by: Wang, Xihan, et al.
Published: (2025)
Dense Dispersed Structured Light for Hyperspectral 3D Imaging of Dynamic Scenes
by: Shin, Suhyun, et al.
Published: (2024)
by: Shin, Suhyun, et al.
Published: (2024)
WaSt-3D: Wasserstein-2 Distance for Scene-to-Scene Stylization on 3D Gaussians
by: Kotovenko, Dmytro, et al.
Published: (2024)
by: Kotovenko, Dmytro, et al.
Published: (2024)
DreamScene: 3D Gaussian-based End-to-end Text-to-3D Scene Generation
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
Gaussian2Scene: 3D Scene Representation Learning via Self-supervised Learning with 3D Gaussian Splatting
by: Liu, Keyi, et al.
Published: (2025)
by: Liu, Keyi, et al.
Published: (2025)
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models
by: Wang, Haoming, et al.
Published: (2026)
by: Wang, Haoming, et al.
Published: (2026)
VesselTok: Tokenizing Vessel-like 3D Biomedical Graph Representations for Reconstruction and Generation
by: Prabhakar, Chinmay, et al.
Published: (2026)
by: Prabhakar, Chinmay, et al.
Published: (2026)
ODE-GS: Latent ODEs for Dynamic Scene Extrapolation with 3D Gaussian Splatting
by: Wang, Daniel, et al.
Published: (2025)
by: Wang, Daniel, et al.
Published: (2025)
CanonicalFusion: Generating Drivable 3D Human Avatars from Multiple Images
by: Shin, Jisu, et al.
Published: (2024)
by: Shin, Jisu, et al.
Published: (2024)
Stylizing Sparse-View 3D Scenes with Hierarchical Neural Representation
by: Wang, Y., et al.
Published: (2024)
by: Wang, Y., et al.
Published: (2024)
JOG3R: Towards 3D-Consistent Video Generators
by: Huang, Chun-Hao Paul, et al.
Published: (2025)
by: Huang, Chun-Hao Paul, et al.
Published: (2025)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
by: Cha, Junuk, et al.
Published: (2024)
by: Cha, Junuk, et al.
Published: (2024)
3D Gaussian Flats: Hybrid 2D/3D Photometric Scene Reconstruction
by: Taktasheva, Maria, et al.
Published: (2025)
by: Taktasheva, Maria, et al.
Published: (2025)
Ilov3Splat: Instance-Level Open-Vocabulary 3D Scene Understanding in Gaussian Splatting
by: Nguyen, Binh Long, et al.
Published: (2026)
by: Nguyen, Binh Long, et al.
Published: (2026)
GlobalSplat: Efficient Feed-Forward 3D Gaussian Splatting via Global Scene Tokens
by: Itkin, Roni, et al.
Published: (2026)
by: Itkin, Roni, et al.
Published: (2026)
L3DG: Latent 3D Gaussian Diffusion
by: Roessle, Barbara, et al.
Published: (2024)
by: Roessle, Barbara, et al.
Published: (2024)
LT3SD: Latent Trees for 3D Scene Diffusion
by: Meng, Quan, et al.
Published: (2024)
by: Meng, Quan, et al.
Published: (2024)
Any 3D Scene is Worth 1K Tokens: 3D-Grounded Representation for Scene Generation at Scale
by: Wei, Dongxu, et al.
Published: (2026)
by: Wei, Dongxu, et al.
Published: (2026)
Zoo3D: Zero-Shot 3D Object Detection at Scene Level
by: Lemeshko, Andrey, et al.
Published: (2025)
by: Lemeshko, Andrey, et al.
Published: (2025)
GS-Scale: Unlocking Large-Scale 3D Gaussian Splatting Training via Host Offloading
by: Lee, Donghyun, et al.
Published: (2025)
by: Lee, Donghyun, et al.
Published: (2025)
V-RGBX: Video Editing with Accurate Controls over Intrinsic Properties
by: Fang, Ye, et al.
Published: (2025)
by: Fang, Ye, et al.
Published: (2025)
DreamScene: 3D Gaussian-based Text-to-3D Scene Generation via Formation Pattern Sampling
by: Li, Haoran, et al.
Published: (2024)
by: Li, Haoran, et al.
Published: (2024)
VastGaussian: Vast 3D Gaussians for Large Scene Reconstruction
by: Lin, Jiaqi, et al.
Published: (2024)
by: Lin, Jiaqi, et al.
Published: (2024)
GaussianGPT: Towards Autoregressive 3D Gaussian Scene Generation
by: von Lützow, Nicolas, et al.
Published: (2026)
by: von Lützow, Nicolas, et al.
Published: (2026)
Similar Items
-
Stochastic Ray Tracing for the Reconstruction of 3D Gaussian Splatting
by: Xu, Peiyu, et al.
Published: (2026) -
SceneTok: A Compressed, Diffusable Token Space for 3D Scenes
by: Asim, Mohammad, et al.
Published: (2026) -
GaussianFlow: Splatting Gaussian Dynamics for 4D Content Creation
by: Gao, Quankai, et al.
Published: (2024) -
3D-Fixup: Advancing Photo Editing with 3D Priors
by: Cheng, Yen-Chi, et al.
Published: (2025) -
Stochastic Ray Tracing of Transparent 3D Gaussians
by: Sun, Xin, et al.
Published: (2025)