SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Yongwei, Lan, Yushi, Zhou, Shangchen, Wang, Tengfei, Pan, Xingang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PnP-U3D: Plug-and-Play 3D Framework Bridging Autoregression and Diffusion for Unified Understanding and Generation
by: Chen, Yongwei, et al.
Published: (2026)
by: Chen, Yongwei, et al.
Published: (2026)
ArtiLatent: Realistic Articulated 3D Object Generation via Structured Latents
by: Chen, Honghua, et al.
Published: (2025)
by: Chen, Honghua, et al.
Published: (2025)
MvDrag3D: Drag-based Creative 3D Editing via Multi-view Generation-Reconstruction Priors
by: Chen, Honghua, et al.
Published: (2024)
by: Chen, Honghua, et al.
Published: (2024)
3DEnhancer: Consistent Multi-View Diffusion for 3D Enhancement
by: Luo, Yihang, et al.
Published: (2024)
by: Luo, Yihang, et al.
Published: (2024)
LN3DIFF++: Scalable Latent Neural Fields Diffusion for Speedy 3D Generation
by: Lan, Yushi, et al.
Published: (2024)
by: Lan, Yushi, et al.
Published: (2024)
4RC: 4D Reconstruction via Conditional Querying Anytime and Anywhere
by: Luo, Yihang, et al.
Published: (2026)
by: Luo, Yihang, et al.
Published: (2026)
GaussianAnything: Interactive Point Cloud Flow Matching For 3D Object Generation
by: Lan, Yushi, et al.
Published: (2024)
by: Lan, Yushi, et al.
Published: (2024)
Textured 3D Regenerative Morphing with 3D Diffusion Prior
by: Yang, Songlin, et al.
Published: (2025)
by: Yang, Songlin, et al.
Published: (2025)
ObjCtrl-2.5D: Training-free Object Control with Camera Poses
by: Wang, Zhouxia, et al.
Published: (2024)
by: Wang, Zhouxia, et al.
Published: (2024)
STream3R: Scalable Sequential 3D Reconstruction with Causal Transformer
by: Lan, Yushi, et al.
Published: (2025)
by: Lan, Yushi, et al.
Published: (2025)
ComboVerse: Compositional 3D Assets Creation Using Spatially-Aware Diffusion Guidance
by: Chen, Yongwei, et al.
Published: (2024)
by: Chen, Yongwei, et al.
Published: (2024)
FastMesh: Efficient Artistic Mesh Generation via Component Decoupling
by: Kim, Jeonghwan, et al.
Published: (2025)
by: Kim, Jeonghwan, et al.
Published: (2025)
MVIP-NeRF: Multi-view 3D Inpainting on NeRF Scenes via Diffusion Prior
by: Chen, Honghua, et al.
Published: (2024)
by: Chen, Honghua, et al.
Published: (2024)
Material Anything: Generating Materials for Any 3D Object via Diffusion
by: Huang, Xin, et al.
Published: (2024)
by: Huang, Xin, et al.
Published: (2024)
3DTopia-XL: Scaling High-quality 3D Asset Generation via Primitive Diffusion
by: Chen, Zhaoxi, et al.
Published: (2024)
by: Chen, Zhaoxi, et al.
Published: (2024)
SHOW3D: Capturing Scenes of 3D Hands and Objects in the Wild
by: Rim, Patrick, et al.
Published: (2026)
by: Rim, Patrick, et al.
Published: (2026)
PI-Light: Physics-Inspired Diffusion for Full-Image Relighting
by: Liang, Zhexin, et al.
Published: (2026)
by: Liang, Zhexin, et al.
Published: (2026)
On the Feasibility and Opportunity of Autoregressive 3D Object Detection
by: Huang, Zanming, et al.
Published: (2026)
by: Huang, Zanming, et al.
Published: (2026)
Bokeh Diffusion: Defocus Blur Control in Text-to-Image Diffusion Models
by: Fortes, Armando, et al.
Published: (2025)
by: Fortes, Armando, et al.
Published: (2025)
ImageNet3D: Towards General-Purpose Object-Level 3D Understanding
by: Ma, Wufei, et al.
Published: (2024)
by: Ma, Wufei, et al.
Published: (2024)
Neural LightRig: Unlocking Accurate Object Normal and Material Estimation with Multi-Light Diffusion
by: He, Zexin, et al.
Published: (2024)
by: He, Zexin, et al.
Published: (2024)
DreamScene: 3D Gaussian-based End-to-end Text-to-3D Scene Generation
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
Style3D: Attention-guided Multi-view Style Transfer for 3D Object Generation
by: Song, Bingjie, et al.
Published: (2024)
by: Song, Bingjie, et al.
Published: (2024)
G3PT: Unleash the power of Autoregressive Modeling in 3D Generation via Cross-scale Querying Transformer
by: Zhang, Jinzhi, et al.
Published: (2024)
by: Zhang, Jinzhi, et al.
Published: (2024)
Stream3D: Sequential Multi-View 3D Generation via Evidential Memory
by: Zhou, Kaichen, et al.
Published: (2026)
by: Zhou, Kaichen, et al.
Published: (2026)
HOT3D: Hand and Object Tracking in 3D from Egocentric Multi-View Videos
by: Banerjee, Prithviraj, et al.
Published: (2024)
by: Banerjee, Prithviraj, et al.
Published: (2024)
Weakly Supervised 3D Object Detection with Multi-Stage Generalization
by: He, Jiawei, et al.
Published: (2023)
by: He, Jiawei, et al.
Published: (2023)
BadFusion: 2D-Oriented Backdoor Attacks against 3D Object Detection
by: Chaturvedi, Saket S., et al.
Published: (2024)
by: Chaturvedi, Saket S., et al.
Published: (2024)
AssetFormer: Modular 3D Assets Generation with Autoregressive Transformer
by: Zhu, Lingting, et al.
Published: (2026)
by: Zhu, Lingting, et al.
Published: (2026)
CCF: Complementary Collaborative Fusion for Domain Generalized Multi-Modal 3D Object Detection
by: Wu, Yuchen, et al.
Published: (2026)
by: Wu, Yuchen, et al.
Published: (2026)
ExCap3D: Expressive 3D Scene Understanding via Object Captioning with Varying Detail
by: Yeshwanth, Chandan, et al.
Published: (2025)
by: Yeshwanth, Chandan, et al.
Published: (2025)
Ouroboros3D: Image-to-3D Generation via 3D-aware Recursive Diffusion
by: Wen, Hao, et al.
Published: (2024)
by: Wen, Hao, et al.
Published: (2024)
ObjFiller3D: Scaling 3D Object Inpainting to Dense Multi-View Consistency
by: Feng, Haitang, et al.
Published: (2025)
by: Feng, Haitang, et al.
Published: (2025)
Repurposing 3D Generative Model for Autoregressive Layout Generation
by: Feng, Haoran, et al.
Published: (2026)
by: Feng, Haoran, et al.
Published: (2026)
DreamCS: Geometry-Aware Text-to-3D Generation with Unpaired 3D Reward Supervision
by: Zou, Xiandong, et al.
Published: (2025)
by: Zou, Xiandong, et al.
Published: (2025)
Adv3D: Generating 3D Adversarial Examples for 3D Object Detection in Driving Scenarios with NeRF
by: Li, Leheng, et al.
Published: (2023)
by: Li, Leheng, et al.
Published: (2023)
RecDreamer: Consistent Text-to-3D Generation via Uniform Score Distillation
by: Zheng, Chenxi, et al.
Published: (2025)
by: Zheng, Chenxi, et al.
Published: (2025)
JM3D & JM3D-LLM: Elevating 3D Understanding with Joint Multi-modal Cues
by: Ji, Jiayi, et al.
Published: (2023)
by: Ji, Jiayi, et al.
Published: (2023)
Eval3D: Interpretable and Fine-grained Evaluation for 3D Generation
by: Duggal, Shivam, et al.
Published: (2025)
by: Duggal, Shivam, et al.
Published: (2025)
3D-WAG: Hierarchical Wavelet-Guided Autoregressive Generation for High-Fidelity 3D Shapes
by: Medi, Tejaswini, et al.
Published: (2024)
by: Medi, Tejaswini, et al.
Published: (2024)
Similar Items
-
PnP-U3D: Plug-and-Play 3D Framework Bridging Autoregression and Diffusion for Unified Understanding and Generation
by: Chen, Yongwei, et al.
Published: (2026) -
ArtiLatent: Realistic Articulated 3D Object Generation via Structured Latents
by: Chen, Honghua, et al.
Published: (2025) -
MvDrag3D: Drag-based Creative 3D Editing via Multi-view Generation-Reconstruction Priors
by: Chen, Honghua, et al.
Published: (2024) -
3DEnhancer: Consistent Multi-View Diffusion for 3D Enhancement
by: Luo, Yihang, et al.
Published: (2024) -
LN3DIFF++: Scalable Latent Neural Fields Diffusion for Speedy 3D Generation
by: Lan, Yushi, et al.
Published: (2024)