TreeSBA: Tree-Transformer for Self-Supervised Sequential Brick Assembly
Fuente:
arXiv
Guardado en:
| Autores principales: | Guo, Mengqi, Li, Chen, Zhao, Yuyang, Lee, Gim Hee |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
X-Ray: A Sequential 3D Representation For Generation
por: Hu, Tao, et al.
Publicado: (2024)
por: Hu, Tao, et al.
Publicado: (2024)
UNIKD: UNcertainty-filtered Incremental Knowledge Distillation for Neural Implicit Representation
por: Guo, Mengqi, et al.
Publicado: (2022)
por: Guo, Mengqi, et al.
Publicado: (2022)
SCAPO: Self-Supervised Category-Level Articulated Pose Estimation from a Single 3D Observation
por: Zhang, Can, et al.
Publicado: (2026)
por: Zhang, Can, et al.
Publicado: (2026)
Segment Any 3D Object with Language
por: Lee, Seungjun, et al.
Publicado: (2024)
por: Lee, Seungjun, et al.
Publicado: (2024)
4D3R: Motion-Aware Neural Reconstruction and Rendering of Dynamic Scenes from Monocular Videos
por: Guo, Mengqi, et al.
Publicado: (2025)
por: Guo, Mengqi, et al.
Publicado: (2025)
URS-NeRF: Unordered Rolling Shutter Bundle Adjustment for Neural Radiance Fields
por: Xu, Bo, et al.
Publicado: (2024)
por: Xu, Bo, et al.
Publicado: (2024)
ComPC: Completing a 3D Point Cloud with 2D Diffusion Priors
por: Huang, Tianxin, et al.
Publicado: (2024)
por: Huang, Tianxin, et al.
Publicado: (2024)
Make-A-Protagonist: Generic Video Editing with An Ensemble of Experts
por: Zhao, Yuyang, et al.
Publicado: (2023)
por: Zhao, Yuyang, et al.
Publicado: (2023)
DOGS: Distributed-Oriented Gaussian Splatting for Large-Scale 3D Reconstruction Via Gaussian Consensus
por: Chen, Yu, et al.
Publicado: (2024)
por: Chen, Yu, et al.
Publicado: (2024)
CLAIR: CLIP-Aided Weakly Supervised Zero-Shot Cross-Domain Image Retrieval
por: Tan, Chor Boon, et al.
Publicado: (2025)
por: Tan, Chor Boon, et al.
Publicado: (2025)
TreeMeshGPT: Artistic Mesh Generation with Autoregressive Tree Sequencing
por: Lionar, Stefan, et al.
Publicado: (2025)
por: Lionar, Stefan, et al.
Publicado: (2025)
DiHuR: Diffusion-Guided Generalizable Human Reconstruction
por: Chen, Jinnan, et al.
Publicado: (2024)
por: Chen, Jinnan, et al.
Publicado: (2024)
Segment Any Events with Language
por: Lee, Seungjun, et al.
Publicado: (2026)
por: Lee, Seungjun, et al.
Publicado: (2026)
MVGSR: Multi-View Consistency Gaussian Splatting for Robust Surface Reconstruction
por: Hou, Chenfeng, et al.
Publicado: (2025)
por: Hou, Chenfeng, et al.
Publicado: (2025)
Animate124: Animating One Image to 4D Dynamic Scene
por: Zhao, Yuyang, et al.
Publicado: (2023)
por: Zhao, Yuyang, et al.
Publicado: (2023)
MVSDet: Multi-View Indoor 3D Object Detection via Efficient Plane Sweeps
por: Xu, Yating, et al.
Publicado: (2024)
por: Xu, Yating, et al.
Publicado: (2024)
DiSR-NeRF: Diffusion-Guided View-Consistent Super-Resolution NeRF
por: Lee, Jie Long, et al.
Publicado: (2024)
por: Lee, Jie Long, et al.
Publicado: (2024)
econSG: Efficient and Multi-view Consistent Open-Vocabulary 3D Semantic Gaussians
por: Zhang, Can, et al.
Publicado: (2025)
por: Zhang, Can, et al.
Publicado: (2025)
Motion4D: Learning 3D-Consistent Motion and Semantics for 4D Scene Understanding
por: Zhou, Haoran, et al.
Publicado: (2025)
por: Zhou, Haoran, et al.
Publicado: (2025)
LLaFEA: Frame-Event Complementary Fusion for Fine-Grained Spatiotemporal Understanding in LMMs
por: Zhou, Hanyu, et al.
Publicado: (2025)
por: Zhou, Hanyu, et al.
Publicado: (2025)
LLaVA-4D: Embedding SpatioTemporal Prompt into LMMs for 4D Scene Understanding
por: Zhou, Hanyu, et al.
Publicado: (2025)
por: Zhou, Hanyu, et al.
Publicado: (2025)
IAAO: Interactive Affordance Learning for Articulated Objects in 3D Environments
por: Zhang, Can, et al.
Publicado: (2025)
por: Zhang, Can, et al.
Publicado: (2025)
MotionScale: Reconstructing Appearance, Geometry, and Motion of Dynamic Scenes with Scalable 4D Gaussian Splatting
por: Zhou, Haoran, et al.
Publicado: (2026)
por: Zhou, Haoran, et al.
Publicado: (2026)
Unified Geometry and Color Compression Framework for Point Clouds via Generative Diffusion Priors
por: Huang, Tianxin, et al.
Publicado: (2025)
por: Huang, Tianxin, et al.
Publicado: (2025)
Flow4DGS-SLAM: Optical Flow-Guided 4D Gaussian Splatting SLAM
por: Wang, Yunsong, et al.
Publicado: (2026)
por: Wang, Yunsong, et al.
Publicado: (2026)
HandMCM: Multi-modal Point Cloud-based Correspondence State Space Model for 3D Hand Pose Estimation
por: Cheng, Wencan, et al.
Publicado: (2026)
por: Cheng, Wencan, et al.
Publicado: (2026)
Uni4D-LLM: A Unified SpatioTemporal-Aware VLM for 4D Understanding and Generation
por: Zhou, Hanyu, et al.
Publicado: (2025)
por: Zhou, Hanyu, et al.
Publicado: (2025)
Syn-to-Real Unsupervised Domain Adaptation for Indoor 3D Object Detection
por: Wang, Yunsong, et al.
Publicado: (2024)
por: Wang, Yunsong, et al.
Publicado: (2024)
Enhancing Generalizability of Representation Learning for Data-Efficient 3D Scene Understanding
por: Wang, Yunsong, et al.
Publicado: (2024)
por: Wang, Yunsong, et al.
Publicado: (2024)
BrickNet: Graph-Backed Generative Brick Assembly
por: Kulits, Peter, et al.
Publicado: (2026)
por: Kulits, Peter, et al.
Publicado: (2026)
GOV-NeSF: Generalizable Open-Vocabulary Neural Semantic Fields
por: Wang, Yunsong, et al.
Publicado: (2024)
por: Wang, Yunsong, et al.
Publicado: (2024)
ChatSplat: 3D Conversational Gaussian Splatting
por: Chen, Hanlin, et al.
Publicado: (2024)
por: Chen, Hanlin, et al.
Publicado: (2024)
NeuSG: Neural Implicit Surface Reconstruction with 3D Gaussian Splatting Guidance
por: Chen, Hanlin, et al.
Publicado: (2023)
por: Chen, Hanlin, et al.
Publicado: (2023)
DiET-GS: Diffusion Prior and Event Stream-Assisted Motion Deblurring 3D Gaussian Splatting
por: Lee, Seungjun, et al.
Publicado: (2025)
por: Lee, Seungjun, et al.
Publicado: (2025)
VLA-4D: Embedding 4D Awareness into Vision-Language-Action Models for SpatioTemporally Coherent Robotic Manipulation
por: Zhou, Hanyu, et al.
Publicado: (2025)
por: Zhou, Hanyu, et al.
Publicado: (2025)
Brick-Diffusion: Generating Long Videos with Brick-to-Wall Denoising
por: Yuan, Yunlong, et al.
Publicado: (2025)
por: Yuan, Yunlong, et al.
Publicado: (2025)
Dynam3D: Dynamic Layered 3D Tokens Empower VLM for Vision-and-Language Navigation
por: Wang, Zihan, et al.
Publicado: (2025)
por: Wang, Zihan, et al.
Publicado: (2025)
FreeSplat: Generalizable 3D Gaussian Splatting Towards Free-View Synthesis of Indoor Scenes
por: Wang, Yunsong, et al.
Publicado: (2024)
por: Wang, Yunsong, et al.
Publicado: (2024)
FreeSplat++: Generalizable 3D Gaussian Splatting for Efficient Indoor Scene Reconstruction
por: Wang, Yunsong, et al.
Publicado: (2025)
por: Wang, Yunsong, et al.
Publicado: (2025)
SmileSplat: Generalizable Gaussian Splats for Unconstrained Sparse Images
por: Li, Yanyan, et al.
Publicado: (2024)
por: Li, Yanyan, et al.
Publicado: (2024)
Ejemplares similares
-
X-Ray: A Sequential 3D Representation For Generation
por: Hu, Tao, et al.
Publicado: (2024) -
UNIKD: UNcertainty-filtered Incremental Knowledge Distillation for Neural Implicit Representation
por: Guo, Mengqi, et al.
Publicado: (2022) -
SCAPO: Self-Supervised Category-Level Articulated Pose Estimation from a Single 3D Observation
por: Zhang, Can, et al.
Publicado: (2026) -
Segment Any 3D Object with Language
por: Lee, Seungjun, et al.
Publicado: (2024) -
4D3R: Motion-Aware Neural Reconstruction and Rendering of Dynamic Scenes from Monocular Videos
por: Guo, Mengqi, et al.
Publicado: (2025)