Swin3D++: Effective Multi-Source Pretraining for 3D Indoor Scene Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Yu-Qi, Guo, Yu-Xiao, Liu, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LPA3D: 3D Room-Level Scene Generation from In-the-Wild Images
von: Yang, Ming-Jia, et al.
Veröffentlicht: (2025)
von: Yang, Ming-Jia, et al.
Veröffentlicht: (2025)
Chorus: Multi-Teacher Pretraining for Holistic 3D Gaussian Scene Encoding
von: Li, Yue, et al.
Veröffentlicht: (2025)
von: Li, Yue, et al.
Veröffentlicht: (2025)
3D Question Answering for City Scene Understanding
von: Sun, Penglei, et al.
Veröffentlicht: (2024)
von: Sun, Penglei, et al.
Veröffentlicht: (2024)
UnScene3D: Unsupervised 3D Instance Segmentation for Indoor Scenes
von: Rozenberszki, David, et al.
Veröffentlicht: (2023)
von: Rozenberszki, David, et al.
Veröffentlicht: (2023)
GaussianGraph: 3D Gaussian-based Scene Graph Generation for Open-world Scene Understanding
von: Wang, Xihan, et al.
Veröffentlicht: (2025)
von: Wang, Xihan, et al.
Veröffentlicht: (2025)
Rein3D: Reinforced 3D Indoor Scene Generation with Panoramic Video Diffusion Models
von: Wang, Dehui, et al.
Veröffentlicht: (2026)
von: Wang, Dehui, et al.
Veröffentlicht: (2026)
Adaptive Multi-Scale Channel-Spatial Attention Aggregation Framework for 3D Indoor Semantic Scene Completion Toward Assisting Visually Impaired
von: He, Qi, et al.
Veröffentlicht: (2026)
von: He, Qi, et al.
Veröffentlicht: (2026)
LLplace: The 3D Indoor Scene Layout Generation and Editing via Large Language Model
von: Yang, Yixuan, et al.
Veröffentlicht: (2024)
von: Yang, Yixuan, et al.
Veröffentlicht: (2024)
CineScene: Implicit 3D as Effective Scene Representation for Cinematic Video Generation
von: Huang, Kaiyi, et al.
Veröffentlicht: (2026)
von: Huang, Kaiyi, et al.
Veröffentlicht: (2026)
MVRoom: Controllable 3D Indoor Scene Generation with Multi-View Diffusion Models
von: Fang, Shaoheng, et al.
Veröffentlicht: (2025)
von: Fang, Shaoheng, et al.
Veröffentlicht: (2025)
SceneReVis: A Self-Reflective Vision-Grounded Framework for 3D Indoor Scene Synthesis via Multi-turn RL
von: Zhao, Yang, et al.
Veröffentlicht: (2026)
von: Zhao, Yang, et al.
Veröffentlicht: (2026)
Inst3D-LMM: Instance-Aware 3D Scene Understanding with Multi-modal Instruction Tuning
von: Yu, Hanxun, et al.
Veröffentlicht: (2025)
von: Yu, Hanxun, et al.
Veröffentlicht: (2025)
Mixed Diffusion for 3D Indoor Scene Synthesis
von: Hu, Siyi, et al.
Veröffentlicht: (2024)
von: Hu, Siyi, et al.
Veröffentlicht: (2024)
SPATIALGEN: Layout-guided 3D Indoor Scene Generation
von: Fang, Chuan, et al.
Veröffentlicht: (2025)
von: Fang, Chuan, et al.
Veröffentlicht: (2025)
DC-Scene: Data-Centric Learning for 3D Scene Understanding
von: Huang, Ting, et al.
Veröffentlicht: (2025)
von: Huang, Ting, et al.
Veröffentlicht: (2025)
Global-Local Tree Search in VLMs for 3D Indoor Scene Generation
von: Deng, Wei, et al.
Veröffentlicht: (2025)
von: Deng, Wei, et al.
Veröffentlicht: (2025)
SAI3D: Segment Any Instance in 3D Scenes
von: Yin, Yingda, et al.
Veröffentlicht: (2023)
von: Yin, Yingda, et al.
Veröffentlicht: (2023)
CommonScenes: Generating Commonsense 3D Indoor Scenes with Scene Graph Diffusion
von: Zhai, Guangyao, et al.
Veröffentlicht: (2023)
von: Zhai, Guangyao, et al.
Veröffentlicht: (2023)
Articulate3D: Holistic Understanding of 3D Scenes as Universal Scene Description
von: Halacheva, Anna-Maria, et al.
Veröffentlicht: (2024)
von: Halacheva, Anna-Maria, et al.
Veröffentlicht: (2024)
Behind the Veil: Enhanced Indoor 3D Scene Reconstruction with Occluded Surfaces Completion
von: Sun, Su, et al.
Veröffentlicht: (2024)
von: Sun, Su, et al.
Veröffentlicht: (2024)
Reg3D: Reconstructive Geometry Instruction Tuning for 3D Scene Understanding
von: Zheng, Hongpei, et al.
Veröffentlicht: (2025)
von: Zheng, Hongpei, et al.
Veröffentlicht: (2025)
SceneCraft: Layout-Guided 3D Scene Generation
von: Yang, Xiuyu, et al.
Veröffentlicht: (2024)
von: Yang, Xiuyu, et al.
Veröffentlicht: (2024)
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding
von: Huang, Wencan, et al.
Veröffentlicht: (2025)
von: Huang, Wencan, et al.
Veröffentlicht: (2025)
Jointly Understand Your Command and Intention:Reciprocal Co-Evolution between Scene-Aware 3D Human Motion Synthesis and Analysis
von: Gao, Xuehao, et al.
Veröffentlicht: (2025)
von: Gao, Xuehao, et al.
Veröffentlicht: (2025)
SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes
von: Huang, Jiaxin, et al.
Veröffentlicht: (2025)
von: Huang, Jiaxin, et al.
Veröffentlicht: (2025)
CSS: Overcoming Pose and Scene Challenges in Crowd-Sourced 3D Gaussian Splatting
von: Chen, Runze, et al.
Veröffentlicht: (2024)
von: Chen, Runze, et al.
Veröffentlicht: (2024)
3D-RE-GEN: 3D Reconstruction of Indoor Scenes with a Generative Framework
von: Sautter, Tobias, et al.
Veröffentlicht: (2025)
von: Sautter, Tobias, et al.
Veröffentlicht: (2025)
Embodied Intelligence for 3D Understanding: A Survey on 3D Scene Question Answering
von: Li, Zechuan, et al.
Veröffentlicht: (2025)
von: Li, Zechuan, et al.
Veröffentlicht: (2025)
IL3D: A Large-Scale Indoor Layout Dataset for LLM-Driven 3D Scene Generation
von: Zhou, Wenxu, et al.
Veröffentlicht: (2025)
von: Zhou, Wenxu, et al.
Veröffentlicht: (2025)
Contrastive Language-Colored Pointmap Pretraining for Unified 3D Scene Understanding
von: Mao, Ye, et al.
Veröffentlicht: (2026)
von: Mao, Ye, et al.
Veröffentlicht: (2026)
Function2Scene: 3D Indoor Scene Layout from Functional Specifications
von: Wang, Ruiqi, et al.
Veröffentlicht: (2026)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2026)
Domain Aware Multi-Task Pretraining of 3D Swin Transformer for T1-weighted Brain MRI
von: Kim, Jonghun, et al.
Veröffentlicht: (2024)
von: Kim, Jonghun, et al.
Veröffentlicht: (2024)
Indoor 3D Reconstruction with an Unknown Camera-Projector Pair
von: Qi, Zhaoshuai, et al.
Veröffentlicht: (2024)
von: Qi, Zhaoshuai, et al.
Veröffentlicht: (2024)
SceneCritic: A Symbolic Evaluator for 3D Indoor Scene Synthesis
von: Sengupta, Kathakoli, et al.
Veröffentlicht: (2026)
von: Sengupta, Kathakoli, et al.
Veröffentlicht: (2026)
ReScene4D: Temporally Consistent Semantic Instance Segmentation of Evolving Indoor 3D Scenes
von: Steiner, Emily, et al.
Veröffentlicht: (2026)
von: Steiner, Emily, et al.
Veröffentlicht: (2026)
ARKit LabelMaker: A New Scale for Indoor 3D Scene Understanding
von: Ji, Guangda, et al.
Veröffentlicht: (2024)
von: Ji, Guangda, et al.
Veröffentlicht: (2024)
M3DLayout: A Multi-Source Dataset of 3D Indoor Layouts and Structured Descriptions for 3D Generation
von: Zhang, Yiheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yiheng, et al.
Veröffentlicht: (2025)
Style-Consistent 3D Indoor Scene Synthesis with Decoupled Objects
von: Zhang, Yunfan, et al.
Veröffentlicht: (2024)
von: Zhang, Yunfan, et al.
Veröffentlicht: (2024)
Predicting 3D representations for Dynamic Scenes
von: Qi, Di, et al.
Veröffentlicht: (2025)
von: Qi, Di, et al.
Veröffentlicht: (2025)
GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation
von: Deng, Tianchen, et al.
Veröffentlicht: (2025)
von: Deng, Tianchen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LPA3D: 3D Room-Level Scene Generation from In-the-Wild Images
von: Yang, Ming-Jia, et al.
Veröffentlicht: (2025) -
Chorus: Multi-Teacher Pretraining for Holistic 3D Gaussian Scene Encoding
von: Li, Yue, et al.
Veröffentlicht: (2025) -
3D Question Answering for City Scene Understanding
von: Sun, Penglei, et al.
Veröffentlicht: (2024) -
UnScene3D: Unsupervised 3D Instance Segmentation for Indoor Scenes
von: Rozenberszki, David, et al.
Veröffentlicht: (2023) -
GaussianGraph: 3D Gaussian-based Scene Graph Generation for Open-world Scene Understanding
von: Wang, Xihan, et al.
Veröffentlicht: (2025)