Hierarchically-Structured Open-Vocabulary Indoor Scene Synthesis with Pre-trained Large Language Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Weilin, Li, Xinran, Li, Manyi, Xu, Kai, Meng, Xiangxu, Meng, Lei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Open-Vocabulary Indoor Object Grounding with 3D Hierarchical Scene Graph
von: Linok, Sergey, et al.
Veröffentlicht: (2025)
von: Linok, Sergey, et al.
Veröffentlicht: (2025)
Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes
von: Zhou, Changqing, et al.
Veröffentlicht: (2026)
von: Zhou, Changqing, et al.
Veröffentlicht: (2026)
Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces
von: Hu, Xinggang, et al.
Veröffentlicht: (2026)
von: Hu, Xinggang, et al.
Veröffentlicht: (2026)
From Pixels to Graphs: Open-Vocabulary Scene Graph Generation with Vision-Language Models
von: Li, Rongjie, et al.
Veröffentlicht: (2024)
von: Li, Rongjie, et al.
Veröffentlicht: (2024)
Open-Vocabulary Semantic Segmentation with Uncertainty Alignment for Robotic Scene Understanding in Indoor Building Environments
von: Xu, Yifan, et al.
Veröffentlicht: (2025)
von: Xu, Yifan, et al.
Veröffentlicht: (2025)
Open-Vocabulary Functional 3D Scene Graphs for Real-World Indoor Spaces
von: Zhang, Chenyangguang, et al.
Veröffentlicht: (2025)
von: Zhang, Chenyangguang, et al.
Veröffentlicht: (2025)
AG$^2$aussian: Anchor-Graph Structured Gaussian Splatting for Instance-Level 3D Scene Understanding and Editing
von: Wang, Zhaonan, et al.
Veröffentlicht: (2025)
von: Wang, Zhaonan, et al.
Veröffentlicht: (2025)
ProtoConNet: Prototypical Augmentation and Alignment for Open-Set Few-Shot Image Classification
von: Shi, Kexuan, et al.
Veröffentlicht: (2025)
von: Shi, Kexuan, et al.
Veröffentlicht: (2025)
Structure-aware World Model for Probe Guidance via Large-scale Self-supervised Pre-train
von: Jiang, Haojun, et al.
Veröffentlicht: (2024)
von: Jiang, Haojun, et al.
Veröffentlicht: (2024)
SceneNAT: Masked Generative Modeling for Language-Guided Indoor Scene Synthesis
von: Choi, Jeongjun, et al.
Veröffentlicht: (2026)
von: Choi, Jeongjun, et al.
Veröffentlicht: (2026)
FreeScene: Mixed Graph Diffusion for 3D Scene Synthesis from Free Prompts
von: Bai, Tongyuan, et al.
Veröffentlicht: (2025)
von: Bai, Tongyuan, et al.
Veröffentlicht: (2025)
LLM-Enabled Style and Content Regularization for Personalized Text-to-Image Generation
von: Yu, Anran, et al.
Veröffentlicht: (2025)
von: Yu, Anran, et al.
Veröffentlicht: (2025)
Forest2Seq: Revitalizing Order Prior for Sequential Indoor Scene Synthesis
von: Sun, Qi, et al.
Veröffentlicht: (2024)
von: Sun, Qi, et al.
Veröffentlicht: (2024)
Empowering Vision Transformers with Multi-Scale Causal Intervention for Long-Tailed Image Classification
von: Yan, Xiaoshuo, et al.
Veröffentlicht: (2025)
von: Yan, Xiaoshuo, et al.
Veröffentlicht: (2025)
Open Vocabulary 3D Scene Understanding via Geometry Guided Self-Distillation
von: Wang, Pengfei, et al.
Veröffentlicht: (2024)
von: Wang, Pengfei, et al.
Veröffentlicht: (2024)
Measuring Image-Relation Alignment: Reference-Free Evaluation of VLMs and Synthetic Pre-training for Open-Vocabulary Scene Graph Generation
von: Neau, Maëlic, et al.
Veröffentlicht: (2025)
von: Neau, Maëlic, et al.
Veröffentlicht: (2025)
Dense Multimodal Alignment for Open-Vocabulary 3D Scene Understanding
von: Li, Ruihuang, et al.
Veröffentlicht: (2024)
von: Li, Ruihuang, et al.
Veröffentlicht: (2024)
Interaction-Centric Knowledge Infusion and Transfer for Open-Vocabulary Scene Graph Generation
von: Li, Lin, et al.
Veröffentlicht: (2025)
von: Li, Lin, et al.
Veröffentlicht: (2025)
Zero-Shot Open-Vocabulary Tracking with Large Pre-Trained Models
von: Chu, Wen-Hsuan, et al.
Veröffentlicht: (2023)
von: Chu, Wen-Hsuan, et al.
Veröffentlicht: (2023)
Unifying Visual and Semantic Feature Spaces with Diffusion Models for Enhanced Cross-Modal Alignment
von: Zheng, Yuze, et al.
Veröffentlicht: (2024)
von: Zheng, Yuze, et al.
Veröffentlicht: (2024)
LLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of Large Language Models
von: Fu, Shenghao, et al.
Veröffentlicht: (2025)
von: Fu, Shenghao, et al.
Veröffentlicht: (2025)
S-INF: Towards Realistic Indoor Scene Synthesis via Scene Implicit Neural Field
von: Liang, Zixi, et al.
Veröffentlicht: (2024)
von: Liang, Zixi, et al.
Veröffentlicht: (2024)
Open-Vocabulary Domain Generalization in Urban-Scene Segmentation
von: Zhao, Dong, et al.
Veröffentlicht: (2026)
von: Zhao, Dong, et al.
Veröffentlicht: (2026)
Open-Vocabulary Octree-Graph for 3D Scene Understanding
von: Wang, Zhigang, et al.
Veröffentlicht: (2024)
von: Wang, Zhigang, et al.
Veröffentlicht: (2024)
CAGS: Open-Vocabulary 3D Scene Understanding with Context-Aware Gaussian Splatting
von: Sun, Wei, et al.
Veröffentlicht: (2025)
von: Sun, Wei, et al.
Veröffentlicht: (2025)
Global Prompt Refinement with Non-Interfering Attention Masking for One-Shot Federated Learning
von: Qi, Zhuang, et al.
Veröffentlicht: (2025)
von: Qi, Zhuang, et al.
Veröffentlicht: (2025)
Bilateral Collaboration with Large Vision-Language Models for Open Vocabulary Human-Object Interaction Detection
von: Hu, Yupeng, et al.
Veröffentlicht: (2025)
von: Hu, Yupeng, et al.
Veröffentlicht: (2025)
OpenOcc: Open Vocabulary 3D Scene Reconstruction via Occupancy Representation
von: Jiang, Haochen, et al.
Veröffentlicht: (2024)
von: Jiang, Haochen, et al.
Veröffentlicht: (2024)
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
von: Wang, Zhaowei, et al.
Veröffentlicht: (2024)
von: Wang, Zhaowei, et al.
Veröffentlicht: (2024)
Edit-As-Act: Goal-Regressive Planning for Open-Vocabulary 3D Indoor Scene Editing
von: Noh, Seongrae, et al.
Veröffentlicht: (2026)
von: Noh, Seongrae, et al.
Veröffentlicht: (2026)
Federated Deconfounding and Debiasing Learning for Out-of-Distribution Generalization
von: Qi, Zhuang, et al.
Veröffentlicht: (2025)
von: Qi, Zhuang, et al.
Veröffentlicht: (2025)
DiffuScene: Denoising Diffusion Models for Generative Indoor Scene Synthesis
von: Tang, Jiapeng, et al.
Veröffentlicht: (2023)
von: Tang, Jiapeng, et al.
Veröffentlicht: (2023)
Open-Vocabulary Camouflaged Object Segmentation with Cascaded Vision Language Models
von: Zhao, Kai, et al.
Veröffentlicht: (2025)
von: Zhao, Kai, et al.
Veröffentlicht: (2025)
Open-Vocabulary SAM3D: Towards Training-free Open-Vocabulary 3D Scene Understanding
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
EgoSplat: Open-Vocabulary Egocentric Scene Understanding with Language Embedded 3D Gaussian Splatting
von: Li, Di, et al.
Veröffentlicht: (2025)
von: Li, Di, et al.
Veröffentlicht: (2025)
Taking A Closer Look at Interacting Objects: Interaction-Aware Open Vocabulary Scene Graph Generation
von: Li, Lin, et al.
Veröffentlicht: (2025)
von: Li, Lin, et al.
Veröffentlicht: (2025)
Open Vocabulary Semantic Scene Sketch Understanding
von: Bourouis, Ahmed, et al.
Veröffentlicht: (2023)
von: Bourouis, Ahmed, et al.
Veröffentlicht: (2023)
Class-wise Balancing Data Replay for Federated Class-Incremental Learning
von: Qi, Zhuang, et al.
Veröffentlicht: (2025)
von: Qi, Zhuang, et al.
Veröffentlicht: (2025)
Exploring the Potential of Large Foundation Models for Open-Vocabulary HOI Detection
von: Lei, Ting, et al.
Veröffentlicht: (2024)
von: Lei, Ting, et al.
Veröffentlicht: (2024)
SceneCritic: A Symbolic Evaluator for 3D Indoor Scene Synthesis
von: Sengupta, Kathakoli, et al.
Veröffentlicht: (2026)
von: Sengupta, Kathakoli, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Open-Vocabulary Indoor Object Grounding with 3D Hierarchical Scene Graph
von: Linok, Sergey, et al.
Veröffentlicht: (2025) -
Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes
von: Zhou, Changqing, et al.
Veröffentlicht: (2026) -
Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces
von: Hu, Xinggang, et al.
Veröffentlicht: (2026) -
From Pixels to Graphs: Open-Vocabulary Scene Graph Generation with Vision-Language Models
von: Li, Rongjie, et al.
Veröffentlicht: (2024) -
Open-Vocabulary Semantic Segmentation with Uncertainty Alignment for Robotic Scene Understanding in Indoor Building Environments
von: Xu, Yifan, et al.
Veröffentlicht: (2025)