MMGDreamer: Mixed-Modality Graph for Geometry-Controllable 3D Indoor Scene Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Zhifei, Lu, Keyang, Zhang, Chao, Qi, Jiaxing, Jiang, Hanqi, Ma, Ruifei, Yin, Shenglin, Xu, Yifan, Xing, Mingzhe, Xiao, Zhen, Long, Jieyi, Zhai, Guangyao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FlowScene: Style-Consistent Indoor Scene Generation with Multimodal Graph Rectified Flow
by: Yang, Zhifei, et al.
Published: (2026)
by: Yang, Zhifei, et al.
Published: (2026)
Yo'City: Personalized and Boundless 3D Realistic City Scene Generation via Self-Critic Expansion
by: Lu, Keyang, et al.
Published: (2025)
by: Lu, Keyang, et al.
Published: (2025)
Spatial 3D-LLM: Exploring Spatial Awareness in 3D Vision-Language Models
by: Wang, Xiaoyan, et al.
Published: (2025)
by: Wang, Xiaoyan, et al.
Published: (2025)
CommonScenes: Generating Commonsense 3D Indoor Scenes with Scene Graph Diffusion
by: Zhai, Guangyao, et al.
Published: (2023)
by: Zhai, Guangyao, et al.
Published: (2023)
EchoScene: Indoor Scene Generation via Information Echo over Scene Graph Diffusion
by: Zhai, Guangyao, et al.
Published: (2024)
by: Zhai, Guangyao, et al.
Published: (2024)
Argus: Leveraging Multiview Images for Improved 3-D Scene Understanding With Large Language Models
by: Xu, Yifan, et al.
Published: (2025)
by: Xu, Yifan, et al.
Published: (2025)
DARA: Few-shot Budget Allocation in Online Advertising via In-Context Decision Making with RL-Finetuned LLMs
by: Song, Mingxuan, et al.
Published: (2026)
by: Song, Mingxuan, et al.
Published: (2026)
Geometry-guided Feature Learning and Fusion for Indoor Scene Reconstruction
by: Yin, Ruihong, et al.
Published: (2024)
by: Yin, Ruihong, et al.
Published: (2024)
GeoGaussian: Geometry-aware Gaussian Splatting for Scene Rendering
by: Li, Yanyan, et al.
Published: (2024)
by: Li, Yanyan, et al.
Published: (2024)
3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding
by: Li, Zeju, et al.
Published: (2024)
by: Li, Zeju, et al.
Published: (2024)
Mixed Diffusion for 3D Indoor Scene Synthesis
by: Hu, Siyi, et al.
Published: (2024)
by: Hu, Siyi, et al.
Published: (2024)
SG-Tailor: Inter-Object Commonsense Relationship Reasoning for Scene Graph Manipulation
by: Shang, Haoliang, et al.
Published: (2025)
by: Shang, Haoliang, et al.
Published: (2025)
SG-Bot: Object Rearrangement via Coarse-to-Fine Robotic Imagination on Scene Graphs
by: Zhai, Guangyao, et al.
Published: (2023)
by: Zhai, Guangyao, et al.
Published: (2023)
Mobile Robot Oriented Large-Scale Indoor Dataset for Dynamic Scene Understanding
by: Tang, Yifan, et al.
Published: (2024)
by: Tang, Yifan, et al.
Published: (2024)
Attention over Scene Graphs: Indoor Scene Representations Toward CSAI Classification
by: Barros, Artur, et al.
Published: (2025)
by: Barros, Artur, et al.
Published: (2025)
GeoSceneGraph: Geometric Scene Graph Diffusion Model for Text-guided 3D Indoor Scene Synthesis
by: Ruiz, Antonio, et al.
Published: (2025)
by: Ruiz, Antonio, et al.
Published: (2025)
Inter-object Discriminative Graph Modeling for Indoor Scene Recognition
by: Song, Chuanxin, et al.
Published: (2023)
by: Song, Chuanxin, et al.
Published: (2023)
Swin3D++: Effective Multi-Source Pretraining for 3D Indoor Scene Understanding
by: Yang, Yu-Qi, et al.
Published: (2024)
by: Yang, Yu-Qi, et al.
Published: (2024)
Expert Knowledge-Guided Decision Calibration for Accurate Fine-Grained Tree Species Classification
by: Long, Chen, et al.
Published: (2026)
by: Long, Chen, et al.
Published: (2026)
Understanding the Weakness of Large Language Model Agents within a Complex Android Environment
by: Xing, Mingzhe, et al.
Published: (2024)
by: Xing, Mingzhe, et al.
Published: (2024)
Unsupervised Radio Map Construction in Mixed LoS/NLoS Indoor Environments
by: Xing, Zheng, et al.
Published: (2025)
by: Xing, Zheng, et al.
Published: (2025)
Intelligent Spatial Perception by Building Hierarchical 3D Scene Graphs for Indoor Scenarios with the Help of LLMs
by: Cheng, Yao, et al.
Published: (2025)
by: Cheng, Yao, et al.
Published: (2025)
InstructScene: Instruction-Driven 3D Indoor Scene Synthesis with Semantic Graph Prior
by: Lin, Chenguo, et al.
Published: (2024)
by: Lin, Chenguo, et al.
Published: (2024)
Open-Vocabulary Semantic Segmentation with Uncertainty Alignment for Robotic Scene Understanding in Indoor Building Environments
by: Xu, Yifan, et al.
Published: (2025)
by: Xu, Yifan, et al.
Published: (2025)
GUME: Graphs and User Modalities Enhancement for Long-Tail Multimodal Recommendation
by: Lin, Guojiao, et al.
Published: (2024)
by: Lin, Guojiao, et al.
Published: (2024)
Autonomous Implicit Indoor Scene Reconstruction with Frontier Exploration
by: Zeng, Jing, et al.
Published: (2024)
by: Zeng, Jing, et al.
Published: (2024)
Multi-Modal Representation Learning for Molecular Property Prediction: Sequence, Graph, Geometry
by: Wang, Zeyu, et al.
Published: (2024)
by: Wang, Zeyu, et al.
Published: (2024)
WHU-STree: A Multi-modal Benchmark Dataset for Street Tree Inventory
by: Ding, Ruifei, et al.
Published: (2025)
by: Ding, Ruifei, et al.
Published: (2025)
Contrast-X: A Multi-Modal Contrast Image Synthesis Benchmark and Universal Modality Flow Matching
by: Chen, Yifan, et al.
Published: (2026)
by: Chen, Yifan, et al.
Published: (2026)
Multi-Modal Scene Graph with Kolmogorov-Arnold Experts for Audio-Visual Question Answering
by: Fu, Zijian, et al.
Published: (2025)
by: Fu, Zijian, et al.
Published: (2025)
A Local Differential Privacy Method With Layer‐Wise Importance Based on Fisher Information in Federated Recommendation Systems
by: Jieyi Yan, et al.
Published: (2025)
by: Jieyi Yan, et al.
Published: (2025)
Infinigen Indoors: Photorealistic Indoor Scenes using Procedural Generation
by: Raistrick, Alexander, et al.
Published: (2024)
by: Raistrick, Alexander, et al.
Published: (2024)
Open-Vocabulary Indoor Object Grounding with 3D Hierarchical Scene Graph
by: Linok, Sergey, et al.
Published: (2025)
by: Linok, Sergey, et al.
Published: (2025)
SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs
by: Miao, Yang, et al.
Published: (2024)
by: Miao, Yang, et al.
Published: (2024)
LumiNet: Latent Intrinsics Meets Diffusion Models for Indoor Scene Relighting
by: Xing, Xiaoyan, et al.
Published: (2024)
by: Xing, Xiaoyan, et al.
Published: (2024)
Global-Local Tree Search in VLMs for 3D Indoor Scene Generation
by: Deng, Wei, et al.
Published: (2025)
by: Deng, Wei, et al.
Published: (2025)
Video Perception Models for 3D Scene Synthesis
by: Huang, Rui, et al.
Published: (2025)
by: Huang, Rui, et al.
Published: (2025)
DebSDF: Delving into the Details and Bias of Neural Indoor Scene Reconstruction
by: Xiao, Yuting, et al.
Published: (2023)
by: Xiao, Yuting, et al.
Published: (2023)
AV-Unified: A Unified Framework for Audio-visual Scene Understanding
by: Li, Guangyao, et al.
Published: (2026)
by: Li, Guangyao, et al.
Published: (2026)
Beyond Cross-Modal Alignment: Measuring and Leveraging Modality Gap in Vision-Language Models
by: Yan, Hanqi, et al.
Published: (2025)
by: Yan, Hanqi, et al.
Published: (2025)
Similar Items
-
FlowScene: Style-Consistent Indoor Scene Generation with Multimodal Graph Rectified Flow
by: Yang, Zhifei, et al.
Published: (2026) -
Yo'City: Personalized and Boundless 3D Realistic City Scene Generation via Self-Critic Expansion
by: Lu, Keyang, et al.
Published: (2025) -
Spatial 3D-LLM: Exploring Spatial Awareness in 3D Vision-Language Models
by: Wang, Xiaoyan, et al.
Published: (2025) -
CommonScenes: Generating Commonsense 3D Indoor Scenes with Scene Graph Diffusion
by: Zhai, Guangyao, et al.
Published: (2023) -
EchoScene: Indoor Scene Generation via Information Echo over Scene Graph Diffusion
by: Zhai, Guangyao, et al.
Published: (2024)