MV-SAM3D: Adaptive Multi-View Fusion for Layout-Aware 3D Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Baicheng, Wu, Dong, Li, Jun, Zhou, Shunkai, Zeng, Zecui, Li, Lusong, Zha, Hongbin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Proactive Scene Decomposition and Reconstruction
by: Li, Baicheng, et al.
Published: (2025)
by: Li, Baicheng, et al.
Published: (2025)
Reflection-Based Task Adaptation for Self-Improving VLA
by: Li, Baicheng, et al.
Published: (2025)
by: Li, Baicheng, et al.
Published: (2025)
Online3R: Online Learning for Consistent Sequential Reconstruction Based on Geometry Foundation Model
by: Zhou, Shunkai, et al.
Published: (2026)
by: Zhou, Shunkai, et al.
Published: (2026)
Learn to Memorize and to Forget: A Continual Learning Perspective of Dynamic SLAM
by: Li, Baicheng, et al.
Published: (2024)
by: Li, Baicheng, et al.
Published: (2024)
MV-MOS: Multi-View Feature Fusion for 3D Moving Object Segmentation
by: Cheng, Jintao, et al.
Published: (2024)
by: Cheng, Jintao, et al.
Published: (2024)
REPARO: Compositional 3D Assets Generation with Differentiable 3D Layout Alignment
by: Han, Haonan, et al.
Published: (2024)
by: Han, Haonan, et al.
Published: (2024)
SDesc3D: Towards Layout-Aware 3D Indoor Scene Generation from Short Descriptions
by: Feng, Jie, et al.
Published: (2026)
by: Feng, Jie, et al.
Published: (2026)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
by: Song, Dan, et al.
Published: (2023)
by: Song, Dan, et al.
Published: (2023)
Layout-your-3D: Controllable and Precise 3D Generation with 2D Blueprint
by: Zhou, Junwei, et al.
Published: (2024)
by: Zhou, Junwei, et al.
Published: (2024)
MV-SSM: Multi-View State Space Modeling for 3D Human Pose Estimation
by: Chharia, Aviral, et al.
Published: (2025)
by: Chharia, Aviral, et al.
Published: (2025)
AdaptiveFusion: Adaptive Multi-Modal Multi-View Fusion for 3D Human Body Reconstruction
by: Chen, Anjun, et al.
Published: (2024)
by: Chen, Anjun, et al.
Published: (2024)
FusionBERT: Multi-View Image-3D Retrieval via Cross-Attention Visual Fusion and Normal-Aware 3D Encoder
by: Li, Wei, et al.
Published: (2026)
by: Li, Wei, et al.
Published: (2026)
3D-Aware Implicit Motion Control for View-Adaptive Human Video Generation
by: Fang, Zhixue, et al.
Published: (2026)
by: Fang, Zhixue, et al.
Published: (2026)
RoboFusion: Towards Robust Multi-Modal 3D Object Detection via SAM
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
SPATIALGEN: Layout-guided 3D Indoor Scene Generation
by: Fang, Chuan, et al.
Published: (2025)
by: Fang, Chuan, et al.
Published: (2025)
Hitem3D 2.0: Multi-View Guided Native 3D Texture Generation
by: He, Huiang, et al.
Published: (2026)
by: He, Huiang, et al.
Published: (2026)
M3DLayout: A Multi-Source Dataset of 3D Indoor Layouts and Structured Descriptions for 3D Generation
by: Zhang, Yiheng, et al.
Published: (2025)
by: Zhang, Yiheng, et al.
Published: (2025)
MV3DIS: Multi-View Mask Matching via 3D Guides for Zero-Shot 3D Instance Segmentation
by: Zhao, Yibo, et al.
Published: (2026)
by: Zhao, Yibo, et al.
Published: (2026)
MV2Cyl: Reconstructing 3D Extrusion Cylinders from Multi-View Images
by: Hong, Eunji, et al.
Published: (2024)
by: Hong, Eunji, et al.
Published: (2024)
TAGS: 3D Tumor-Adaptive Guidance for SAM
by: Li, Sirui, et al.
Published: (2025)
by: Li, Sirui, et al.
Published: (2025)
Ref-SAM3D: Bridging SAM3D with Text for Reference 3D Reconstruction
by: Zhou, Yun, et al.
Published: (2025)
by: Zhou, Yun, et al.
Published: (2025)
StereoMV2D: A Sparse Temporal Stereo-Enhanced Framework for Robust Multi-View 3D Object Detection
by: Wu, Di, et al.
Published: (2025)
by: Wu, Di, et al.
Published: (2025)
WildFusion: Learning 3D-Aware Latent Diffusion Models in View Space
by: Schwarz, Katja, et al.
Published: (2023)
by: Schwarz, Katja, et al.
Published: (2023)
ForgeDreamer: Industrial Text-to-3D Generation with Multi-Expert LoRA and Cross-View Hypergraph
by: Cai, Junhao, et al.
Published: (2026)
by: Cai, Junhao, et al.
Published: (2026)
SpatialFusion: Endowing Unified Image Generation with Intrinsic 3D Geometric Awareness
by: Qiu, Haiyi, et al.
Published: (2026)
by: Qiu, Haiyi, et al.
Published: (2026)
SceneCraft: Layout-Guided 3D Scene Generation
by: Yang, Xiuyu, et al.
Published: (2024)
by: Yang, Xiuyu, et al.
Published: (2024)
CLIP3D-AD: Extending CLIP for 3D Few-Shot Anomaly Detection with Multi-View Images Generation
by: Zuo, Zuo, et al.
Published: (2024)
by: Zuo, Zuo, et al.
Published: (2024)
LAYOUTDREAMER: Physics-guided Layout for Text-to-3D Compositional Scene Generation
by: Zhou, Yang, et al.
Published: (2025)
by: Zhou, Yang, et al.
Published: (2025)
ToLL: Topological Layout Learning with Asymmetric Cross-View Structural Distillation for 3D Scene Graph Generation Pretraining
by: Huang, Yucheng, et al.
Published: (2026)
by: Huang, Yucheng, et al.
Published: (2026)
Multi-View Attentive Contextualization for Multi-View 3D Object Detection
by: Liu, Xianpeng, et al.
Published: (2024)
by: Liu, Xianpeng, et al.
Published: (2024)
SAM 3D for 3D Object Reconstruction from Remote Sensing Images
by: Yao, Junsheng, et al.
Published: (2025)
by: Yao, Junsheng, et al.
Published: (2025)
NeRF-DetS: Enhanced Adaptive Spatial-wise Sampling and View-wise Fusion Strategies for NeRF-based Indoor Multi-view 3D Object Detection
by: Huang, Chi, et al.
Published: (2024)
by: Huang, Chi, et al.
Published: (2024)
AutoProSAM: Automated Prompting SAM for 3D Multi-Organ Segmentation
by: Li, Chengyin, et al.
Published: (2023)
by: Li, Chengyin, et al.
Published: (2023)
Embodied Tree of Thoughts: Deliberate Manipulation Planning with Embodied World Model
by: Xu, Wenjiang, et al.
Published: (2025)
by: Xu, Wenjiang, et al.
Published: (2025)
LabBuilder: Protocol-Grounded 3D Layout Generation for Interactable and Safe Laboratory
by: Cao, Jianbao, et al.
Published: (2026)
by: Cao, Jianbao, et al.
Published: (2026)
IL3D: A Large-Scale Indoor Layout Dataset for LLM-Driven 3D Scene Generation
by: Zhou, Wenxu, et al.
Published: (2025)
by: Zhou, Wenxu, et al.
Published: (2025)
Shape from Semantics: 3D Shape Generation from Multi-View Semantics
by: Li, Liangchen, et al.
Published: (2025)
by: Li, Liangchen, et al.
Published: (2025)
Pose-Aware Diffusion for 3D Generation
by: Zhou, Zihan, et al.
Published: (2026)
by: Zhou, Zihan, et al.
Published: (2026)
Perceive-then-Plan: Layout-as-Policy for Monocular 3D Scene Layout Estimation
by: Zhou, Junwei, et al.
Published: (2026)
by: Zhou, Junwei, et al.
Published: (2026)
UniDA3D: A Unified Domain-Adaptive Framework for Multi-View 3D Object Detection
by: Wu, Hongjing, et al.
Published: (2026)
by: Wu, Hongjing, et al.
Published: (2026)
Similar Items
-
Proactive Scene Decomposition and Reconstruction
by: Li, Baicheng, et al.
Published: (2025) -
Reflection-Based Task Adaptation for Self-Improving VLA
by: Li, Baicheng, et al.
Published: (2025) -
Online3R: Online Learning for Consistent Sequential Reconstruction Based on Geometry Foundation Model
by: Zhou, Shunkai, et al.
Published: (2026) -
Learn to Memorize and to Forget: A Continual Learning Perspective of Dynamic SLAM
by: Li, Baicheng, et al.
Published: (2024) -
MV-MOS: Multi-View Feature Fusion for 3D Moving Object Segmentation
by: Cheng, Jintao, et al.
Published: (2024)