From Part to Whole: 3D Generative World Model with an Adaptive Structural Hierarchy
Fuente:
arXiv
Guardado en:
| Autores principales: | Du, Bi'an, Liu, Daizong, Li, Pufan, Hu, Wei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Generative 3D Part Assembly via Part-Whole-Hierarchy Message Passing
por: Du, Bi'an, et al.
Publicado: (2024)
por: Du, Bi'an, et al.
Publicado: (2024)
HierOctFusion: Multi-scale Octree-based 3D Shape Generation via Part-Whole-Hierarchy Message Passing
por: Gao, Xinjie, et al.
Publicado: (2025)
por: Gao, Xinjie, et al.
Publicado: (2025)
Geometry and Perception Guided Gaussians for Multiview-consistent 3D Generation from a Single Image
por: Li, Pufan, et al.
Publicado: (2025)
por: Li, Pufan, et al.
Publicado: (2025)
View-Consistent 3D Scene Editing via Dual-Path Structural Correspondense and Semantic Continuity
por: Li, Pufan, et al.
Publicado: (2026)
por: Li, Pufan, et al.
Publicado: (2026)
Multi-scale Latent Point Consistency Models for 3D Shape Generation
por: Du, Bi'an, et al.
Publicado: (2024)
por: Du, Bi'an, et al.
Publicado: (2024)
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding
por: Huang, Wencan, et al.
Publicado: (2025)
por: Huang, Wencan, et al.
Publicado: (2025)
AugGS: Self-augmented Gaussians with Structural Masks for Sparse-view 3D Reconstruction
por: Du, Bi'an, et al.
Publicado: (2024)
por: Du, Bi'an, et al.
Publicado: (2024)
Joint Top-Down and Bottom-Up Frameworks for 3D Visual Grounding
por: Liu, Yang, et al.
Publicado: (2024)
por: Liu, Yang, et al.
Publicado: (2024)
Improving the Transferability of 3D Point Cloud Attack via Spectral-aware Admix and Optimization Designs
por: Hu, Shiyu, et al.
Publicado: (2024)
por: Hu, Shiyu, et al.
Publicado: (2024)
A Survey on Text-guided 3D Visual Grounding: Elements, Recent Advances, and Future Directions
por: Liu, Daizong, et al.
Publicado: (2024)
por: Liu, Daizong, et al.
Publicado: (2024)
AdaCo: Overcoming Visual Foundation Model Noise in 3D Semantic Segmentation via Adaptive Label Correction
por: Zou, Pufan, et al.
Publicado: (2024)
por: Zou, Pufan, et al.
Publicado: (2024)
PanoWorld: A Generative Spatial World Model for Consistent Whole-House Panorama Synthesis
por: Jia, Jinrang, et al.
Publicado: (2026)
por: Jia, Jinrang, et al.
Publicado: (2026)
Hard-Label Black-Box Attacks on 3D Point Clouds
por: Liu, Daizong, et al.
Publicado: (2024)
por: Liu, Daizong, et al.
Publicado: (2024)
Representing Part-Whole Hierarchies in Foundation Models by Learning Localizability, Composability, and Decomposability from Anatomy via Self-Supervision
por: Taher, Mohammad Reza Hosseinzadeh, et al.
Publicado: (2024)
por: Taher, Mohammad Reza Hosseinzadeh, et al.
Publicado: (2024)
UniPart: Part-Level 3D Generation with Unified 3D Geom-Seg Latents
por: He, Xufan, et al.
Publicado: (2025)
por: He, Xufan, et al.
Publicado: (2025)
Recursive Neural Programs: Variational Learning of Image Grammars and Part-Whole Hierarchies
por: Fisher, Ares, et al.
Publicado: (2022)
por: Fisher, Ares, et al.
Publicado: (2022)
A Survey of Attacks on Large Vision-Language Models: Resources, Advances, and Future Trends
por: Liu, Daizong, et al.
Publicado: (2024)
por: Liu, Daizong, et al.
Publicado: (2024)
OmniPart: Part-Aware 3D Generation with Semantic Decoupling and Structural Cohesion
por: Yang, Yunhan, et al.
Publicado: (2025)
por: Yang, Yunhan, et al.
Publicado: (2025)
Chart Specification: Structural Representations for Incentivizing VLM Reasoning in Chart-to-Code Generation
por: He, Minggui, et al.
Publicado: (2026)
por: He, Minggui, et al.
Publicado: (2026)
From Parts to Whole: A Unified Reference Framework for Controllable Human Image Generation
por: Huang, Zehuan, et al.
Publicado: (2024)
por: Huang, Zehuan, et al.
Publicado: (2024)
SegviGen: Repurposing 3D Generative Model for Part Segmentation
por: Li, Lin, et al.
Publicado: (2026)
por: Li, Lin, et al.
Publicado: (2026)
From 2D to 3D Cognition: A Brief Survey of General World Models
por: Xie, Ningwei, et al.
Publicado: (2025)
por: Xie, Ningwei, et al.
Publicado: (2025)
PAFUSE: Part-based Diffusion for 3D Whole-Body Pose Estimation
por: Samet, Nermin, et al.
Publicado: (2024)
por: Samet, Nermin, et al.
Publicado: (2024)
OneWorld: Taming Scene Generation with 3D Unified Representation Autoencoder
por: Gao, Sensen, et al.
Publicado: (2026)
por: Gao, Sensen, et al.
Publicado: (2026)
HoloPart: Generative 3D Part Amodal Segmentation
por: Yang, Yunhan, et al.
Publicado: (2025)
por: Yang, Yunhan, et al.
Publicado: (2025)
UrbanWorld: An Urban World Model for 3D City Generation
por: Shang, Yu, et al.
Publicado: (2024)
por: Shang, Yu, et al.
Publicado: (2024)
From One to More: Contextual Part Latents for 3D Generation
por: Dong, Shaocong, et al.
Publicado: (2025)
por: Dong, Shaocong, et al.
Publicado: (2025)
SP3D: Boosting Sparsely-Supervised 3D Object Detection via Accurate Cross-Modal Semantic Prompts
por: Zhao, Shijia, et al.
Publicado: (2025)
por: Zhao, Shijia, et al.
Publicado: (2025)
CoSMo3D: Open-World Promptable 3D Semantic Part Segmentation through LLM-Guided Canonical Spatial Modeling
por: Jin, Li, et al.
Publicado: (2026)
por: Jin, Li, et al.
Publicado: (2026)
Behave Your Motion: Habit-preserved Cross-category Animal Motion Transfer
por: Zhang, Zhimin, et al.
Publicado: (2025)
por: Zhang, Zhimin, et al.
Publicado: (2025)
Part-Whole Relational Fusion Towards Multi-Modal Scene Understanding
por: Liu, Yi, et al.
Publicado: (2024)
por: Liu, Yi, et al.
Publicado: (2024)
HOLODECK 2.0: Vision-Language-Guided 3D World Generation with Editing
por: Bian, Zixuan, et al.
Publicado: (2025)
por: Bian, Zixuan, et al.
Publicado: (2025)
HumanOrbit: 3D Human Reconstruction as 360° Orbit Generation
por: Suzuki, Keito, et al.
Publicado: (2026)
por: Suzuki, Keito, et al.
Publicado: (2026)
An Image Is Worth Ten Thousand Words: Verbose-Text Induction Attacks on VLMs
por: Luo, Zhi, et al.
Publicado: (2025)
por: Luo, Zhi, et al.
Publicado: (2025)
Rethinking Video-Language Model from the Language Input Perspective
por: Fang, Xiang, et al.
Publicado: (2026)
por: Fang, Xiang, et al.
Publicado: (2026)
StructLDM: Structured Latent Diffusion for 3D Human Generation
por: Hu, Tao, et al.
Publicado: (2024)
por: Hu, Tao, et al.
Publicado: (2024)
MotionCraft: Crafting Whole-Body Motion with Plug-and-Play Multimodal Controls
por: Bian, Yuxuan, et al.
Publicado: (2024)
por: Bian, Yuxuan, et al.
Publicado: (2024)
OmniMotion-X: Versatile Multimodal Whole-Body Motion Generation
por: Xu, Guowei, et al.
Publicado: (2025)
por: Xu, Guowei, et al.
Publicado: (2025)
PartRAG: Retrieval-Augmented Part-Level 3D Generation and Editing
por: Li, Peize, et al.
Publicado: (2026)
por: Li, Peize, et al.
Publicado: (2026)
WorldGrow: Generating Infinite 3D World
por: Li, Sikuang, et al.
Publicado: (2025)
por: Li, Sikuang, et al.
Publicado: (2025)
Ejemplares similares
-
Generative 3D Part Assembly via Part-Whole-Hierarchy Message Passing
por: Du, Bi'an, et al.
Publicado: (2024) -
HierOctFusion: Multi-scale Octree-based 3D Shape Generation via Part-Whole-Hierarchy Message Passing
por: Gao, Xinjie, et al.
Publicado: (2025) -
Geometry and Perception Guided Gaussians for Multiview-consistent 3D Generation from a Single Image
por: Li, Pufan, et al.
Publicado: (2025) -
View-Consistent 3D Scene Editing via Dual-Path Structural Correspondense and Semantic Continuity
por: Li, Pufan, et al.
Publicado: (2026) -
Multi-scale Latent Point Consistency Models for 3D Shape Generation
por: Du, Bi'an, et al.
Publicado: (2024)