HumanCrafter: Synergizing Generalizable Human Reconstruction and Semantic 3D Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Pan, Panwang, Shen, Tingting, Li, Chenxin, Lin, Yunlong, Wen, Kairun, Zhao, Jingjing, Yuan, Yixuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diff4Splat: Controllable 4D Scene Generation with Latent Dynamic Reconstruction Models
by: Pan, Panwang, et al.
Published: (2025)
by: Pan, Panwang, et al.
Published: (2025)
ID-Crafter: VLM-Grounded Online RL for Compositional Multi-Subject Video Generation
by: Pan, Panwang, et al.
Published: (2025)
by: Pan, Panwang, et al.
Published: (2025)
GaussianStego: A Generalizable Stenography Pipeline for Generative 3D Gaussians Splatting
by: Li, Chenxin, et al.
Published: (2024)
by: Li, Chenxin, et al.
Published: (2024)
HumanSplat: Generalizable Single-Image Human Gaussian Splatting with Structure Priors
by: Pan, Panwang, et al.
Published: (2024)
by: Pan, Panwang, et al.
Published: (2024)
PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion Transformers
by: Lin, Yuchen, et al.
Published: (2025)
by: Lin, Yuchen, et al.
Published: (2025)
Harnessing Lightweight Transformer with Contextual Synergic Enhancement for Efficient 3D Medical Image Segmentation
by: Liu, Xinyu, et al.
Published: (2026)
by: Liu, Xinyu, et al.
Published: (2026)
JarvisArt: Liberating Human Artistic Creativity via an Intelligent Photo Retouching Agent
by: Lin, Yunlong, et al.
Published: (2025)
by: Lin, Yunlong, et al.
Published: (2025)
InstructLayout: Instruction-Driven 2D and 3D Layout Synthesis with Semantic Graph Prior
by: Lin, Chenguo, et al.
Published: (2024)
by: Lin, Chenguo, et al.
Published: (2024)
MonoSplat: Generalizable 3D Gaussian Splatting from Monocular Depth Foundation Models
by: Liu, Yifan, et al.
Published: (2025)
by: Liu, Yifan, et al.
Published: (2025)
Thinking in Dynamics: How Multimodal Large Language Models Perceive, Track, and Reason Dynamics in Physical 4D World
by: Huang, Yuzhi, et al.
Published: (2026)
by: Huang, Yuzhi, et al.
Published: (2026)
GeneMAN: Generalizable Single-Image 3D Human Reconstruction from Multi-Source Human Data
by: Wang, Wentao, et al.
Published: (2024)
by: Wang, Wentao, et al.
Published: (2024)
JarvisIR: Elevating Autonomous Driving Perception with Intelligent Image Restoration
by: Lin, Yunlong, et al.
Published: (2025)
by: Lin, Yunlong, et al.
Published: (2025)
DynamicVerse: A Physically-Aware Multimodal Framework for 4D World Modeling
by: Wen, Kairun, et al.
Published: (2025)
by: Wen, Kairun, et al.
Published: (2025)
LGS: A Light-weight 4D Gaussian Splatting for Efficient Surgical Scene Reconstruction
by: Liu, Hengyu, et al.
Published: (2024)
by: Liu, Hengyu, et al.
Published: (2024)
Sitcom-Crafter: A Plot-Driven Human Motion Generation System in 3D Scenes
by: Chen, Jianqi, et al.
Published: (2024)
by: Chen, Jianqi, et al.
Published: (2024)
EndoGaussian: Real-time Gaussian Splatting for Dynamic Endoscopic Scene Reconstruction
by: Liu, Yifan, et al.
Published: (2024)
by: Liu, Yifan, et al.
Published: (2024)
Open-Vocabulary Semantic Part Segmentation of 3D Human
by: Suzuki, Keito, et al.
Published: (2025)
by: Suzuki, Keito, et al.
Published: (2025)
X$^{2}$-Gaussian: 4D Radiative Gaussian Splatting for Continuous-time Tomographic Reconstruction
by: Yu, Weihao, et al.
Published: (2025)
by: Yu, Weihao, et al.
Published: (2025)
SyncHuman: Synchronizing 2D and 3D Generative Models for Single-view Human Reconstruction
by: Chen, Wenyue, et al.
Published: (2025)
by: Chen, Wenyue, et al.
Published: (2025)
WonderHuman: Hallucinating Unseen Parts in Dynamic 3D Human Reconstruction
by: Wang, Zilong, et al.
Published: (2025)
by: Wang, Zilong, et al.
Published: (2025)
EfficientHuman: Efficient Training and Reconstruction of Moving Human using Articulated 2D Gaussian
by: Tian, Hao, et al.
Published: (2025)
by: Tian, Hao, et al.
Published: (2025)
SegMo: Segment-aligned Text to 3D Human Motion Generation
by: Dang, Bowen, et al.
Published: (2025)
by: Dang, Bowen, et al.
Published: (2025)
DiHuR: Diffusion-Guided Generalizable Human Reconstruction
by: Chen, Jinnan, et al.
Published: (2024)
by: Chen, Jinnan, et al.
Published: (2024)
Unlocking 3D Affordance Segmentation with 2D Semantic Knowledge
by: Huang, Yu, et al.
Published: (2025)
by: Huang, Yu, et al.
Published: (2025)
Uplifting Range-View-based 3D Semantic Segmentation in Real-Time with Multi-Sensor Fusion
by: Tan, Shiqi, et al.
Published: (2024)
by: Tan, Shiqi, et al.
Published: (2024)
Semantic Human Mesh Reconstruction with Textures
by: Zhan, Xiaoyu, et al.
Published: (2024)
by: Zhan, Xiaoyu, et al.
Published: (2024)
DiffHuman: Probabilistic Photorealistic 3D Reconstruction of Humans
by: Sengupta, Akash, et al.
Published: (2024)
by: Sengupta, Akash, et al.
Published: (2024)
CityCraft: A Real Crafter for 3D City Generation
by: Deng, Jie, et al.
Published: (2024)
by: Deng, Jie, et al.
Published: (2024)
LightGaussian: Unbounded 3D Gaussian Compression with 15x Reduction and 200+ FPS
by: Fan, Zhiwen, et al.
Published: (2023)
by: Fan, Zhiwen, et al.
Published: (2023)
UniSem: Generalizable Semantic 3D Reconstruction from Sparse Unposed Images
by: Liao, Guibiao, et al.
Published: (2026)
by: Liao, Guibiao, et al.
Published: (2026)
GeoT: Geometry-guided Instance-dependent Transition Matrix for Semi-supervised Tooth Point Cloud Segmentation
by: Yu, Weihao, et al.
Published: (2025)
by: Yu, Weihao, et al.
Published: (2025)
Generalizable Human Gaussian Splatting via Multi-view Semantic Consistency
by: Kim, Jingi, et al.
Published: (2026)
by: Kim, Jingi, et al.
Published: (2026)
AniCrafter: Customizing Realistic Human-Centric Animation via Avatar-Background Conditioning in Video Diffusion Models
by: Niu, Muyao, et al.
Published: (2025)
by: Niu, Muyao, et al.
Published: (2025)
HuPrior3R: Incorporating Human Priors for Better 3D Dynamic Reconstruction from Monocular Videos
by: Xiong, Weitao, et al.
Published: (2025)
by: Xiong, Weitao, et al.
Published: (2025)
HumanOrbit: 3D Human Reconstruction as 360° Orbit Generation
by: Suzuki, Keito, et al.
Published: (2026)
by: Suzuki, Keito, et al.
Published: (2026)
Towards Human-Level 3D Relative Pose Estimation: Generalizable, Training-Free, with Single Reference
by: Gao, Yuan, et al.
Published: (2024)
by: Gao, Yuan, et al.
Published: (2024)
Pan-LUT: Efficient Pan-sharpening via Learnable Look-Up Tables
by: Cai, Zhongnan, et al.
Published: (2025)
by: Cai, Zhongnan, et al.
Published: (2025)
Bridge the Gap Between Visual and Linguistic Comprehension for Generalized Zero-shot Semantic Segmentation
by: Guo, Xiaoqing, et al.
Published: (2025)
by: Guo, Xiaoqing, et al.
Published: (2025)
GenDP: 3D Semantic Fields for Category-Level Generalizable Diffusion Policy
by: Wang, Yixuan, et al.
Published: (2024)
by: Wang, Yixuan, et al.
Published: (2024)
CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction
by: Xie, Xianghui, et al.
Published: (2025)
by: Xie, Xianghui, et al.
Published: (2025)
Similar Items
-
Diff4Splat: Controllable 4D Scene Generation with Latent Dynamic Reconstruction Models
by: Pan, Panwang, et al.
Published: (2025) -
ID-Crafter: VLM-Grounded Online RL for Compositional Multi-Subject Video Generation
by: Pan, Panwang, et al.
Published: (2025) -
GaussianStego: A Generalizable Stenography Pipeline for Generative 3D Gaussians Splatting
by: Li, Chenxin, et al.
Published: (2024) -
HumanSplat: Generalizable Single-Image Human Gaussian Splatting with Structure Priors
by: Pan, Panwang, et al.
Published: (2024) -
PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion Transformers
by: Lin, Yuchen, et al.
Published: (2025)