IPoD: Implicit Field Learning with Point Diffusion for Generalizable 3D Object Reconstruction from Single RGB-D Images
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Yushuang, Shi, Luyue, Cai, Junhao, Yuan, Weihao, Qiu, Lingteng, Dong, Zilong, Bo, Liefeng, Cui, Shuguang, Han, Xiaoguang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MVImgNet2.0: A Larger-scale Dataset of Multi-view Images
by: Han, Xiaoguang, et al.
Published: (2024)
by: Han, Xiaoguang, et al.
Published: (2024)
DreamDissector: Learning Disentangled Text-to-3D Generation from 2D Diffusion Priors
by: Yan, Zizheng, et al.
Published: (2024)
by: Yan, Zizheng, et al.
Published: (2024)
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
by: Cai, Junhao, et al.
Published: (2024)
by: Cai, Junhao, et al.
Published: (2024)
StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal
by: Ye, Chongjie, et al.
Published: (2024)
by: Ye, Chongjie, et al.
Published: (2024)
LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
by: Qiu, Lingteng, et al.
Published: (2025)
by: Qiu, Lingteng, et al.
Published: (2025)
PartNerFace: Part-based Neural Radiance Fields for Animatable Facial Avatar Reconstruction
by: Yu, Xianggang, et al.
Published: (2026)
by: Yu, Xianggang, et al.
Published: (2026)
AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
by: Qiu, Lingteng, et al.
Published: (2024)
by: Qiu, Lingteng, et al.
Published: (2024)
LHM++: An Efficient Large Human Reconstruction Model for Pose-free Images to 3D
by: Qiu, Lingteng, et al.
Published: (2025)
by: Qiu, Lingteng, et al.
Published: (2025)
SAMPro3D: Locating SAM Prompts in 3D for Zero-Shot Instance Segmentation
by: Xu, Mutian, et al.
Published: (2023)
by: Xu, Mutian, et al.
Published: (2023)
GIC: Gaussian-Informed Continuum for Physical Property Identification and Simulation
by: Cai, Junhao, et al.
Published: (2024)
by: Cai, Junhao, et al.
Published: (2024)
Sketch2NeRF: Multi-view Sketch-guided Text-to-3D Generation
by: Chen, Minglin, et al.
Published: (2024)
by: Chen, Minglin, et al.
Published: (2024)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
by: Zuo, Qi, et al.
Published: (2024)
by: Zuo, Qi, et al.
Published: (2024)
Hyperbolic-constraint Point Cloud Reconstruction from Single RGB-D Images
by: Li, Wenrui, et al.
Published: (2024)
by: Li, Wenrui, et al.
Published: (2024)
ReconViaGen: Towards Accurate Multi-view 3D Object Reconstruction via Generation
by: Chang, Jiahao, et al.
Published: (2025)
by: Chang, Jiahao, et al.
Published: (2025)
An Optimization Framework to Enforce Multi-View Consistency for Texturing 3D Meshes
by: Zhao, Zhengyi, et al.
Published: (2024)
by: Zhao, Zhengyi, et al.
Published: (2024)
Condition Matters in Full-head 3D GANs
by: Li, Heyuan, et al.
Published: (2026)
by: Li, Heyuan, et al.
Published: (2026)
GaussianIP: Identity-Preserving Realistic 3D Human Generation via Human-Centric Diffusion Prior
by: Tang, Zichen, et al.
Published: (2025)
by: Tang, Zichen, et al.
Published: (2025)
RayletDF: Raylet Distance Fields for Generalizable 3D Surface Reconstruction from Point Clouds or Gaussians
by: Wei, Shenxing, et al.
Published: (2025)
by: Wei, Shenxing, et al.
Published: (2025)
RGB2Point: 3D Point Cloud Generation from Single RGB Images
by: Lee, Jae Joong, et al.
Published: (2024)
by: Lee, Jae Joong, et al.
Published: (2024)
Towards Flexible 3D Perception: Object-Centric Occupancy Completion Augments 3D Object Detection
by: Zheng, Chaoda, et al.
Published: (2024)
by: Zheng, Chaoda, et al.
Published: (2024)
HyPlaneHead: Rethinking Tri-plane-like Representations in Full-Head Image Synthesis
by: Li, Heyuan, et al.
Published: (2025)
by: Li, Heyuan, et al.
Published: (2025)
4DRecons: 4D Neural Implicit Deformable Objects Reconstruction from a single RGB-D Camera with Geometrical and Topological Regularizations
by: Cong, Xiaoyan, et al.
Published: (2024)
by: Cong, Xiaoyan, et al.
Published: (2024)
Human as Points: Explicit Point-based 3D Human Reconstruction from Single-view RGB Images
by: Tang, Yingzhi, et al.
Published: (2023)
by: Tang, Yingzhi, et al.
Published: (2023)
TexSpot: 3D Texture Enhancement with Spatially-uniform Point Latent Representation
by: Lu, Ziteng, et al.
Published: (2026)
by: Lu, Ziteng, et al.
Published: (2026)
HIVE: HIerarchical Volume Encoding for Neural Implicit Surface Reconstruction
by: Gu, Xiaodong, et al.
Published: (2024)
by: Gu, Xiaodong, et al.
Published: (2024)
Stable-Sim2Real: Exploring Simulation of Real-Captured 3D Data with Two-Stage Depth Diffusion
by: Xu, Mutian, et al.
Published: (2025)
by: Xu, Mutian, et al.
Published: (2025)
Freditor: High-Fidelity and Transferable NeRF Editing by Frequency Decomposition
by: He, Yisheng, et al.
Published: (2024)
by: He, Yisheng, et al.
Published: (2024)
ASGrasp: Generalizable Transparent Object Reconstruction and 6-DoF Grasp Detection from RGB-D Active Stereo Camera
by: Shi, Jun, et al.
Published: (2024)
by: Shi, Jun, et al.
Published: (2024)
Forge4D: Feed-Forward 4D Human Reconstruction and Interpolation from Uncalibrated Sparse-view Videos
by: Hu, Yingdong, et al.
Published: (2025)
by: Hu, Yingdong, et al.
Published: (2025)
ViSA: 3D-Aware Video Shading for Real-Time Upper-Body Avatar Creation
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
by: Huang, Zixuan, et al.
Published: (2025)
by: Huang, Zixuan, et al.
Published: (2025)
Textoon: Generating Vivid 2D Cartoon Characters from Text Descriptions
by: He, Chao, et al.
Published: (2025)
by: He, Chao, et al.
Published: (2025)
Towards Unified 3D Hair Reconstruction from Single-View Portraits
by: Zheng, Yujian, et al.
Published: (2024)
by: Zheng, Yujian, et al.
Published: (2024)
Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation
by: Chen, Yingjie, et al.
Published: (2025)
by: Chen, Yingjie, et al.
Published: (2025)
HiSplat: Hierarchical 3D Gaussian Splatting for Generalizable Sparse-View Reconstruction
by: Tang, Shengji, et al.
Published: (2024)
by: Tang, Shengji, et al.
Published: (2024)
Fully Test-Time Adaptation for Monocular 3D Object Detection
by: Lin, Hongbin, et al.
Published: (2024)
by: Lin, Hongbin, et al.
Published: (2024)
TransDiff: Diffusion-Based Method for Manipulating Transparent Objects Using a Single RGB-D Image
by: Wang, Haoxiao, et al.
Published: (2025)
by: Wang, Haoxiao, et al.
Published: (2025)
Category Level 6D Object Pose Estimation from a Single RGB Image using Diffusion
by: Bethell, Adam, et al.
Published: (2024)
by: Bethell, Adam, et al.
Published: (2024)
FlowTrack: Point-level Flow Network for 3D Single Object Tracking
by: Li, Shuo, et al.
Published: (2024)
by: Li, Shuo, et al.
Published: (2024)
ANIM: Accurate Neural Implicit Model for Human Reconstruction from a single RGB-D image
by: Pesavento, Marco, et al.
Published: (2024)
by: Pesavento, Marco, et al.
Published: (2024)
Similar Items
-
MVImgNet2.0: A Larger-scale Dataset of Multi-view Images
by: Han, Xiaoguang, et al.
Published: (2024) -
DreamDissector: Learning Disentangled Text-to-3D Generation from 2D Diffusion Priors
by: Yan, Zizheng, et al.
Published: (2024) -
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
by: Cai, Junhao, et al.
Published: (2024) -
StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal
by: Ye, Chongjie, et al.
Published: (2024) -
LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
by: Qiu, Lingteng, et al.
Published: (2025)