DepR: Depth Guided Single-view Scene Reconstruction with Instance-level Diffusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Qingcheng, Zhang, Xiang, Xu, Haiyang, Chen, Zeyuan, Xie, Jianwen, Gao, Yuan, Tu, Zhuowen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CVP: Central-Peripheral Vision-Inspired Multimodal Model for Spatial Reasoning
von: Chen, Zeyuan, et al.
Veröffentlicht: (2025)
von: Chen, Zeyuan, et al.
Veröffentlicht: (2025)
PixARMesh: Autoregressive Mesh-Native Single-View Scene Reconstruction
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
Bayesian Diffusion Models for 3D Shape Reconstruction
von: Xu, Haiyang, et al.
Veröffentlicht: (2024)
von: Xu, Haiyang, et al.
Veröffentlicht: (2024)
OverLayBench: A Benchmark for Layout-to-Image Generation with Dense Overlaps
von: Li, Bingnan, et al.
Veröffentlicht: (2025)
von: Li, Bingnan, et al.
Veröffentlicht: (2025)
Lay-Your-Scene: Natural Scene Layout Generation with Diffusion Transformers
von: Srivastava, Divyansh, et al.
Veröffentlicht: (2025)
von: Srivastava, Divyansh, et al.
Veröffentlicht: (2025)
YOLO-Count: Differentiable Object Counting for Text-to-Image Generation
von: Zeng, Guanning, et al.
Veröffentlicht: (2025)
von: Zeng, Guanning, et al.
Veröffentlicht: (2025)
OmniControlNet: Dual-stage Integration for Conditional Image Generation
von: Wang, Yilin, et al.
Veröffentlicht: (2024)
von: Wang, Yilin, et al.
Veröffentlicht: (2024)
Exploring the Equivalence of Closed-Set Generative and Real Data Augmentation in Image Classification
von: Wang, Haowen, et al.
Veröffentlicht: (2025)
von: Wang, Haowen, et al.
Veröffentlicht: (2025)
Gaussian Swaying: Surface-Based Framework for Aerodynamic Simulation with 3D Gaussians
von: Yan, Hongru, et al.
Veröffentlicht: (2025)
von: Yan, Hongru, et al.
Veröffentlicht: (2025)
Soft Tail-dropping for Adaptive Visual Tokenization
von: Chen, Zeyuan, et al.
Veröffentlicht: (2026)
von: Chen, Zeyuan, et al.
Veröffentlicht: (2026)
VideoNSA: Native Sparse Attention Scales Video Understanding
von: Song, Enxin, et al.
Veröffentlicht: (2025)
von: Song, Enxin, et al.
Veröffentlicht: (2025)
C3Editor: Achieving Controllable Consistency in 2D Model for 3D Editing
von: Tao, Zeng, et al.
Veröffentlicht: (2025)
von: Tao, Zeng, et al.
Veröffentlicht: (2025)
MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation
von: Huang, Zehuan, et al.
Veröffentlicht: (2024)
von: Huang, Zehuan, et al.
Veröffentlicht: (2024)
TokenCompose: Text-to-Image Diffusion with Token-level Supervision
von: Wang, Zirui, et al.
Veröffentlicht: (2023)
von: Wang, Zirui, et al.
Veröffentlicht: (2023)
DepGAN: Leveraging Depth Maps for Handling Occlusions and Transparency in Image Composition
von: Ghoneim, Amr, et al.
Veröffentlicht: (2024)
von: Ghoneim, Amr, et al.
Veröffentlicht: (2024)
SplatSSC: Decoupled Depth-Guided Gaussian Splatting for Semantic Scene Completion
von: Qian, Rui, et al.
Veröffentlicht: (2025)
von: Qian, Rui, et al.
Veröffentlicht: (2025)
Depth-Guided Semi-Supervised Instance Segmentation
von: Chen, Xin, et al.
Veröffentlicht: (2024)
von: Chen, Xin, et al.
Veröffentlicht: (2024)
Deep Feature Gaussian Processes for Single-Scene Aerosol Optical Depth Reconstruction
von: Liu, Shengjie, et al.
Veröffentlicht: (2024)
von: Liu, Shengjie, et al.
Veröffentlicht: (2024)
CyCLeGen: Cycle-Consistent Layout Prediction and Image Generation in Vision Foundation Models
von: Shan, Xiaojun, et al.
Veröffentlicht: (2026)
von: Shan, Xiaojun, et al.
Veröffentlicht: (2026)
RGE-GS: Reward-Guided Expansive Driving Scene Reconstruction via Diffusion Priors
von: Du, Sicong, et al.
Veröffentlicht: (2025)
von: Du, Sicong, et al.
Veröffentlicht: (2025)
Gaussian Scenes: Pose-Free Sparse-View Scene Reconstruction using Depth-Enhanced Diffusion Priors
von: Paul, Soumava, et al.
Veröffentlicht: (2024)
von: Paul, Soumava, et al.
Veröffentlicht: (2024)
DragScene: Interactive 3D Scene Editing with Single-view Drag Instructions
von: Gu, Chenghao, et al.
Veröffentlicht: (2024)
von: Gu, Chenghao, et al.
Veröffentlicht: (2024)
DrivingForward: Feed-forward 3D Gaussian Splatting for Driving Scene Reconstruction from Flexible Surround-view Input
von: Tian, Qijian, et al.
Veröffentlicht: (2024)
von: Tian, Qijian, et al.
Veröffentlicht: (2024)
SketchSplat: 3D Edge Reconstruction via Differentiable Multi-view Sketch Splatting
von: Ying, Haiyang, et al.
Veröffentlicht: (2025)
von: Ying, Haiyang, et al.
Veröffentlicht: (2025)
Clutt3R-Seg: Sparse-view 3D Instance Segmentation for Language-grounded Grasping in Cluttered Scenes
von: Noh, Jeongho, et al.
Veröffentlicht: (2026)
von: Noh, Jeongho, et al.
Veröffentlicht: (2026)
MonoInstance: Enhancing Monocular Priors via Multi-view Instance Alignment for Neural Rendering and Reconstruction
von: Zhang, Wenyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Wenyuan, et al.
Veröffentlicht: (2025)
Multi-view Reconstruction via SfM-guided Monocular Depth Estimation
von: Guo, Haoyu, et al.
Veröffentlicht: (2025)
von: Guo, Haoyu, et al.
Veröffentlicht: (2025)
SiTH: Single-view Textured Human Reconstruction with Image-Conditioned Diffusion
von: Ho, Hsuan-I, et al.
Veröffentlicht: (2023)
von: Ho, Hsuan-I, et al.
Veröffentlicht: (2023)
SceneTransporter: Optimal Transport-Guided Compositional Latent Diffusion for Single-Image Structured 3D Scene Generation
von: Wang, Ling, et al.
Veröffentlicht: (2026)
von: Wang, Ling, et al.
Veröffentlicht: (2026)
DepTR-MOT: Unveiling the Potential of Depth-Informed Trajectory Refinement for Multi-Object Tracking
von: Deng, Buyin, et al.
Veröffentlicht: (2025)
von: Deng, Buyin, et al.
Veröffentlicht: (2025)
InstaScene: Towards Complete 3D Instance Decomposition and Reconstruction from Cluttered Scenes
von: Yang, Zesong, et al.
Veröffentlicht: (2025)
von: Yang, Zesong, et al.
Veröffentlicht: (2025)
Sp2360: Sparse-view 360 Scene Reconstruction using Cascaded 2D Diffusion Priors
von: Paul, Soumava, et al.
Veröffentlicht: (2024)
von: Paul, Soumava, et al.
Veröffentlicht: (2024)
BLADE: Single-view Body Mesh Learning through Accurate Depth Estimation
von: Wang, Shengze, et al.
Veröffentlicht: (2024)
von: Wang, Shengze, et al.
Veröffentlicht: (2024)
Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
From Local Matches to Global Masks: Template-Guided Instance Detection and Segmentation in Open-World Scenes
von: Zhang, Qifan, et al.
Veröffentlicht: (2026)
von: Zhang, Qifan, et al.
Veröffentlicht: (2026)
UnIRe: Unsupervised Instance Decomposition for Dynamic Urban Scene Reconstruction
von: Mao, Yunxuan, et al.
Veröffentlicht: (2025)
von: Mao, Yunxuan, et al.
Veröffentlicht: (2025)
Training-Free Instance-Aware 3D Scene Reconstruction and Diffusion-Based View Synthesis from Sparse Images
von: Xia, Jiatong, et al.
Veröffentlicht: (2026)
von: Xia, Jiatong, et al.
Veröffentlicht: (2026)
Part123: Part-aware 3D Reconstruction from a Single-view Image
von: Liu, Anran, et al.
Veröffentlicht: (2024)
von: Liu, Anran, et al.
Veröffentlicht: (2024)
GeoDiff: Geometry-Guided Diffusion for Metric Depth Estimation
von: Pham, Tuan, et al.
Veröffentlicht: (2025)
von: Pham, Tuan, et al.
Veröffentlicht: (2025)
DepMicroDiff: Diffusion-Based Dependency-Aware Multimodal Imputation for Microbiome Data
von: Sadia, Rabeya Tus, et al.
Veröffentlicht: (2025)
von: Sadia, Rabeya Tus, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CVP: Central-Peripheral Vision-Inspired Multimodal Model for Spatial Reasoning
von: Chen, Zeyuan, et al.
Veröffentlicht: (2025) -
PixARMesh: Autoregressive Mesh-Native Single-View Scene Reconstruction
von: Zhang, Xiang, et al.
Veröffentlicht: (2026) -
Bayesian Diffusion Models for 3D Shape Reconstruction
von: Xu, Haiyang, et al.
Veröffentlicht: (2024) -
OverLayBench: A Benchmark for Layout-to-Image Generation with Dense Overlaps
von: Li, Bingnan, et al.
Veröffentlicht: (2025) -
Lay-Your-Scene: Natural Scene Layout Generation with Diffusion Transformers
von: Srivastava, Divyansh, et al.
Veröffentlicht: (2025)