SharpDepth: Sharpening Metric Depth Predictions Using Diffusion Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Pham, Duc-Hai, Do, Tung, Nguyen, Phong, Hua, Binh-Son, Nguyen, Khoi, Nguyen, Rang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Semi-supervised 3D Semantic Scene Completion with 2D Vision Foundation Model Guidance
by: Pham, Duc-Hai, et al.
Published: (2024)
by: Pham, Duc-Hai, et al.
Published: (2024)
ModeDreamer: Mode Guiding Score Distillation for Text-to-3D Generation using Reference Image Prompts
by: Tran, Uy Dieu, et al.
Published: (2024)
by: Tran, Uy Dieu, et al.
Published: (2024)
LiftRefine: Progressively Refined View Synthesis from 3D Lifting with Volume-Triplane Representations
by: Do, Tung, et al.
Published: (2024)
by: Do, Tung, et al.
Published: (2024)
Text-to-3D Generation using Jensen-Shannon Score Distillation
by: Do, Khoi, et al.
Published: (2025)
by: Do, Khoi, et al.
Published: (2025)
InverFill: One-Step Inversion for Enhanced Few-Step Diffusion Inpainting
by: Vu, Duc, et al.
Published: (2026)
by: Vu, Duc, et al.
Published: (2026)
DiverseDream: Diverse Text-to-3D Synthesis with Augmented Text Embedding
by: Tran, Uy Dieu, et al.
Published: (2023)
by: Tran, Uy Dieu, et al.
Published: (2023)
SwiftTailor: Efficient 3D Garment Generation with Geometry Image Representation
by: Pham, Phuc, et al.
Published: (2026)
by: Pham, Phuc, et al.
Published: (2026)
SwiftEdit: Lightning Fast Text-Guided Image Editing via One-Step Diffusion
by: Nguyen, Trong-Tung, et al.
Published: (2024)
by: Nguyen, Trong-Tung, et al.
Published: (2024)
SwiftTry: Fast and Consistent Video Virtual Try-On with Diffusion Models
by: Nguyen, Hung, et al.
Published: (2024)
by: Nguyen, Hung, et al.
Published: (2024)
Self-supervised Video Object Segmentation with Distillation Learning of Deformable Attention
by: Truong, Quang-Trung, et al.
Published: (2024)
by: Truong, Quang-Trung, et al.
Published: (2024)
Color Alignment in Diffusion
by: Shum, Ka Chun, et al.
Published: (2025)
by: Shum, Ka Chun, et al.
Published: (2025)
FlexEdit: Flexible and Controllable Diffusion-based Object-centric Image Editing
by: Nguyen, Trong-Tung, et al.
Published: (2024)
by: Nguyen, Trong-Tung, et al.
Published: (2024)
GeoDiff: Geometry-Guided Diffusion for Metric Depth Estimation
by: Pham, Tuan, et al.
Published: (2025)
by: Pham, Tuan, et al.
Published: (2025)
SwiftPie: Lightning-fast Subject-driven Image Personalization via One step Diffusion
by: Duong, Huy, et al.
Published: (2026)
by: Duong, Huy, et al.
Published: (2026)
Toward Fine-Grained Speech Inpainting Forensics:A Dataset, Method, and Metric for Multi-Region Tampering Localization
by: Vu, Tung, et al.
Published: (2026)
by: Vu, Tung, et al.
Published: (2026)
Blur2Blur: Blur Conversion for Unsupervised Image Deblurring on Unknown Domains
by: Pham, Bang-Dang, et al.
Published: (2024)
by: Pham, Bang-Dang, et al.
Published: (2024)
h-Edit: Effective and Flexible Diffusion-Based Editing via Doob's h-Transform
by: Nguyen, Toan, et al.
Published: (2025)
by: Nguyen, Toan, et al.
Published: (2025)
MMAP: A Multi-Magnification and Prototype-Aware Architecture for Predicting Spatial Gene Expression
by: Nguyen, Hai Dang, et al.
Published: (2025)
by: Nguyen, Hai Dang, et al.
Published: (2025)
Supercharged One-step Text-to-Image Diffusion Models with Negative Prompts
by: Nguyen, Viet, et al.
Published: (2024)
by: Nguyen, Viet, et al.
Published: (2024)
Bidirectional Diffusion Bridge Models
by: Kieu, Duc, et al.
Published: (2025)
by: Kieu, Duc, et al.
Published: (2025)
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
by: Nguyen, Quang-Binh, et al.
Published: (2025)
by: Nguyen, Quang-Binh, et al.
Published: (2025)
An End-to-End Depth-Based Pipeline for Selfie Image Rectification
by: Alhawwary, Ahmed, et al.
Published: (2024)
by: Alhawwary, Ahmed, et al.
Published: (2024)
LP-OVOD: Open-Vocabulary Object Detection by Linear Probing
by: Pham, Chau, et al.
Published: (2023)
by: Pham, Chau, et al.
Published: (2023)
Adaptive Cache Enhancement for Test-Time Adaptation of Vision-Language Models
by: Nguyen, Khanh-Binh, et al.
Published: (2025)
by: Nguyen, Khanh-Binh, et al.
Published: (2025)
Advancing Vietnamese Visual Question Answering with Transformer and Convolutional Integration
by: Nguyen, Ngoc Son, et al.
Published: (2024)
by: Nguyen, Ngoc Son, et al.
Published: (2024)
Dream-in-Style: Text-to-3D Generation Using Stylized Score Distillation
by: Kompanowski, Hubert, et al.
Published: (2024)
by: Kompanowski, Hubert, et al.
Published: (2024)
SwiftBrush v2: Make Your One-step Diffusion Model Better Than Its Teacher
by: Dao, Trung, et al.
Published: (2024)
by: Dao, Trung, et al.
Published: (2024)
PAT: Pixel-wise Adaptive Training for Long-tailed Segmentation
by: Do, Khoi, et al.
Published: (2024)
by: Do, Khoi, et al.
Published: (2024)
CAKE: Real-time Action Detection via Motion Distillation and Background-aware Contrastive Learning
by: Hoang, Hieu, et al.
Published: (2026)
by: Hoang, Hieu, et al.
Published: (2026)
Any3DIS: Class-Agnostic 3D Instance Segmentation by 2D Mask Tracking
by: Nguyen, Phuc, et al.
Published: (2024)
by: Nguyen, Phuc, et al.
Published: (2024)
Stable Messenger: Steganography for Message-Concealed Image Generation
by: Nguyen, Quang, et al.
Published: (2023)
by: Nguyen, Quang, et al.
Published: (2023)
PADM: A Physics-aware Diffusion Model for Attenuation Correction
by: Pham, Trung Kien, et al.
Published: (2025)
by: Pham, Trung Kien, et al.
Published: (2025)
Depth-aware Panoptic Segmentation
by: Nguyen, Tuan, et al.
Published: (2024)
by: Nguyen, Tuan, et al.
Published: (2024)
Language-driven Object Fusion into Neural Radiance Fields with Pose-Conditioned Dataset Updates
by: Shum, Ka Chun, et al.
Published: (2023)
by: Shum, Ka Chun, et al.
Published: (2023)
EditScout: Locating Forged Regions from Diffusion-based Edited Images with Multimodal LLM
by: Nguyen, Quang, et al.
Published: (2024)
by: Nguyen, Quang, et al.
Published: (2024)
Depth Pro: Sharp Monocular Metric Depth in Less Than a Second
by: Bochkovskii, Aleksei, et al.
Published: (2024)
by: Bochkovskii, Aleksei, et al.
Published: (2024)
Open3DIS: Open-Vocabulary 3D Instance Segmentation with 2D Mask Guidance
by: Nguyen, Phuc D. A., et al.
Published: (2023)
by: Nguyen, Phuc D. A., et al.
Published: (2023)
VEIGAR: View-consistent Explicit Inpainting and Geometry Alignment for 3D object Removal
by: Do, Pham Khai Nguyen, et al.
Published: (2025)
by: Do, Pham Khai Nguyen, et al.
Published: (2025)
ScaleDepth: Decomposing Metric Depth Estimation into Scale Prediction and Relative Depth Estimation
by: Zhu, Ruijie, et al.
Published: (2024)
by: Zhu, Ruijie, et al.
Published: (2024)
Universal Multi-Domain Translation via Diffusion Routers
by: Kieu, Duc, et al.
Published: (2025)
by: Kieu, Duc, et al.
Published: (2025)
Similar Items
-
Semi-supervised 3D Semantic Scene Completion with 2D Vision Foundation Model Guidance
by: Pham, Duc-Hai, et al.
Published: (2024) -
ModeDreamer: Mode Guiding Score Distillation for Text-to-3D Generation using Reference Image Prompts
by: Tran, Uy Dieu, et al.
Published: (2024) -
LiftRefine: Progressively Refined View Synthesis from 3D Lifting with Volume-Triplane Representations
by: Do, Tung, et al.
Published: (2024) -
Text-to-3D Generation using Jensen-Shannon Score Distillation
by: Do, Khoi, et al.
Published: (2025) -
InverFill: One-Step Inversion for Enhanced Few-Step Diffusion Inpainting
by: Vu, Duc, et al.
Published: (2026)