Iris: Bringing Real-World Priors into Diffusion Model for Monocular Depth Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Xinhao, Pei, Gensheng, Sun, Zeren, Yao, Yazhou, Shen, Fumin, Wang, Wenguan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PKINet-v2: Towards Powerful and Efficient Poly-Kernel Remote Sensing Object Detection
by: Cai, Xinhao, et al.
Published: (2026)
by: Cai, Xinhao, et al.
Published: (2026)
Poly Kernel Inception Network for Remote Sensing Detection
by: Cai, Xinhao, et al.
Published: (2024)
by: Cai, Xinhao, et al.
Published: (2024)
Unbiased Object Detection Beyond Frequency with Visually Prompted Image Synthesis
by: Cai, Xinhao, et al.
Published: (2025)
by: Cai, Xinhao, et al.
Published: (2025)
Towards Remote Sensing Change Detection with Neural Memory
by: Yang, Zhenyu, et al.
Published: (2026)
by: Yang, Zhenyu, et al.
Published: (2026)
Taming SAM3 in the Wild: A Concept Bank for Open-Vocabulary Segmentation
by: Pei, Gensheng, et al.
Published: (2026)
by: Pei, Gensheng, et al.
Published: (2026)
Dynamic in Static: Hybrid Visual Correspondence for Self-Supervised Video Object Segmentation
by: Pei, Gensheng, et al.
Published: (2024)
by: Pei, Gensheng, et al.
Published: (2024)
VideoMAC: Video Masked Autoencoders Meet ConvNets
by: Pei, Gensheng, et al.
Published: (2024)
by: Pei, Gensheng, et al.
Published: (2024)
Knowledge Transfer with Simulated Inter-Image Erasing for Weakly Supervised Semantic Segmentation
by: Chen, Tao, et al.
Published: (2024)
by: Chen, Tao, et al.
Published: (2024)
PEARL: Geometry Aligns Semantics for Training-Free Open-Vocabulary Semantic Segmentation
by: Pei, Gensheng, et al.
Published: (2026)
by: Pei, Gensheng, et al.
Published: (2026)
Efficiency Follows Global-Local Decoupling
by: Yang, Zhenyu, et al.
Published: (2026)
by: Yang, Zhenyu, et al.
Published: (2026)
PCA-Seg: Revisiting Cost Aggregation for Open-Vocabulary Semantic and Part Segmentation
by: Yin, Jianjian, et al.
Published: (2026)
by: Yin, Jianjian, et al.
Published: (2026)
Seeing What Matters: Empowering CLIP with Patch Generation-to-Selection
by: Pei, Gensheng, et al.
Published: (2025)
by: Pei, Gensheng, et al.
Published: (2025)
Combating Noisy Labels through Fostering Self- and Neighbor-Consistency
by: Sun, Zeren, et al.
Published: (2026)
by: Sun, Zeren, et al.
Published: (2026)
Relating CNN-Transformer Fusion Network for Change Detection
by: Gao, Yuhao, et al.
Published: (2024)
by: Gao, Yuhao, et al.
Published: (2024)
Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images
by: Zhou, Bo, et al.
Published: (2026)
by: Zhou, Bo, et al.
Published: (2026)
Iris: Integrating Language into Diffusion-based Monocular Depth Estimation
by: Zeng, Ziyao, et al.
Published: (2024)
by: Zeng, Ziyao, et al.
Published: (2024)
Guided Diffusion-based Generation of Adversarial Objects for Real-World Monocular Depth Estimation Attacks
by: Chen, Yongtao, et al.
Published: (2025)
by: Chen, Yongtao, et al.
Published: (2025)
Stealing Stable Diffusion Prior for Robust Monocular Depth Estimation
by: Mao, Yifan, et al.
Published: (2024)
by: Mao, Yifan, et al.
Published: (2024)
Underwater Monocular Metric Depth Estimation: Real-World Benchmarks and Synthetic Fine-Tuning with Vision Foundation Models
by: Cai, Zijie, et al.
Published: (2025)
by: Cai, Zijie, et al.
Published: (2025)
DepthMaster: Taming Diffusion Models for Monocular Depth Estimation
by: Song, Ziyang, et al.
Published: (2025)
by: Song, Ziyang, et al.
Published: (2025)
High-Precision Self-Supervised Monocular Depth Estimation with Rich-Resource Prior
by: Han, Wencheng, et al.
Published: (2024)
by: Han, Wencheng, et al.
Published: (2024)
Multi-view Reconstruction via SfM-guided Monocular Depth Estimation
by: Guo, Haoyu, et al.
Published: (2025)
by: Guo, Haoyu, et al.
Published: (2025)
LMDepth: Lightweight Mamba-based Monocular Depth Estimation for Real-World Deployment
by: Long, Jiahuan, et al.
Published: (2025)
by: Long, Jiahuan, et al.
Published: (2025)
OmniGaze: Reward-inspired Generalizable Gaze Estimation In The Wild
by: Qu, Hongyu, et al.
Published: (2025)
by: Qu, Hongyu, et al.
Published: (2025)
WEDepth: Efficient Adaptation of World Knowledge for Monocular Depth Estimation
by: Wang, Gongshu, et al.
Published: (2025)
by: Wang, Gongshu, et al.
Published: (2025)
Language as Prior, Vision as Calibration: Metric Scale Recovery for Monocular Depth Estimation
by: Zhan, Mingxia, et al.
Published: (2026)
by: Zhan, Mingxia, et al.
Published: (2026)
A Light-weight Transformer-based Self-supervised Matching Network for Heterogeneous Images
by: Zhang, Wang, et al.
Published: (2024)
by: Zhang, Wang, et al.
Published: (2024)
Diffusion Models for Monocular Depth Estimation: Overcoming Challenging Conditions
by: Tosi, Fabio, et al.
Published: (2024)
by: Tosi, Fabio, et al.
Published: (2024)
BadDepth: Backdoor Attacks Against Monocular Depth Estimation in the Physical World
by: Guo, Ji, et al.
Published: (2025)
by: Guo, Ji, et al.
Published: (2025)
Relative Pose Estimation through Affine Corrections of Monocular Depth Priors
by: Yu, Yifan, et al.
Published: (2025)
by: Yu, Yifan, et al.
Published: (2025)
FlowDepth: Decoupling Optical Flow for Self-Supervised Monocular Depth Estimation
by: Sun, Yiyang, et al.
Published: (2024)
by: Sun, Yiyang, et al.
Published: (2024)
Real-time Monocular Depth Estimation on Embedded Systems
by: Feng, Cheng, et al.
Published: (2023)
by: Feng, Cheng, et al.
Published: (2023)
FreeReg: Image-to-Point Cloud Registration Leveraging Pretrained Diffusion Models and Monocular Depth Estimators
by: Wang, Haiping, et al.
Published: (2023)
by: Wang, Haiping, et al.
Published: (2023)
MultiDepth: Multi-Sample Priors for Refining Monocular Metric Depth Estimations in Indoor Scenes
by: Byun, Sanghyun, et al.
Published: (2024)
by: Byun, Sanghyun, et al.
Published: (2024)
Semi-supervised Semantic Segmentation with Multi-Constraint Consistency Learning
by: Yin, Jianjian, et al.
Published: (2025)
by: Yin, Jianjian, et al.
Published: (2025)
RTS-Mono: A Real-Time Self-Supervised Monocular Depth Estimation Method for Real-World Deployment
by: Cheng, Zeyu, et al.
Published: (2025)
by: Cheng, Zeyu, et al.
Published: (2025)
DCPI-Depth: Explicitly Infusing Dense Correspondence Prior to Unsupervised Monocular Depth Estimation
by: Zhang, Mengtan, et al.
Published: (2024)
by: Zhang, Mengtan, et al.
Published: (2024)
PrimeDepth: Efficient Monocular Depth Estimation with a Stable Diffusion Preimage
by: Zavadski, Denis, et al.
Published: (2024)
by: Zavadski, Denis, et al.
Published: (2024)
CDPR: Cross-modal Diffusion with Polarization for Reliable Monocular Depth Estimation
by: Yu, Rongjia, et al.
Published: (2026)
by: Yu, Rongjia, et al.
Published: (2026)
Advancing Depth Anything Model for Unsupervised Monocular Depth Estimation in Endoscopy
by: Li, Bojian, et al.
Published: (2024)
by: Li, Bojian, et al.
Published: (2024)
Similar Items
-
PKINet-v2: Towards Powerful and Efficient Poly-Kernel Remote Sensing Object Detection
by: Cai, Xinhao, et al.
Published: (2026) -
Poly Kernel Inception Network for Remote Sensing Detection
by: Cai, Xinhao, et al.
Published: (2024) -
Unbiased Object Detection Beyond Frequency with Visually Prompted Image Synthesis
by: Cai, Xinhao, et al.
Published: (2025) -
Towards Remote Sensing Change Detection with Neural Memory
by: Yang, Zhenyu, et al.
Published: (2026) -
Taming SAM3 in the Wild: A Concept Bank for Open-Vocabulary Segmentation
by: Pei, Gensheng, et al.
Published: (2026)