Bridging Spectral-wise and Multi-spectral Depth Estimation via Geometry-guided Contrastive Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Shin, Ukcheol, Lee, Kyunghyun, Oh, Jean |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Complementary Random Masking for RGB-Thermal Semantic Segmentation
by: Shin, Ukcheol, et al.
Published: (2023)
by: Shin, Ukcheol, et al.
Published: (2023)
Learning to Control Camera Exposure via Reinforcement Learning
by: Lee, Kyunghyun, et al.
Published: (2024)
by: Lee, Kyunghyun, et al.
Published: (2024)
FIReStereo: Forest InfraRed Stereo Dataset for UAS Depth Perception in Visually Degraded Environments
by: Dhrafani, Devansh, et al.
Published: (2024)
by: Dhrafani, Devansh, et al.
Published: (2024)
Deep Depth Estimation from Thermal Image: Dataset, Benchmark, and Challenges
by: Shin, Ukcheol, et al.
Published: (2025)
by: Shin, Ukcheol, et al.
Published: (2025)
All-day Depth Completion via Thermal-LiDAR Fusion
by: Kim, Janghyun, et al.
Published: (2025)
by: Kim, Janghyun, et al.
Published: (2025)
VPOcc: Exploiting Vanishing Point for 3D Semantic Occupancy Prediction
by: Kim, Junsu, et al.
Published: (2024)
by: Kim, Junsu, et al.
Published: (2024)
Depth Any Camera: Zero-Shot Metric Depth Estimation from Any Camera
by: Guo, Yuliang, et al.
Published: (2025)
by: Guo, Yuliang, et al.
Published: (2025)
A Systematic Literature Review on Deep Learning-based Depth Estimation in Computer Vision
by: Rohan, Ali, et al.
Published: (2025)
by: Rohan, Ali, et al.
Published: (2025)
MrGS: Multi-modal Radiance Fields with 3D Gaussian Splatting for RGB-Thermal Novel View Synthesis
by: Kweon, Minseong, et al.
Published: (2025)
by: Kweon, Minseong, et al.
Published: (2025)
Towards Sharper Object Boundaries in Self-Supervised Depth Estimation
by: Cecille, Aurélien, et al.
Published: (2025)
by: Cecille, Aurélien, et al.
Published: (2025)
Gaze on the Prize: Shaping Visual Attention with Return-Guided Contrastive Learning
by: Lee, Andrew, et al.
Published: (2025)
by: Lee, Andrew, et al.
Published: (2025)
Helvipad: A Real-World Dataset for Omnidirectional Stereo Depth Estimation
by: Zayene, Mehdi, et al.
Published: (2024)
by: Zayene, Mehdi, et al.
Published: (2024)
Adaptive Discrete Disparity Volume for Self-supervised Monocular Depth Estimation
by: Ren, Jianwei
Published: (2024)
by: Ren, Jianwei
Published: (2024)
Learning Point Cloud Representations with Pose Continuity for Depth-Based Category-Level 6D Object Pose Estimation
by: Li, Zhujun, et al.
Published: (2025)
by: Li, Zhujun, et al.
Published: (2025)
Towards Deviation-Robust Agent Navigation via Perturbation-Aware Contrastive Learning
by: Lin, Bingqian, et al.
Published: (2024)
by: Lin, Bingqian, et al.
Published: (2024)
FuzzRisk: Online Collision Risk Estimation for Autonomous Vehicles based on Depth-Aware Object Detection via Fuzzy Inference
by: Liao, Brian Hsuan-Cheng, et al.
Published: (2024)
by: Liao, Brian Hsuan-Cheng, et al.
Published: (2024)
RVN-Bench: A Benchmark for Reactive Visual Navigation
by: Lee, Jaewon, et al.
Published: (2026)
by: Lee, Jaewon, et al.
Published: (2026)
MonoPP: Metric-Scaled Self-Supervised Monocular Depth Estimation by Planar-Parallax Geometry in Automotive Applications
by: Elazab, Gasser, et al.
Published: (2024)
by: Elazab, Gasser, et al.
Published: (2024)
ShapeICP: Iterative Category-level Object Pose and Shape Estimation from Depth
by: Zhang, Yihao, et al.
Published: (2024)
by: Zhang, Yihao, et al.
Published: (2024)
GVDepth: Zero-Shot Monocular Depth Estimation for Ground Vehicles based on Probabilistic Cue Fusion
by: Koledić, Karlo, et al.
Published: (2024)
by: Koledić, Karlo, et al.
Published: (2024)
{S\textsuperscript{2}M\textsuperscript{2}}: Scalable Stereo Matching Model for Reliable Depth Estimation
by: Min, Junhong, et al.
Published: (2025)
by: Min, Junhong, et al.
Published: (2025)
EGSA-PT:Edge-Guided Spatial Attention with Progressive Training for Monocular Depth Estimation and Segmentation of Transparent Objects
by: Omotara, Gbenga, et al.
Published: (2025)
by: Omotara, Gbenga, et al.
Published: (2025)
RAPID: Robust and Agile Planner Using Inverse Reinforcement Learning for Vision-Based Drone Navigation
by: Kim, Minwoo, et al.
Published: (2025)
by: Kim, Minwoo, et al.
Published: (2025)
ViTaS: Visual Tactile Soft Fusion Contrastive Learning for Visuomotor Learning
by: Tian, Yufeng, et al.
Published: (2026)
by: Tian, Yufeng, et al.
Published: (2026)
Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation
by: Xing, Eliot, et al.
Published: (2024)
by: Xing, Eliot, et al.
Published: (2024)
Beyond Frame-wise Tracking: A Trajectory-based Paradigm for Efficient Point Cloud Tracking
by: Fan, BaiChen, et al.
Published: (2025)
by: Fan, BaiChen, et al.
Published: (2025)
Rule-VLN: Bridging Perception and Compliance via Semantic Reasoning and Geometric Rectification
by: Wen, Jiawen, et al.
Published: (2026)
by: Wen, Jiawen, et al.
Published: (2026)
MCRL4OR: Multimodal Contrastive Representation Learning for Off-Road Environmental Perception
by: Yang, Yi, et al.
Published: (2025)
by: Yang, Yi, et al.
Published: (2025)
MetricGold: Leveraging Text-To-Image Latent Diffusion Models for Metric Depth Estimation
by: Shah, Ansh, et al.
Published: (2024)
by: Shah, Ansh, et al.
Published: (2024)
LaB-CL: Localized and Balanced Contrastive Learning for improving parking slot detection
by: Jeong, U Jin, et al.
Published: (2024)
by: Jeong, U Jin, et al.
Published: (2024)
V$^2$-SfMLearner: Learning Monocular Depth and Ego-motion for Multimodal Wireless Capsule Endoscopy
by: Bai, Long, et al.
Published: (2024)
by: Bai, Long, et al.
Published: (2024)
Toward Aligning Human and Robot Actions via Multi-Modal Demonstration Learning
by: Zahid, Azizul, et al.
Published: (2025)
by: Zahid, Azizul, et al.
Published: (2025)
VIKI-R: Coordinating Embodied Multi-Agent Cooperation via Reinforcement Learning
by: Kang, Li, et al.
Published: (2025)
by: Kang, Li, et al.
Published: (2025)
MM3DGS SLAM: Multi-modal 3D Gaussian Splatting for SLAM Using Vision, Depth, and Inertial Measurements
by: Sun, Lisong C., et al.
Published: (2024)
by: Sun, Lisong C., et al.
Published: (2024)
CL3R: 3D Reconstruction and Contrastive Learning for Enhanced Robotic Manipulation Representations
by: Cui, Wenbo, et al.
Published: (2025)
by: Cui, Wenbo, et al.
Published: (2025)
RoBridge: A Hierarchical Architecture Bridging Cognition and Execution for General Robotic Manipulation
by: Zhang, Kaidong, et al.
Published: (2025)
by: Zhang, Kaidong, et al.
Published: (2025)
DVGT: Driving Visual Geometry Transformer
by: Zuo, Sicheng, et al.
Published: (2025)
by: Zuo, Sicheng, et al.
Published: (2025)
UNIC: Learning Unified Multimodal Extrinsic Contact Estimation
by: Xu, Zhengtong, et al.
Published: (2026)
by: Xu, Zhengtong, et al.
Published: (2026)
Diffusion-guided Generalizable Enhancer for Urban Scene Reconstruction
by: Che, Henry, et al.
Published: (2026)
by: Che, Henry, et al.
Published: (2026)
Look, Focus, Act: Efficient and Robust Robot Learning via Human Gaze and Foveated Vision Transformers
by: Chuang, Ian, et al.
Published: (2025)
by: Chuang, Ian, et al.
Published: (2025)
Similar Items
-
Complementary Random Masking for RGB-Thermal Semantic Segmentation
by: Shin, Ukcheol, et al.
Published: (2023) -
Learning to Control Camera Exposure via Reinforcement Learning
by: Lee, Kyunghyun, et al.
Published: (2024) -
FIReStereo: Forest InfraRed Stereo Dataset for UAS Depth Perception in Visually Degraded Environments
by: Dhrafani, Devansh, et al.
Published: (2024) -
Deep Depth Estimation from Thermal Image: Dataset, Benchmark, and Challenges
by: Shin, Ukcheol, et al.
Published: (2025) -
All-day Depth Completion via Thermal-LiDAR Fusion
by: Kim, Janghyun, et al.
Published: (2025)