Rethinking RGB-D Fusion for Semantic Segmentation in Surgical Datasets
Fuente:
arXiv
Saved in:
| Main Authors: | Jamal, Muhammad Abdullah, Mohareri, Omid |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VidLPRO: A $\underline{Vid}$eo-$\underline{L}$anguage $\underline{P}$re-training Framework for $\underline{Ro}$botic and Laparoscopic Surgery
by: Honarmand, Mohammadmahdi, et al.
Published: (2024)
by: Honarmand, Mohammadmahdi, et al.
Published: (2024)
SUREON: A Benchmark and Vision-Language-Model for Surgical Reasoning
by: Perez, Alejandra, et al.
Published: (2026)
by: Perez, Alejandra, et al.
Published: (2026)
AdaEmbed: Semi-supervised Domain Adaptation in the Embedding Space
by: Mottaghi, Ali, et al.
Published: (2024)
by: Mottaghi, Ali, et al.
Published: (2024)
A Two-Stage Progressive Pre-training using Multi-Modal Contrastive Masked Autoencoders
by: Jamal, Muhammad Abdullah, et al.
Published: (2024)
by: Jamal, Muhammad Abdullah, et al.
Published: (2024)
SurgLaVi: Large-Scale Hierarchical Dataset for Surgical Vision-Language Representation Learning
by: Perez, Alejandra, et al.
Published: (2025)
by: Perez, Alejandra, et al.
Published: (2025)
On the Role of Depth in Surgical Vision Foundation Models: An Empirical Study of RGB-D Pre-training
by: Han, John J., et al.
Published: (2026)
by: Han, John J., et al.
Published: (2026)
Multi-view Video-Pose Pretraining for Operating Room Surgical Activity Recognition
by: Hamoud, Idris, et al.
Published: (2025)
by: Hamoud, Idris, et al.
Published: (2025)
Handling Geometric Domain Shifts in Semantic Segmentation of Surgical RGB and Hyperspectral Images
by: Seidlitz, Silvia, et al.
Published: (2024)
by: Seidlitz, Silvia, et al.
Published: (2024)
Surgical Tattoos in Infrared: A Dataset for Quantifying Tissue Tracking and Mapping
by: Schmidt, Adam, et al.
Published: (2023)
by: Schmidt, Adam, et al.
Published: (2023)
Rethinking Unsupervised Domain Adaptation for Semantic Segmentation
by: Wang, Zhijie, et al.
Published: (2022)
by: Wang, Zhijie, et al.
Published: (2022)
Rethinking Alignment and Uniformity in Unsupervised Semantic Segmentation
by: Zhang, Daoan, et al.
Published: (2022)
by: Zhang, Daoan, et al.
Published: (2022)
Complementary Random Masking for RGB-Thermal Semantic Segmentation
by: Shin, Ukcheol, et al.
Published: (2023)
by: Shin, Ukcheol, et al.
Published: (2023)
D3S2: Diffusion-Guided Dataset Distillation for Semantic Segmentation
by: Zheng, Wenjie, et al.
Published: (2026)
by: Zheng, Wenjie, et al.
Published: (2026)
Deep Learning-Based 3D Instance and Semantic Segmentation: A Review
by: Yasir, Siddiqui Muhammad, et al.
Published: (2024)
by: Yasir, Siddiqui Muhammad, et al.
Published: (2024)
RGB-Event based Pedestrian Attribute Recognition: A Benchmark Dataset and An Asymmetric RWKV Fusion Framework
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
Predicting Depth Maps from Single RGB Images and Addressing Missing Information in Depth Estimation
by: Chaar, Mohamad Mofeed, et al.
Published: (2025)
by: Chaar, Mohamad Mofeed, et al.
Published: (2025)
Image Synthesis with Class-Aware Semantic Diffusion Models for Surgical Scene Segmentation
by: Zhou, Yihang, et al.
Published: (2024)
by: Zhou, Yihang, et al.
Published: (2024)
Rethinking Data Augmentation for Robust LiDAR Semantic Segmentation in Adverse Weather
by: Park, Junsung, et al.
Published: (2024)
by: Park, Junsung, et al.
Published: (2024)
Rethinking Evaluation of Multiple Sclerosis (MS) Lesion Segmentation Models
by: Basit, Abdul, et al.
Published: (2026)
by: Basit, Abdul, et al.
Published: (2026)
Forest Inspection Dataset for Aerial Semantic Segmentation and Depth Estimation
by: Blaga, Bianca-Cerasela-Zelia, et al.
Published: (2024)
by: Blaga, Bianca-Cerasela-Zelia, et al.
Published: (2024)
Enhanced Semantic Segmentation Pipeline for WeatherProof Dataset Challenge
by: Zhang, Nan, et al.
Published: (2024)
by: Zhang, Nan, et al.
Published: (2024)
Cross-Modal Purification and Fusion for Small-Object RGB-D Transmission-Line Defect Detection
by: Cui, Jiaming, et al.
Published: (2026)
by: Cui, Jiaming, et al.
Published: (2026)
Seeing Through Smoke: Surgical Desmoking for Improved Visual Perception
by: Lu, Jingpei, et al.
Published: (2026)
by: Lu, Jingpei, et al.
Published: (2026)
Explainable Parkinsons Disease Gait Recognition Using Multimodal RGB-D Fusion and Large Language Models
by: Alnaasan, Manar, et al.
Published: (2025)
by: Alnaasan, Manar, et al.
Published: (2025)
D-PLS: Decoupled Semantic Segmentation for 4D-Panoptic-LiDAR-Segmentation
by: Steinhauser, Maik, et al.
Published: (2025)
by: Steinhauser, Maik, et al.
Published: (2025)
LFA-Net: A Lightweight Network with LiteFusion Attention for Retinal Vessel Segmentation
by: Mehmood, Mehwish, et al.
Published: (2025)
by: Mehmood, Mehwish, et al.
Published: (2025)
Segment Any RGB-Thermal Model with Language-aided Distillation
by: Xing, Dong, et al.
Published: (2025)
by: Xing, Dong, et al.
Published: (2025)
Knowledge-Guided Brain Tumor Segmentation via Synchronized Visual-Semantic-Topological Prior Fusion
by: Zhang, Mingda, et al.
Published: (2025)
by: Zhang, Mingda, et al.
Published: (2025)
REL-SF4PASS: Panoramic Semantic Segmentation with REL Depth Representation and Spherical Fusion
by: Li, Xuewei, et al.
Published: (2026)
by: Li, Xuewei, et al.
Published: (2026)
The Surprising Effectiveness of Canonical Knowledge Distillation for Semantic Segmentation
by: Ali, Muhammad, et al.
Published: (2026)
by: Ali, Muhammad, et al.
Published: (2026)
Semantic Segmentation by Semantic Proportions
by: Aysel, Halil Ibrahim, et al.
Published: (2023)
by: Aysel, Halil Ibrahim, et al.
Published: (2023)
RDFC-GAN: RGB-Depth Fusion CycleGAN for Indoor Depth Completion
by: Wang, Haowen, et al.
Published: (2023)
by: Wang, Haowen, et al.
Published: (2023)
Event-Adaptive State Transition and Gated Fusion for RGB-Event Object Tracking
by: You, Jinlin, et al.
Published: (2026)
by: You, Jinlin, et al.
Published: (2026)
SERNet-Former: Semantic Segmentation by Efficient Residual Network with Attention-Boosting Gates and Attention-Fusion Networks
by: Erisen, Serdar
Published: (2024)
by: Erisen, Serdar
Published: (2024)
Multi-encoder ConvNeXt Network with Smooth Attentional Feature Fusion for Multispectral Semantic Segmentation
by: Ramos, Leo Thomas, et al.
Published: (2026)
by: Ramos, Leo Thomas, et al.
Published: (2026)
TowerDataset: A Heterogeneous Benchmark for Transmission Corridor Segmentation with a Global-Local Fusion Framework
by: Cui, Xu, et al.
Published: (2026)
by: Cui, Xu, et al.
Published: (2026)
Rethinking FID Through the Geometry of the Reference Dataset
by: Lee, Yunghee, et al.
Published: (2026)
by: Lee, Yunghee, et al.
Published: (2026)
FusionVision: A comprehensive approach of 3D object reconstruction and segmentation from RGB-D cameras using YOLO and fast segment anything
by: Ghazouali, Safouane El, et al.
Published: (2024)
by: Ghazouali, Safouane El, et al.
Published: (2024)
V3D-SLAM: Robust RGB-D SLAM in Dynamic Environments with 3D Semantic Geometry Voting
by: Dang, Tuan, et al.
Published: (2024)
by: Dang, Tuan, et al.
Published: (2024)
Rethinking Normalization Strategies and Convolutional Kernels for Multimodal Image Fusion
by: He, Dan, et al.
Published: (2024)
by: He, Dan, et al.
Published: (2024)
Similar Items
-
VidLPRO: A $\underline{Vid}$eo-$\underline{L}$anguage $\underline{P}$re-training Framework for $\underline{Ro}$botic and Laparoscopic Surgery
by: Honarmand, Mohammadmahdi, et al.
Published: (2024) -
SUREON: A Benchmark and Vision-Language-Model for Surgical Reasoning
by: Perez, Alejandra, et al.
Published: (2026) -
AdaEmbed: Semi-supervised Domain Adaptation in the Embedding Space
by: Mottaghi, Ali, et al.
Published: (2024) -
A Two-Stage Progressive Pre-training using Multi-Modal Contrastive Masked Autoencoders
by: Jamal, Muhammad Abdullah, et al.
Published: (2024) -
SurgLaVi: Large-Scale Hierarchical Dataset for Surgical Vision-Language Representation Learning
by: Perez, Alejandra, et al.
Published: (2025)