LEMON: A Large Endoscopic MONocular Dataset and Foundation Model for Perception in Surgical Settings
Fuente:
arXiv
Saved in:
| Main Authors: | Che, Chengan, Wang, Chao, Vercauteren, Tom, Tsoka, Sophia, Garcia-Peraza-Herrera, Luis C. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Back to the Feature: Explaining Video Classifiers with Video Counterfactual Explanations
by: Wang, Chao, et al.
Published: (2025)
by: Wang, Chao, et al.
Published: (2025)
A Stitch in Time: Learning Procedural Workflow via Self-Supervised Plackett-Luce Ranking
by: Che, Chengan, et al.
Published: (2025)
by: Che, Chengan, et al.
Published: (2025)
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training?
by: Che, Chengan, et al.
Published: (2026)
by: Che, Chengan, et al.
Published: (2026)
LoViT: Long Video Transformer for Surgical Phase Recognition
by: Liu, Yang, et al.
Published: (2023)
by: Liu, Yang, et al.
Published: (2023)
Average Calibration Error: A Differentiable Loss for Improved Reliability in Image Segmentation
by: Barfoot, Theodore, et al.
Published: (2024)
by: Barfoot, Theodore, et al.
Published: (2024)
Average Calibration Losses for Reliable Uncertainty in Medical Image Segmentation
by: Barfoot, Theodore, et al.
Published: (2025)
by: Barfoot, Theodore, et al.
Published: (2025)
Transferring Relative Monocular Depth to Surgical Vision with Temporal Consistency
by: Budd, Charlie, et al.
Published: (2024)
by: Budd, Charlie, et al.
Published: (2024)
Multitask Learning in Minimally Invasive Surgical Vision: A Review
by: Alabi, Oluwatosin, et al.
Published: (2024)
by: Alabi, Oluwatosin, et al.
Published: (2024)
Grounding Surgical Action Triplets with Instrument Instance Segmentation: A Dataset and Target-Aware Fusion Approach
by: Alabi, Oluwatosin, et al.
Published: (2025)
by: Alabi, Oluwatosin, et al.
Published: (2025)
SegMatch: A semi-supervised learning method for surgical instrument segmentation
by: Wei, Meng, et al.
Published: (2023)
by: Wei, Meng, et al.
Published: (2023)
Text Promptable Surgical Instrument Segmentation with Vision-Language Models
by: Zhou, Zijian, et al.
Published: (2023)
by: Zhou, Zijian, et al.
Published: (2023)
ROBUST-MIPS: A Combined Skeletal Pose and Instance Segmentation Dataset for Laparoscopic Surgical Instruments
by: Han, Zhe, et al.
Published: (2025)
by: Han, Zhe, et al.
Published: (2025)
Enhancing Generalized Few-Shot Semantic Segmentation via Effective Knowledge Transfer
by: Chen, Xinyue, et al.
Published: (2024)
by: Chen, Xinyue, et al.
Published: (2024)
SurgPIS: Surgical-instrument-level Instances and Part-level Semantics for Weakly-supervised Part-aware Instance Segmentation
by: Wei, Meng, et al.
Published: (2025)
by: Wei, Meng, et al.
Published: (2025)
Surgical-DINO: Adapter Learning of Foundation Models for Depth Estimation in Endoscopic Surgery
by: Cui, Beilei, et al.
Published: (2024)
by: Cui, Beilei, et al.
Published: (2024)
Large Model driven Radiology Report Generation with Clinical Quality Reinforcement Learning
by: Zhou, Zijian, et al.
Published: (2024)
by: Zhou, Zijian, et al.
Published: (2024)
CoPESD: A Multi-Level Surgical Motion Dataset for Training Large Vision-Language Models to Co-Pilot Endoscopic Submucosal Dissection
by: Wang, Guankun, et al.
Published: (2024)
by: Wang, Guankun, et al.
Published: (2024)
UltraFlwr -- An Efficient Federated Surgical Object Detection Framework
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
Benchmarking Endoscopic Surgical Image Restoration and Beyond
by: Pei, Jialun, et al.
Published: (2025)
by: Pei, Jialun, et al.
Published: (2025)
DARES: Depth Anything in Robotic Endoscopic Surgery with Self-supervised Vector-LoRA of the Foundation Model
by: Zeinoddin, Mona Sheikh, et al.
Published: (2024)
by: Zeinoddin, Mona Sheikh, et al.
Published: (2024)
Label merge-and-split: A graph-colouring approach for memory-efficient brain parcellation
by: Kujawa, Aaron, et al.
Published: (2024)
by: Kujawa, Aaron, et al.
Published: (2024)
LEMON: Localized Editing with Mesh Optimization and Neural Shaders
by: Algan, Furkan Mert, et al.
Published: (2024)
by: Algan, Furkan Mert, et al.
Published: (2024)
TEMSET-24K: Densely Annotated Dataset for Indexing Multipart Endoscopic Videos using Surgical Timeline Segmentation
by: Bilal, Muhammad, et al.
Published: (2025)
by: Bilal, Muhammad, et al.
Published: (2025)
Tree-based Semantic Losses: Application to Sparsely-supervised Large Multi-class Hyperspectral Segmentation
by: Wang, Junwen, et al.
Published: (2025)
by: Wang, Junwen, et al.
Published: (2025)
SimuScope: Realistic Endoscopic Synthetic Dataset Generation through Surgical Simulation and Diffusion Models
by: Martyniak, Sabina, et al.
Published: (2024)
by: Martyniak, Sabina, et al.
Published: (2024)
CholecInstanceSeg: A Tool Instance Segmentation Dataset for Laparoscopic Surgery
by: Alabi, Oluwatosin, et al.
Published: (2024)
by: Alabi, Oluwatosin, et al.
Published: (2024)
LEMON: a foundation model for nuclear morphology in Computational Pathology
by: Chadoutaud, Loïc, et al.
Published: (2026)
by: Chadoutaud, Loïc, et al.
Published: (2026)
Industrial Language-Image Dataset (ILID): Adapting Vision Foundation Models for Industrial Settings
by: Moenck, Keno, et al.
Published: (2024)
by: Moenck, Keno, et al.
Published: (2024)
Beyond one-hot encoding? Journey into compact encoding for large multi-class segmentation
by: Kujawa, Aaron, et al.
Published: (2025)
by: Kujawa, Aaron, et al.
Published: (2025)
Learning-based sound speed estimation and aberration correction in linear-array photoacoustic imaging
by: Shi, Mengjie, et al.
Published: (2023)
by: Shi, Mengjie, et al.
Published: (2023)
SPRMamba: Surgical Phase Recognition for Endoscopic Submucosal Dissection with Mamba
by: Zhang, Xiangning, et al.
Published: (2024)
by: Zhang, Xiangning, et al.
Published: (2024)
Scaling Video Pretraining for Surgical Foundation Models
by: Lu, Sicheng, et al.
Published: (2026)
by: Lu, Sicheng, et al.
Published: (2026)
Surgical Depth Anything: Depth Estimation for Surgical Scenes using Foundation Models
by: Lou, Ange, et al.
Published: (2024)
by: Lou, Ange, et al.
Published: (2024)
Insect-Foundation: A Foundation Model and Large Multimodal Dataset for Vision-Language Insect Understanding
by: Truong, Thanh-Dat, et al.
Published: (2025)
by: Truong, Thanh-Dat, et al.
Published: (2025)
Insect-Foundation: A Foundation Model and Large-scale 1M Dataset for Visual Insect Understanding
by: Nguyen, Hoang-Quan, et al.
Published: (2023)
by: Nguyen, Hoang-Quan, et al.
Published: (2023)
CPKD: Clinical Prior Knowledge-Constrained Diffusion Models for Surgical Phase Recognition in Endoscopic Submucosal Dissection
by: Zhang, Xiangning, et al.
Published: (2025)
by: Zhang, Xiangning, et al.
Published: (2025)
PERSEUS: Perception with Semantic Endoscopic Understanding and SLAM
by: Acar, Ayberk, et al.
Published: (2025)
by: Acar, Ayberk, et al.
Published: (2025)
Where It Moves, It Matters: Referring Surgical Instrument Segmentation via Motion
by: Wei, Meng, et al.
Published: (2026)
by: Wei, Meng, et al.
Published: (2026)
Open Set Recognition for Endoscopic Image Classification: A Deep Learning Approach on the Kvasir Dataset
by: Moazzami, Kasra, et al.
Published: (2025)
by: Moazzami, Kasra, et al.
Published: (2025)
Leveraging Generic Foundation Models for Multimodal Surgical Data Analysis
by: Pezold, Simon, et al.
Published: (2025)
by: Pezold, Simon, et al.
Published: (2025)
Similar Items
-
Back to the Feature: Explaining Video Classifiers with Video Counterfactual Explanations
by: Wang, Chao, et al.
Published: (2025) -
A Stitch in Time: Learning Procedural Workflow via Self-Supervised Plackett-Luce Ranking
by: Che, Chengan, et al.
Published: (2025) -
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training?
by: Che, Chengan, et al.
Published: (2026) -
LoViT: Long Video Transformer for Surgical Phase Recognition
by: Liu, Yang, et al.
Published: (2023) -
Average Calibration Error: A Differentiable Loss for Improved Reliability in Image Segmentation
by: Barfoot, Theodore, et al.
Published: (2024)