MedDINOv3: How to adapt vision foundation models for medical image segmentation?
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yuheng, Wu, Yizhou, Lai, Yuxiang, Hu, Mingzhe, Yang, Xiaofeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting
by: Li, Yuheng, et al.
Published: (2025)
by: Li, Yuheng, et al.
Published: (2025)
Are foundation models efficient for medical image segmentation?
by: Ferreira, Danielle, et al.
Published: (2023)
by: Ferreira, Danielle, et al.
Published: (2023)
ENSAM: an efficient foundation model for interactive segmentation of 3D medical images
by: Stenhede, Elias, et al.
Published: (2025)
by: Stenhede, Elias, et al.
Published: (2025)
Med-R1: Reinforcement Learning for Generalizable Medical Reasoning in Vision-Language Models
by: Lai, Yuxiang, et al.
Published: (2025)
by: Lai, Yuxiang, et al.
Published: (2025)
DINOv3 with Test-Time Training for Medical Image Registration
by: Wang, Shansong, et al.
Published: (2025)
by: Wang, Shansong, et al.
Published: (2025)
Extending 2D foundational DINOv3 representations to 3D segmentation of neonatal brain MR images
by: Usman, Annayah, et al.
Published: (2026)
by: Usman, Annayah, et al.
Published: (2026)
MedVL-SAM2: A unified 3D medical vision-language model for multimodal reasoning and prompt-driven segmentation
by: Xing, Yang, et al.
Published: (2026)
by: Xing, Yang, et al.
Published: (2026)
Fine-tuning vision foundation model for crack segmentation in civil infrastructures
by: Ge, Kang, et al.
Published: (2023)
by: Ge, Kang, et al.
Published: (2023)
Are Video Models Emerging as Zero-Shot Learners and Reasoners in Medical Imaging?
by: Lai, Yuxiang, et al.
Published: (2025)
by: Lai, Yuxiang, et al.
Published: (2025)
Zero-shot segmentation of skin tumors in whole-slide images with vision-language foundation models
by: Moreno, Santiago, et al.
Published: (2025)
by: Moreno, Santiago, et al.
Published: (2025)
Dynamically evolving segment anything model with continuous learning for medical image segmentation
by: Liu, Zhaori, et al.
Published: (2025)
by: Liu, Zhaori, et al.
Published: (2025)
How to build the best medical image segmentation algorithm using foundation models: a comprehensive empirical study with Segment Anything Model
by: Gu, Hanxue, et al.
Published: (2024)
by: Gu, Hanxue, et al.
Published: (2024)
RoMedFormer: A Rotary-Embedding Transformer Foundation Model for 3D Genito-Pelvic Structure Segmentation in MRI and CT
by: Li, Yuheng, et al.
Published: (2025)
by: Li, Yuheng, et al.
Published: (2025)
MedLVR: Latent Visual Reasoning for Reliable Medical Visual Question Answering
by: Xi, Suyang, et al.
Published: (2026)
by: Xi, Suyang, et al.
Published: (2026)
AnatoMask: Enhancing Medical Image Segmentation with Reconstruction-guided Self-masking
by: Li, Yuheng, et al.
Published: (2024)
by: Li, Yuheng, et al.
Published: (2024)
Object segmentation in the wild with foundation models: application to vision assisted neuro-prostheses for upper limbs
by: Atoki, Bolutife, et al.
Published: (2025)
by: Atoki, Bolutife, et al.
Published: (2025)
BrainDINO: A Brain MRI Foundation Model for Generalizable Clinical Representation Learning
by: Wu, Yizhou, et al.
Published: (2026)
by: Wu, Yizhou, et al.
Published: (2026)
Self-adaptive vision-language model for 3D segmentation of pulmonary artery and vein
by: Guo, Xiaotong, et al.
Published: (2025)
by: Guo, Xiaotong, et al.
Published: (2025)
VISTA-PATH: An interactive foundation model for pathology image segmentation and quantitative analysis in computational pathology
by: Liang, Peixian, et al.
Published: (2026)
by: Liang, Peixian, et al.
Published: (2026)
Hi-End-MAE: Hierarchical encoder-driven masked autoencoders are stronger vision learners for medical image segmentation
by: Tang, Fenghe, et al.
Published: (2025)
by: Tang, Fenghe, et al.
Published: (2025)
Revisiting foundation models for cell instance segmentation
by: Archit, Anwai, et al.
Published: (2026)
by: Archit, Anwai, et al.
Published: (2026)
SemiSAM-O1: How far can we push the boundary of annotation-efficient medical image segmentation?
by: Zhang, Yichi, et al.
Published: (2026)
by: Zhang, Yichi, et al.
Published: (2026)
ProFound: A moderate-sized vision foundation model for multi-task prostate imaging
by: Wang, Yipei, et al.
Published: (2026)
by: Wang, Yipei, et al.
Published: (2026)
Upsampling DINOv2 features for unsupervised vision tasks and weakly supervised materials segmentation
by: Docherty, Ronan, et al.
Published: (2024)
by: Docherty, Ronan, et al.
Published: (2024)
Cross-head mutual Mean-Teaching for semi-supervised medical image segmentation
by: Li, Wei, et al.
Published: (2023)
by: Li, Wei, et al.
Published: (2023)
Novel class discovery meets foundation models for 3D semantic segmentation
by: Riz, Luigi, et al.
Published: (2023)
by: Riz, Luigi, et al.
Published: (2023)
DINOv3
by: Siméoni, Oriane, et al.
Published: (2025)
by: Siméoni, Oriane, et al.
Published: (2025)
Impact of domain adaptation in deep learning for medical image classifications
by: Wu, Yihang, et al.
Published: (2026)
by: Wu, Yihang, et al.
Published: (2026)
Towards Universal Text-driven CT Image Segmentation
by: Li, Yuheng, et al.
Published: (2025)
by: Li, Yuheng, et al.
Published: (2025)
Context Matters: Learning Global Semantics via Object-Centric Representation
by: Zhong, Jike, et al.
Published: (2025)
by: Zhong, Jike, et al.
Published: (2025)
Semantic and structural image segmentation for prosthetic vision
by: Sanchez-Garcia, Melani, et al.
Published: (2018)
by: Sanchez-Garcia, Melani, et al.
Published: (2018)
Universal CT Representations from Anatomy to Disease Phenotype through Agglomerative Pretraining
by: Li, Yuheng, et al.
Published: (2026)
by: Li, Yuheng, et al.
Published: (2026)
EXACT: an explainable anomaly-aware vision foundation model for analysis of 3D chest CT
by: Bai, Xuguang, et al.
Published: (2026)
by: Bai, Xuguang, et al.
Published: (2026)
PARF-Net: integrating pixel-wise adaptive receptive fields into hybrid Transformer-CNN network for medical image segmentation
by: Ma, Xu, et al.
Published: (2025)
by: Ma, Xu, et al.
Published: (2025)
U-RWKV: Lightweight medical image segmentation with direction-adaptive RWKV
by: Ye, Hongbo, et al.
Published: (2025)
by: Ye, Hongbo, et al.
Published: (2025)
AD-DINOv3: Enhancing DINOv3 for Zero-Shot Anomaly Detection with Anomaly-Aware Calibration
by: Yuan, Jingyi, et al.
Published: (2025)
by: Yuan, Jingyi, et al.
Published: (2025)
MedIAnomaly: A comparative study of anomaly detection in medical images
by: Cai, Yu, et al.
Published: (2024)
by: Cai, Yu, et al.
Published: (2024)
Synergistic Foundation Models for Semi-Supervised Fetal Cardiac Ultrasound Analysis: SAM-Med2D Boundary Refinement and DINOv3 Semantic Enhancement
by: Zhuang, Tonghao, et al.
Published: (2026)
by: Zhuang, Tonghao, et al.
Published: (2026)
Are foundation models for computer vision good conformal predictors?
by: Fillioux, Leo, et al.
Published: (2024)
by: Fillioux, Leo, et al.
Published: (2024)
Rethinking Cross-Generator Image Forgery Detection through DINOv3
by: Huang, Zhenglin, et al.
Published: (2025)
by: Huang, Zhenglin, et al.
Published: (2025)
Similar Items
-
MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting
by: Li, Yuheng, et al.
Published: (2025) -
Are foundation models efficient for medical image segmentation?
by: Ferreira, Danielle, et al.
Published: (2023) -
ENSAM: an efficient foundation model for interactive segmentation of 3D medical images
by: Stenhede, Elias, et al.
Published: (2025) -
Med-R1: Reinforcement Learning for Generalizable Medical Reasoning in Vision-Language Models
by: Lai, Yuxiang, et al.
Published: (2025) -
DINOv3 with Test-Time Training for Medical Image Registration
by: Wang, Shansong, et al.
Published: (2025)