DeViDe: Faceted medical knowledge for improved medical vision-language pre-training
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Haozhe, Zhou, Ziyu, Royer, Corentin, Sekuboyina, Anjany, Menze, Bjoern |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MultiMedEval: A Benchmark and a Toolkit for Evaluating Medical Vision-Language Models
by: Royer, Corentin, et al.
Published: (2024)
by: Royer, Corentin, et al.
Published: (2024)
RadDiagSeg-M: A Vision Language Model for Joint Diagnosis and Multi-Target Segmentation in Radiology
by: Li, Chengrun, et al.
Published: (2025)
by: Li, Chengrun, et al.
Published: (2025)
MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation
by: Xu, Lijian, et al.
Published: (2024)
by: Xu, Lijian, et al.
Published: (2024)
Comprehensive language-image pre-training for 3D medical image understanding
by: Wald, Tassilo, et al.
Published: (2025)
by: Wald, Tassilo, et al.
Published: (2025)
Anatomical grounding pre-training for medical phrase grounding
by: Zhang, Wenjun, et al.
Published: (2025)
by: Zhang, Wenjun, et al.
Published: (2025)
Enhancing Interpretability of Vertebrae Fracture Grading using Human-interpretable Prototypes
by: Sinhamahapatra, Poulami, et al.
Published: (2024)
by: Sinhamahapatra, Poulami, et al.
Published: (2024)
Beyond the Hype: A dispassionate look at vision-language models in medical scenario
by: Nan, Yang, et al.
Published: (2024)
by: Nan, Yang, et al.
Published: (2024)
More performant and scalable: Rethinking contrastive vision-language pre-training of radiology in the LLM era
by: Li, Yingtai, et al.
Published: (2025)
by: Li, Yingtai, et al.
Published: (2025)
Enhancing medical vision-language contrastive learning via inter-matching relation modelling
by: Li, Mingjian, et al.
Published: (2024)
by: Li, Mingjian, et al.
Published: (2024)
Deep classification algorithm for De-identification of DICOM medical images
by: Michele, Bufano, et al.
Published: (2025)
by: Michele, Bufano, et al.
Published: (2025)
VERIDAH: Solving Enumeration Anomaly Aware Vertebra Labeling across Imaging Sequences
by: Möller, Hendrik, et al.
Published: (2026)
by: Möller, Hendrik, et al.
Published: (2026)
GenerateCT: Text-Conditional Generation of 3D Chest CT Volumes
by: Hamamci, Ibrahim Ethem, et al.
Published: (2023)
by: Hamamci, Ibrahim Ethem, et al.
Published: (2023)
Can language-guided unsupervised adaptation improve medical image classification using unpaired images and texts?
by: Rahman, Umaima, et al.
Published: (2024)
by: Rahman, Umaima, et al.
Published: (2024)
MIMO: A medical vision language model with visual referring multimodal input and pixel grounding multimodal output
by: Chen, Yanyuan, et al.
Published: (2025)
by: Chen, Yanyuan, et al.
Published: (2025)
Revisiting MAE pre-training for 3D medical image segmentation
by: Wald, Tassilo, et al.
Published: (2024)
by: Wald, Tassilo, et al.
Published: (2024)
Mobile U-ViT: Revisiting large kernel and U-shaped ViT for efficient medical image segmentation
by: Tang, Fenghe, et al.
Published: (2025)
by: Tang, Fenghe, et al.
Published: (2025)
OmDet: Large-scale vision-language multi-dataset pre-training with multimodal detection network
by: Zhao, Tiancheng, et al.
Published: (2022)
by: Zhao, Tiancheng, et al.
Published: (2022)
When are radiology reports useful for training medical image classifiers?
by: Bergström, Herman, et al.
Published: (2025)
by: Bergström, Herman, et al.
Published: (2025)
LLaVAction: evaluating and training multi-modal large language models for action understanding
by: Qi, Haozhe, et al.
Published: (2025)
by: Qi, Haozhe, et al.
Published: (2025)
CT2Rep: Automated Radiology Report Generation for 3D Medical Imaging
by: Hamamci, Ibrahim Ethem, et al.
Published: (2024)
by: Hamamci, Ibrahim Ethem, et al.
Published: (2024)
Enhancing Cardiovascular Disease Prediction through Multi-Modal Self-Supervised Learning
by: Girlanda, Francesco, et al.
Published: (2024)
by: Girlanda, Francesco, et al.
Published: (2024)
BUSClean: Open-source software for breast ultrasound image pre-processing and knowledge extraction for medical AI
by: Bunnell, Arianna, et al.
Published: (2024)
by: Bunnell, Arianna, et al.
Published: (2024)
VariViT: A Vision Transformer for Variable Image Sizes
by: Varma, Aswathi, et al.
Published: (2026)
by: Varma, Aswathi, et al.
Published: (2026)
PTQ4ViT: Post-training quantization for vision transformers with twin uniform quantization
by: Yuan, Zhihang, et al.
Published: (2021)
by: Yuan, Zhihang, et al.
Published: (2021)
Semantically Consistent Discrete Diffusion for 3D Biological Graph Modeling
by: Prabhakar, Chinmay, et al.
Published: (2025)
by: Prabhakar, Chinmay, et al.
Published: (2025)
Link Prediction for Flow-Driven Spatial Networks
by: Wittmann, Bastian, et al.
Published: (2023)
by: Wittmann, Bastian, et al.
Published: (2023)
KEPIL: Knowledge-Enhanced Prompt-Image Learning for Prompt-Robust Disease Detection
by: Luo, Haozhe, et al.
Published: (2026)
by: Luo, Haozhe, et al.
Published: (2026)
Hi-End-MAE: Hierarchical encoder-driven masked autoencoders are stronger vision learners for medical image segmentation
by: Tang, Fenghe, et al.
Published: (2025)
by: Tang, Fenghe, et al.
Published: (2025)
MedDINOv3: How to adapt vision foundation models for medical image segmentation?
by: Li, Yuheng, et al.
Published: (2025)
by: Li, Yuheng, et al.
Published: (2025)
ViSTa Dataset: Do vision-language models understand sequential tasks?
by: Wybitul, Evžen, et al.
Published: (2024)
by: Wybitul, Evžen, et al.
Published: (2024)
SPINEPS -- Automatic Whole Spine Segmentation of T2-weighted MR images using a Two-Phase Approach to Multi-class Semantic and Instance Segmentation
by: Möller, Hendrik, et al.
Published: (2024)
by: Möller, Hendrik, et al.
Published: (2024)
Enhancing the vision-language foundation model with key semantic knowledge-emphasized report refinement
by: Huang, Weijian, et al.
Published: (2024)
by: Huang, Weijian, et al.
Published: (2024)
Learning Anatomically Consistent Embedding for Chest Radiography
by: Zhou, Ziyu, et al.
Published: (2023)
by: Zhou, Ziyu, et al.
Published: (2023)
UCell: rethinking generalizability and scaling of bio-medical vision models
by: Kuang, Nicholas, et al.
Published: (2026)
by: Kuang, Nicholas, et al.
Published: (2026)
Shape-intensity knowledge distillation for robust medical image segmentation
by: Dong, Wenhui, et al.
Published: (2024)
by: Dong, Wenhui, et al.
Published: (2024)
Predicting Stroke through Retinal Graphs and Multimodal Self-supervised Learning
by: Huang, Yuqing, et al.
Published: (2024)
by: Huang, Yuqing, et al.
Published: (2024)
MedVL-SAM2: A unified 3D medical vision-language model for multimodal reasoning and prompt-driven segmentation
by: Xing, Yang, et al.
Published: (2026)
by: Xing, Yang, et al.
Published: (2026)
Hyperlocal disaster damage assessment using bi-temporal street-view imagery and pre-trained vision models
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
Circle of Willis Centerline Graphs: A Dataset and Baseline Algorithm
by: Musio, Fabio, et al.
Published: (2025)
by: Musio, Fabio, et al.
Published: (2025)
Combining inherent knowledge of vision-language models with unsupervised domain adaptation through strong-weak guidance
by: Westfechtel, Thomas, et al.
Published: (2023)
by: Westfechtel, Thomas, et al.
Published: (2023)
Similar Items
-
MultiMedEval: A Benchmark and a Toolkit for Evaluating Medical Vision-Language Models
by: Royer, Corentin, et al.
Published: (2024) -
RadDiagSeg-M: A Vision Language Model for Joint Diagnosis and Multi-Target Segmentation in Radiology
by: Li, Chengrun, et al.
Published: (2025) -
MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation
by: Xu, Lijian, et al.
Published: (2024) -
Comprehensive language-image pre-training for 3D medical image understanding
by: Wald, Tassilo, et al.
Published: (2025) -
Anatomical grounding pre-training for medical phrase grounding
by: Zhang, Wenjun, et al.
Published: (2025)