Adapting Foundation Vision-Language Models to Medical Diagnosis via Query-Driven Expert Bridging
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yitong, Ghahremani, Morteza, Wachinger, Christian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DiaMond: Dementia Diagnosis with Multi-Modal Vision Transformers Using MRI and PET
by: Li, Yitong, et al.
Published: (2024)
by: Li, Yitong, et al.
Published: (2024)
Stable-Pose: Leveraging Transformers for Pose-Guided Text-to-Image Generation
by: Wang, Jiajun, et al.
Published: (2024)
by: Wang, Jiajun, et al.
Published: (2024)
Conditional Diffusion for 3D CT Volume Reconstruction from 2D X-rays
by: Rath, Martin, et al.
Published: (2026)
by: Rath, Martin, et al.
Published: (2026)
Generating Reports or Repeating Templates? Measuring and Mitigating Template Collapse in 3D CT Report Generation
by: Maye-Lasserre, Tom, et al.
Published: (2026)
by: Maye-Lasserre, Tom, et al.
Published: (2026)
3D Shape-to-Image Brownian Bridge Diffusion for Brain MRI Synthesis from Cortical Surfaces
by: Bongratz, Fabian, et al.
Published: (2025)
by: Bongratz, Fabian, et al.
Published: (2025)
Disentangling Progress in Medical Image Registration: Beyond Trend-Driven Architectures towards Domain-Specific Strategies
by: Jian, Bailiang, et al.
Published: (2025)
by: Jian, Bailiang, et al.
Published: (2025)
Mamba? Catch The Hype Or Rethink What Really Helps for Image Registration
by: Jian, Bailiang, et al.
Published: (2024)
by: Jian, Bailiang, et al.
Published: (2024)
Adapting Vision-Language Foundation Model for Next Generation Medical Ultrasound Image Analysis
by: Qu, Jingguo, et al.
Published: (2025)
by: Qu, Jingguo, et al.
Published: (2025)
Translating MRI to PET through Conditional Diffusion Models with Enhanced Pathology Awareness
by: Li, Yitong, et al.
Published: (2026)
by: Li, Yitong, et al.
Published: (2026)
Uncertainty-Driven Expert Control: Enhancing the Reliability of Medical Vision-Language Models
by: Liang, Xiao, et al.
Published: (2025)
by: Liang, Xiao, et al.
Published: (2025)
Cortex-Grounded Diffusion Models for Brain Image Generation
by: Bongratz, Fabian, et al.
Published: (2026)
by: Bongratz, Fabian, et al.
Published: (2026)
PASTA: Pathology-Aware MRI to PET Cross-Modal Translation with Diffusion Models
by: Li, Yitong, et al.
Published: (2024)
by: Li, Yitong, et al.
Published: (2024)
Adapt-As-You-Walk Through the Clouds: Training-Free Online Test-Time Adaptation of 3D Vision-Language Foundation Models
by: Tamjidi, Mehran, et al.
Published: (2025)
by: Tamjidi, Mehran, et al.
Published: (2025)
Adversarial Distortion Learning for Medical Image Denoising
by: Ghahremani, Morteza, et al.
Published: (2022)
by: Ghahremani, Morteza, et al.
Published: (2022)
Diffusion Bridge Networks Simulate Clinical-grade PET from MRI for Dementia Diagnostics
by: Li, Yitong, et al.
Published: (2025)
by: Li, Yitong, et al.
Published: (2025)
X-SiT: Inherently Interpretable Surface Vision Transformers for Dementia Diagnosis
by: Bongratz, Fabian, et al.
Published: (2025)
by: Bongratz, Fabian, et al.
Published: (2025)
WeMMU: Enhanced Bridging of Vision-Language Models and Diffusion Models via Noisy Query Tokens
by: Yang, Jian, et al.
Published: (2025)
by: Yang, Jian, et al.
Published: (2025)
Adapting a Segmentation Foundation Model for Medical Image Classification
by: Gu, Pengfei, et al.
Published: (2025)
by: Gu, Pengfei, et al.
Published: (2025)
Parameters as Experts: Adapting Vision Models with Dynamic Parameter Routing
by: Lou, Meng, et al.
Published: (2026)
by: Lou, Meng, et al.
Published: (2026)
MoIIE: Mixture of Intra- and Inter-Modality Experts for Large Vision Language Models
by: Wang, Dianyi, et al.
Published: (2025)
by: Wang, Dianyi, et al.
Published: (2025)
Learning to Adapt Foundation Model DINOv2 for Capsule Endoscopy Diagnosis
by: Zhang, Bowen, et al.
Published: (2024)
by: Zhang, Bowen, et al.
Published: (2024)
Adapting Foundation Models for Few-Shot Medical Image Segmentation: Actively and Sequentially
by: Yang, Jingyun, et al.
Published: (2025)
by: Yang, Jingyun, et al.
Published: (2025)
Industrial Language-Image Dataset (ILID): Adapting Vision Foundation Models for Industrial Settings
by: Moenck, Keno, et al.
Published: (2024)
by: Moenck, Keno, et al.
Published: (2024)
From Barlow Twins to Triplet Training: Differentiating Dementia with Limited Data
by: Li, Yitong, et al.
Published: (2024)
by: Li, Yitong, et al.
Published: (2024)
Spherical Brownian Bridge Diffusion Models for Conditional Cortical Thickness Forecasting
by: Stoyanov, Ivan, et al.
Published: (2025)
by: Stoyanov, Ivan, et al.
Published: (2025)
ViLReF: An Expert Knowledge Enabled Vision-Language Retinal Foundation Model
by: Yang, Shengzhu, et al.
Published: (2024)
by: Yang, Shengzhu, et al.
Published: (2024)
MoVA: Adapting Mixture of Vision Experts to Multimodal Context
by: Zong, Zhuofan, et al.
Published: (2024)
by: Zong, Zhuofan, et al.
Published: (2024)
VILA-M3: Enhancing Vision-Language Models with Medical Expert Knowledge
by: Nath, Vishwesh, et al.
Published: (2024)
by: Nath, Vishwesh, et al.
Published: (2024)
AdaptSplat: Adapting Vision Foundation Models for Feed-Forward 3D Gaussian Splatting
by: Xing, Mingwei, et al.
Published: (2026)
by: Xing, Mingwei, et al.
Published: (2026)
Fusion of Domain-Adapted Vision and Language Models for Medical Visual Question Answering
by: Ha, Cuong Nhat, et al.
Published: (2024)
by: Ha, Cuong Nhat, et al.
Published: (2024)
Enhancing Representation in Medical Vision-Language Foundation Models via Multi-Scale Information Extraction Techniques
by: Huang, Weijian, et al.
Published: (2024)
by: Huang, Weijian, et al.
Published: (2024)
Vision-Language Enhanced Foundation Model for Semi-supervised Medical Image Segmentation
by: Guo, Jiaqi, et al.
Published: (2025)
by: Guo, Jiaqi, et al.
Published: (2025)
Demographic Bias of Expert-Level Vision-Language Foundation Models in Medical Imaging
by: Yang, Yuzhe, et al.
Published: (2024)
by: Yang, Yuzhe, et al.
Published: (2024)
See-in-Pairs: Reference Image-Guided Comparative Vision-Language Models for Medical Diagnosis
by: Jin, Ruinan, et al.
Published: (2025)
by: Jin, Ruinan, et al.
Published: (2025)
Med-LEGO: Editing and Adapting toward Generalist Medical Image Diagnosis
by: Zhu, Yitao, et al.
Published: (2025)
by: Zhu, Yitao, et al.
Published: (2025)
Adapting Vision Foundation Models for Robust Cloud Segmentation in Remote Sensing Images
by: Zou, Xuechao, et al.
Published: (2024)
by: Zou, Xuechao, et al.
Published: (2024)
Tell2Adapt: A Unified Framework for Source Free Unsupervised Domain Adaptation via Vision Foundation Model
by: Shi, Yulong, et al.
Published: (2026)
by: Shi, Yulong, et al.
Published: (2026)
Individualized Mapping of Aberrant Cortical Thickness via Stochastic Cortical Self-Reconstruction
by: Wachinger, Christian, et al.
Published: (2024)
by: Wachinger, Christian, et al.
Published: (2024)
Adapting Vision Foundation Models for Real-time Ultrasound Image Segmentation
by: Zhang, Xiaoran, et al.
Published: (2025)
by: Zhang, Xiaoran, et al.
Published: (2025)
Controllable Complex Human Motion Video Generation via Text-to-Skeleton Cascades
by: Taghipour, Ashkan, et al.
Published: (2026)
by: Taghipour, Ashkan, et al.
Published: (2026)
Similar Items
-
DiaMond: Dementia Diagnosis with Multi-Modal Vision Transformers Using MRI and PET
by: Li, Yitong, et al.
Published: (2024) -
Stable-Pose: Leveraging Transformers for Pose-Guided Text-to-Image Generation
by: Wang, Jiajun, et al.
Published: (2024) -
Conditional Diffusion for 3D CT Volume Reconstruction from 2D X-rays
by: Rath, Martin, et al.
Published: (2026) -
Generating Reports or Repeating Templates? Measuring and Mitigating Template Collapse in 3D CT Report Generation
by: Maye-Lasserre, Tom, et al.
Published: (2026) -
3D Shape-to-Image Brownian Bridge Diffusion for Brain MRI Synthesis from Cortical Surfaces
by: Bongratz, Fabian, et al.
Published: (2025)