A New Era in Computational Pathology: A Survey on Foundation and Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Chanda, Dibaloke, Aryal, Milan, Soltani, Nasim Yahya, Ganji, Masoud |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Heterogeneous Graph-Based Multi-Task Learning for Fault Event Diagnosis in Smart Grid
by: Chanda, Dibaloke, et al.
Published: (2023)
by: Chanda, Dibaloke, et al.
Published: (2023)
VisionUnite: A Vision-Language Foundation Model for Ophthalmology Enhanced with Clinical Knowledge
by: Li, Zihan, et al.
Published: (2024)
by: Li, Zihan, et al.
Published: (2024)
Survey: Transformer-based Models in Data Modality Conversion
by: Rashno, Elyas, et al.
Published: (2024)
by: Rashno, Elyas, et al.
Published: (2024)
Simplifying Multimodality: Unimodal Approach to Multimodal Challenges in Radiology with General-Domain Large Language Model
by: Cho, Seonhee, et al.
Published: (2024)
by: Cho, Seonhee, et al.
Published: (2024)
THOR: A Versatile Foundation Model for Earth Observation Climate and Society Applications
by: Forgaard, Theodor, et al.
Published: (2026)
by: Forgaard, Theodor, et al.
Published: (2026)
Understanding Task Aggregation for Generalizable Ultrasound Foundation Models
by: Wang, Fangyijie, et al.
Published: (2026)
by: Wang, Fangyijie, et al.
Published: (2026)
Assessment of Cell Nuclei AI Foundation Models in Kidney Pathology
by: Guo, Junlin, et al.
Published: (2024)
by: Guo, Junlin, et al.
Published: (2024)
Prototype-Guided Diffusion for Digital Pathology: Achieving Foundation Model Performance with Minimal Clinical Data
by: Redekop, Ekaterina, et al.
Published: (2025)
by: Redekop, Ekaterina, et al.
Published: (2025)
A Disease-Centric Vision-Language Foundation Model for Precision Oncology in Kidney Cancer
by: Tao, Yuhui, et al.
Published: (2025)
by: Tao, Yuhui, et al.
Published: (2025)
Multispectral to Hyperspectral using Pretrained Foundational model
by: Gonzalez, Ruben, et al.
Published: (2025)
by: Gonzalez, Ruben, et al.
Published: (2025)
Evaluating Cell AI Foundation Models in Kidney Pathology with Human-in-the-Loop Enrichment
by: Guo, Junlin, et al.
Published: (2024)
by: Guo, Junlin, et al.
Published: (2024)
Structural Entities Extraction and Patient Indications Incorporation for Chest X-ray Report Generation
by: Liu, Kang, et al.
Published: (2024)
by: Liu, Kang, et al.
Published: (2024)
OpenVision 3: A Family of Unified Visual Encoder for Both Understanding and Generation
by: Zhang, Letian, et al.
Published: (2026)
by: Zhang, Letian, et al.
Published: (2026)
Beyond Calibration: Confounding Pathology Limits Foundation Model Specificity in Abdominal Trauma CT
by: Raythatha, Jineel H, et al.
Published: (2026)
by: Raythatha, Jineel H, et al.
Published: (2026)
Ensemble of Pathology Foundation Models for MIDOG 2025 Track 2: Atypical Mitosis Classification
by: Ochi, Mieko, et al.
Published: (2025)
by: Ochi, Mieko, et al.
Published: (2025)
Computational Pathology for Accurate Prediction of Breast Cancer Recurrence: Development and Validation of a Deep Learning-based Tool
by: Su, Ziyu, et al.
Published: (2024)
by: Su, Ziyu, et al.
Published: (2024)
Vision Language Models in Medicine
by: Kalpelbe, Beria Chingnabe, et al.
Published: (2025)
by: Kalpelbe, Beria Chingnabe, et al.
Published: (2025)
Subspecialty-Specific Foundation Model for Intelligent Gastrointestinal Pathology
by: Zhu, Lianghui, et al.
Published: (2025)
by: Zhu, Lianghui, et al.
Published: (2025)
A Versatile Pathology Co-pilot via Reasoning Enhanced Multimodal Large Language Model
by: Xu, Zhe, et al.
Published: (2025)
by: Xu, Zhe, et al.
Published: (2025)
A Survey of Multimodal Ophthalmic Diagnostics: From Task-Specific Approaches to Foundational Models
by: Luo, Xiaoling, et al.
Published: (2025)
by: Luo, Xiaoling, et al.
Published: (2025)
A Scalable AI Driven, IoT Integrated Cognitive Digital Twin for Multi-Modal Neuro-Oncological Prognostics and Tumor Kinetics Prediction using Enhanced Vision Transformer and XAI
by: Banerjee, Saptarshi, et al.
Published: (2025)
by: Banerjee, Saptarshi, et al.
Published: (2025)
A Short Survey on Set-Based Aggregation Techniques for Single-Vector WSI Representation in Digital Pathology
by: Hemati, S., et al.
Published: (2024)
by: Hemati, S., et al.
Published: (2024)
Multimodal Whole Slide Foundation Model for Pathology
by: Ding, Tong, et al.
Published: (2024)
by: Ding, Tong, et al.
Published: (2024)
Vision Transformers for End-to-End Vision-Based Quadrotor Obstacle Avoidance
by: Bhattacharya, Anish, et al.
Published: (2024)
by: Bhattacharya, Anish, et al.
Published: (2024)
An Inclusive Foundation Model for Generalizable Cytogenetics in Precision Oncology
by: Yang, Changchun, et al.
Published: (2025)
by: Yang, Changchun, et al.
Published: (2025)
Computer aided diagnosis system for Alzheimers disease using principal component analysis and machine learning based approaches
by: Lazli, Lilia
Published: (2024)
by: Lazli, Lilia
Published: (2024)
Are Vision Foundation Models Ready for Out-of-the-Box Medical Image Registration?
by: Gu, Hanxue, et al.
Published: (2025)
by: Gu, Hanxue, et al.
Published: (2025)
A Disease Labeler for Chinese Chest X-Ray Report Generation
by: Wang, Mengwei, et al.
Published: (2024)
by: Wang, Mengwei, et al.
Published: (2024)
Enabling Ultra-Fast Cardiovascular Imaging Across Heterogeneous Clinical Environments with A Generalist Foundation Model and Multimodal Database
by: Wang, Zi, et al.
Published: (2025)
by: Wang, Zi, et al.
Published: (2025)
FetalCLIP: A Visual-Language Foundation Model for Fetal Ultrasound Image Analysis
by: Maani, Fadillah, et al.
Published: (2025)
by: Maani, Fadillah, et al.
Published: (2025)
RAVEN: Radar Adaptive Vision Encoders for Efficient Chirp-wise Object Detection and Segmentation
by: Sen, Anuvab, et al.
Published: (2026)
by: Sen, Anuvab, et al.
Published: (2026)
High Efficiency Image Compression for Large Visual-Language Models
by: Li, Binzhe, et al.
Published: (2024)
by: Li, Binzhe, et al.
Published: (2024)
MANTA -- Model Adapter Native generations that's Affordable
by: Chaurasia, Ansh
Published: (2024)
by: Chaurasia, Ansh
Published: (2024)
Foundation Models and Information Retrieval in Digital Pathology
by: Tizhoosh, H. R.
Published: (2024)
by: Tizhoosh, H. R.
Published: (2024)
Towards Robust Foundation Models for Digital Pathology
by: Kömen, Jonah, et al.
Published: (2025)
by: Kömen, Jonah, et al.
Published: (2025)
Artificial Intelligence for Digital and Computational Pathology
by: Song, Andrew H., et al.
Published: (2023)
by: Song, Andrew H., et al.
Published: (2023)
Leveraging Multimodal Models for Enhanced Neuroimaging Diagnostics in Alzheimer's Disease
by: Chiumento, Francesco, et al.
Published: (2024)
by: Chiumento, Francesco, et al.
Published: (2024)
A Target-Free Harmonization Method for MRI
by: Kim, Minjun, et al.
Published: (2026)
by: Kim, Minjun, et al.
Published: (2026)
Pathological MRI Segmentation by Synthetic Pathological Data Generation in Fetuses and Neonates
by: Kaandorp, Misha P. T, et al.
Published: (2025)
by: Kaandorp, Misha P. T, et al.
Published: (2025)
Efficient Convolutional Forward Model for Passive Acoustic Mapping and Temporal Monitoring
by: Gelvez-Barrera, Tatiana, et al.
Published: (2026)
by: Gelvez-Barrera, Tatiana, et al.
Published: (2026)
Similar Items
-
A Heterogeneous Graph-Based Multi-Task Learning for Fault Event Diagnosis in Smart Grid
by: Chanda, Dibaloke, et al.
Published: (2023) -
VisionUnite: A Vision-Language Foundation Model for Ophthalmology Enhanced with Clinical Knowledge
by: Li, Zihan, et al.
Published: (2024) -
Survey: Transformer-based Models in Data Modality Conversion
by: Rashno, Elyas, et al.
Published: (2024) -
Simplifying Multimodality: Unimodal Approach to Multimodal Challenges in Radiology with General-Domain Large Language Model
by: Cho, Seonhee, et al.
Published: (2024) -
THOR: A Versatile Foundation Model for Earth Observation Climate and Society Applications
by: Forgaard, Theodor, et al.
Published: (2026)