Multimodal Machine Learning in Image-Based and Clinical Biomedicine: Survey and Prospects
Fuente:
arXiv
Saved in:
| Main Authors: | Warner, Elisa, Lee, Joonsang, Hsu, William, Syeda-Mahmood, Tanveer, Kahn, Charles, Gevaert, Olivier, Rao, Arvind |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HNOSeg-XS: Extremely Small Hartley Neural Operator for Efficient and Resolution-Robust 3D Image Segmentation
by: Wong, Ken C. L., et al.
Published: (2025)
by: Wong, Ken C. L., et al.
Published: (2025)
Comparative Analysis of Machine Learning Models for Lung Cancer Mutation Detection and Staging Using 3D CT Scans
by: Li, Yiheng, et al.
Published: (2025)
by: Li, Yiheng, et al.
Published: (2025)
Evaluating Automated Radiology Report Quality through Fine-Grained Phrasal Grounding of Clinical Findings
by: Mahmood, Razi, et al.
Published: (2024)
by: Mahmood, Razi, et al.
Published: (2024)
MAM-CLIP: Vision-Language Pretraining on Mammography Atlases for BI-RADS Classification
by: Gulluk, Halil Ibrahim, et al.
Published: (2026)
by: Gulluk, Halil Ibrahim, et al.
Published: (2026)
Improving Medical VQA through Trajectory-Aware Process Supervision
by: Gulluk, Halil Ibrahim, et al.
Published: (2026)
by: Gulluk, Halil Ibrahim, et al.
Published: (2026)
ECLIPSE: Efficient Continual Learning in Panoptic Segmentation with Visual Prompt Tuning
by: Kim, Beomyoung, et al.
Published: (2024)
by: Kim, Beomyoung, et al.
Published: (2024)
Prior Knowledge Injection into Deep Learning Models Predicting Gene Expression from Whole Slide Images
by: Hallemeesch, Max, et al.
Published: (2025)
by: Hallemeesch, Max, et al.
Published: (2025)
A Refer-and-Ground Multimodal Large Language Model for Biomedicine
by: Huang, Xiaoshuang, et al.
Published: (2024)
by: Huang, Xiaoshuang, et al.
Published: (2024)
A Survey on 3D Skeleton Based Person Re-Identification: Taxonomy, Advances, Challenges, and Interdisciplinary Prospects
by: Rao, Haocong, et al.
Published: (2024)
by: Rao, Haocong, et al.
Published: (2024)
Parkinson's Disease Classification Using Contrastive Graph Cross-View Learning with Multimodal Fusion of SPECT Images and Clinical Features
by: Ding, Jun-En, et al.
Published: (2023)
by: Ding, Jun-En, et al.
Published: (2023)
ZIM: Zero-Shot Image Matting for Anything
by: Kim, Beomyoung, et al.
Published: (2024)
by: Kim, Beomyoung, et al.
Published: (2024)
Learning Complex Non-Rigid Image Edits from Multimodal Conditioning
by: Warner, Nikolai, et al.
Published: (2024)
by: Warner, Nikolai, et al.
Published: (2024)
Towards a Multimodal Large Language Model with Pixel-Level Insight for Biomedicine
by: Huang, Xiaoshuang, et al.
Published: (2024)
by: Huang, Xiaoshuang, et al.
Published: (2024)
Multi-modality fusion using canonical correlation analysis methods: Application in breast cancer survival prediction from histology and genomics
by: Subramanian, Vaishnavi, et al.
Published: (2021)
by: Subramanian, Vaishnavi, et al.
Published: (2021)
Modern Hopfield Networks meet Encoded Neural Representations -- Addressing Practical Considerations
by: Kashyap, Satyananda, et al.
Published: (2024)
by: Kashyap, Satyananda, et al.
Published: (2024)
Geo-UNet: A Geometrically Constrained Neural Framework for Clinical-Grade Lumen Segmentation in Intravascular Ultrasound
by: Chen, Yiming, et al.
Published: (2024)
by: Chen, Yiming, et al.
Published: (2024)
Gaussian Mixture Proposals with Pull-Push Learning Scheme to Capture Diverse Events for Weakly Supervised Temporal Video Grounding
by: Kim, Sunoh, et al.
Published: (2023)
by: Kim, Sunoh, et al.
Published: (2023)
Phrase-grounded Fact-checking for Automatically Generated Chest X-ray Reports
by: Mahmood, Razi, et al.
Published: (2025)
by: Mahmood, Razi, et al.
Published: (2025)
Context Matters: Leveraging Spatiotemporal Metadata for Semi-Supervised Learning on Remote Sensing Images
by: Bernhard, Maximilian, et al.
Published: (2024)
by: Bernhard, Maximilian, et al.
Published: (2024)
Revitalizing Electoral Trust: Enhancing Transparency and Efficiency through Automated Voter Counting with Machine Learning
by: Faris, Mir, et al.
Published: (2024)
by: Faris, Mir, et al.
Published: (2024)
SemEnrich: Self-Supervised Semantic Enrichment of Radiology Reports for Vision-Language Learning
by: Gulluk, Halil Ibrahim, et al.
Published: (2026)
by: Gulluk, Halil Ibrahim, et al.
Published: (2026)
Multimodal Deep Learning for Stroke Prediction and Detection using Retinal Imaging and Clinical Data
by: Shurrab, Saeed, et al.
Published: (2025)
by: Shurrab, Saeed, et al.
Published: (2025)
Towards Label-Efficient Human Matting: A Simple Baseline for Weakly Semi-Supervised Trimap-Free Human Matting
by: Kim, Beomyoung, et al.
Published: (2024)
by: Kim, Beomyoung, et al.
Published: (2024)
Machine Unlearning for Image-to-Image Generative Models
by: Li, Guihong, et al.
Published: (2024)
by: Li, Guihong, et al.
Published: (2024)
AgriChain Visually Grounded Expert Verified Reasoning for Interpretable Agricultural Vision Language Models
by: Mahmood, Hazza, et al.
Published: (2026)
by: Mahmood, Hazza, et al.
Published: (2026)
Fine-Grained Zero-Shot Learning: Advances, Challenges, and Prospects
by: Guo, Jingcai, et al.
Published: (2024)
by: Guo, Jingcai, et al.
Published: (2024)
AI-Derived Structural Building Intelligence for Urban Resilience: An Application in Saint Vincent and the Grenadines
by: Tingzon, Isabelle, et al.
Published: (2025)
by: Tingzon, Isabelle, et al.
Published: (2025)
Advancements in Machine Learning for Microrobotics in Biomedicine
by: Amar Salehi, et al.
Published: (2024)
by: Amar Salehi, et al.
Published: (2024)
Research, Applications and Prospects of Event-Based Pedestrian Detection: A Survey
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
EPEE: Towards Efficient and Effective Foundation Models in Biomedicine
by: Zhan, Zaifu, et al.
Published: (2025)
by: Zhan, Zaifu, et al.
Published: (2025)
Evaluating Feature Attribution Methods in the Image Domain
by: Gevaert, Arne, et al.
Published: (2022)
by: Gevaert, Arne, et al.
Published: (2022)
A Survey on Self-supervised Contrastive Learning for Multimodal Text-Image Analysis
by: Khan, Asifullah, et al.
Published: (2025)
by: Khan, Asifullah, et al.
Published: (2025)
Better, Not Just More: Data-Centric Machine Learning for Earth Observation
by: Roscher, Ribana, et al.
Published: (2023)
by: Roscher, Ribana, et al.
Published: (2023)
UMCL: Unimodal-generated Multimodal Contrastive Learning for Cross-compression-rate Deepfake Detection
by: Lai, Ching-Yi, et al.
Published: (2025)
by: Lai, Ching-Yi, et al.
Published: (2025)
A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models
by: Shuai, Xincheng, et al.
Published: (2024)
by: Shuai, Xincheng, et al.
Published: (2024)
Advancing High Resolution Vision-Language Models in Biomedicine
by: Chen, Zekai, et al.
Published: (2024)
by: Chen, Zekai, et al.
Published: (2024)
Learning-Based Image Compression for Machines
by: Gupta, Kartik, et al.
Published: (2024)
by: Gupta, Kartik, et al.
Published: (2024)
Redesigning Traffic Signs to Mitigate Machine-Learning Patch Attacks
by: Shua, Tsufit, et al.
Published: (2024)
by: Shua, Tsufit, et al.
Published: (2024)
Deep Learning Empowered Super-Resolution: A Comprehensive Survey and Future Prospects
by: Zhang, Le, et al.
Published: (2025)
by: Zhang, Le, et al.
Published: (2025)
GestaltMML: Enhancing Rare Genetic Disease Diagnosis through Multimodal Machine Learning Combining Facial Images and Clinical Texts
by: Wu, Da, et al.
Published: (2023)
by: Wu, Da, et al.
Published: (2023)
Similar Items
-
HNOSeg-XS: Extremely Small Hartley Neural Operator for Efficient and Resolution-Robust 3D Image Segmentation
by: Wong, Ken C. L., et al.
Published: (2025) -
Comparative Analysis of Machine Learning Models for Lung Cancer Mutation Detection and Staging Using 3D CT Scans
by: Li, Yiheng, et al.
Published: (2025) -
Evaluating Automated Radiology Report Quality through Fine-Grained Phrasal Grounding of Clinical Findings
by: Mahmood, Razi, et al.
Published: (2024) -
MAM-CLIP: Vision-Language Pretraining on Mammography Atlases for BI-RADS Classification
by: Gulluk, Halil Ibrahim, et al.
Published: (2026) -
Improving Medical VQA through Trajectory-Aware Process Supervision
by: Gulluk, Halil Ibrahim, et al.
Published: (2026)