Multimodal Carotid Risk Stratification with Large Vision-Language Models: Benchmarking, Fine-Tuning, and Clinical Insights
Fuente:
arXiv
Saved in:
| Main Authors: | Tsolissou, Daphne, Ganitidis, Theofanis, Mitsis, Konstantinos, CHristodoulidis, Stergios, Vakalopoulou, Maria, Nikita, Konstantina |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sustaining model performance for covid-19 detection from dynamic audio data: Development and evaluation of a comprehensive drift-adaptive framework
by: Ganitidis, Theofanis, et al.
Published: (2024)
by: Ganitidis, Theofanis, et al.
Published: (2024)
On the Risk of Misleading Reports: Diagnosing Textual Biases in Multimodal Clinical AI
by: Restrepo, David, et al.
Published: (2025)
by: Restrepo, David, et al.
Published: (2025)
Medical Context Distorts Decisions in Clinical Vision Language Models
by: Restrepo, David, et al.
Published: (2026)
by: Restrepo, David, et al.
Published: (2026)
Full Conformal Adaptation of Medical Vision-Language Models
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
SoC: Semantic Orthogonal Calibration for Test-Time Prompt Tuning
by: Fillioux, Leo, et al.
Published: (2026)
by: Fillioux, Leo, et al.
Published: (2026)
SGPMIL: Sparse Gaussian Process Multiple Instance Learning
by: Lolos, Andreas, et al.
Published: (2025)
by: Lolos, Andreas, et al.
Published: (2025)
OCTOPUS: Enhancing the Spatial-Awareness of Vision SSMs with Multi-Dimensional Scans and Traversal Selection
by: Mahatha, Kunal, et al.
Published: (2026)
by: Mahatha, Kunal, et al.
Published: (2026)
Mask-HybridGNet: Graph-based segmentation with emergent anatomical correspondence from pixel-level supervision
by: Gaggion, Nicolás, et al.
Published: (2026)
by: Gaggion, Nicolás, et al.
Published: (2026)
Fairness and Robustness of CLIP-Based Models for Chest X-rays
by: Sourget, Théo, et al.
Published: (2025)
by: Sourget, Théo, et al.
Published: (2025)
CAPRMIL: Context-Aware Patch Representations for Multiple Instance Learning
by: Lolos, Andreas, et al.
Published: (2025)
by: Lolos, Andreas, et al.
Published: (2025)
BayesAdapter: enhanced uncertainty estimation in CLIP few-shot adaptation
by: Morales-Álvarez, Pablo, et al.
Published: (2024)
by: Morales-Álvarez, Pablo, et al.
Published: (2024)
PVeRA: Probabilistic Vector-Based Random Matrix Adaptation
by: Fillioux, Leo, et al.
Published: (2025)
by: Fillioux, Leo, et al.
Published: (2025)
FedVLMBench: Benchmarking Federated Fine-Tuning of Vision-Language Models
by: Zheng, Weiying, et al.
Published: (2025)
by: Zheng, Weiying, et al.
Published: (2025)
Exploring Real-Time Super-Resolution: Benchmarking and Fine-Tuning for Streaming Content
by: Bogatyrev, Evgeney, et al.
Published: (2026)
by: Bogatyrev, Evgeney, et al.
Published: (2026)
Controllable Latent Space Augmentation for Digital Pathology
by: Boutaj, Sofiène, et al.
Published: (2025)
by: Boutaj, Sofiène, et al.
Published: (2025)
ViG-Bias: Visually Grounded Bias Discovery and Mitigation
by: Marani, Badr-Eddine, et al.
Published: (2024)
by: Marani, Badr-Eddine, et al.
Published: (2024)
ViSurf: Visual Supervised-and-Reinforcement Fine-Tuning for Large Vision-and-Language Models
by: Liu, Yuqi, et al.
Published: (2025)
by: Liu, Yuqi, et al.
Published: (2025)
MERGETUNE: Continued Fine-Tuning of Vision-Language Models
by: Wang, Wenqing, et al.
Published: (2026)
by: Wang, Wenqing, et al.
Published: (2026)
Language Integration in Fine-Tuning Multimodal Large Language Models for Image-Based Regression
by: Jennings, Roy H., et al.
Published: (2025)
by: Jennings, Roy H., et al.
Published: (2025)
LMOD: A Large Multimodal Ophthalmology Dataset and Benchmark for Large Vision-Language Models
by: Qin, Zhenyue, et al.
Published: (2024)
by: Qin, Zhenyue, et al.
Published: (2024)
Temporal Gains, Spatial Costs: Revisiting Video Fine-Tuning in Multimodal Large Language Models
by: Zhang, Linghao, et al.
Published: (2026)
by: Zhang, Linghao, et al.
Published: (2026)
Multimodal Large Language Models as Image Classifiers
by: Kisel, Nikita, et al.
Published: (2026)
by: Kisel, Nikita, et al.
Published: (2026)
VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models
by: Ren, Yufan, et al.
Published: (2025)
by: Ren, Yufan, et al.
Published: (2025)
Attention-driven GUI Grounding: Leveraging Pretrained Multimodal Large Language Models without Fine-Tuning
by: Xu, Hai-Ming, et al.
Published: (2024)
by: Xu, Hai-Ming, et al.
Published: (2024)
Mamba-driven MRI-to-CT Synthesis for MRI-only Radiotherapy Planning
by: Barmpounakis, Konstantinos, et al.
Published: (2026)
by: Barmpounakis, Konstantinos, et al.
Published: (2026)
A modular framework for automated evaluation of procedural content generation in serious games with deep reinforcement learning agents
by: Kalafatis, Eleftherios, et al.
Published: (2025)
by: Kalafatis, Eleftherios, et al.
Published: (2025)
Learning Domain Knowledge in Multimodal Large Language Models through Reinforcement Fine-Tuning
by: Cao, Qinglong, et al.
Published: (2026)
by: Cao, Qinglong, et al.
Published: (2026)
Efficient Prompt Tuning of Large Vision-Language Model for Fine-Grained Ship Classification
by: Lan, Long, et al.
Published: (2024)
by: Lan, Long, et al.
Published: (2024)
Beyond Accuracy Optimization: Computer Vision Losses for Large Language Model Fine-Tuning
by: Cambrin, Daniele Rege, et al.
Published: (2024)
by: Cambrin, Daniele Rege, et al.
Published: (2024)
Benchmarking Large Vision-Language Models on Fine-Grained Image Tasks: A Comprehensive Evaluation
by: Yu, Hong-Tao, et al.
Published: (2025)
by: Yu, Hong-Tao, et al.
Published: (2025)
MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models
by: Ren, Xiyu, et al.
Published: (2026)
by: Ren, Xiyu, et al.
Published: (2026)
Hierarchy-Aware Fine-Tuning of Vision-Language Models
by: Li, Jiayu, et al.
Published: (2025)
by: Li, Jiayu, et al.
Published: (2025)
Remodeling Semantic Relationships in Vision-Language Fine-Tuning
by: Wu, Xiangyang, et al.
Published: (2025)
by: Wu, Xiangyang, et al.
Published: (2025)
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models
by: Sun, Haoyuan, et al.
Published: (2025)
by: Sun, Haoyuan, et al.
Published: (2025)
Fine-Tuning a Large Vision-Language Model for Artwork's Scoring and Critique
by: Zhang, Zhehan, et al.
Published: (2026)
by: Zhang, Zhehan, et al.
Published: (2026)
PPGL-Swarm: Integrated Multimodal Risk Stratification and Hereditary Syndrome Detection in Pheochromocytoma and Paraganglioma
by: Liu, Zelin, et al.
Published: (2026)
by: Liu, Zelin, et al.
Published: (2026)
Parameter-Efficient Fine-Tuning Medical Multimodal Large Language Models for Medical Visual Grounding
by: He, Jinlong, et al.
Published: (2024)
by: He, Jinlong, et al.
Published: (2024)
Class Adaptive Conformal Training
by: Marani, Badr-Eddine, et al.
Published: (2026)
by: Marani, Badr-Eddine, et al.
Published: (2026)
Are foundation models for computer vision good conformal predictors?
by: Fillioux, Leo, et al.
Published: (2024)
by: Fillioux, Leo, et al.
Published: (2024)
To See is Not to Learn: Protecting Multimodal Data from Unauthorized Fine-Tuning of Large Vision-Language Model
by: Zhao, Chengshuai, et al.
Published: (2026)
by: Zhao, Chengshuai, et al.
Published: (2026)
Similar Items
-
Sustaining model performance for covid-19 detection from dynamic audio data: Development and evaluation of a comprehensive drift-adaptive framework
by: Ganitidis, Theofanis, et al.
Published: (2024) -
On the Risk of Misleading Reports: Diagnosing Textual Biases in Multimodal Clinical AI
by: Restrepo, David, et al.
Published: (2025) -
Medical Context Distorts Decisions in Clinical Vision Language Models
by: Restrepo, David, et al.
Published: (2026) -
Full Conformal Adaptation of Medical Vision-Language Models
by: Silva-Rodríguez, Julio, et al.
Published: (2025) -
SoC: Semantic Orthogonal Calibration for Test-Time Prompt Tuning
by: Fillioux, Leo, et al.
Published: (2026)