Are Vision Language Models Ready for Clinical Diagnosis? A 3D Medical Benchmark for Tumor-centric Visual Question Answering
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Yixiong, Xiao, Wenjie, Bassi, Pedro R. A. S., Zhou, Xinze, Er, Sezgin, Hamamci, Ibrahim Ethem, Zhou, Zongwei, Yuille, Alan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents
by: Chen, Yixiong, et al.
Published: (2026)
by: Chen, Yixiong, et al.
Published: (2026)
CT2Rep: Automated Radiology Report Generation for 3D Medical Imaging
by: Hamamci, Ibrahim Ethem, et al.
Published: (2024)
by: Hamamci, Ibrahim Ethem, et al.
Published: (2024)
Quality Sentinel: Estimating Label Quality and Errors in Medical Segmentation Datasets
by: Chen, Yixiong, et al.
Published: (2024)
by: Chen, Yixiong, et al.
Published: (2024)
Large-Scale Label Quality Assessment for Medical Segmentation via a Vision-Language Judge and Synthetic Data
by: Chen, Yixiong, et al.
Published: (2026)
by: Chen, Yixiong, et al.
Published: (2026)
CRG Score: A Distribution-Aware Clinical Metric for Radiology Report Generation
by: Hamamci, Ibrahim Ethem, et al.
Published: (2025)
by: Hamamci, Ibrahim Ethem, et al.
Published: (2025)
RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology
by: Li, Wenxuan, et al.
Published: (2026)
by: Li, Wenxuan, et al.
Published: (2026)
Meissa: Multi-modal Medical Agentic Intelligence
by: Chen, Yixiong, et al.
Published: (2026)
by: Chen, Yixiong, et al.
Published: (2026)
Better Tokens for Better 3D: Advancing Vision-Language Modeling in 3D Medical Imaging
by: Hamamci, Ibrahim Ethem, et al.
Published: (2025)
by: Hamamci, Ibrahim Ethem, et al.
Published: (2025)
SigVLP: Sigmoid Volume-Language Pre-Training for Self-Supervised CT-Volume Adaptive Representation Learning
by: Wang, Jiayi, et al.
Published: (2026)
by: Wang, Jiayi, et al.
Published: (2026)
How Well Do Supervised 3D Models Transfer to Medical Imaging Tasks?
by: Li, Wenxuan, et al.
Published: (2025)
by: Li, Wenxuan, et al.
Published: (2025)
Embracing Massive Medical Data
by: Chou, Yu-Cheng, et al.
Published: (2024)
by: Chou, Yu-Cheng, et al.
Published: (2024)
Beyond Masks: The Case for Medical Image Parsing
by: Gupta, Siddharth, et al.
Published: (2026)
by: Gupta, Siddharth, et al.
Published: (2026)
See More, Change Less: Anatomy-Aware Diffusion for Contrast Enhancement
by: Liu, Junqi, et al.
Published: (2025)
by: Liu, Junqi, et al.
Published: (2025)
RadDiagSeg-M: A Vision Language Model for Joint Diagnosis and Multi-Target Segmentation in Radiology
by: Li, Chengrun, et al.
Published: (2025)
by: Li, Chengrun, et al.
Published: (2025)
Auditing Significance, Metric Choice, and Demographic Fairness in Medical AI Challenges
by: Lubonja, Ariel, et al.
Published: (2025)
by: Lubonja, Ariel, et al.
Published: (2025)
Analyzing Tumors by Synthesis
by: Chen, Qi, et al.
Published: (2024)
by: Chen, Qi, et al.
Published: (2024)
Acquiring Weak Annotations for Tumor Localization in Temporal and Volumetric Data
by: Chou, Yu-Cheng, et al.
Published: (2023)
by: Chou, Yu-Cheng, et al.
Published: (2023)
Hallucination Benchmark in Medical Visual Question Answering
by: Wu, Jinge, et al.
Published: (2024)
by: Wu, Jinge, et al.
Published: (2024)
Scaling Artificial Intelligence for Multi-Tumor Early Detection with More Reports, Fewer Masks
by: Bassi, Pedro R. A. S., et al.
Published: (2025)
by: Bassi, Pedro R. A. S., et al.
Published: (2025)
Label Critic: Design Data Before Models
by: Bassi, Pedro R. A. S., et al.
Published: (2024)
by: Bassi, Pedro R. A. S., et al.
Published: (2024)
Towards Generalizable Tumor Synthesis
by: Chen, Qi, et al.
Published: (2024)
by: Chen, Qi, et al.
Published: (2024)
RadGPT: Constructing 3D Image-Text Tumor Datasets
by: Bassi, Pedro R. A. S., et al.
Published: (2025)
by: Bassi, Pedro R. A. S., et al.
Published: (2025)
Object-centric Video Question Answering with Visual Grounding and Referring
by: Wang, Haochen, et al.
Published: (2025)
by: Wang, Haochen, et al.
Published: (2025)
The Ouroboros of Benchmarking: Reasoning Evaluation in an Era of Saturation
by: Deveci, İbrahim Ethem, et al.
Published: (2025)
by: Deveci, İbrahim Ethem, et al.
Published: (2025)
PanTS: The Pancreatic Tumor Segmentation Dataset
by: Li, Wenxuan, et al.
Published: (2025)
by: Li, Wenxuan, et al.
Published: (2025)
From Pixel to Cancer: Cellular Automata in Computed Tomography
by: Lai, Yuxiang, et al.
Published: (2024)
by: Lai, Yuxiang, et al.
Published: (2024)
Leveraging AI Predicted and Expert Revised Annotations in Interactive Segmentation: Continual Tuning or Full Training?
by: Zhang, Tiezheng, et al.
Published: (2024)
by: Zhang, Tiezheng, et al.
Published: (2024)
Structure-Aware Sparse-View X-ray 3D Reconstruction
by: Cai, Yuanhao, et al.
Published: (2023)
by: Cai, Yuanhao, et al.
Published: (2023)
Hierarchical Modeling for Medical Visual Question Answering with Cross-Attention Fusion
by: Zhang, Junkai, et al.
Published: (2025)
by: Zhang, Junkai, et al.
Published: (2025)
Benchmarking Real-Time Question Answering via Executable Code Workflows
by: Zhou, Wenjie, et al.
Published: (2026)
by: Zhou, Wenjie, et al.
Published: (2026)
Medical World Model: Generative Simulation of Tumor Evolution for Treatment Planning
by: Yang, Yijun, et al.
Published: (2025)
by: Yang, Yijun, et al.
Published: (2025)
Text-Driven Tumor Synthesis
by: Li, Xinran, et al.
Published: (2024)
by: Li, Xinran, et al.
Published: (2024)
Scaling Tumor Segmentation: Best Lessons from Real and Synthetic Data
by: Chen, Qi, et al.
Published: (2025)
by: Chen, Qi, et al.
Published: (2025)
Object Retrieval for Visual Question Answering with Outside Knowledge
by: Kan, Shichao, et al.
Published: (2024)
by: Kan, Shichao, et al.
Published: (2024)
Letter to Editor: Comments on Impact of Peri‐Procedural Antibiotics on Post‐ERCP Infectious Adverse Events With Distal Malignant Biliary Obstruction
by: İbrahim Ethem Güven
Published: (2026)
by: İbrahim Ethem Güven
Published: (2026)
How Good LLMs Are at Answering Bangla Medical Visual Questions? Dataset and Benchmarking
by: Ahmed, Rafid, et al.
Published: (2026)
by: Ahmed, Rafid, et al.
Published: (2026)
IMB: An Italian Medical Benchmark for Question Answering
by: Romano, Antonio, et al.
Published: (2025)
by: Romano, Antonio, et al.
Published: (2025)
Fusion of Domain-Adapted Vision and Language Models for Medical Visual Question Answering
by: Ha, Cuong Nhat, et al.
Published: (2024)
by: Ha, Cuong Nhat, et al.
Published: (2024)
Targeted Visual Prompting for Medical Visual Question Answering
by: Tascon-Morales, Sergio, et al.
Published: (2024)
by: Tascon-Morales, Sergio, et al.
Published: (2024)
Medical Vision Generalist: Unifying Medical Imaging Tasks in Context
by: Ren, Sucheng, et al.
Published: (2024)
by: Ren, Sucheng, et al.
Published: (2024)
Similar Items
-
DeepTumorVQA: A Hierarchical 3D CT Benchmark for Stage-Wise Evaluation of Medical VLMs and Tool-Augmented Agents
by: Chen, Yixiong, et al.
Published: (2026) -
CT2Rep: Automated Radiology Report Generation for 3D Medical Imaging
by: Hamamci, Ibrahim Ethem, et al.
Published: (2024) -
Quality Sentinel: Estimating Label Quality and Errors in Medical Segmentation Datasets
by: Chen, Yixiong, et al.
Published: (2024) -
Large-Scale Label Quality Assessment for Medical Segmentation via a Vision-Language Judge and Synthetic Data
by: Chen, Yixiong, et al.
Published: (2026) -
CRG Score: A Distribution-Aware Clinical Metric for Radiology Report Generation
by: Hamamci, Ibrahim Ethem, et al.
Published: (2025)