Is ChatGPT-5 Ready for Mammogram VQA?
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Qiang, Wang, Shansong, Hu, Mingzhe, Safari, Mojtaba, Eidex, Zachary, Yang, Xiaofeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Performance of GPT-5 in Brain Tumor MRI Reasoning
by: Safari, Mojtaba, et al.
Published: (2025)
by: Safari, Mojtaba, et al.
Published: (2025)
Evaluating GPT-5 as a Multimodal Clinical Reasoner: A Landscape Commentary
by: Florea, Alexandru, et al.
Published: (2026)
by: Florea, Alexandru, et al.
Published: (2026)
Benchmarking GPT-5 for Zero-Shot Multimodal Medical Reasoning in Radiology and Radiation Oncology
by: Hu, Mingzhe, et al.
Published: (2025)
by: Hu, Mingzhe, et al.
Published: (2025)
DINOv3 with Test-Time Training for Medical Image Registration
by: Wang, Shansong, et al.
Published: (2025)
by: Wang, Shansong, et al.
Published: (2025)
Capabilities of GPT-5 on Multimodal Medical Reasoning
by: Wang, Shansong, et al.
Published: (2025)
by: Wang, Shansong, et al.
Published: (2025)
Foundation Models in Medical Image Analysis: A Systematic Review and Meta-Analysis
by: Rajendran, Praveenbalaji, et al.
Published: (2025)
by: Rajendran, Praveenbalaji, et al.
Published: (2025)
BrainDINO: A Brain MRI Foundation Model for Generalizable Clinical Representation Learning
by: Wu, Yizhou, et al.
Published: (2026)
by: Wu, Yizhou, et al.
Published: (2026)
Triad: Vision Foundation Model for 3D Magnetic Resonance Imaging
by: Wang, Shansong, et al.
Published: (2025)
by: Wang, Shansong, et al.
Published: (2025)
Unifying Biomedical Vision-Language Expertise: Towards a Generalist Foundation Model via Multi-CLIP Knowledge Distillation
by: Wang, Shansong, et al.
Published: (2025)
by: Wang, Shansong, et al.
Published: (2025)
Advancing MRI Reconstruction: A Systematic Review of Deep Learning and Compressed Sensing Integration
by: Safari, Mojtaba, et al.
Published: (2025)
by: Safari, Mojtaba, et al.
Published: (2025)
MRI super-resolution reconstruction using efficient diffusion probabilistic model with residual shifting
by: Safari, Mojtaba, et al.
Published: (2025)
by: Safari, Mojtaba, et al.
Published: (2025)
IQAGPT: Image Quality Assessment with Vision-language and ChatGPT Models
by: Chen, Zhihao, et al.
Published: (2023)
by: Chen, Zhihao, et al.
Published: (2023)
GPTDrawer: Enhancing Visual Synthesis through ChatGPT
by: Li, Kun, et al.
Published: (2024)
by: Li, Kun, et al.
Published: (2024)
A Physics-Informed Deep Learning Model for MRI Brain Motion Correction
by: Safari, Mojtaba, et al.
Published: (2025)
by: Safari, Mojtaba, et al.
Published: (2025)
Bidirectional Mammogram View Translation with Column-Aware and Implicit 3D Conditional Diffusion
by: Li, Xin, et al.
Published: (2025)
by: Li, Xin, et al.
Published: (2025)
T1-contrast Enhanced MRI Generation from Multi-parametric MRI for Glioma Patients with Latent Tumor Conditioning
by: Eidex, Zach, et al.
Published: (2024)
by: Eidex, Zach, et al.
Published: (2024)
Res-MoCoDiff: Residual-guided diffusion models for motion artifact correction in brain MRI
by: Safari, Mojtaba, et al.
Published: (2025)
by: Safari, Mojtaba, et al.
Published: (2025)
PhD: A ChatGPT-Prompted Visual hallucination Evaluation Dataset
by: Liu, Jiazhen, et al.
Published: (2024)
by: Liu, Jiazhen, et al.
Published: (2024)
Demystifying the Potential of ChatGPT-4 Vision for Construction Progress Monitoring
by: Ersoz, Ahmet Bahaddin
Published: (2024)
by: Ersoz, Ahmet Bahaddin
Published: (2024)
ChatGPT and biometrics: an assessment of face recognition, gender detection, and age estimation capabilities
by: Hassanpour, Ahmad, et al.
Published: (2024)
by: Hassanpour, Ahmad, et al.
Published: (2024)
Intelligent Director: An Automatic Framework for Dynamic Visual Composition using ChatGPT
by: Zheng, Sixiao, et al.
Published: (2024)
by: Zheng, Sixiao, et al.
Published: (2024)
Visual Reasoning Evaluation of Grok, Deepseek Janus, Gemini, Qwen, Mistral, and ChatGPT
by: Jegham, Nidhal, et al.
Published: (2025)
by: Jegham, Nidhal, et al.
Published: (2025)
Evaluating ChatGPT's Performance in Classifying Pneumonia from Chest X-Ray Images
by: Prahallad, Pragna, et al.
Published: (2025)
by: Prahallad, Pragna, et al.
Published: (2025)
Explainable Cross-Disease Reasoning for Cardiovascular Risk Assessment from Low-Dose Computed Tomography
by: Zhang, Yifei, et al.
Published: (2025)
by: Zhang, Yifei, et al.
Published: (2025)
MedLVR: Latent Visual Reasoning for Reliable Medical Visual Question Answering
by: Xi, Suyang, et al.
Published: (2026)
by: Xi, Suyang, et al.
Published: (2026)
Prompt fidelity of ChatGPT4o / Dall-E3 text-to-image visualisations
by: Spennemann, Dirk HR
Published: (2025)
by: Spennemann, Dirk HR
Published: (2025)
AI-Generated Content Enhanced Computer-Aided Diagnosis Model for Thyroid Nodules: A ChatGPT-Style Assistant
by: Yao, Jincao, et al.
Published: (2024)
by: Yao, Jincao, et al.
Published: (2024)
Joint Holistic and Lesion Controllable Mammogram Synthesis via Gated Conditional Diffusion Model
by: Li, Xin, et al.
Published: (2025)
by: Li, Xin, et al.
Published: (2025)
An Evaluation of GPT-4V and Gemini in Online VQA
by: Liu, Mengchen, et al.
Published: (2023)
by: Liu, Mengchen, et al.
Published: (2023)
How Good is ChatGPT at Audiovisual Deepfake Detection: A Comparative Study of ChatGPT, AI Models and Human Perception
by: Shahzad, Sahibzada Adil, et al.
Published: (2024)
by: Shahzad, Sahibzada Adil, et al.
Published: (2024)
Can ChatGPT Perform Image Splicing Detection? A Preliminary Study
by: Nath, Souradip
Published: (2025)
by: Nath, Souradip
Published: (2025)
Panoptic Segmentation of Mammograms with Text-To-Image Diffusion Model
by: Zhao, Kun, et al.
Published: (2024)
by: Zhao, Kun, et al.
Published: (2024)
MechVQA: Benchmarking and Enhancing Multimodal LLMs on Comprehensive Mechanical Drawing Understanding
by: Kou, Qian, et al.
Published: (2026)
by: Kou, Qian, et al.
Published: (2026)
Assessing Greenspace Attractiveness with ChatGPT, Claude, and Gemini: Do AI Models Reflect Human Perceptions?
by: Malekzadeh, Milad, et al.
Published: (2025)
by: Malekzadeh, Milad, et al.
Published: (2025)
Leveraging ChatGPT's Multimodal Vision Capabilities to Rank Satellite Images by Poverty Level: Advancing Tools for Social Science Research
by: Sarmadi, Hamid, et al.
Published: (2025)
by: Sarmadi, Hamid, et al.
Published: (2025)
GRIP-VLM: Group-Relative Importance Pruning for Efficient Vision-Language Models
by: Huang, Mingzhe, et al.
Published: (2026)
by: Huang, Mingzhe, et al.
Published: (2026)
MV-Swin-T: Mammogram Classification with Multi-view Swin Transformer
by: Sarker, Sushmita, et al.
Published: (2024)
by: Sarker, Sushmita, et al.
Published: (2024)
Can ChatGPT Learn My Life From a Week of First-Person Video?
by: Harris, Keegan
Published: (2025)
by: Harris, Keegan
Published: (2025)
Self-Supervised Adversarial Diffusion Models for Fast MRI Reconstruction
by: Safari, Mojtaba, et al.
Published: (2024)
by: Safari, Mojtaba, et al.
Published: (2024)
Generalizable 7T T1-map Synthesis from 1.5T and 3T T1 MRI with an Efficient Transformer Model
by: Eidex, Zach, et al.
Published: (2025)
by: Eidex, Zach, et al.
Published: (2025)
Similar Items
-
Performance of GPT-5 in Brain Tumor MRI Reasoning
by: Safari, Mojtaba, et al.
Published: (2025) -
Evaluating GPT-5 as a Multimodal Clinical Reasoner: A Landscape Commentary
by: Florea, Alexandru, et al.
Published: (2026) -
Benchmarking GPT-5 for Zero-Shot Multimodal Medical Reasoning in Radiology and Radiation Oncology
by: Hu, Mingzhe, et al.
Published: (2025) -
DINOv3 with Test-Time Training for Medical Image Registration
by: Wang, Shansong, et al.
Published: (2025) -
Capabilities of GPT-5 on Multimodal Medical Reasoning
by: Wang, Shansong, et al.
Published: (2025)