Simulating Clinical AI Assistance using Multimodal LLMs: A Case Study in Diabetic Retinopathy
Fuente:
arXiv
Saved in:
| Main Authors: | Barakat, Nadim, Lotter, William |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vision-Integrated LLMs for Autonomous Driving Assistance : Human Performance Comparison and Trust Evaluation
by: Kim, Namhee, et al.
Published: (2025)
by: Kim, Namhee, et al.
Published: (2025)
Joining Forces for Pathology Diagnostics with AI Assistance: The EMPAIA Initiative
by: Zerbe, Norman, et al.
Published: (2023)
by: Zerbe, Norman, et al.
Published: (2023)
Scene-Aware Conversational ADAS with Generative AI for Real-Time Driver Assistance
by: Han, Kyungtae, et al.
Published: (2025)
by: Han, Kyungtae, et al.
Published: (2025)
GazeLLM: Multimodal LLMs incorporating Human Visual Attention
by: Rekimoto, Jun
Published: (2025)
by: Rekimoto, Jun
Published: (2025)
Towards a Multimodal Document-grounded Conversational AI System for Education
by: Taneja, Karan, et al.
Published: (2025)
by: Taneja, Karan, et al.
Published: (2025)
Lightweight Structured Multimodal Reasoning for Clinical Scene Understanding in Robotics
by: Jha, Saurav, et al.
Published: (2025)
by: Jha, Saurav, et al.
Published: (2025)
Pencils to Pixels: A Systematic Study of Creative Drawings across Children, Adults and AI
by: Nath, Surabhi S, et al.
Published: (2025)
by: Nath, Surabhi S, et al.
Published: (2025)
Zero-Shot Pupil Segmentation with SAM 2: A Case Study of Over 14 Million Images
by: Maquiling, Virmarie, et al.
Published: (2024)
by: Maquiling, Virmarie, et al.
Published: (2024)
CG-MER: A Card Game-based Multimodal dataset for Emotion Recognition
by: Farhat, Nessrine, et al.
Published: (2025)
by: Farhat, Nessrine, et al.
Published: (2025)
Generative AI for Cel-Animation: A Survey
by: Tang, Yolo Y., et al.
Published: (2025)
by: Tang, Yolo Y., et al.
Published: (2025)
Automated Visual Attention Detection using Mobile Eye Tracking in Behavioral Classroom Studies
by: Bozkir, Efe, et al.
Published: (2025)
by: Bozkir, Efe, et al.
Published: (2025)
See-Control: A Multimodal Agent Framework for Smartphone Interaction with a Robotic Arm
by: Zhao, Haoyu, et al.
Published: (2025)
by: Zhao, Haoyu, et al.
Published: (2025)
EEG-based Multimodal Representation Learning for Emotion Recognition
by: Yin, Kang, et al.
Published: (2024)
by: Yin, Kang, et al.
Published: (2024)
AutoTour: Automatic Photo Tour Guide with Smartphones and LLMs
by: Xu, Huatao, et al.
Published: (2026)
by: Xu, Huatao, et al.
Published: (2026)
AI-based Multimodal Biometrics for Detecting Smartphone Distractions: Application to Online Learning
by: Becerra, Alvaro, et al.
Published: (2025)
by: Becerra, Alvaro, et al.
Published: (2025)
OmniResponse: Online Multimodal Conversational Response Generation in Dyadic Interactions
by: Luo, Cheng, et al.
Published: (2025)
by: Luo, Cheng, et al.
Published: (2025)
GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents
by: Ouyang, Mingyu, et al.
Published: (2026)
by: Ouyang, Mingyu, et al.
Published: (2026)
ViSIL: Unified Evaluation of Information Loss in Multimodal Video Captioning
by: Li, Po-han, et al.
Published: (2026)
by: Li, Po-han, et al.
Published: (2026)
Enhancing Online Learning by Integrating Biosensors and Multimodal Learning Analytics for Detecting and Predicting Student Behavior: A Review
by: Becerra, Alvaro, et al.
Published: (2025)
by: Becerra, Alvaro, et al.
Published: (2025)
Milmer: a Framework for Multiple Instance Learning based Multimodal Emotion Recognition
by: Wang, Zaitian, et al.
Published: (2025)
by: Wang, Zaitian, et al.
Published: (2025)
A User-Centric Analysis of Explainability in AI-Based Medical Image Diagnosis
by: Wagner, Julia, et al.
Published: (2026)
by: Wagner, Julia, et al.
Published: (2026)
Not There Yet: Evaluating Vision Language Models in Simulating the Visual Perception of People with Low Vision
by: Natalie, Rosiana, et al.
Published: (2025)
by: Natalie, Rosiana, et al.
Published: (2025)
Visual Evaluative AI: A Hypothesis-Driven Tool with Concept-Based Explanations and Weight of Evidence
by: Le, Thao, et al.
Published: (2024)
by: Le, Thao, et al.
Published: (2024)
Automated ARAT Scoring Using Multimodal Video Analysis, Multi-View Fusion, and Hierarchical Bayesian Models: A Clinician Study
by: Ahmed, Tamim, et al.
Published: (2025)
by: Ahmed, Tamim, et al.
Published: (2025)
AltCanvas: A Tile-Based Image Editor with Generative AI for Blind or Visually Impaired People
by: Lee, Seonghee, et al.
Published: (2024)
by: Lee, Seonghee, et al.
Published: (2024)
Enabling Collaborative Clinical Diagnosis of Infectious Keratitis by Integrating Expert Knowledge and Interpretable Data-driven Intelligence
by: Fang, Zhengqing, et al.
Published: (2024)
by: Fang, Zhengqing, et al.
Published: (2024)
How to Distinguish AI-Generated Images from Authentic Photographs
by: Kamali, Negar, et al.
Published: (2024)
by: Kamali, Negar, et al.
Published: (2024)
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
by: Yang, Boyin, et al.
Published: (2025)
by: Yang, Boyin, et al.
Published: (2025)
ASAP: Interpretable Analysis and Summarization of AI-generated Image Patterns at Scale
by: Huang, Jinbin, et al.
Published: (2024)
by: Huang, Jinbin, et al.
Published: (2024)
How good are humans at detecting AI-generated images? Learnings from an experiment
by: Roca, Thomas, et al.
Published: (2025)
by: Roca, Thomas, et al.
Published: (2025)
TraitSpaces: Towards Interpretable Visual Creativity for Human-AI Co-Creation
by: Luthra, Prerna
Published: (2025)
by: Luthra, Prerna
Published: (2025)
Sketch2Prototype: Rapid Conceptual Design Exploration and Prototyping with Generative AI
by: Edwards, Kristen M., et al.
Published: (2024)
by: Edwards, Kristen M., et al.
Published: (2024)
Towards user-centered interactive medical image segmentation in VR with an assistive AI agent
by: Spiegler, Pascal, et al.
Published: (2025)
by: Spiegler, Pascal, et al.
Published: (2025)
Dermatologist-like explainable AI enhances melanoma diagnosis accuracy: eye-tracking study
by: Chanda, Tirtha, et al.
Published: (2024)
by: Chanda, Tirtha, et al.
Published: (2024)
PedaCo-Gen: Scaffolding Pedagogical Agency in Human-AI Collaborative Video Authoring
by: Baek, Injun, et al.
Published: (2026)
by: Baek, Injun, et al.
Published: (2026)
Toward a Universal Color Naming System: A Clustering-Based Approach using Multisource Data
by: Sabitkyzy, Aruzhan, et al.
Published: (2026)
by: Sabitkyzy, Aruzhan, et al.
Published: (2026)
LEyes: A Lightweight Framework for Deep Learning-Based Eye Tracking using Synthetic Eye Images
by: Byrne, Sean Anthony, et al.
Published: (2023)
by: Byrne, Sean Anthony, et al.
Published: (2023)
GarmentLab: A Unified Simulation and Benchmark for Garment Manipulation
by: Lu, Haoran, et al.
Published: (2024)
by: Lu, Haoran, et al.
Published: (2024)
AI-Based Facial Emotion Recognition Solutions for Education: A Study of Teacher-User and Other Categories
by: Ravenor, R. Yamamoto
Published: (2023)
by: Ravenor, R. Yamamoto
Published: (2023)
Magma: A Foundation Model for Multimodal AI Agents
by: Yang, Jianwei, et al.
Published: (2025)
by: Yang, Jianwei, et al.
Published: (2025)
Similar Items
-
Vision-Integrated LLMs for Autonomous Driving Assistance : Human Performance Comparison and Trust Evaluation
by: Kim, Namhee, et al.
Published: (2025) -
Joining Forces for Pathology Diagnostics with AI Assistance: The EMPAIA Initiative
by: Zerbe, Norman, et al.
Published: (2023) -
Scene-Aware Conversational ADAS with Generative AI for Real-Time Driver Assistance
by: Han, Kyungtae, et al.
Published: (2025) -
GazeLLM: Multimodal LLMs incorporating Human Visual Attention
by: Rekimoto, Jun
Published: (2025) -
Towards a Multimodal Document-grounded Conversational AI System for Education
by: Taneja, Karan, et al.
Published: (2025)