Insight: A Multi-Modal Diagnostic Pipeline using LLMs for Ocular Surface Disease Diagnosis
Fuente:
arXiv
Saved in:
| Main Authors: | Yeh, Chun-Hsiao, Wang, Jiayun, Graham, Andrew D., Liu, Andrea J., Tan, Bo, Chen, Yubei, Ma, Yi, Lin, Meng C. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Gen4Gen: Generative Data Pipeline for Generative Multi-Concept Composition
by: Yeh, Chun-Hsiao, et al.
Published: (2024)
by: Yeh, Chun-Hsiao, et al.
Published: (2024)
Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs
by: Yeh, Chun-Hsiao, et al.
Published: (2025)
by: Yeh, Chun-Hsiao, et al.
Published: (2025)
Pose-Aware Self-Supervised Learning with Viewpoint Trajectory Regularization
by: Wang, Jiayun, et al.
Published: (2024)
by: Wang, Jiayun, et al.
Published: (2024)
Ocular Authentication: Fusion of Gaze and Periocular Modalities
by: Lohr, Dillon, et al.
Published: (2025)
by: Lohr, Dillon, et al.
Published: (2025)
Incomplete Modality Disentangled Representation for Ophthalmic Disease Grading and Diagnosis
by: Liu, Chengzhi, et al.
Published: (2025)
by: Liu, Chengzhi, et al.
Published: (2025)
Training an LLM-as-a-Judge Model: Pipeline, Insights, and Practical Lessons
by: Hu, Renjun, et al.
Published: (2025)
by: Hu, Renjun, et al.
Published: (2025)
MoRA: LoRA Guided Multi-Modal Disease Diagnosis with Missing Modality
by: Shi, Zhiyi, et al.
Published: (2024)
by: Shi, Zhiyi, et al.
Published: (2024)
Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning
by: Yeh, Chun-Hsiao, et al.
Published: (2026)
by: Yeh, Chun-Hsiao, et al.
Published: (2026)
A Progressive Single-Modality to Multi-Modality Classification Framework for Alzheimer's Disease Sub-type Diagnosis
by: Liu, Yuxiao, et al.
Published: (2024)
by: Liu, Yuxiao, et al.
Published: (2024)
TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models
by: Hsiao, Teng-Fang, et al.
Published: (2025)
by: Hsiao, Teng-Fang, et al.
Published: (2025)
Enhancing Meme Emotion Understanding with Multi-Level Modality Enhancement and Dual-Stage Modal Fusion
by: Shi, Yi, et al.
Published: (2025)
by: Shi, Yi, et al.
Published: (2025)
Beyond Simple Edits: X-Planner for Complex Instruction-Based Image Editing
by: Yeh, Chun-Hsiao, et al.
Published: (2025)
by: Yeh, Chun-Hsiao, et al.
Published: (2025)
Unveiling the Relationship Between Undergraduate Students' Emergent Roles and Learning Performance in Collaborative Argumentation‐Based Learning: Insights From Sequence Clustering and Entropy Analysis
by: Qingtang Liu, et al.
Published: (2025)
by: Qingtang Liu, et al.
Published: (2025)
Deep Learning-Based Prediction of Suspension Dynamics Performance in Multi-Axle Vehicles
by: Lin, Kai Chun, et al.
Published: (2024)
by: Lin, Kai Chun, et al.
Published: (2024)
Advancing Histopathology-Based Breast Cancer Diagnosis: Insights into Multi-Modality and Explainability
by: Abdullakutty, Faseela, et al.
Published: (2024)
by: Abdullakutty, Faseela, et al.
Published: (2024)
DialogCC: An Automated Pipeline for Creating High-Quality Multi-Modal Dialogue Dataset
by: Lee, Young-Jun, et al.
Published: (2022)
by: Lee, Young-Jun, et al.
Published: (2022)
DOSE: Data Selection for Multi-Modal LLMs via Off-the-Shelf Models
by: Wu, Biao, et al.
Published: (2026)
by: Wu, Biao, et al.
Published: (2026)
Robust Incomplete-Modality Alignment for Ophthalmic Disease Grading and Diagnosis via Labeled Optimal Transport
by: Yu, Qinkai, et al.
Published: (2025)
by: Yu, Qinkai, et al.
Published: (2025)
MegaRAG: Multimodal Knowledge Graph-Based Retrieval Augmented Generation
by: Hsiao, Chi-Hsiang, et al.
Published: (2025)
by: Hsiao, Chi-Hsiang, et al.
Published: (2025)
Copiloting Diagnosis of Autism in Real Clinical Scenarios via LLMs
by: Jiang, Yi, et al.
Published: (2024)
by: Jiang, Yi, et al.
Published: (2024)
Prompt Highlighter: Interactive Control for Multi-Modal LLMs
by: Zhang, Yuechen, et al.
Published: (2023)
by: Zhang, Yuechen, et al.
Published: (2023)
WAFFLE: Finetuning Multi-Modal Models for Automated Front-End Development
by: Liang, Shanchao, et al.
Published: (2024)
by: Liang, Shanchao, et al.
Published: (2024)
Ocular-Induced Abnormal Head Posture: Diagnosis and Missing Data Imputation
by: Al-Dabet, Saja, et al.
Published: (2025)
by: Al-Dabet, Saja, et al.
Published: (2025)
Textureless Deformable Surface Reconstruction with Invisible Markers
by: Li, Xinyuan, et al.
Published: (2023)
by: Li, Xinyuan, et al.
Published: (2023)
Learning Contrastive Multimodal Fusion with Improved Modality Dropout for Disease Detection and Prediction
by: Gu, Yi, et al.
Published: (2025)
by: Gu, Yi, et al.
Published: (2025)
EMMA: Efficient Visual Alignment in Multi-Modal LLMs
by: Ghazanfari, Sara, et al.
Published: (2024)
by: Ghazanfari, Sara, et al.
Published: (2024)
Pathology Context Recalibration Network for Ocular Disease Recognition
by: Xiao, Zunjie, et al.
Published: (2025)
by: Xiao, Zunjie, et al.
Published: (2025)
LLMs as Workers in Human-Computational Algorithms? Replicating Crowdsourcing Pipelines with LLMs
by: Wu, Tongshuang, et al.
Published: (2023)
by: Wu, Tongshuang, et al.
Published: (2023)
EyeAI: AI-Assisted Ocular Disease Detection for Equitable Healthcare Access
by: Garg, Shiv, et al.
Published: (2025)
by: Garg, Shiv, et al.
Published: (2025)
OrthoInsight: Rib Fracture Diagnosis and Report Generation Based on Multi-Modal Large Models
by: Wu, Ningyong, et al.
Published: (2025)
by: Wu, Ningyong, et al.
Published: (2025)
TransFair: Transferring Fairness from Ocular Disease Classification to Progression Prediction
by: Gheisi, Leila, et al.
Published: (2024)
by: Gheisi, Leila, et al.
Published: (2024)
MRI to PET Cross-Modality Translation using Globally and Locally Aware GAN (GLA-GAN) for Multi-Modal Diagnosis of Alzheimer's Disease
by: Sikka, Apoorva, et al.
Published: (2021)
by: Sikka, Apoorva, et al.
Published: (2021)
S3-CoT: Self-Sampled Succinct Reasoning Enables Efficient Chain-of-Thought LLMs
by: Du, Yanrui, et al.
Published: (2026)
by: Du, Yanrui, et al.
Published: (2026)
Calibrating Uncertainty Quantification of Multi-Modal LLMs using Grounding
by: Padhi, Trilok, et al.
Published: (2025)
by: Padhi, Trilok, et al.
Published: (2025)
Reliable Multi-Modal Object Re-Identification via Modality-Aware Graph Reasoning
by: Wan, Xixi, et al.
Published: (2025)
by: Wan, Xixi, et al.
Published: (2025)
Narrative-to-Scene Generation: An LLM-Driven Pipeline for 2D Game Environments
by: Chen, Yi-Chun, et al.
Published: (2025)
by: Chen, Yi-Chun, et al.
Published: (2025)
Toward Fairness Through Fair Multi-Exit Framework for Dermatological Disease Diagnosis
by: Chiu, Ching-Hao, et al.
Published: (2023)
by: Chiu, Ching-Hao, et al.
Published: (2023)
DCMM-SQL: Automated Data-Centric Pipeline and Multi-Model Collaboration Training for Text-to-SQL Model
by: Xie, Yuanzhen, et al.
Published: (2025)
by: Xie, Yuanzhen, et al.
Published: (2025)
Differentiable Modal Logic for Multi-Agent Diagnosis, Orchestration and Communication
by: Sulc, Antonin
Published: (2026)
by: Sulc, Antonin
Published: (2026)
From Test-taking to Cognitive Scaffolding: A Pedagogical Diagnostic Benchmark for LLMs on English Standardized Tests
by: Tang, Luoxi, et al.
Published: (2025)
by: Tang, Luoxi, et al.
Published: (2025)
Similar Items
-
Gen4Gen: Generative Data Pipeline for Generative Multi-Concept Composition
by: Yeh, Chun-Hsiao, et al.
Published: (2024) -
Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs
by: Yeh, Chun-Hsiao, et al.
Published: (2025) -
Pose-Aware Self-Supervised Learning with Viewpoint Trajectory Regularization
by: Wang, Jiayun, et al.
Published: (2024) -
Ocular Authentication: Fusion of Gaze and Periocular Modalities
by: Lohr, Dillon, et al.
Published: (2025) -
Incomplete Modality Disentangled Representation for Ophthalmic Disease Grading and Diagnosis
by: Liu, Chengzhi, et al.
Published: (2025)