Can OpenAI o1 Reason Well in Ophthalmology? A 6,990-Question Head-to-Head Evaluation Study
Fuente:
arXiv
Saved in:
| Main Authors: | Srinivasan, Sahana, Ai, Xuguang, Zou, Minjie, Zou, Ke, Kim, Hyunjae, Lo, Thaddaeus Wai Soon, Pushpanathan, Krithi, Kong, Yiming, Li, Anran, Singer, Maxwell, Jin, Kai, Antaki, Fares, Chen, David Ziyou, Liu, Dianbo, Adelman, Ron A., Chen, Qingyu, Tham, Yih Chung |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmarking Next-Generation Reasoning-Focused Large Language Models in Ophthalmology: A Head-to-Head Evaluation on 5,888 Items
by: Zou, Minjie, et al.
Published: (2025)
by: Zou, Minjie, et al.
Published: (2025)
BEnchmarking LLMs for Ophthalmology (BELO) for Ophthalmological Knowledge and Reasoning
by: Srinivasan, Sahana, et al.
Published: (2025)
by: Srinivasan, Sahana, et al.
Published: (2025)
LEME: Open Large Language Models for Ophthalmology with Advanced Reasoning and Clinical Validation
by: Kim, Hyunjae, et al.
Published: (2024)
by: Kim, Hyunjae, et al.
Published: (2024)
VOLMO: Versatile and Open Large Models for Ophthalmology
by: Qin, Zhenyue, et al.
Published: (2026)
by: Qin, Zhenyue, et al.
Published: (2026)
Enhancing Large Language Models with Domain-specific Retrieval Augment Generation: A Case Study on Long-form Consumer Health Question Answering in Ophthalmology
by: Gilson, Aidan, et al.
Published: (2024)
by: Gilson, Aidan, et al.
Published: (2024)
LMOD: A Large Multimodal Ophthalmology Dataset and Benchmark for Large Vision-Language Models
by: Qin, Zhenyue, et al.
Published: (2024)
by: Qin, Zhenyue, et al.
Published: (2024)
Performance of GPT-5 Frontier Models in Ophthalmology Question Answering
by: Antaki, Fares, et al.
Published: (2025)
by: Antaki, Fares, et al.
Published: (2025)
LMOD+: A Comprehensive Multimodal Dataset and Benchmark for Developing and Evaluating Multimodal Large Language Models in Ophthalmology
by: Qin, Zhenyue, et al.
Published: (2025)
by: Qin, Zhenyue, et al.
Published: (2025)
L-HYDRA: Multi-Head Physics-Informed Neural Networks
by: Zou, Zongren, et al.
Published: (2023)
by: Zou, Zongren, et al.
Published: (2023)
A Federated and Parameter-Efficient Framework for Large Language Model Training in Medicine
by: Li, Anran, et al.
Published: (2026)
by: Li, Anran, et al.
Published: (2026)
Deliberative multi-agent large language models improve clinical reasoning in ophthalmology
by: Misaghi, Ehsan, et al.
Published: (2026)
by: Misaghi, Ehsan, et al.
Published: (2026)
Personalized Head-Related Transfer Function Prediction Based on Spatial Grouping
by: Chang, Keng-Wei, et al.
Published: (2024)
by: Chang, Keng-Wei, et al.
Published: (2024)
Effective Bibliographic Instruction Programs: A Comparison of Coordinators and Reference Heads in ARL Libraries.
by: Blazek, Ron
Published: (1985)
by: Blazek, Ron
Published: (1985)
Accuracy and Precision of Random Walk with Barrier Model Fitting: Simulations and Applications in Head and Neck Cancers
by: Zou, Jiaren, et al.
Published: (2025)
by: Zou, Jiaren, et al.
Published: (2025)
Ethics in Practice: Immunotherapy at End of Life in Head and Neck Cancer
by: Sholem Hack, et al.
Published: (2026)
by: Sholem Hack, et al.
Published: (2026)
Co‐Expression Pattern Analysis of Head‐to‐Head NLR Gene Pair Pik‐H4
by: Fengwei Gu, et al.
Published: (2025)
by: Fengwei Gu, et al.
Published: (2025)
Memorization in Large Language Models in Medicine: Prevalence, Characteristics, and Implications
by: Li, Anran, et al.
Published: (2025)
by: Li, Anran, et al.
Published: (2025)
En el banquete de Platón / Ikram Antaki
by: Antaki, Ikram
by: Antaki, Ikram
Después del naufragio: la literatura en este fin de siglo
by: Antaki, Ikram
Published: (1991)
by: Antaki, Ikram
Published: (1991)
De la sumisión a la autodeterminación: Nomad: From Islam to America: A Personal Journey Through the Clash of Civilizations, por Ayaan Hirsi Ali
by: Vivian Antaki
Published: (2015)
by: Vivian Antaki
Published: (2015)
Nichoas D. Kristof and Sheryl WuDunn, half the sky: turning oppression into opportunity for Women Worldwide [Mitad del cielo: transformando opresión en oportunidad para las mujeres del mundo], Knopf, Nueva York, 2009.
by: Vivian Antaki
Published: (2014)
by: Vivian Antaki
Published: (2014)
Los problemas reales de la enseñanza. ¿Formación humanística a los ingenieros?
by: Ikram Antaki
Published: (2001)
by: Ikram Antaki
Published: (2001)
Reseña de "Adán en Edén" de Carlos Fuentes
by: Vivian Antaki
Published: (2011)
by: Vivian Antaki
Published: (2011)
RESEñA DE THE GULF, MY SHOULDER... NORTH WAS THE HORIZON, DE CLARK MURRAY: ACERCAMIENTO VORAZ A UNA REALIDAD DESMEDIDA
by: Vivian Antaki
Published: (2013)
by: Vivian Antaki
Published: (2013)
El Análisis del discurso implica analizar: Crítica de seis atajos analíticos
by: Charles Antaki
Published: (2003)
by: Charles Antaki
Published: (2003)
Accelerating Audio Research with Robotic Dummy Heads
by: Lu, Austin, et al.
Published: (2025)
by: Lu, Austin, et al.
Published: (2025)
A Benchmark for End-to-End Zero-Shot Biomedical Relation Extraction with LLMs: Experiments with OpenAI Models
by: Brokman, Aviv, et al.
Published: (2025)
by: Brokman, Aviv, et al.
Published: (2025)
Human-like Content Analysis for Generative AI with Language-Grounded Sparse Encoders
by: Tang, Yiming, et al.
Published: (2025)
by: Tang, Yiming, et al.
Published: (2025)
EyeAgent: An Agentic AI System for Multimodal Clinical Decision Support in Ophthalmology
by: Shi, Danli, et al.
Published: (2025)
by: Shi, Danli, et al.
Published: (2025)
RecurFormer: Not All Transformer Heads Need Self-Attention
by: Yan, Ruiqing, et al.
Published: (2024)
by: Yan, Ruiqing, et al.
Published: (2024)
Primary Aldosteronism Presenting as Dropped Head Syndrome With Hypokalemic Rhabdomyolysis: A Case Report
by: Ya-Chen Kao, et al.
Published: (2026)
by: Ya-Chen Kao, et al.
Published: (2026)
How Transformers Utilize Multi-Head Attention in In-Context Learning? A Case Study on Sparse Linear Regression
by: Chen, Xingwu, et al.
Published: (2024)
by: Chen, Xingwu, et al.
Published: (2024)
Complementary Human-AI Clinical Reasoning in Ophthalmology
by: Sevgi, Mertcan, et al.
Published: (2025)
by: Sevgi, Mertcan, et al.
Published: (2025)
Representation Collapsing Problems in Vector Quantization
by: Zhao, Wenhao, et al.
Published: (2024)
by: Zhao, Wenhao, et al.
Published: (2024)
Mitigating Premature Discretization with Progressive Quantization for Robust Vector Tokenization
by: Zhao, Wenhao, et al.
Published: (2026)
by: Zhao, Wenhao, et al.
Published: (2026)
Acute Gout Flare After Carboplatin/5‐Fluorouracil for Locally Advanced Head and Neck Squamous Cell Carcinoma
by: Henry Zou, et al.
Published: (2026)
by: Henry Zou, et al.
Published: (2026)
Induction Head Toxicity Mechanistically Explains Repetition Curse in Large Language Models
by: Wang, Shuxun, et al.
Published: (2025)
by: Wang, Shuxun, et al.
Published: (2025)
Beyond Nation‐State Climate Governance: Cuerpo‐Territorio and Decolonial Feminist Pathways to Justice
by: Miriam Gay‐Antaki
Published: (2025)
by: Miriam Gay‐Antaki
Published: (2025)
EH-Benchmark Ophthalmic Hallucination Benchmark and Agent-Driven Top-Down Traceable Reasoning Workflow
by: Pan, Xiaoyu, et al.
Published: (2025)
by: Pan, Xiaoyu, et al.
Published: (2025)
Head-Aware KV Cache Compression for Efficient Visual Autoregressive Modeling
by: Qin, Ziran, et al.
Published: (2025)
by: Qin, Ziran, et al.
Published: (2025)
Similar Items
-
Benchmarking Next-Generation Reasoning-Focused Large Language Models in Ophthalmology: A Head-to-Head Evaluation on 5,888 Items
by: Zou, Minjie, et al.
Published: (2025) -
BEnchmarking LLMs for Ophthalmology (BELO) for Ophthalmological Knowledge and Reasoning
by: Srinivasan, Sahana, et al.
Published: (2025) -
LEME: Open Large Language Models for Ophthalmology with Advanced Reasoning and Clinical Validation
by: Kim, Hyunjae, et al.
Published: (2024) -
VOLMO: Versatile and Open Large Models for Ophthalmology
by: Qin, Zhenyue, et al.
Published: (2026) -
Enhancing Large Language Models with Domain-specific Retrieval Augment Generation: A Case Study on Long-form Consumer Health Question Answering in Ophthalmology
by: Gilson, Aidan, et al.
Published: (2024)