LMOD: A Large Multimodal Ophthalmology Dataset and Benchmark for Large Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qin, Zhenyue, Yin, Yu, Campbell, Dylan, Wu, Xuansheng, Zou, Ke, Tham, Yih-Chung, Liu, Ninghao, Zhang, Xiuzhen, Chen, Qingyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LMOD+: A Comprehensive Multimodal Dataset and Benchmark for Developing and Evaluating Multimodal Large Language Models in Ophthalmology
von: Qin, Zhenyue, et al.
Veröffentlicht: (2025)
von: Qin, Zhenyue, et al.
Veröffentlicht: (2025)
VOLMO: Versatile and Open Large Models for Ophthalmology
von: Qin, Zhenyue, et al.
Veröffentlicht: (2026)
von: Qin, Zhenyue, et al.
Veröffentlicht: (2026)
HandCraft: Anatomically Correct Restoration of Malformed Hands in Diffusion Generated Images
von: Qin, Zhenyue, et al.
Veröffentlicht: (2024)
von: Qin, Zhenyue, et al.
Veröffentlicht: (2024)
Benchmarking Next-Generation Reasoning-Focused Large Language Models in Ophthalmology: A Head-to-Head Evaluation on 5,888 Items
von: Zou, Minjie, et al.
Veröffentlicht: (2025)
von: Zou, Minjie, et al.
Veröffentlicht: (2025)
EH-Benchmark Ophthalmic Hallucination Benchmark and Agent-Driven Top-Down Traceable Reasoning Workflow
von: Pan, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Pan, Xiaoyu, et al.
Veröffentlicht: (2025)
RefRef: A Synthetic Dataset and Benchmark for Reconstructing Refractive and Reflective Objects
von: Yin, Yue, et al.
Veröffentlicht: (2025)
von: Yin, Yue, et al.
Veröffentlicht: (2025)
Applying Large Language Models and Chain-of-Thought for Automatic Scoring
von: Lee, Gyeong-Geon, et al.
Veröffentlicht: (2023)
von: Lee, Gyeong-Geon, et al.
Veröffentlicht: (2023)
Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language Models
von: Yin, Jianghao, et al.
Veröffentlicht: (2026)
von: Yin, Jianghao, et al.
Veröffentlicht: (2026)
SPARK: Multi-Vision Sensor Perception and Reasoning Benchmark for Large-scale Vision-Language Models
von: Yu, Youngjoon, et al.
Veröffentlicht: (2024)
von: Yu, Youngjoon, et al.
Veröffentlicht: (2024)
LEME: Open Large Language Models for Ophthalmology with Advanced Reasoning and Clinical Validation
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
Benchmarking Direct Preference Optimization for Medical Large Vision-Language Models
von: Kim, Dain, et al.
Veröffentlicht: (2026)
von: Kim, Dain, et al.
Veröffentlicht: (2026)
MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
von: Fu, Chaoyou, et al.
Veröffentlicht: (2023)
von: Fu, Chaoyou, et al.
Veröffentlicht: (2023)
Benchmarking Large Vision-Language Models on CFMME: A Comprehensive Chinese Financial Multimodal Evaluation Dataset
von: Chen, Qian, et al.
Veröffentlicht: (2026)
von: Chen, Qian, et al.
Veröffentlicht: (2026)
Vision-DeepResearch: Incentivizing DeepResearch Capability in Multimodal Large Language Models
von: Huang, Wenxuan, et al.
Veröffentlicht: (2026)
von: Huang, Wenxuan, et al.
Veröffentlicht: (2026)
SurgMLLMBench: A Multimodal Large Language Model Benchmark Dataset for Surgical Scene Understanding
von: Choi, Tae-Min, et al.
Veröffentlicht: (2025)
von: Choi, Tae-Min, et al.
Veröffentlicht: (2025)
Detecting Latin in Historical Books with Large Language Models: A Multimodal Benchmark
von: Wu, Yu, et al.
Veröffentlicht: (2025)
von: Wu, Yu, et al.
Veröffentlicht: (2025)
Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space
von: Trinh, Quoc-Huy, et al.
Veröffentlicht: (2026)
von: Trinh, Quoc-Huy, et al.
Veröffentlicht: (2026)
DDFAV: Remote Sensing Large Vision Language Models Dataset and Evaluation Benchmark
von: Li, Haodong, et al.
Veröffentlicht: (2024)
von: Li, Haodong, et al.
Veröffentlicht: (2024)
Ranked from Within: Ranking Large Multimodal Models Without Labels
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
Jagle: Building a Large-Scale Japanese Multimodal Post-Training Dataset for Vision-Language Models
von: Sugiura, Issa, et al.
Veröffentlicht: (2026)
von: Sugiura, Issa, et al.
Veröffentlicht: (2026)
MMR-AD: A Large-Scale Multimodal Dataset for Benchmarking General Anomaly Detection with Multimodal Large Language Models
von: Yao, Xincheng, et al.
Veröffentlicht: (2026)
von: Yao, Xincheng, et al.
Veröffentlicht: (2026)
Rethinking Facial Expression Recognition in the Era of Multimodal Large Language Models: Benchmark, Datasets, and Beyond
von: Zhang, Fan, et al.
Veröffentlicht: (2025)
von: Zhang, Fan, et al.
Veröffentlicht: (2025)
AutoSCORE: Enhancing Automated Scoring with Multi-Agent Large Language Models via Structured Component Recognition
von: Wang, Yun, et al.
Veröffentlicht: (2025)
von: Wang, Yun, et al.
Veröffentlicht: (2025)
EyecareGPT: Boosting Comprehensive Ophthalmology Understanding with Tailored Dataset, Benchmark and Model
von: Li, Sijing, et al.
Veröffentlicht: (2025)
von: Li, Sijing, et al.
Veröffentlicht: (2025)
Concept-Centric Token Interpretation for Vector-Quantized Generative Models
von: Yang, Tianze, et al.
Veröffentlicht: (2025)
von: Yang, Tianze, et al.
Veröffentlicht: (2025)
PEBench: A Fictitious Dataset to Benchmark Machine Unlearning for Multimodal Large Language Models
von: Xu, Zhaopan, et al.
Veröffentlicht: (2025)
von: Xu, Zhaopan, et al.
Veröffentlicht: (2025)
Self-Regularization with Sparse Autoencoders for Controllable LLM-based Classification
von: Wu, Xuansheng, et al.
Veröffentlicht: (2025)
von: Wu, Xuansheng, et al.
Veröffentlicht: (2025)
MFC-Bench: Benchmarking Multimodal Fact-Checking with Large Vision-Language Models
von: Wang, Shengkang, et al.
Veröffentlicht: (2024)
von: Wang, Shengkang, et al.
Veröffentlicht: (2024)
MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models
von: Ren, Xiyu, et al.
Veröffentlicht: (2026)
von: Ren, Xiyu, et al.
Veröffentlicht: (2026)
Gastric-X: A Multimodal Multi-Phase Benchmark Dataset for Advancing Vision-Language Models in Gastric Cancer Analysis
von: Lu, Sheng, et al.
Veröffentlicht: (2026)
von: Lu, Sheng, et al.
Veröffentlicht: (2026)
A Survey on Sparse Autoencoders: Interpreting the Internal Mechanisms of Large Language Models
von: Shu, Dong, et al.
Veröffentlicht: (2025)
von: Shu, Dong, et al.
Veröffentlicht: (2025)
Vision-DeepResearch Benchmark: Rethinking Visual and Textual Search for Multimodal Large Language Models
von: Zeng, Yu, et al.
Veröffentlicht: (2026)
von: Zeng, Yu, et al.
Veröffentlicht: (2026)
ERVQA: A Dataset to Benchmark the Readiness of Large Vision Language Models in Hospital Environments
von: Ray, Sourjyadip, et al.
Veröffentlicht: (2024)
von: Ray, Sourjyadip, et al.
Veröffentlicht: (2024)
MMIE: Massive Multimodal Interleaved Comprehension Benchmark for Large Vision-Language Models
von: Xia, Peng, et al.
Veröffentlicht: (2024)
von: Xia, Peng, et al.
Veröffentlicht: (2024)
Room Envelopes: A Synthetic Dataset for Indoor Layout Reconstruction from Images
von: Bahrami, Sam, et al.
Veröffentlicht: (2025)
von: Bahrami, Sam, et al.
Veröffentlicht: (2025)
Insect-Foundation: A Foundation Model and Large Multimodal Dataset for Vision-Language Insect Understanding
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2025)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2025)
A Touch, Vision, and Language Dataset for Multimodal Alignment
von: Fu, Letian, et al.
Veröffentlicht: (2024)
von: Fu, Letian, et al.
Veröffentlicht: (2024)
OMIBench: Benchmarking Olympiad-Level Multi-Image Reasoning in Large Vision-Language Model
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
Hallucination-Aware Multimodal Benchmark for Gastrointestinal Image Analysis with Large Vision-Language Models
von: Khanal, Bidur, et al.
Veröffentlicht: (2025)
von: Khanal, Bidur, et al.
Veröffentlicht: (2025)
RS5M and GeoRSCLIP: A Large Scale Vision-Language Dataset and A Large Vision-Language Model for Remote Sensing
von: Zhang, Zilun, et al.
Veröffentlicht: (2023)
von: Zhang, Zilun, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
LMOD+: A Comprehensive Multimodal Dataset and Benchmark for Developing and Evaluating Multimodal Large Language Models in Ophthalmology
von: Qin, Zhenyue, et al.
Veröffentlicht: (2025) -
VOLMO: Versatile and Open Large Models for Ophthalmology
von: Qin, Zhenyue, et al.
Veröffentlicht: (2026) -
HandCraft: Anatomically Correct Restoration of Malformed Hands in Diffusion Generated Images
von: Qin, Zhenyue, et al.
Veröffentlicht: (2024) -
Benchmarking Next-Generation Reasoning-Focused Large Language Models in Ophthalmology: A Head-to-Head Evaluation on 5,888 Items
von: Zou, Minjie, et al.
Veröffentlicht: (2025) -
EH-Benchmark Ophthalmic Hallucination Benchmark and Agent-Driven Top-Down Traceable Reasoning Workflow
von: Pan, Xiaoyu, et al.
Veröffentlicht: (2025)