FunBench: Benchmarking Fundus Reading Skills of MLLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Wei, Qijie, Qian, Kaiheng, Li, Xirong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Fundus-R1: Training a Fundus-Reading MLLM with Knowledge-Aware Reasoning on Public Data
di: Deng, Yuchuan, et al.
Pubblicazione: (2026)
di: Deng, Yuchuan, et al.
Pubblicazione: (2026)
Cross-modal Fundus Image Registration under Large FoV Disparity
di: Li, Hongyang, et al.
Pubblicazione: (2025)
di: Li, Hongyang, et al.
Pubblicazione: (2025)
EI: Early Intervention for Multimodal Imaging based Disease Recognition
di: Wei, Qijie, et al.
Pubblicazione: (2026)
di: Wei, Qijie, et al.
Pubblicazione: (2026)
Convolutional Prompting for Broad-Domain Retinal Vessel Segmentation
di: Wei, Qijie, et al.
Pubblicazione: (2024)
di: Wei, Qijie, et al.
Pubblicazione: (2024)
Co-Teaching for Unsupervised Domain Adaptation and Expansion
di: Lin, Hailan, et al.
Pubblicazione: (2022)
di: Lin, Hailan, et al.
Pubblicazione: (2022)
EventBench: Towards Comprehensive Benchmarking of Event-based MLLMs
di: Liu, Shaoyu, et al.
Pubblicazione: (2025)
di: Liu, Shaoyu, et al.
Pubblicazione: (2025)
FunOTTA: On-the-Fly Adaptation on Cross-Domain Fundus Image via Stable Test-time Training
di: Zeng, Qian, et al.
Pubblicazione: (2024)
di: Zeng, Qian, et al.
Pubblicazione: (2024)
PhysToolBench: Benchmarking Physical Tool Understanding for MLLMs
di: Zhang, Zixin, et al.
Pubblicazione: (2025)
di: Zhang, Zixin, et al.
Pubblicazione: (2025)
PunchBench: Benchmarking MLLMs in Multimodal Punchline Comprehension
di: Ouyang, Kun, et al.
Pubblicazione: (2024)
di: Ouyang, Kun, et al.
Pubblicazione: (2024)
IF-Bench: Benchmarking and Enhancing MLLMs for Infrared Images with Generative Visual Prompting
di: Zhang, Tao, et al.
Pubblicazione: (2025)
di: Zhang, Tao, et al.
Pubblicazione: (2025)
CameraBench: Benchmarking Visual Reasoning in MLLMs via Photography
di: Fang, I-Sheng, et al.
Pubblicazione: (2025)
di: Fang, I-Sheng, et al.
Pubblicazione: (2025)
MC-Bench: A Benchmark for Multi-Context Visual Grounding in the Era of MLLMs
di: Xu, Yunqiu, et al.
Pubblicazione: (2024)
di: Xu, Yunqiu, et al.
Pubblicazione: (2024)
BADet: Boundary-Aware 3D Object Detection from Point Clouds
di: Qian, Rui, et al.
Pubblicazione: (2021)
di: Qian, Rui, et al.
Pubblicazione: (2021)
3D Object Detection for Autonomous Driving: A Survey
di: Qian, Rui, et al.
Pubblicazione: (2021)
di: Qian, Rui, et al.
Pubblicazione: (2021)
RoadBench: Benchmarking MLLMs on Fine-Grained Spatial Understanding and Reasoning under Urban Road Scenarios
di: Zhang, Jun, et al.
Pubblicazione: (2025)
di: Zhang, Jun, et al.
Pubblicazione: (2025)
VER-Bench: Evaluating MLLMs on Reasoning with Fine-Grained Visual Evidence
di: Qiang, Chenhui, et al.
Pubblicazione: (2025)
di: Qiang, Chenhui, et al.
Pubblicazione: (2025)
MileBench: Benchmarking MLLMs in Long Context
di: Song, Dingjie, et al.
Pubblicazione: (2024)
di: Song, Dingjie, et al.
Pubblicazione: (2024)
EgoExoBench: A Benchmark for First- and Third-person View Video Understanding in MLLMs
di: He, Yuping, et al.
Pubblicazione: (2025)
di: He, Yuping, et al.
Pubblicazione: (2025)
Benchmarking Large and Small MLLMs
di: Feng, Xuelu, et al.
Pubblicazione: (2025)
di: Feng, Xuelu, et al.
Pubblicazione: (2025)
Cube Bench: A Benchmark for Spatial Visual Reasoning in MLLMs
di: Anand, Dhruv, et al.
Pubblicazione: (2025)
di: Anand, Dhruv, et al.
Pubblicazione: (2025)
ODI-Bench: Can MLLMs Understand Immersive Omnidirectional Environments?
di: Yang, Liu, et al.
Pubblicazione: (2025)
di: Yang, Liu, et al.
Pubblicazione: (2025)
FaceBench: A Multi-View Multi-Level Facial Attribute VQA Dataset for Benchmarking Face Perception MLLMs
di: Wang, Xiaoqin, et al.
Pubblicazione: (2025)
di: Wang, Xiaoqin, et al.
Pubblicazione: (2025)
Bridging Restoration and Diagnosis: A Comprehensive Benchmark for Retinal Fundus Enhancement
di: Dong, Xuanzhao, et al.
Pubblicazione: (2026)
di: Dong, Xuanzhao, et al.
Pubblicazione: (2026)
STI-Bench: Are MLLMs Ready for Precise Spatial-Temporal World Understanding?
di: Li, Yun, et al.
Pubblicazione: (2025)
di: Li, Yun, et al.
Pubblicazione: (2025)
EOC-Bench: Can MLLMs Identify, Recall, and Forecast Objects in an Egocentric World?
di: Yuan, Yuqian, et al.
Pubblicazione: (2025)
di: Yuan, Yuqian, et al.
Pubblicazione: (2025)
Artifact-Bench: Evaluating MLLMs on Detecting and Assessing the Artifacts of AI-Generated Videos
di: Tang, Yuqi, et al.
Pubblicazione: (2026)
di: Tang, Yuqi, et al.
Pubblicazione: (2026)
ViC-Bench: Benchmarking Visual-Interleaved Chain-of-Thought Capability in MLLMs with Free-Style Intermediate State Representations
di: Wu, Xuecheng, et al.
Pubblicazione: (2025)
di: Wu, Xuecheng, et al.
Pubblicazione: (2025)
SpatialViz-Bench: A Cognitively-Grounded Benchmark for Diagnosing Spatial Visualization in MLLMs
di: Wang, Siting, et al.
Pubblicazione: (2025)
di: Wang, Siting, et al.
Pubblicazione: (2025)
Towards Trustworthy Dermatology MLLMs: A Benchmark and Multimodal Evaluator for Diagnostic Narratives
di: Shen, Yuhao, et al.
Pubblicazione: (2025)
di: Shen, Yuhao, et al.
Pubblicazione: (2025)
Can MLLMs Read the Room? A Multimodal Benchmark for Assessing Deception in Multi-Party Social Interactions
di: Kang, Caixin, et al.
Pubblicazione: (2025)
di: Kang, Caixin, et al.
Pubblicazione: (2025)
MedQ-Bench: Evaluating and Exploring Medical Image Quality Assessment Abilities in MLLMs
di: Liu, Jiyao, et al.
Pubblicazione: (2025)
di: Liu, Jiyao, et al.
Pubblicazione: (2025)
MMDG-Bench: A Benchmark for Multimodal Domain Generalization
di: Zhan, Qianshan, et al.
Pubblicazione: (2026)
di: Zhan, Qianshan, et al.
Pubblicazione: (2026)
EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs
di: Dai, Yang, et al.
Pubblicazione: (2026)
di: Dai, Yang, et al.
Pubblicazione: (2026)
ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs
di: Wu, Xin, et al.
Pubblicazione: (2026)
di: Wu, Xin, et al.
Pubblicazione: (2026)
MultiEYE: Dataset and Benchmark for OCT-Enhanced Retinal Disease Recognition from Fundus Images
di: Wang, Lehan, et al.
Pubblicazione: (2024)
di: Wang, Lehan, et al.
Pubblicazione: (2024)
VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning?
di: Liu, Yuanxin, et al.
Pubblicazione: (2025)
di: Liu, Yuanxin, et al.
Pubblicazione: (2025)
RefereeBench: Are Video MLLMs Ready to be Multi-Sport Referees
di: Xu, Yichen, et al.
Pubblicazione: (2026)
di: Xu, Yichen, et al.
Pubblicazione: (2026)
OST-Bench: Evaluating the Capabilities of MLLMs in Online Spatio-temporal Scene Understanding
di: Lin, Jingli, et al.
Pubblicazione: (2025)
di: Lin, Jingli, et al.
Pubblicazione: (2025)
InstructionBench: An Instructional Video Understanding Benchmark
di: Wei, Haiwan, et al.
Pubblicazione: (2025)
di: Wei, Haiwan, et al.
Pubblicazione: (2025)
Video-MSR: Benchmarking Multi-hop Spatial Reasoning Capabilities of MLLMs
di: Zhu, Rui, et al.
Pubblicazione: (2026)
di: Zhu, Rui, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Fundus-R1: Training a Fundus-Reading MLLM with Knowledge-Aware Reasoning on Public Data
di: Deng, Yuchuan, et al.
Pubblicazione: (2026) -
Cross-modal Fundus Image Registration under Large FoV Disparity
di: Li, Hongyang, et al.
Pubblicazione: (2025) -
EI: Early Intervention for Multimodal Imaging based Disease Recognition
di: Wei, Qijie, et al.
Pubblicazione: (2026) -
Convolutional Prompting for Broad-Domain Retinal Vessel Segmentation
di: Wei, Qijie, et al.
Pubblicazione: (2024) -
Co-Teaching for Unsupervised Domain Adaptation and Expansion
di: Lin, Hailan, et al.
Pubblicazione: (2022)