PathMMU: A Massive Multimodal Expert-Level Benchmark for Understanding and Reasoning in Pathology
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Yuxuan, Wu, Hao, Zhu, Chenglu, Zheng, Sunyi, Chen, Qizi, Zhang, Kai, Zhang, Yunlong, Wan, Dan, Lan, Xiaoxiao, Zheng, Mengyue, Li, Jingxiong, Lyu, Xinheng, Lin, Tao, Yang, Lin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmarking PathCLIP for Pathology Image Analysis
by: Zheng, Sunyi, et al.
Published: (2024)
by: Zheng, Sunyi, et al.
Published: (2024)
PathAsst: A Generative Foundation AI Assistant Towards Artificial General Intelligence of Pathology
by: Sun, Yuxuan, et al.
Published: (2023)
by: Sun, Yuxuan, et al.
Published: (2023)
PathGen-1.6M: 1.6 Million Pathology Image-text Pairs Generation through Multi-agent Collaboration
by: Sun, Yuxuan, et al.
Published: (2024)
by: Sun, Yuxuan, et al.
Published: (2024)
Attention-Challenging Multiple Instance Learning for Whole Slide Image Classification
by: Zhang, Yunlong, et al.
Published: (2023)
by: Zhang, Yunlong, et al.
Published: (2023)
Unleashing the Power of Prompt-driven Nucleus Instance Segmentation
by: Shui, Zhongyi, et al.
Published: (2023)
by: Shui, Zhongyi, et al.
Published: (2023)
WsiCaption: Multiple Instance Generation of Pathology Reports for Gigapixel Whole-Slide Images
by: Chen, Pingyi, et al.
Published: (2023)
by: Chen, Pingyi, et al.
Published: (2023)
Large-scale cervical precancerous screening via AI-assisted cytology whole slide image analysis
by: Li, Honglin, et al.
Published: (2024)
by: Li, Honglin, et al.
Published: (2024)
PathVQ: Reforming Computational Pathology Foundation Model for Whole Slide Image Analysis via Vector Quantization
by: Li, Honglin, et al.
Published: (2025)
by: Li, Honglin, et al.
Published: (2025)
WSI-VQA: Interpreting Whole Slide Images by Generative Visual Question Answering
by: Chen, Pingyi, et al.
Published: (2024)
by: Chen, Pingyi, et al.
Published: (2024)
AEM: Attention Entropy Maximization for Multiple Instance Learning based Whole Slide Image Classification
by: Zhang, Yunlong, et al.
Published: (2024)
by: Zhang, Yunlong, et al.
Published: (2024)
AgMMU: A Comprehensive Agricultural Multimodal Understanding Benchmark
by: Gauba, Aruna, et al.
Published: (2025)
by: Gauba, Aruna, et al.
Published: (2025)
CPath-Omni: A Unified Multimodal Foundation Model for Patch and Whole Slide Image Analysis in Computational Pathology
by: Sun, Yuxuan, et al.
Published: (2024)
by: Sun, Yuxuan, et al.
Published: (2024)
MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI
by: Yue, Xiang, et al.
Published: (2023)
by: Yue, Xiang, et al.
Published: (2023)
WebMMU: A Benchmark for Multimodal Multilingual Website Understanding and Code Generation
by: Awal, Rabiul, et al.
Published: (2025)
by: Awal, Rabiul, et al.
Published: (2025)
Multi-modal Learning with Missing Modality in Predicting Axillary Lymph Node Metastasis
by: Zhang, Shichuan, et al.
Published: (2024)
by: Zhang, Shichuan, et al.
Published: (2024)
ReinPath: A Multimodal Reinforcement Learning Approach for Pathology
by: Zhou, Kangcheng, et al.
Published: (2026)
by: Zhou, Kangcheng, et al.
Published: (2026)
CPathAgent: An Agent-based Foundation Model for Interpretable High-Resolution Pathology Image Analysis Mimicking Pathologists' Diagnostic Logic
by: Sun, Yuxuan, et al.
Published: (2025)
by: Sun, Yuxuan, et al.
Published: (2025)
IMPROVEMENT OF ACCOUNTING FOR TAXES AND MANDATORY PAYMENTS IN THE STRUCTURE OF PERIOD EXPENSES
by: Maxkamova Dilshodabonu Shavkat Qizi, et al.
Published: (2025)
by: Maxkamova Dilshodabonu Shavkat Qizi, et al.
Published: (2025)
WeMMU: Enhanced Bridging of Vision-Language Models and Diffusion Models via Noisy Query Tokens
by: Yang, Jian, et al.
Published: (2025)
by: Yang, Jian, et al.
Published: (2025)
HBridge: H-Shape Bridging of Heterogeneous Experts for Unified Multimodal Understanding and Generation
by: Wang, Xiang, et al.
Published: (2025)
by: Wang, Xiang, et al.
Published: (2025)
Rethinking Transformer for Long Contextual Histopathology Whole Slide Image Analysis
by: Li, Honglin, et al.
Published: (2024)
by: Li, Honglin, et al.
Published: (2024)
AW-MoE: All-Weather Mixture of Experts for Robust Multi-Modal 3D Object Detection
by: Lin, Hongwei, et al.
Published: (2026)
by: Lin, Hongwei, et al.
Published: (2026)
Multimodal Table Understanding
by: Zheng, Mingyu, et al.
Published: (2024)
by: Zheng, Mingyu, et al.
Published: (2024)
CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark
by: Zhang, Ge, et al.
Published: (2024)
by: Zhang, Ge, et al.
Published: (2024)
Comment on “Assessment of a New Tool to Monitor Oral Hydration and Dry Mouth: FishburneTabs”
by: Chenglu Ruan, et al.
Published: (2026)
by: Chenglu Ruan, et al.
Published: (2026)
Comment on “Patient and System Barriers to Early Diagnosis of Oral Cancer in the UK ”
by: Chenglu Ruan, et al.
Published: (2026)
by: Chenglu Ruan, et al.
Published: (2026)
Patho-R1: A Multimodal Reinforcement Learning-Based Pathology Expert Reasoner
by: Zhang, Wenchuan, et al.
Published: (2025)
by: Zhang, Wenchuan, et al.
Published: (2025)
Editorial for “Use of Blood Oxygenation Level‐Dependent MRI to Predict Clinical Outcomes After Endovascular Revascularization in Peripheral Artery Disease”
by: Xinheng Zhang, et al.
Published: (2025)
by: Xinheng Zhang, et al.
Published: (2025)
LoC-Path: Learning to Compress for Pathology Multimodal Large Language Models
by: Hu, Qingqiao, et al.
Published: (2025)
by: Hu, Qingqiao, et al.
Published: (2025)
PathMR: Multimodal Visual Reasoning for Interpretable Pathology Diagnosis
by: Zhang, Ye, et al.
Published: (2025)
by: Zhang, Ye, et al.
Published: (2025)
THE SPEECH ERRORS IN TEACHING FOREIGN LANGUAGES
by: Dusmatova Laylo Ikrom Qizi
Published: (2025)
by: Dusmatova Laylo Ikrom Qizi
Published: (2025)
GIVING FEEDBACK TO THE MIDDLE SCHOOL LEARNERS TO DEVELOP SOCIOCULTURAL COMPETENCE IN PRIMARY SCHOOL LEARNERS
by: Rashidova Rayhona Toshtemir Qizi
Published: (2025)
by: Rashidova Rayhona Toshtemir Qizi
Published: (2025)
TUG'MA TANGLAY YORIQLIKLARIDA DAVOLOVCHI VA PROFILAKTIK YORDAMNI TASHKIL ETISH USULLARI
by: Shukurova Ozoda Nodir Qizi
Published: (2025)
by: Shukurova Ozoda Nodir Qizi
Published: (2025)
MOTOR ALALIYA KAMCHILIGIGA EGA BOLALARNING SHAXS XUSUSIYATLARI
by: Shukurova Ozoda Nodir Qizi
Published: (2025)
by: Shukurova Ozoda Nodir Qizi
Published: (2025)
TiMi: Empower Time Series Transformers with Multimodal Mixture of Experts
by: Lin, Jiafeng, et al.
Published: (2026)
by: Lin, Jiafeng, et al.
Published: (2026)
Living Arrangements and Women's Household Decision‐Making Power in China
by: Xinheng Li
Published: (2026)
by: Xinheng Li
Published: (2026)
Towards Effective and Efficient Context-aware Nucleus Detection in Histopathology Whole Slide Images
by: Shui, Zhongyi, et al.
Published: (2025)
by: Shui, Zhongyi, et al.
Published: (2025)
PathAR: Structure-First Autoregressive Synthesis of Multimodal Pathology Images
by: Zhang, Yuan, et al.
Published: (2026)
by: Zhang, Yuan, et al.
Published: (2026)
PathFound: An Agentic Multimodal Model Activating Evidence-seeking Pathological Diagnosis
by: Hua, Shengyi, et al.
Published: (2025)
by: Hua, Shengyi, et al.
Published: (2025)
RTMC: Step-Level Credit Assignment via Rollout Trees
by: Wang, Tao, et al.
Published: (2026)
by: Wang, Tao, et al.
Published: (2026)
Similar Items
-
Benchmarking PathCLIP for Pathology Image Analysis
by: Zheng, Sunyi, et al.
Published: (2024) -
PathAsst: A Generative Foundation AI Assistant Towards Artificial General Intelligence of Pathology
by: Sun, Yuxuan, et al.
Published: (2023) -
PathGen-1.6M: 1.6 Million Pathology Image-text Pairs Generation through Multi-agent Collaboration
by: Sun, Yuxuan, et al.
Published: (2024) -
Attention-Challenging Multiple Instance Learning for Whole Slide Image Classification
by: Zhang, Yunlong, et al.
Published: (2023) -
Unleashing the Power of Prompt-driven Nucleus Instance Segmentation
by: Shui, Zhongyi, et al.
Published: (2023)