ReX-MLE: The Autonomous Agent Benchmark for Medical Imaging Challenges
Fuente:
arXiv
Saved in:
| Main Authors: | Kenia, Roshan, Zhang, Xiaoman, Rajpurkar, Pranav |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding
by: Wang, Xucheng, et al.
Published: (2026)
by: Wang, Xucheng, et al.
Published: (2026)
FactCheXcker: Mitigating Measurement Hallucinations in Chest X-ray Report Generation Models
by: Heiman, Alice, et al.
Published: (2024)
by: Heiman, Alice, et al.
Published: (2024)
ReXGradient-160K: A Large-Scale Publicly Available Dataset of Chest Radiographs with Free-text Reports
by: Zhang, Xiaoman, et al.
Published: (2025)
by: Zhang, Xiaoman, et al.
Published: (2025)
MedAutoCorrect: Image-Conditioned Autocorrection in Medical Reporting
by: Asiimwe, Arnold Caleb, et al.
Published: (2024)
by: Asiimwe, Arnold Caleb, et al.
Published: (2024)
ReXInTheWild: A Unified Benchmark for Medical Photograph Understanding
by: Banerjee, Oishi, et al.
Published: (2026)
by: Banerjee, Oishi, et al.
Published: (2026)
Uncovering Knowledge Gaps in Radiology Report Generation Models through Knowledge Graphs
by: Zhang, Xiaoman, et al.
Published: (2024)
by: Zhang, Xiaoman, et al.
Published: (2024)
ReXVQA: A Large-scale Visual Question Answering Benchmark for Generalist Chest X-ray Understanding
by: Pal, Ankit, et al.
Published: (2025)
by: Pal, Ankit, et al.
Published: (2025)
Learning Generalized Medical Image Representations through Image-Graph Contrastive Pretraining
by: Khanna, Sameer, et al.
Published: (2024)
by: Khanna, Sameer, et al.
Published: (2024)
Evaluating Contextual Intelligence in Recyclability: A Comprehensive Study of Image-Based Reasoning Systems
by: Park, Eliot, et al.
Published: (2025)
by: Park, Eliot, et al.
Published: (2025)
ReXrank: A Public Leaderboard for AI-Powered Radiology Report Generation
by: Zhang, Xiaoman, et al.
Published: (2024)
by: Zhang, Xiaoman, et al.
Published: (2024)
3DReasonKnee: Advancing Grounded Reasoning in Medical Vision Language Models
by: Sambara, Sraavya, et al.
Published: (2025)
by: Sambara, Sraavya, et al.
Published: (2025)
MedVersa: A Generalist Foundation Model for Medical Image Interpretation
by: Zhou, Hong-Yu, et al.
Published: (2024)
by: Zhou, Hong-Yu, et al.
Published: (2024)
Multimodal Foundation Models Exploit Text to Make Medical Image Predictions
by: Buckley, Thomas, et al.
Published: (2023)
by: Buckley, Thomas, et al.
Published: (2023)
3D ReX: Causal Explanations in 3D Neuroimaging Classification
by: Navaratnarajah, Melane, et al.
Published: (2025)
by: Navaratnarajah, Melane, et al.
Published: (2025)
MedFrameQA: A Multi-Image Medical VQA Benchmark for Clinical Reasoning
by: Yu, Suhao, et al.
Published: (2025)
by: Yu, Suhao, et al.
Published: (2025)
RadFlag: A Black-Box Hallucination Detection Method for Medical Vision Language Models
by: Zhang, Serena, et al.
Published: (2024)
by: Zhang, Serena, et al.
Published: (2024)
ColonCrafter: A Depth Estimation Model for Colonoscopy Videos Using Diffusion Priors
by: Hardy, Romain, et al.
Published: (2025)
by: Hardy, Romain, et al.
Published: (2025)
Medical Image De-Identification Benchmark Challenge
by: Pei, Linmin, et al.
Published: (2025)
by: Pei, Linmin, et al.
Published: (2025)
Benchmarking Real-World Medical Image Classification with Noisy Labels: Challenges, Practice, and Outlook
by: Ma, Yuan, et al.
Published: (2025)
by: Ma, Yuan, et al.
Published: (2025)
Towards All-in-One Medical Image Re-Identification
by: Tian, Yuan, et al.
Published: (2025)
by: Tian, Yuan, et al.
Published: (2025)
X-Mark: Saliency-Guided Robust Dataset Ownership Verification for Medical Imaging
by: Kulkarni, Pranav, et al.
Published: (2026)
by: Kulkarni, Pranav, et al.
Published: (2026)
AI-CNet3D: An Anatomically-Informed Cross-Attention Network with Multi-Task Consistency Fine-tuning for 3D Glaucoma Classification
by: Kenia, Roshan, et al.
Published: (2025)
by: Kenia, Roshan, et al.
Published: (2025)
Exploring Intrinsic Properties of Medical Images for Self-Supervised Binary Semantic Segmentation
by: Singh, Pranav, et al.
Published: (2024)
by: Singh, Pranav, et al.
Published: (2024)
a2z-1 for Multi-Disease Detection in Abdomen-Pelvis CT: External Validation and Performance Analysis Across 21 Conditions
by: Rajpurkar, Pranav, et al.
Published: (2024)
by: Rajpurkar, Pranav, et al.
Published: (2024)
PhysMLE: Generalizable and Priors-Inclusive Multi-task Remote Physiological Measurement
by: Wang, Jiyao, et al.
Published: (2024)
by: Wang, Jiyao, et al.
Published: (2024)
COOOL: Challenge Of Out-Of-Label A Novel Benchmark for Autonomous Driving
by: AlShami, Ali K., et al.
Published: (2024)
by: AlShami, Ali K., et al.
Published: (2024)
Large-Vocabulary Segmentation for Medical Images with Text Prompts
by: Zhao, Ziheng, et al.
Published: (2023)
by: Zhao, Ziheng, et al.
Published: (2023)
PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
by: Zhang, Xiaoman, et al.
Published: (2023)
by: Zhang, Xiaoman, et al.
Published: (2023)
A New Benchmark and Model for Challenging Image Manipulation Detection
by: Zhang, Zhenfei, et al.
Published: (2023)
by: Zhang, Zhenfei, et al.
Published: (2023)
MLE-UVAD: Minimal Latent Entropy Autoencoder for Fully Unsupervised Video Anomaly Detection
by: Geng, Yuang, et al.
Published: (2026)
by: Geng, Yuang, et al.
Published: (2026)
CoRe: Joint Optimization with Contrastive Learning for Medical Image Registration
by: Kats, Eytan, et al.
Published: (2026)
by: Kats, Eytan, et al.
Published: (2026)
MedSG-Bench: A Benchmark for Medical Image Sequences Grounding
by: Yue, Jingkun, et al.
Published: (2025)
by: Yue, Jingkun, et al.
Published: (2025)
CleanStyle: Plug-and-Play Style Conditioning Purification for Text-to-Image Stylization
by: Feng, Xiaoman, et al.
Published: (2026)
by: Feng, Xiaoman, et al.
Published: (2026)
Direct Preference Optimization for Suppressing Hallucinated Prior Exams in Radiology Report Generation
by: Banerjee, Oishi, et al.
Published: (2024)
by: Banerjee, Oishi, et al.
Published: (2024)
Open Challenges on Fairness of Artificial Intelligence in Medical Imaging Applications
by: Ferrante, Enzo, et al.
Published: (2024)
by: Ferrante, Enzo, et al.
Published: (2024)
Interactive Medical Image Segmentation: A Benchmark Dataset and Baseline
by: Cheng, Junlong, et al.
Published: (2024)
by: Cheng, Junlong, et al.
Published: (2024)
Development and Enhancement of Text-to-Image Diffusion Models
by: Sahu, Rajdeep Roshan
Published: (2025)
by: Sahu, Rajdeep Roshan
Published: (2025)
Image Segmentation using Chan-Vese Active Contours
by: P, Pranav Shenoy K.
Published: (2025)
by: P, Pranav Shenoy K.
Published: (2025)
4Seasons: Benchmarking Visual SLAM and Long-Term Localization for Autonomous Driving in Challenging Conditions
by: Wenzel, Patrick, et al.
Published: (2022)
by: Wenzel, Patrick, et al.
Published: (2022)
BenchX: A Unified Benchmark Framework for Medical Vision-Language Pretraining on Chest X-Rays
by: Zhou, Yang, et al.
Published: (2024)
by: Zhou, Yang, et al.
Published: (2024)
Similar Items
-
ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding
by: Wang, Xucheng, et al.
Published: (2026) -
FactCheXcker: Mitigating Measurement Hallucinations in Chest X-ray Report Generation Models
by: Heiman, Alice, et al.
Published: (2024) -
ReXGradient-160K: A Large-Scale Publicly Available Dataset of Chest Radiographs with Free-text Reports
by: Zhang, Xiaoman, et al.
Published: (2025) -
MedAutoCorrect: Image-Conditioned Autocorrection in Medical Reporting
by: Asiimwe, Arnold Caleb, et al.
Published: (2024) -
ReXInTheWild: A Unified Benchmark for Medical Photograph Understanding
by: Banerjee, Oishi, et al.
Published: (2026)