ReXErr: Synthesizing Clinically Meaningful Errors in Diagnostic Radiology Reports
Fuente:
arXiv
Saved in:
| Main Authors: | Rao, Vishwanatha M., Zhang, Serena, Acosta, Julian N., Adithan, Subathra, Rajpurkar, Pranav |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Direct Preference Optimization for Suppressing Hallucinated Prior Exams in Radiology Report Generation
by: Banerjee, Oishi, et al.
Published: (2024)
by: Banerjee, Oishi, et al.
Published: (2024)
The Impact of AI Assistance on Radiology Reporting: A Pilot Study Using Simulated AI Draft Reports
by: Acosta, Julián N., et al.
Published: (2024)
by: Acosta, Julián N., et al.
Published: (2024)
Uncovering Knowledge Gaps in Radiology Report Generation Models through Knowledge Graphs
by: Zhang, Xiaoman, et al.
Published: (2024)
by: Zhang, Xiaoman, et al.
Published: (2024)
ReXrank: A Public Leaderboard for AI-Powered Radiology Report Generation
by: Zhang, Xiaoman, et al.
Published: (2024)
by: Zhang, Xiaoman, et al.
Published: (2024)
MedVersa: A Generalist Foundation Model for Medical Image Interpretation
by: Zhou, Hong-Yu, et al.
Published: (2024)
by: Zhou, Hong-Yu, et al.
Published: (2024)
ReXTrust: A Model for Fine-Grained Hallucination Detection in AI-Generated Radiology Reports
by: Hardy, Romain, et al.
Published: (2024)
by: Hardy, Romain, et al.
Published: (2024)
FineRadScore: A Radiology Report Line-by-Line Evaluation Technique Generating Corrections with Severity Scores
by: Huang, Alyssa, et al.
Published: (2024)
by: Huang, Alyssa, et al.
Published: (2024)
CRIMSON: A Clinically-Grounded LLM-Based Metric for Generative Radiology Report Evaluation
by: Baharoon, Mohammed, et al.
Published: (2026)
by: Baharoon, Mohammed, et al.
Published: (2026)
ReXGradient-160K: A Large-Scale Publicly Available Dataset of Chest Radiographs with Free-text Reports
by: Zhang, Xiaoman, et al.
Published: (2025)
by: Zhang, Xiaoman, et al.
Published: (2025)
ReXamine-Global: A Framework for Uncovering Inconsistencies in Radiology Report Generation Metrics
by: Banerjee, Oishi, et al.
Published: (2024)
by: Banerjee, Oishi, et al.
Published: (2024)
Do Mixed-Vendor Multi-Agent LLMs Improve Clinical Diagnosis?
by: Yuan, Grace Chang, et al.
Published: (2026)
by: Yuan, Grace Chang, et al.
Published: (2026)
Voice-guided Orchestrated Intelligence for Clinical Evaluation (VOICE): A Voice AI Agent System for Prehospital Stroke Assessment
by: Acosta, Julian, et al.
Published: (2025)
by: Acosta, Julian, et al.
Published: (2025)
GREEN: Generative Radiology Report Evaluation and Error Notation
by: Ostmeier, Sophie, et al.
Published: (2024)
by: Ostmeier, Sophie, et al.
Published: (2024)
Error Correction in Radiology Reports: A Knowledge Distillation-Based Multi-Stage Framework
by: Wu, Jinge, et al.
Published: (2024)
by: Wu, Jinge, et al.
Published: (2024)
ReX-MLE: The Autonomous Agent Benchmark for Medical Imaging Challenges
by: Kenia, Roshan, et al.
Published: (2025)
by: Kenia, Roshan, et al.
Published: (2025)
Comparative Analysis of Abstractive Summarization Models for Clinical Radiology Reports
by: Bhattacharya, Anindita, et al.
Published: (2025)
by: Bhattacharya, Anindita, et al.
Published: (2025)
EvidenceOutcomes: a Dataset of Clinical Trial Publications with Clinically Meaningful Outcomes
by: Zhou, Yiliang, et al.
Published: (2025)
by: Zhou, Yiliang, et al.
Published: (2025)
Exploring Multimodal Large Language Models for Radiology Report Error-checking
by: Wu, Jinge, et al.
Published: (2023)
by: Wu, Jinge, et al.
Published: (2023)
Generative Large Language Models Trained for Detecting Errors in Radiology Reports
by: Sun, Cong, et al.
Published: (2025)
by: Sun, Cong, et al.
Published: (2025)
CCS: Clinical Consensus Selection for Radiology Report Generation
by: Zhang, Xi, et al.
Published: (2026)
by: Zhang, Xi, et al.
Published: (2026)
RadFlag: A Black-Box Hallucination Detection Method for Medical Vision Language Models
by: Zhang, Serena, et al.
Published: (2024)
by: Zhang, Serena, et al.
Published: (2024)
Calibrated Confidence Expression for Radiology Report Generation
by: Bani-Harouni, David, et al.
Published: (2026)
by: Bani-Harouni, David, et al.
Published: (2026)
CLEAR: A Clinically-Grounded Tabular Framework for Radiology Report Evaluation
by: Jiang, Yuyang, et al.
Published: (2025)
by: Jiang, Yuyang, et al.
Published: (2025)
RaTEScore: A Metric for Radiology Report Generation
by: Zhao, Weike, et al.
Published: (2024)
by: Zhao, Weike, et al.
Published: (2024)
ReXVQA: A Large-scale Visual Question Answering Benchmark for Generalist Chest X-ray Understanding
by: Pal, Ankit, et al.
Published: (2025)
by: Pal, Ankit, et al.
Published: (2025)
Clinically Grounded Agent-based Report Evaluation: An Interpretable Metric for Radiology Report Generation
by: Dua, Radhika, et al.
Published: (2025)
by: Dua, Radhika, et al.
Published: (2025)
CRG Score: A Distribution-Aware Clinical Metric for Radiology Report Generation
by: Hamamci, Ibrahim Ethem, et al.
Published: (2025)
by: Hamamci, Ibrahim Ethem, et al.
Published: (2025)
CLARIFID: Improving Radiology Report Generation by Reinforcing Clinically Accurate Impressions and Enforcing Detailed Findings
by: Lee, Kyeongkyu, et al.
Published: (2025)
by: Lee, Kyeongkyu, et al.
Published: (2025)
Modeling Clinical Uncertainty in Radiology Reports: from Explicit Uncertainty Markers to Implicit Reasoning Pathways
by: Rabaey, Paloma, et al.
Published: (2025)
by: Rabaey, Paloma, et al.
Published: (2025)
ReFINE: A Reward-Based Framework for Interpretable and Nuanced Evaluation of Radiology Report Generation
by: Liu, Yunyi, et al.
Published: (2024)
by: Liu, Yunyi, et al.
Published: (2024)
Automated Structured Radiology Report Generation
by: Delbrouck, Jean-Benoit, et al.
Published: (2025)
by: Delbrouck, Jean-Benoit, et al.
Published: (2025)
LLM-GLOBE: A Benchmark Evaluating the Cultural Values Embedded in LLM Output
by: Karinshak, Elise, et al.
Published: (2024)
by: Karinshak, Elise, et al.
Published: (2024)
a2z-1 for Multi-Disease Detection in Abdomen-Pelvis CT: External Validation and Performance Analysis Across 21 Conditions
by: Rajpurkar, Pranav, et al.
Published: (2024)
by: Rajpurkar, Pranav, et al.
Published: (2024)
Multimodal Foundation Models Exploit Text to Make Medical Image Predictions
by: Buckley, Thomas, et al.
Published: (2023)
by: Buckley, Thomas, et al.
Published: (2023)
BLUEmed: Retrieval-Augmented Multi-Agent Debate for Clinical Error Detection
by: You, Saukun Thika, et al.
Published: (2026)
by: You, Saukun Thika, et al.
Published: (2026)
Clinical Context-aware Radiology Report Generation from Medical Images using Transformers
by: Singh, Sonit
Published: (2024)
by: Singh, Sonit
Published: (2024)
DiReCT: Diagnostic Reasoning for Clinical Notes via Large Language Models
by: Wang, Bowen, et al.
Published: (2024)
by: Wang, Bowen, et al.
Published: (2024)
A Multi-Pass Large Language Model Framework for Precise and Efficient Radiology Report Error Detection
by: Kim, Songsoo, et al.
Published: (2025)
by: Kim, Songsoo, et al.
Published: (2025)
Hidden Error Awareness in Chain-of-Thought Reasoning: The Signal Is Diagnostic, Not Causal
by: Yuan, Aojie, et al.
Published: (2026)
by: Yuan, Aojie, et al.
Published: (2026)
Semantic Similarity in Radiology Reports via LLMs and NER
by: Pearson, Beth, et al.
Published: (2025)
by: Pearson, Beth, et al.
Published: (2025)
Similar Items
-
Direct Preference Optimization for Suppressing Hallucinated Prior Exams in Radiology Report Generation
by: Banerjee, Oishi, et al.
Published: (2024) -
The Impact of AI Assistance on Radiology Reporting: A Pilot Study Using Simulated AI Draft Reports
by: Acosta, Julián N., et al.
Published: (2024) -
Uncovering Knowledge Gaps in Radiology Report Generation Models through Knowledge Graphs
by: Zhang, Xiaoman, et al.
Published: (2024) -
ReXrank: A Public Leaderboard for AI-Powered Radiology Report Generation
by: Zhang, Xiaoman, et al.
Published: (2024) -
MedVersa: A Generalist Foundation Model for Medical Image Interpretation
by: Zhou, Hong-Yu, et al.
Published: (2024)