Process Reward Models for Sentence-Level Verification of LVLM Radiology Reports
Fuente:
arXiv
Saved in:
| Main Authors: | Thomas, Alois, Varma, Maya, Delbrouck, Jean-Benoit, Langlotz, Curtis P. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GREEN: Generative Radiology Report Evaluation and Error Notation
by: Ostmeier, Sophie, et al.
Published: (2024)
by: Ostmeier, Sophie, et al.
Published: (2024)
CheXpert Plus: Augmenting a Large Chest X-ray Dataset with Text Radiology Reports, Patient Demographics and Additional Image Formats
by: Chambon, Pierre, et al.
Published: (2024)
by: Chambon, Pierre, et al.
Published: (2024)
RaVL: Discovering and Mitigating Spurious Correlations in Fine-Tuned Vision-Language Models
by: Varma, Maya, et al.
Published: (2024)
by: Varma, Maya, et al.
Published: (2024)
Toward expanding the scope of radiology report summarization to multiple anatomies and modalities
by: Chen, Zhihong, et al.
Published: (2022)
by: Chen, Zhihong, et al.
Published: (2022)
Automated Structured Radiology Report Generation
by: Delbrouck, Jean-Benoit, et al.
Published: (2025)
by: Delbrouck, Jean-Benoit, et al.
Published: (2025)
Activation Matters: Test-time Activated Negative Labels for OOD Detection with Vision-Language Models
by: Zhang, Yabin, et al.
Published: (2026)
by: Zhang, Yabin, et al.
Published: (2026)
Structuring Radiology Reports: Challenging LLMs with Lightweight Models
by: Moll, Johannes, et al.
Published: (2025)
by: Moll, Johannes, et al.
Published: (2025)
RadDiff: Describing Differences in Radiology Image Sets with Natural Language
by: Shen, Xiaoxian, et al.
Published: (2026)
by: Shen, Xiaoxian, et al.
Published: (2026)
Improving the Performance of Radiology Report De-identification with Large-Scale Training and Benchmarking Against Cloud Vendor Methods
by: Prakash, Eva, et al.
Published: (2025)
by: Prakash, Eva, et al.
Published: (2025)
TRoVe: Discovering Error-Inducing Static Feature Biases in Temporal Vision-Language Models
by: Varma, Maya, et al.
Published: (2025)
by: Varma, Maya, et al.
Published: (2025)
From Detection to Mitigation: Addressing Bias in Deep Learning Models for Chest X-Ray Diagnosis
by: Mottez, Clemence, et al.
Published: (2025)
by: Mottez, Clemence, et al.
Published: (2025)
MRScore: Evaluating Radiology Report Generation with LLM-based Reward System
by: Liu, Yunyi, et al.
Published: (2024)
by: Liu, Yunyi, et al.
Published: (2024)
Beyond Outcome Verification: Verifiable Process Reward Models for Structured Reasoning
by: Pronesti, Massimiliano, et al.
Published: (2026)
by: Pronesti, Massimiliano, et al.
Published: (2026)
Multilingual Natural Language Processing Model for Radiology Reports -- The Summary is all you need!
by: Lindo, Mariana, et al.
Published: (2023)
by: Lindo, Mariana, et al.
Published: (2023)
ReFINE: A Reward-Based Framework for Interpretable and Nuanced Evaluation of Radiology Report Generation
by: Liu, Yunyi, et al.
Published: (2024)
by: Liu, Yunyi, et al.
Published: (2024)
VERT: Reliable LLM Judges for Radiology Report Evaluation
by: Bologna, Federica, et al.
Published: (2026)
by: Bologna, Federica, et al.
Published: (2026)
Comparing Human and Language Models Sentence Processing Difficulties on Complex Structures
by: Amouyal, Samuel Joseph, et al.
Published: (2025)
by: Amouyal, Samuel Joseph, et al.
Published: (2025)
Summarizing Radiology Reports Findings into Impressions
by: de Padua, Raul Salles, et al.
Published: (2024)
by: de Padua, Raul Salles, et al.
Published: (2024)
Radiology-GPT: A Large Language Model for Radiology
by: Liu, Zhengliang, et al.
Published: (2023)
by: Liu, Zhengliang, et al.
Published: (2023)
PARROT: An Open Multilingual Radiology Reports Dataset
by: Guellec, Bastien Le, et al.
Published: (2025)
by: Guellec, Bastien Le, et al.
Published: (2025)
Contextual Refinement of Translations: Large Language Models for Sentence and Document-Level Post-Editing
by: Koneru, Sai, et al.
Published: (2023)
by: Koneru, Sai, et al.
Published: (2023)
Generative Large Language Models Trained for Detecting Errors in Radiology Reports
by: Sun, Cong, et al.
Published: (2025)
by: Sun, Cong, et al.
Published: (2025)
Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities
by: Liu, Zhichen, et al.
Published: (2026)
by: Liu, Zhichen, et al.
Published: (2026)
Standardizing Longitudinal Radiology Report Evaluation via Large Language Model Annotation
by: Wang, Xinyi, et al.
Published: (2026)
by: Wang, Xinyi, et al.
Published: (2026)
Improving Radiology Report Conciseness and Structure via Local Large Language Models
by: Hartsock, Iryna, et al.
Published: (2024)
by: Hartsock, Iryna, et al.
Published: (2024)
PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models
by: Song, Mingyang, et al.
Published: (2025)
by: Song, Mingyang, et al.
Published: (2025)
Detecting Sexual Content at the Sentence Level in First Millennium Latin Texts
by: Clérice, Thibault
Published: (2023)
by: Clérice, Thibault
Published: (2023)
Process-based Self-Rewarding Language Models
by: Zhang, Shimao, et al.
Published: (2025)
by: Zhang, Shimao, et al.
Published: (2025)
Process Reward Model with Q-Value Rankings
by: Li, Wendi, et al.
Published: (2024)
by: Li, Wendi, et al.
Published: (2024)
Intertwining CP and NLP: The Generation of Unreasonably Constrained Sentences
by: Bonlarron, Alexandre, et al.
Published: (2024)
by: Bonlarron, Alexandre, et al.
Published: (2024)
SentenceKV: Efficient LLM Inference via Sentence-Level Semantic KV Caching
by: Zhu, Yuxuan, et al.
Published: (2025)
by: Zhu, Yuxuan, et al.
Published: (2025)
Rewarding the Scientific Process: Process-Level Reward Modeling for Agentic Data Analysis
by: Qiu, Zhisong, et al.
Published: (2026)
by: Qiu, Zhisong, et al.
Published: (2026)
MGH Radiology Llama: A Llama 3 70B Model for Radiology
by: Shi, Yucheng, et al.
Published: (2024)
by: Shi, Yucheng, et al.
Published: (2024)
On the Relationship between Sentence Analogy Identification and Sentence Structure Encoding in Large Language Models
by: Wijesiriwardene, Thilini, et al.
Published: (2023)
by: Wijesiriwardene, Thilini, et al.
Published: (2023)
Improving Automatic Summarization of Radiology Reports through Mid-Training of Large Language Models
by: Lyu, Mengxian, et al.
Published: (2026)
by: Lyu, Mengxian, et al.
Published: (2026)
Development and Validation of a Large Language Model for Generating Fully-Structured Radiology Reports
by: Niu, Chuang, et al.
Published: (2024)
by: Niu, Chuang, et al.
Published: (2024)
Fine-Grained Detection of AI-Generated Text Using Sentence-Level Segmentation
by: Teja, Lekkala Sai, et al.
Published: (2025)
by: Teja, Lekkala Sai, et al.
Published: (2025)
Rubric-Guided Process Reward for Stepwise Model Routing
by: Ye, Shenghao, et al.
Published: (2026)
by: Ye, Shenghao, et al.
Published: (2026)
Coarse-to-Fine Personalized LLM Impressions for Streamlined Radiology Reports
by: Sun, Chengbo, et al.
Published: (2025)
by: Sun, Chengbo, et al.
Published: (2025)
Error Typing for Smarter Rewards: Improving Process Reward Models with Error-Aware Hierarchical Supervision
by: Pala, Tej Deep, et al.
Published: (2025)
by: Pala, Tej Deep, et al.
Published: (2025)
Similar Items
-
GREEN: Generative Radiology Report Evaluation and Error Notation
by: Ostmeier, Sophie, et al.
Published: (2024) -
CheXpert Plus: Augmenting a Large Chest X-ray Dataset with Text Radiology Reports, Patient Demographics and Additional Image Formats
by: Chambon, Pierre, et al.
Published: (2024) -
RaVL: Discovering and Mitigating Spurious Correlations in Fine-Tuned Vision-Language Models
by: Varma, Maya, et al.
Published: (2024) -
Toward expanding the scope of radiology report summarization to multiple anatomies and modalities
by: Chen, Zhihong, et al.
Published: (2022) -
Automated Structured Radiology Report Generation
by: Delbrouck, Jean-Benoit, et al.
Published: (2025)