Enhancing Radiology Report Generation and Visual Grounding using Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Gundersen, Benjamin, Deperrois, Nicolas, Ruiperez-Campillo, Samuel, Sutter, Thomas M., Vogt, Julia E., Moor, Michael, Nooralahzadeh, Farhad, Krauthammer, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Structure is Supervision: Multiview Masked Autoencoders for Radiology
by: Laguna, Sonia, et al.
Published: (2025)
by: Laguna, Sonia, et al.
Published: (2025)
Leveraging the Structure of Medical Data for Improved Representation Learning
by: Agostini, Andrea, et al.
Published: (2025)
by: Agostini, Andrea, et al.
Published: (2025)
RadVLM: A Multitask Conversational Vision-Language Model for Radiology
by: Deperrois, Nicolas, et al.
Published: (2025)
by: Deperrois, Nicolas, et al.
Published: (2025)
Agentic Systems in Radiology: Design, Applications, Evaluation, and Challenges
by: Bluethgen, Christian, et al.
Published: (2025)
by: Bluethgen, Christian, et al.
Published: (2025)
Towards Scalable and Cross-Lingual Specialist Language Models for Oncology
by: Rohanian, Morteza, et al.
Published: (2025)
by: Rohanian, Morteza, et al.
Published: (2025)
Predicting Pulmonary Hypertension in Newborns: A Multi-view VAE Approach
by: Erlacher, Lucas, et al.
Published: (2025)
by: Erlacher, Lucas, et al.
Published: (2025)
Foundation Model for Cardiac Time Series via Masked Latent Attention
by: Vandenhirtz, Moritz, et al.
Published: (2026)
by: Vandenhirtz, Moritz, et al.
Published: (2026)
Uncertainty Modeling in Multimodal Speech Analysis Across the Psychosis Spectrum
by: Rohanian, Morteza, et al.
Published: (2025)
by: Rohanian, Morteza, et al.
Published: (2025)
Arbitration Failure, Not Perceptual Blindness: How Vision-Language Models Resolve Visual-Linguistic Conflicts
by: Nooralahzadeh, Farhad, et al.
Published: (2026)
by: Nooralahzadeh, Farhad, et al.
Published: (2026)
Beyond Independent Frames: Latent Attention Masked Autoencoders for Multi-View Echocardiography
by: Böhi, Simon, et al.
Published: (2026)
by: Böhi, Simon, et al.
Published: (2026)
MAIRA-2: Grounded Radiology Report Generation
by: Bannur, Shruthi, et al.
Published: (2024)
by: Bannur, Shruthi, et al.
Published: (2024)
Multi-Modal Data Exploration via Language Agents
by: Nooralahzadeh, Farhad, et al.
Published: (2024)
by: Nooralahzadeh, Farhad, et al.
Published: (2024)
Visual Alignment of Medical Vision-Language Models for Grounded Radiology Report Generation
by: Bose, Sarosij, et al.
Published: (2025)
by: Bose, Sarosij, et al.
Published: (2025)
Grounding Chest X-Ray Visual Question Answering with Generated Radiology Reports
by: Serra, Francesco Dalla, et al.
Published: (2025)
by: Serra, Francesco Dalla, et al.
Published: (2025)
From Pixels to Components: Eigenvector Masking for Visual Representation Learning
by: Bizeul, Alice, et al.
Published: (2025)
by: Bizeul, Alice, et al.
Published: (2025)
Optimizing Speech Language Models for Acoustic Consistency
by: Rohanian, Morteza, et al.
Published: (2025)
by: Rohanian, Morteza, et al.
Published: (2025)
AgentRxiv: Towards Collaborative Autonomous Research
by: Schmidgall, Samuel, et al.
Published: (2025)
by: Schmidgall, Samuel, et al.
Published: (2025)
Temporal Representation Learning for Real-Time Ultrasound Analysis
by: Stebler, Yves, et al.
Published: (2025)
by: Stebler, Yves, et al.
Published: (2025)
Rethinking the Efficiency and Effectiveness of Reinforcement Learning for Radiology Report Generation
by: Lu, Zilin, et al.
Published: (2026)
by: Lu, Zilin, et al.
Published: (2026)
Enhancing Reinforcement Learning for Radiology Report Generation with Evidence-aware Rewards and Self-correcting Preference Learning
by: Zhou, Qin, et al.
Published: (2026)
by: Zhou, Qin, et al.
Published: (2026)
CLEAR: A Clinically-Grounded Tabular Framework for Radiology Report Evaluation
by: Jiang, Yuyang, et al.
Published: (2025)
by: Jiang, Yuyang, et al.
Published: (2025)
Evaluating the Data Model Robustness of Text-to-SQL Systems Based on Real User Queries
by: Fürst, Jonathan, et al.
Published: (2024)
by: Fürst, Jonathan, et al.
Published: (2024)
Grounded Reinforcement Learning for Visual Reasoning
by: Sarch, Gabriel, et al.
Published: (2025)
by: Sarch, Gabriel, et al.
Published: (2025)
Multi-Modal Multi-Agent Reinforcement Learning for Radiology Report Generation
by: Baba, Kaito, et al.
Published: (2026)
by: Baba, Kaito, et al.
Published: (2026)
Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation
by: Zhou, Qin, et al.
Published: (2025)
by: Zhou, Qin, et al.
Published: (2025)
Rethinking Molecular Design: Integrating Latent Variable and Auto-Regressive Models for Goal Directed Generation
by: Arthur-Loui, Heath, et al.
Published: (2024)
by: Arthur-Loui, Heath, et al.
Published: (2024)
The point of itall : a lifetime of great loves and endeavors / Charles Krauthammer ; edited by Daniel Krauthammer
by: Krauthammer, Charles
Published: (2018)
by: Krauthammer, Charles
Published: (2018)
RL-ACRGNet: Reinforcement Learning-Based Chest Radiology Report Generation Network
by: Meena, Yogesh Kumar, et al.
Published: (2026)
by: Meena, Yogesh Kumar, et al.
Published: (2026)
Large Model driven Radiology Report Generation with Clinical Quality Reinforcement Learning
by: Zhou, Zijian, et al.
Published: (2024)
by: Zhou, Zijian, et al.
Published: (2024)
S2D-ALIGN: Shallow-to-Deep Auxiliary Learning for Anatomically-Grounded Radiology Report Generation
by: Gao, Jiechao, et al.
Published: (2025)
by: Gao, Jiechao, et al.
Published: (2025)
Clinically Grounded Agent-based Report Evaluation: An Interpretable Metric for Radiology Report Generation
by: Dua, Radhika, et al.
Published: (2025)
by: Dua, Radhika, et al.
Published: (2025)
Two Is Better Than One: Aligned Representation Pairs for Anomaly Detection
by: Ryser, Alain, et al.
Published: (2024)
by: Ryser, Alain, et al.
Published: (2024)
TAMER: A Test-Time Adaptive MoE-Driven Framework for EHR Representation Learning
by: Zhu, Yinghao, et al.
Published: (2025)
by: Zhu, Yinghao, et al.
Published: (2025)
Improving Medical Visual Representations via Radiology Report Generation
by: Quigley, Keegan, et al.
Published: (2023)
by: Quigley, Keegan, et al.
Published: (2023)
Visual Grounding for Object-Level Generalization in Reinforcement Learning
by: Jiang, Haobin, et al.
Published: (2024)
by: Jiang, Haobin, et al.
Published: (2024)
The Role of Metacognitive Strategies in Blended Learning: Study Habits and Reading Comprehension
by: Beatriz Ortega Ruipérez
Published: (2022)
by: Beatriz Ortega Ruipérez
Published: (2022)
Dual-Phase Cross-Modal Contrastive Learning for CMR-Guided ECG Representations for Cardiovascular Disease Assessment
by: Alvarez-Florez, Laura, et al.
Published: (2026)
by: Alvarez-Florez, Laura, et al.
Published: (2026)
A Denoising VAE for Intracardiac Time Series in Ischemic Cardiomyopathy
by: Ruipérez-Campillo, Samuel, et al.
Published: (2025)
by: Ruipérez-Campillo, Samuel, et al.
Published: (2025)
OraPO: Oracle-educated Reinforcement Learning for Data-efficient and Factual Radiology Report Generation
by: Chen, Zhuoxiao, et al.
Published: (2025)
by: Chen, Zhuoxiao, et al.
Published: (2025)
Soft-Masked Diffusion Language Models
by: Hersche, Michael, et al.
Published: (2025)
by: Hersche, Michael, et al.
Published: (2025)
Similar Items
-
Structure is Supervision: Multiview Masked Autoencoders for Radiology
by: Laguna, Sonia, et al.
Published: (2025) -
Leveraging the Structure of Medical Data for Improved Representation Learning
by: Agostini, Andrea, et al.
Published: (2025) -
RadVLM: A Multitask Conversational Vision-Language Model for Radiology
by: Deperrois, Nicolas, et al.
Published: (2025) -
Agentic Systems in Radiology: Design, Applications, Evaluation, and Challenges
by: Bluethgen, Christian, et al.
Published: (2025) -
Towards Scalable and Cross-Lingual Specialist Language Models for Oncology
by: Rohanian, Morteza, et al.
Published: (2025)