CheXmix: Unified Generative Pretraining for Vision Language Models in Medical Imaging
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Ashwin, Holland, Robbie, Barrett, Corey, Kim, Jangwon, Varma, Maya, Chen, Zhihong, Gao, Yunhe, Zaharchuk, Greg, Taghavi, Tara, Kenthapadi, Krishnaram, Chaudhari, Akshay |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RadAgents: Multimodal Agentic Reasoning for Chest X-ray Interpretation with Radiologist-like Workflows
by: Zhang, Kai, et al.
Published: (2025)
by: Zhang, Kai, et al.
Published: (2025)
Optimizing Long-Form Clinical Text Generation with Claim-Based Rewards
by: Jhaveri, Samyak, et al.
Published: (2025)
by: Jhaveri, Samyak, et al.
Published: (2025)
OG-Rank: Learning to Rank Fast and Slow with Uncertainty and Reward-Trend Guided Adaptive Exploration
by: Singh, Praphul, et al.
Published: (2025)
by: Singh, Praphul, et al.
Published: (2025)
RaVL: Discovering and Mitigating Spurious Correlations in Fine-Tuned Vision-Language Models
by: Varma, Maya, et al.
Published: (2024)
by: Varma, Maya, et al.
Published: (2024)
Grounding and Evaluation for Large Language Models: Practical Challenges and Lessons Learned (Survey)
by: Kenthapadi, Krishnaram, et al.
Published: (2024)
by: Kenthapadi, Krishnaram, et al.
Published: (2024)
JEDA: Query-Free Clinical Order Search from Ambient Dialogues
by: Singh, Praphul, et al.
Published: (2025)
by: Singh, Praphul, et al.
Published: (2025)
RedactOR: An LLM-Powered Framework for Automatic Clinical Data De-Identification
by: Singh, Praphul, et al.
Published: (2025)
by: Singh, Praphul, et al.
Published: (2025)
Learning Generalizable 3D Medical Image Representations from Mask-Guided Self-Supervision
by: Gao, Yunhe, et al.
Published: (2026)
by: Gao, Yunhe, et al.
Published: (2026)
CheXalign: Preference fine-tuning in chest X-ray interpretation models without human feedback
by: Hein, Dennis, et al.
Published: (2024)
by: Hein, Dennis, et al.
Published: (2024)
TRoVe: Discovering Error-Inducing Static Feature Biases in Temporal Vision-Language Models
by: Varma, Maya, et al.
Published: (2025)
by: Varma, Maya, et al.
Published: (2025)
MedVAE: Efficient Automated Interpretation of Medical Images with Large-Scale Generalizable Autoencoders
by: Varma, Maya, et al.
Published: (2025)
by: Varma, Maya, et al.
Published: (2025)
Permissioned LLMs: Enforcing Access Control in Large Language Models
by: Jayaraman, Bargav, et al.
Published: (2025)
by: Jayaraman, Bargav, et al.
Published: (2025)
From Attribution to Abstention: Training-Free Attention-Based Auditing for Clinical Summarization
by: Yan, Qianqi, et al.
Published: (2026)
by: Yan, Qianqi, et al.
Published: (2026)
Foundation Models in Radiology: What, How, When, Why and Why Not
by: Paschali, Magdalini, et al.
Published: (2024)
by: Paschali, Magdalini, et al.
Published: (2024)
Balancing Safety and Helpfulness in Healthcare AI Assistants through Iterative Preference Alignment
by: Nghiem, Huy, et al.
Published: (2025)
by: Nghiem, Huy, et al.
Published: (2025)
LieRE: Lie Rotational Positional Encodings
by: Ostmeier, Sophie, et al.
Published: (2024)
by: Ostmeier, Sophie, et al.
Published: (2024)
Prompt Triage: Structured Optimization Enhances Vision-Language Model Performance on Medical Imaging Benchmarks
by: Singhvi, Arnav, et al.
Published: (2025)
by: Singhvi, Arnav, et al.
Published: (2025)
Time-to-Event Pretraining for 3D Medical Imaging
by: Huo, Zepeng, et al.
Published: (2024)
by: Huo, Zepeng, et al.
Published: (2024)
Sparse Autoencoders for Interpretable Medical Image Representation Learning
by: Wesp, Philipp, et al.
Published: (2026)
by: Wesp, Philipp, et al.
Published: (2026)
Evaluation of Image‐Level Harmonization Methods for Multi‐Center MR Neuroimaging
by: Brandon C. Ho, et al.
Published: (2026)
by: Brandon C. Ho, et al.
Published: (2026)
CheXpert Plus: Augmenting a Large Chest X-ray Dataset with Text Radiology Reports, Patient Demographics and Additional Image Formats
by: Chambon, Pierre, et al.
Published: (2024)
by: Chambon, Pierre, et al.
Published: (2024)
Activation Matters: Test-time Activated Negative Labels for OOD Detection with Vision-Language Models
by: Zhang, Yabin, et al.
Published: (2026)
by: Zhang, Yabin, et al.
Published: (2026)
Merlin: A Computed Tomography Vision-Language Foundation Model and Dataset
by: Blankemeier, Louis, et al.
Published: (2024)
by: Blankemeier, Louis, et al.
Published: (2024)
GREEN: Generative Radiology Report Evaluation and Error Notation
by: Ostmeier, Sophie, et al.
Published: (2024)
by: Ostmeier, Sophie, et al.
Published: (2024)
A data- and compute-efficient chest X-ray foundation model beyond aggressive scaling
by: Wang, Chong, et al.
Published: (2026)
by: Wang, Chong, et al.
Published: (2026)
Even universal sums of triangular numbers
by: Ju, Jangwon
Published: (2024)
by: Ju, Jangwon
Published: (2024)
Landscape changes in a coastal system undergoing tourism development: implications for Barra de Navidad Lagoon, Jalisco, Mexico
by: Tara L. Holland
Published: (2011)
by: Tara L. Holland
Published: (2011)
Attention Head Entropy of LLMs Predicts Answer Correctness
by: Ostmeier, Sophie, et al.
Published: (2026)
by: Ostmeier, Sophie, et al.
Published: (2026)
Score-based Generative Diffusion Models to Synthesize Full-dose FDG Brain PET from MRI in Epilepsy Patients
by: Wu, Jiaqi, et al.
Published: (2025)
by: Wu, Jiaqi, et al.
Published: (2025)
Deep Learning-Based Prediction of PET Amyloid Status Using Multi-Contrast MRI
by: Kim, Donghoon, et al.
Published: (2024)
by: Kim, Donghoon, et al.
Published: (2024)
Medical Vision Language Models as Policies for Robotic Surgery
by: Muppidi, Akshay, et al.
Published: (2025)
by: Muppidi, Akshay, et al.
Published: (2025)
Toward expanding the scope of radiology report summarization to multiple anatomies and modalities
by: Chen, Zhihong, et al.
Published: (2022)
by: Chen, Zhihong, et al.
Published: (2022)
BenchX: A Unified Benchmark Framework for Medical Vision-Language Pretraining on Chest X-Rays
by: Zhou, Yang, et al.
Published: (2024)
by: Zhou, Yang, et al.
Published: (2024)
Anatomy-VLM: A Fine-grained Vision-Language Model for Medical Interpretation
by: Gu, Difei, et al.
Published: (2025)
by: Gu, Difei, et al.
Published: (2025)
Towards the Use of Saliency Maps for Explaining Low-Quality Electrocardiograms to End Users
by: Lucic, Ana, et al.
Published: (2022)
by: Lucic, Ana, et al.
Published: (2022)
Daily large hail probability on a global scale (1979 to 2015), Version 2, link to netCDF files
by: Prein, Andreas F, et al.
Published: (2018)
by: Prein, Andreas F, et al.
Published: (2018)
Daily gridded hail risk estimates on a global scale (1979 to 2015), link to netCDF files
by: Prein, Andreas F, et al.
Published: (2018)
by: Prein, Andreas F, et al.
Published: (2018)
Unified Local and Global Attention Interaction Modeling for Vision Transformers
by: Nguyen, Tan, et al.
Published: (2024)
by: Nguyen, Tan, et al.
Published: (2024)
Bounding-Box Trajectories Matter for Video Anomaly Detection
by: Song, Inpyo, et al.
Published: (2026)
by: Song, Inpyo, et al.
Published: (2026)
Real-time Traffic Accident Anticipation with Feature Reuse
by: Song, Inpyo, et al.
Published: (2025)
by: Song, Inpyo, et al.
Published: (2025)
Similar Items
-
RadAgents: Multimodal Agentic Reasoning for Chest X-ray Interpretation with Radiologist-like Workflows
by: Zhang, Kai, et al.
Published: (2025) -
Optimizing Long-Form Clinical Text Generation with Claim-Based Rewards
by: Jhaveri, Samyak, et al.
Published: (2025) -
OG-Rank: Learning to Rank Fast and Slow with Uncertainty and Reward-Trend Guided Adaptive Exploration
by: Singh, Praphul, et al.
Published: (2025) -
RaVL: Discovering and Mitigating Spurious Correlations in Fine-Tuned Vision-Language Models
by: Varma, Maya, et al.
Published: (2024) -
Grounding and Evaluation for Large Language Models: Practical Challenges and Lessons Learned (Survey)
by: Kenthapadi, Krishnaram, et al.
Published: (2024)