An Investigation of Memorization Risk in Healthcare Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Tonekaboni, Sana, Stempfle, Lena, Fallahpour, Adibvafa, Gerych, Walter, Ghassemi, Marzyeh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When Style Breaks Safety: Defending LLMs Against Superficial Style Alignment
by: Xiao, Yuxin, et al.
Published: (2025)
by: Xiao, Yuxin, et al.
Published: (2025)
MaskMedPaint: Masked Medical Image Inpainting with Diffusion Models for Mitigation of Spurious Correlations
by: Jin, Qixuan, et al.
Published: (2024)
by: Jin, Qixuan, et al.
Published: (2024)
Learning under Temporal Label Noise
by: Nagaraj, Sujay, et al.
Published: (2024)
by: Nagaraj, Sujay, et al.
Published: (2024)
EHRMamba: Towards Generalizable and Scalable Foundation Models for Electronic Health Records
by: Fallahpour, Adibvafa, et al.
Published: (2024)
by: Fallahpour, Adibvafa, et al.
Published: (2024)
BendVLM: Test-Time Debiasing of Vision-Language Embeddings
by: Gerych, Walter, et al.
Published: (2024)
by: Gerych, Walter, et al.
Published: (2024)
Evaluation without Generation: Non-Generative Assessment of Harmful Model Specialization with Applications to CSAM
by: Suriyakumar, Vinith M., et al.
Published: (2026)
by: Suriyakumar, Vinith M., et al.
Published: (2026)
MedRAX: Medical Reasoning Agent for Chest X-ray
by: Fallahpour, Adibvafa, et al.
Published: (2025)
by: Fallahpour, Adibvafa, et al.
Published: (2025)
The MedPerturb Dataset: What Non-Content Perturbations Reveal About Human and Clinical LLM Decision Making
by: Gourabathina, Abinitha, et al.
Published: (2025)
by: Gourabathina, Abinitha, et al.
Published: (2025)
Measuring Stochastic Data Complexity with Boltzmann Influence Functions
by: Ng, Nathan, et al.
Published: (2024)
by: Ng, Nathan, et al.
Published: (2024)
dnaHNet: A Scalable and Hierarchical Foundation Model for Genomic Sequence Learning
by: Shah, Arnav, et al.
Published: (2026)
by: Shah, Arnav, et al.
Published: (2026)
Identifying Implicit Social Biases in Vision-Language Models
by: Hamidieh, Kimia, et al.
Published: (2024)
by: Hamidieh, Kimia, et al.
Published: (2024)
Complementing Self-Consistency with Cross-Model Disagreement for Uncertainty Quantification
by: Hamidieh, Kimia, et al.
Published: (2026)
by: Hamidieh, Kimia, et al.
Published: (2026)
Prediction Models That Learn to Avoid Missing Values
by: Stempfle, Lena, et al.
Published: (2025)
by: Stempfle, Lena, et al.
Published: (2025)
DynaSubVAE: Adaptive Subgrouping for Scalable and Robust OOD Detection
by: Behrouzi, Tina, et al.
Published: (2025)
by: Behrouzi, Tina, et al.
Published: (2025)
Asymmetry in Low-Rank Adapters of Foundation Models
by: Zhu, Jiacheng, et al.
Published: (2024)
by: Zhu, Jiacheng, et al.
Published: (2024)
Aggregation Hides Out-of-Distribution Generalization Failures from Spurious Correlations
by: Salaudeen, Olawale, et al.
Published: (2025)
by: Salaudeen, Olawale, et al.
Published: (2025)
Robustness Beyond Known Groups with Low-rank Adaptation
by: Gourabathina, Abinitha, et al.
Published: (2026)
by: Gourabathina, Abinitha, et al.
Published: (2026)
Views Can Be Deceiving: Improved SSL Through Feature Space Augmentation
by: Hamidieh, Kimia, et al.
Published: (2024)
by: Hamidieh, Kimia, et al.
Published: (2024)
SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe
by: Xiao, Yuxin, et al.
Published: (2024)
by: Xiao, Yuxin, et al.
Published: (2024)
What's in a Query: Polarity-Aware Distribution-Based Fair Ranking
by: Balagopalan, Aparna, et al.
Published: (2025)
by: Balagopalan, Aparna, et al.
Published: (2025)
Handling missing values in clinical machine learning: Insights from an expert study
by: Stempfle, Lena, et al.
Published: (2024)
by: Stempfle, Lena, et al.
Published: (2024)
Speak Easy: Eliciting Harmful Jailbreaks from LLMs with Simple Interactions
by: Chan, Yik Siu, et al.
Published: (2025)
by: Chan, Yik Siu, et al.
Published: (2025)
In the Name of Fairness: Assessing the Bias in Clinical Record De-identification
by: Xiao, Yuxin, et al.
Published: (2023)
by: Xiao, Yuxin, et al.
Published: (2023)
KScope: A Framework for Characterizing the Knowledge Status of Language Models
by: Xiao, Yuxin, et al.
Published: (2025)
by: Xiao, Yuxin, et al.
Published: (2025)
Reaching Beyond the Mode: RL for Distributional Reasoning in Language Models
by: Puri, Isha, et al.
Published: (2026)
by: Puri, Isha, et al.
Published: (2026)
Improving Mutual Information Estimation with Annealed and Energy-Based Bounds
by: Brekelmans, Rob, et al.
Published: (2023)
by: Brekelmans, Rob, et al.
Published: (2023)
Data Debiasing with Datamodels (D3M): Improving Subgroup Robustness via Data Selection
by: Jain, Saachi, et al.
Published: (2024)
by: Jain, Saachi, et al.
Published: (2024)
Event-Based Contrastive Learning for Medical Time Series
by: Jeong, Hyewon, et al.
Published: (2023)
by: Jeong, Hyewon, et al.
Published: (2023)
BioReason: Incentivizing Multimodal Biological Reasoning within a DNA-LLM Model
by: Fallahpour, Adibvafa, et al.
Published: (2025)
by: Fallahpour, Adibvafa, et al.
Published: (2025)
How Should We Represent History in Interpretable Models of Clinical Policies?
by: Matsson, Anton, et al.
Published: (2024)
by: Matsson, Anton, et al.
Published: (2024)
Improving Black-box Robustness with In-Context Rewriting
by: O'Brien, Kyle, et al.
Published: (2024)
by: O'Brien, Kyle, et al.
Published: (2024)
LEMoN: Label Error Detection using Multimodal Neighbors
by: Zhang, Haoran, et al.
Published: (2024)
by: Zhang, Haoran, et al.
Published: (2024)
Quantifying Memorization and Privacy Risks in Genomic Language Models
by: Nemecek, Alexander, et al.
Published: (2026)
by: Nemecek, Alexander, et al.
Published: (2026)
An Information Criterion for Controlled Disentanglement of Multimodal Data
by: Wang, Chenyu, et al.
Published: (2024)
by: Wang, Chenyu, et al.
Published: (2024)
Machine Learning for Health symposium 2024 -- Findings track
by: Hegselmann, Stefan, et al.
Published: (2025)
by: Hegselmann, Stefan, et al.
Published: (2025)
Batch Normalization Amplifies Memorization and Privacy Risks
by: Doan, Ngoc Phu, et al.
Published: (2026)
by: Doan, Ngoc Phu, et al.
Published: (2026)
On the Edge of Memorization in Diffusion Models
by: Buchanan, Sam, et al.
Published: (2025)
by: Buchanan, Sam, et al.
Published: (2025)
TransConv-DDPM: Enhanced Diffusion Model for Generating Time-Series Data in Healthcare
by: Kabir, Md Shahriar, et al.
Published: (2026)
by: Kabir, Md Shahriar, et al.
Published: (2026)
Advancing Medical Representation Learning Through High-Quality Data
by: Baghbanzadeh, Negin, et al.
Published: (2025)
by: Baghbanzadeh, Negin, et al.
Published: (2025)
Train Long, Think Short: Curriculum Learning for Efficient Reasoning
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
Similar Items
-
When Style Breaks Safety: Defending LLMs Against Superficial Style Alignment
by: Xiao, Yuxin, et al.
Published: (2025) -
MaskMedPaint: Masked Medical Image Inpainting with Diffusion Models for Mitigation of Spurious Correlations
by: Jin, Qixuan, et al.
Published: (2024) -
Learning under Temporal Label Noise
by: Nagaraj, Sujay, et al.
Published: (2024) -
EHRMamba: Towards Generalizable and Scalable Foundation Models for Electronic Health Records
by: Fallahpour, Adibvafa, et al.
Published: (2024) -
BendVLM: Test-Time Debiasing of Vision-Language Embeddings
by: Gerych, Walter, et al.
Published: (2024)