DPF-CM: A Data Processing Framework with Privacy-Preserving Vector Databases for Chinese Medical LLMs Training and Deployment
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Wei, Cheng, Anda, Zhang, Zhao, Wang, Yinggui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLM-AutoDP: Automatic Data Processing via LLM Agents for Model Fine-tuning
by: Huang, Wei, et al.
Published: (2026)
by: Huang, Wei, et al.
Published: (2026)
GradPruner: Gradient-Guided Layer Pruning Enabling Efficient Fine-Tuning and Inference for LLMs
by: Huang, Wei, et al.
Published: (2026)
by: Huang, Wei, et al.
Published: (2026)
DaMoC: Efficiently Selecting the Optimal Large Language Model for Fine-tuning Domain Tasks Based on Data and Model Compression
by: Huang, Wei, et al.
Published: (2025)
by: Huang, Wei, et al.
Published: (2025)
Mitigating Catastrophic Forgetting in Large Language Models with Forgetting-aware Pruning
by: Huang, Wei, et al.
Published: (2025)
by: Huang, Wei, et al.
Published: (2025)
Data-Free Privacy-Preserving for LLMs via Model Inversion and Selective Unlearning
by: Zhou, Xinjie, et al.
Published: (2026)
by: Zhou, Xinjie, et al.
Published: (2026)
RL-Finetuned LLMs for Privacy-Preserving Synthetic Rewriting
by: Shi, Zhan, et al.
Published: (2025)
by: Shi, Zhan, et al.
Published: (2025)
Improving Epidemic Analyses with Privacy-Preserving Integration of Sensitive Data
by: Guan, Zihan, et al.
Published: (2025)
by: Guan, Zihan, et al.
Published: (2025)
Privacy-Preserving Federated Learning Framework for Distributed Chemical Process Optimization
by: Pipattaratonchai, Teetat, et al.
Published: (2026)
by: Pipattaratonchai, Teetat, et al.
Published: (2026)
Losing is for Cherishing: Data Valuation Based on Machine Unlearning and Shapley Value
by: Ma, Le, et al.
Published: (2025)
by: Ma, Le, et al.
Published: (2025)
A Middle Path for On-Premises LLM Deployment: Preserving Privacy Without Sacrificing Model Confidentiality
by: Huang, Hanbo, et al.
Published: (2024)
by: Huang, Hanbo, et al.
Published: (2024)
Preserving Vector Space Properties in Dimensionality Reduction: A Relationship Preserving Loss Framework
by: Weinwurm, Eddi, et al.
Published: (2025)
by: Weinwurm, Eddi, et al.
Published: (2025)
A Fast, Performant, Secure Distributed Training Framework For Large Language Model
by: Huang, Wei, et al.
Published: (2024)
by: Huang, Wei, et al.
Published: (2024)
Operationalizing Data Minimization for Privacy-Preserving LLM Prompting
by: Zhou, Jijie, et al.
Published: (2025)
by: Zhou, Jijie, et al.
Published: (2025)
VecLSTM: Trajectory Data Processing and Management for Activity Recognition through LSTM Vectorization and Database Integration
by: Monir, Solmaz Seyed, et al.
Published: (2024)
by: Monir, Solmaz Seyed, et al.
Published: (2024)
DISCO-TAB: A Hierarchical Reinforcement Learning Framework for Privacy-Preserving Synthesis of Complex Clinical Data
by: Ilaty, Arshia, et al.
Published: (2026)
by: Ilaty, Arshia, et al.
Published: (2026)
PAC Privacy Preserving Diffusion Models
by: Xu, Qipan, et al.
Published: (2023)
by: Xu, Qipan, et al.
Published: (2023)
Privacy-Preserving End-to-End Spoken Language Understanding
by: Wang, Yinggui, et al.
Published: (2024)
by: Wang, Yinggui, et al.
Published: (2024)
Federated Learning for Privacy-Preserving Medical AI
by: Hoang, Tin
Published: (2026)
by: Hoang, Tin
Published: (2026)
Edge-FIT: Federated Instruction Tuning of Quantized LLMs for Privacy-Preserving Smart Home Environments
by: Venkatesh, Vinay, et al.
Published: (2025)
by: Venkatesh, Vinay, et al.
Published: (2025)
Privacy-Preserving Heterogeneous Federated Learning for Sensitive Healthcare Data
by: Xu, Yukai, et al.
Published: (2024)
by: Xu, Yukai, et al.
Published: (2024)
Privacy-Preserving Personalized Federated Learning for Distributed Photovoltaic Disaggregation under Statistical Heterogeneity
by: Chen, Xiaolu, et al.
Published: (2025)
by: Chen, Xiaolu, et al.
Published: (2025)
Efficiently Deploying LLMs with Controlled Risk
by: Zellinger, Michael J., et al.
Published: (2024)
by: Zellinger, Michael J., et al.
Published: (2024)
Privacy-Preserving Dynamic Assortment Selection
by: Cho, Young Hyun, et al.
Published: (2024)
by: Cho, Young Hyun, et al.
Published: (2024)
How to Train Data-Efficient LLMs
by: Sachdeva, Noveen, et al.
Published: (2024)
by: Sachdeva, Noveen, et al.
Published: (2024)
Deploying Privacy Guardrails for LLMs: A Comparative Analysis of Real-World Applications
by: Asthana, Shubhi, et al.
Published: (2025)
by: Asthana, Shubhi, et al.
Published: (2025)
A Survey of Personalized Federated Foundation Models for Privacy-Preserving Recommendation
by: Li, Zhiwei, et al.
Published: (2025)
by: Li, Zhiwei, et al.
Published: (2025)
Adaptive Clipping for Privacy-Preserving Few-Shot Learning: Enhancing Generalization with Limited Data
by: Ranaweera, Kanishka, et al.
Published: (2025)
by: Ranaweera, Kanishka, et al.
Published: (2025)
Anonymous-by-Construction: An LLM-Driven Framework for Privacy-Preserving Text
by: Albanese, Federico, et al.
Published: (2026)
by: Albanese, Federico, et al.
Published: (2026)
Near-Optimal Online Deployment and Routing for Streaming LLMs
by: Li, Shaoang, et al.
Published: (2025)
by: Li, Shaoang, et al.
Published: (2025)
Open-Medical-R1: How to Choose Data for RLVR Training at Medicine Domain
by: Qiu, Zhongxi, et al.
Published: (2025)
by: Qiu, Zhongxi, et al.
Published: (2025)
KIPPS: Knowledge infusion in Privacy Preserving Synthetic Data Generation
by: Kotal, Anantaa, et al.
Published: (2024)
by: Kotal, Anantaa, et al.
Published: (2024)
Towards Privacy-Preserving Relational Data Synthesis via Probabilistic Relational Models
by: Luttermann, Malte, et al.
Published: (2024)
by: Luttermann, Malte, et al.
Published: (2024)
Deploying Atmospheric and Oceanic AI Models on Chinese Hardware and Framework: Migration Strategies, Performance Optimization and Analysis
by: Sun, Yuze, et al.
Published: (2025)
by: Sun, Yuze, et al.
Published: (2025)
Efficient Edge LLMs Deployment via HessianAware Quantization and CPU GPU Collaborative
by: Zhang, Tuo, et al.
Published: (2025)
by: Zhang, Tuo, et al.
Published: (2025)
SmallThinker: A Family of Efficient Large Language Models Natively Trained for Local Deployment
by: Song, Yixin, et al.
Published: (2025)
by: Song, Yixin, et al.
Published: (2025)
Privacy Preserving Reinforcement Learning with One-Sided Feedback
by: Cong, Lin William, et al.
Published: (2026)
by: Cong, Lin William, et al.
Published: (2026)
Understanding and Preserving Safety in Fine-Tuned LLMs
by: Zhang, Jiawen, et al.
Published: (2026)
by: Zhang, Jiawen, et al.
Published: (2026)
AL-GNN: Privacy-Preserving and Replay-Free Continual Graph Learning via Analytic Learning
by: Zhang, Xuling, et al.
Published: (2025)
by: Zhang, Xuling, et al.
Published: (2025)
CADRE: Customizable Assurance of Data Readiness in Privacy-Preserving Federated Learning
by: Hiniduma, Kaveen, et al.
Published: (2025)
by: Hiniduma, Kaveen, et al.
Published: (2025)
Towards Privacy-Preserving Data-Driven Education: The Potential of Federated Learning
by: Khalil, Mohammad, et al.
Published: (2025)
by: Khalil, Mohammad, et al.
Published: (2025)
Similar Items
-
LLM-AutoDP: Automatic Data Processing via LLM Agents for Model Fine-tuning
by: Huang, Wei, et al.
Published: (2026) -
GradPruner: Gradient-Guided Layer Pruning Enabling Efficient Fine-Tuning and Inference for LLMs
by: Huang, Wei, et al.
Published: (2026) -
DaMoC: Efficiently Selecting the Optimal Large Language Model for Fine-tuning Domain Tasks Based on Data and Model Compression
by: Huang, Wei, et al.
Published: (2025) -
Mitigating Catastrophic Forgetting in Large Language Models with Forgetting-aware Pruning
by: Huang, Wei, et al.
Published: (2025) -
Data-Free Privacy-Preserving for LLMs via Model Inversion and Selective Unlearning
by: Zhou, Xinjie, et al.
Published: (2026)