Rethinking Weight-Averaged Model-merging
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Hu, Ma, Congbo, Almakky, Ibrahim, Reid, Ian, Carneiro, Gustavo, Yaqub, Mohammad |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
In-Model Merging for Enhancing the Robustness of Medical Imaging Classification Models
by: Wang, Hu, et al.
Published: (2025)
by: Wang, Hu, et al.
Published: (2025)
MedNNS: Supernet-based Medical Task-Adaptive Neural Network Search
by: Mecharbat, Lotfi Abdelkrim, et al.
Published: (2025)
by: Mecharbat, Lotfi Abdelkrim, et al.
Published: (2025)
Learnable Cross-modal Knowledge Distillation for Multi-modal Learning with Missing Modality
by: Wang, Hu, et al.
Published: (2023)
by: Wang, Hu, et al.
Published: (2023)
MAFM^3: Modular Adaptation of Foundation Models for Multi-Modal Medical AI
by: Qazi, Mohammad Areeb, et al.
Published: (2025)
by: Qazi, Mohammad Areeb, et al.
Published: (2025)
T3: Test-Time Model Merging in VLMs for Zero-Shot Medical Imaging Analysis
by: Imam, Raza, et al.
Published: (2025)
by: Imam, Raza, et al.
Published: (2025)
MedMerge: Merging Models for Effective Transfer Learning to Medical Imaging Tasks
by: Almakky, Ibrahim, et al.
Published: (2024)
by: Almakky, Ibrahim, et al.
Published: (2024)
DynaMMo: Dynamic Model Merging for Efficient Class Incremental Learning for Medical Images
by: Qazi, Mohammad Areeb, et al.
Published: (2024)
by: Qazi, Mohammad Areeb, et al.
Published: (2024)
Meta-Learned Modality-Weighted Knowledge Distillation for Robust Multi-Modal Learning with Missing Data
by: Wang, Hu, et al.
Published: (2024)
by: Wang, Hu, et al.
Published: (2024)
UNICON: UNIfied CONtinual Learning for Medical Foundational Models
by: Qazi, Mohammad Areeb, et al.
Published: (2025)
by: Qazi, Mohammad Areeb, et al.
Published: (2025)
TransPrune: Token Transition Pruning for Efficient Large Vision-Language Model
by: Li, Ao, et al.
Published: (2025)
by: Li, Ao, et al.
Published: (2025)
Kalman Filter Enhanced GRPO for Reinforcement Learning-Based Language Model Reasoning
by: Wang, Hu, et al.
Published: (2025)
by: Wang, Hu, et al.
Published: (2025)
SALT: Parameter-Efficient Fine-Tuning via Singular Value Adaptation with Low-Rank Transformation
by: Elsayed, Abdelrahman, et al.
Published: (2025)
by: Elsayed, Abdelrahman, et al.
Published: (2025)
Multi-modal Learning with Missing Modality via Shared-Specific Feature Modelling
by: Wang, Hu, et al.
Published: (2023)
by: Wang, Hu, et al.
Published: (2023)
Automatic Quality Assessment of First Trimester Crown-Rump-Length Ultrasound Images
by: Cengiz, Sevim, et al.
Published: (2025)
by: Cengiz, Sevim, et al.
Published: (2025)
Kernel Adversarial Learning for Real-world Image Super-resolution
by: Wang, Hu, et al.
Published: (2021)
by: Wang, Hu, et al.
Published: (2021)
Risk Estimation of Knee Osteoarthritis Progression via Predictive Multi-task Modelling from Efficient Diffusion Model using X-ray Images
by: Butler, David, et al.
Published: (2025)
by: Butler, David, et al.
Published: (2025)
Weight Averaging for Out-of-Distribution Generalization and Few-Shot Domain Adaptation
by: Xu, Shijian
Published: (2025)
by: Xu, Shijian
Published: (2025)
Rethinking Weight Decay for Robust Fine-Tuning of Foundation Models
by: Tian, Junjiao, et al.
Published: (2024)
by: Tian, Junjiao, et al.
Published: (2024)
Continual Learning in Medical Imaging: A Survey and Practical Analysis
by: Qazi, Mohammad Areeb, et al.
Published: (2024)
by: Qazi, Mohammad Areeb, et al.
Published: (2024)
FissionFusion: Fast Geometric Generation and Hierarchical Souping for Medical Image Analysis
by: Sanjeev, Santosh, et al.
Published: (2024)
by: Sanjeev, Santosh, et al.
Published: (2024)
Efficient Parameter Adaptation for Multi-Modal Medical Image Segmentation and Prognosis
by: Saeed, Numan, et al.
Published: (2025)
by: Saeed, Numan, et al.
Published: (2025)
TiBiX: Leveraging Temporal Information for Bidirectional X-ray and Report Generation
by: Sanjeev, Santosh, et al.
Published: (2024)
by: Sanjeev, Santosh, et al.
Published: (2024)
DARK: Diagonal-Anchored Repulsive Knowledge Distillation for Vision-Language Models under Extreme Compression
by: Saeed, Numan, et al.
Published: (2026)
by: Saeed, Numan, et al.
Published: (2026)
Evaluating Multiple Instance Learning Strategies for Automated Sebocyte Droplet Counting
by: Adelipour, Maryam, et al.
Published: (2025)
by: Adelipour, Maryam, et al.
Published: (2025)
CoReEcho: Continuous Representation Learning for 2D+time Echocardiography Analysis
by: Maani, Fadillah Adamsyah, et al.
Published: (2024)
by: Maani, Fadillah Adamsyah, et al.
Published: (2024)
Rethinking Model Selection in VLM Through the Lens of Gromov-Wasserstein Distance
by: Li, Muyang, et al.
Published: (2026)
by: Li, Muyang, et al.
Published: (2026)
Transition Models: Rethinking the Generative Learning Objective
by: Wang, Zidong, et al.
Published: (2025)
by: Wang, Zidong, et al.
Published: (2025)
PEMMA: Parameter-Efficient Multi-Modal Adaptation for Medical Image Segmentation
by: Saadi, Nada, et al.
Published: (2024)
by: Saadi, Nada, et al.
Published: (2024)
Forget-MI: Machine Unlearning for Forgetting Multimodal Information in Healthcare Settings
by: Hardan, Shahad, et al.
Published: (2025)
by: Hardan, Shahad, et al.
Published: (2025)
Bridging Generative and Discriminative Noisy-Label Learning via Direction-Agnostic EM Formulation
by: Liu, Fengbei, et al.
Published: (2023)
by: Liu, Fengbei, et al.
Published: (2023)
XReal: Realistic Anatomy and Pathology-Aware X-ray Generation via Controllable Diffusion Model
by: Hashmi, Anees Ur Rehman, et al.
Published: (2024)
by: Hashmi, Anees Ur Rehman, et al.
Published: (2024)
Deep Multimodal Learning with Missing Modality: A Survey
by: Wu, Renjie, et al.
Published: (2024)
by: Wu, Renjie, et al.
Published: (2024)
Set a Thief to Catch a Thief: Combating Label Noise through Noisy Meta Learning
by: Wang, Hanxuan, et al.
Published: (2025)
by: Wang, Hanxuan, et al.
Published: (2025)
WASH: Train your Ensemble with Communication-Efficient Weight Shuffling, then Average
by: Fournier, Louis, et al.
Published: (2024)
by: Fournier, Louis, et al.
Published: (2024)
Predicting Traffic Flow with Federated Learning and Graph Neural with Asynchronous Computations Network
by: Yaqub, Muhammad, et al.
Published: (2024)
by: Yaqub, Muhammad, et al.
Published: (2024)
Concepts or Skills? Rethinking Instruction Selection for Multi-modal Models
by: Bai, Andrew, et al.
Published: (2025)
by: Bai, Andrew, et al.
Published: (2025)
Model and Feature Diversity for Bayesian Neural Networks in Mutual Learning
by: Pham, Cuong, et al.
Published: (2024)
by: Pham, Cuong, et al.
Published: (2024)
Rethinking Fine-Tuning: Unlocking Hidden Capabilities in Vision-Language Models
by: Zhang, Mingyuan, et al.
Published: (2025)
by: Zhang, Mingyuan, et al.
Published: (2025)
An Element-Wise Weights Aggregation Method for Federated Learning
by: Hu, Yi, et al.
Published: (2024)
by: Hu, Yi, et al.
Published: (2024)
DiC: Rethinking Conv3x3 Designs in Diffusion Models
by: Tian, Yuchuan, et al.
Published: (2024)
by: Tian, Yuchuan, et al.
Published: (2024)
Similar Items
-
In-Model Merging for Enhancing the Robustness of Medical Imaging Classification Models
by: Wang, Hu, et al.
Published: (2025) -
MedNNS: Supernet-based Medical Task-Adaptive Neural Network Search
by: Mecharbat, Lotfi Abdelkrim, et al.
Published: (2025) -
Learnable Cross-modal Knowledge Distillation for Multi-modal Learning with Missing Modality
by: Wang, Hu, et al.
Published: (2023) -
MAFM^3: Modular Adaptation of Foundation Models for Multi-Modal Medical AI
by: Qazi, Mohammad Areeb, et al.
Published: (2025) -
T3: Test-Time Model Merging in VLMs for Zero-Shot Medical Imaging Analysis
by: Imam, Raza, et al.
Published: (2025)