Guided Model Merging for Hybrid Data Learning: Leveraging Centralized Data to Refine Decentralized Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Junyi, Yao, Ruicong, Ceritli, Taha, Ozkan, Savas, Blaschko, Matthew B., Noh, Eunchung, Min, Jeongwon, Min, Cho Jung, Ozay, Mete |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient and Accurate Scene Text Recognition with Cascaded-Transformers
by: Ozkan, Savas, et al.
Published: (2025)
by: Ozkan, Savas, et al.
Published: (2025)
Accurate Scene Text Recognition with Efficient Model Scaling and Cloze Self-Distillation
by: Maracani, Andrea, et al.
Published: (2025)
by: Maracani, Andrea, et al.
Published: (2025)
K-Merge: Online Continual Merging of Adapters for On-device Large Language Models
by: Shenaj, Donald, et al.
Published: (2025)
by: Shenaj, Donald, et al.
Published: (2025)
Data-driven Clustering and Merging of Adapters for On-device Large Language Models
by: Bohdal, Ondrej, et al.
Published: (2026)
by: Bohdal, Ondrej, et al.
Published: (2026)
Multi-Task Pre-Finetuning of Lightweight Transformer Encoders for Text Classification and NER
by: Zhu, Junyi, et al.
Published: (2025)
by: Zhu, Junyi, et al.
Published: (2025)
Clustering-driven Memory Compression for On-device Large Language Models
by: Bohdal, Ondrej, et al.
Published: (2026)
by: Bohdal, Ondrej, et al.
Published: (2026)
CG-TTRL: Context-Guided Test-Time Reinforcement Learning for On-Device Large Language Models
by: Hosseini, Peyman, et al.
Published: (2025)
by: Hosseini, Peyman, et al.
Published: (2025)
Decoding Text Spans for Efficient and Accurate Named-Entity Recognition
by: Maracani, Andrea, et al.
Published: (2026)
by: Maracani, Andrea, et al.
Published: (2026)
HydraOpt: Navigating the Efficiency-Performance Trade-off of Adapter Merging
by: Ceritli, Taha, et al.
Published: (2025)
by: Ceritli, Taha, et al.
Published: (2025)
Diffusion Alignment Beyond KL: Variance Minimisation as Effective Policy Optimiser
by: Ou, Zijing, et al.
Published: (2026)
by: Ou, Zijing, et al.
Published: (2026)
MemLoRA: Distilling Expert Adapters for On-Device Memory Systems
by: Bini, Massimo, et al.
Published: (2025)
by: Bini, Massimo, et al.
Published: (2025)
Geometrically Consistent Multi-View Scene Generation from Freehand Sketches
by: Bourouis, Ahmed, et al.
Published: (2026)
by: Bourouis, Ahmed, et al.
Published: (2026)
Responsible Federated LLMs via Safety Filtering and Constitutional AI
by: Noh, Eunchung, et al.
Published: (2025)
by: Noh, Eunchung, et al.
Published: (2025)
Efficient 3D Full-Body Motion Generation from Sparse Tracking Inputs with Temporal Windows
by: Angelis, Georgios Fotios, et al.
Published: (2025)
by: Angelis, Georgios Fotios, et al.
Published: (2025)
Mem-MLP: Real-Time 3D Human Motion Generation from Sparse Inputs
by: Mutlu, Sinan, et al.
Published: (2025)
by: Mutlu, Sinan, et al.
Published: (2025)
Randomized Asymmetric Chain of LoRA: The First Meaningful Theoretical Framework for Low-Rank Adaptation
by: Malinovsky, Grigory, et al.
Published: (2024)
by: Malinovsky, Grigory, et al.
Published: (2024)
A Model for Every User and Budget: Label-Free and Personalized Mixed-Precision Quantization
by: Fish, Edward, et al.
Published: (2023)
by: Fish, Edward, et al.
Published: (2023)
Model Merging and Safety Alignment: One Bad Model Spoils the Bunch
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2024)
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2024)
HOP to the Next Tasks and Domains for Continual Learning in NLP
by: Michieli, Umberto, et al.
Published: (2024)
by: Michieli, Umberto, et al.
Published: (2024)
DALDA: Data Augmentation Leveraging Diffusion Model and LLM with Adaptive Guidance Scaling
by: Jung, Kyuheon, et al.
Published: (2024)
by: Jung, Kyuheon, et al.
Published: (2024)
LoRA.rar: Learning to Merge LoRAs via Hypernetworks for Subject-Style Conditioned Image Generation
by: Shenaj, Donald, et al.
Published: (2024)
by: Shenaj, Donald, et al.
Published: (2024)
Feature-Space Generative Models for One-Shot Class-Incremental Learning
by: Foster, Jack, et al.
Published: (2026)
by: Foster, Jack, et al.
Published: (2026)
Enhanced Model Robustness to Input Corruptions by Per-corruption Adaptation of Normalization Statistics
by: Camuffo, Elena, et al.
Published: (2024)
by: Camuffo, Elena, et al.
Published: (2024)
LoRA-Guard: Parameter-Efficient Guardrail Adaptation for Content Moderation of Large Language Models
by: Elesedy, Hayder, et al.
Published: (2024)
by: Elesedy, Hayder, et al.
Published: (2024)
ValSub: Subsampling Validation Data to Mitigate Forgetting during ASR Personalization
by: Mehmood, Haaris, et al.
Published: (2025)
by: Mehmood, Haaris, et al.
Published: (2025)
Pretrained-Guided Conditional Diffusion Models for Microbiome Data Analysis
by: Shi, Xinyuan, et al.
Published: (2024)
by: Shi, Xinyuan, et al.
Published: (2024)
DP-LAC: Lightweight Adaptive Clipping for Differentially Private Federated Fine-tuning of Language Models
by: Mehmood, Haaris, et al.
Published: (2026)
by: Mehmood, Haaris, et al.
Published: (2026)
persoDA: Personalized Data Augmentation for Personalized ASR
by: Parada, Pablo Peso, et al.
Published: (2025)
by: Parada, Pablo Peso, et al.
Published: (2025)
Let the Void Be Void: Robust Open-Set Semi-Supervised Learning via Selective Non-Alignment
by: Choi, You Rim, et al.
Published: (2025)
by: Choi, You Rim, et al.
Published: (2025)
Object-conditioned Bag of Instances for Few-Shot Personalized Instance Recognition
by: Michieli, Umberto, et al.
Published: (2024)
by: Michieli, Umberto, et al.
Published: (2024)
SoftCFG: Uncertainty-guided Stable Guidance for Visual Autoregressive Model
by: Xu, Dongli, et al.
Published: (2025)
by: Xu, Dongli, et al.
Published: (2025)
Video Summarization with Large Language Models
by: Lee, Min Jung, et al.
Published: (2025)
by: Lee, Min Jung, et al.
Published: (2025)
Wrapper Boxes: Faithful Attribution of Model Predictions to Training Data
by: Su, Yiheng, et al.
Published: (2023)
by: Su, Yiheng, et al.
Published: (2023)
Deep Neural Network Models Trained With A Fixed Random Classifier Transfer Better Across Domains
by: Ali, Hafiz Tiomoko, et al.
Published: (2024)
by: Ali, Hafiz Tiomoko, et al.
Published: (2024)
Efficient Compositional Multi-tasking for On-device Large Language Models
by: Bohdal, Ondrej, et al.
Published: (2025)
by: Bohdal, Ondrej, et al.
Published: (2025)
Merging Smarter, Generalizing Better: Enhancing Model Merging on OOD Data
by: Zhang, Bingjie, et al.
Published: (2025)
by: Zhang, Bingjie, et al.
Published: (2025)
ACE-Merging: Data-Free Model Merging with Adaptive Covariance Estimation
by: Xu, Bo, et al.
Published: (2026)
by: Xu, Bo, et al.
Published: (2026)
A Hybrid Rule‐Based and Large Language Model Artificial Intelligence Systems for Electrodiagnostic Reporting: Two‐Center Retrospective Evaluation
by: Yesung Jung, et al.
Published: (2026)
by: Yesung Jung, et al.
Published: (2026)
NAB: Neural Adaptive Binning for Sparse-View CT reconstruction
by: Xie, Wangduo, et al.
Published: (2026)
by: Xie, Wangduo, et al.
Published: (2026)
DC-Merge: Improving Model Merging with Directional Consistency
by: Zhang, Han-Chen, et al.
Published: (2026)
by: Zhang, Han-Chen, et al.
Published: (2026)
Similar Items
-
Efficient and Accurate Scene Text Recognition with Cascaded-Transformers
by: Ozkan, Savas, et al.
Published: (2025) -
Accurate Scene Text Recognition with Efficient Model Scaling and Cloze Self-Distillation
by: Maracani, Andrea, et al.
Published: (2025) -
K-Merge: Online Continual Merging of Adapters for On-device Large Language Models
by: Shenaj, Donald, et al.
Published: (2025) -
Data-driven Clustering and Merging of Adapters for On-device Large Language Models
by: Bohdal, Ondrej, et al.
Published: (2026) -
Multi-Task Pre-Finetuning of Lightweight Transformer Encoders for Text Classification and NER
by: Zhu, Junyi, et al.
Published: (2025)