DivMerge: A divergence-based model merging method for multi-tasking
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Touayouch, Brahim, Fosse, Loïc, Damnati, Géraldine, Lecorvé, Gwénolé |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Statistical Deficiency for Task Inclusion Estimation
von: Fosse, Loïc, et al.
Veröffentlicht: (2025)
von: Fosse, Loïc, et al.
Veröffentlicht: (2025)
TelcoLM: collecting data, adapting, and benchmarking language models for the telecommunication domain
von: Barboule, Camille, et al.
Veröffentlicht: (2024)
von: Barboule, Camille, et al.
Veröffentlicht: (2024)
Factual Knowledge in Language Models: Robustness and Anomalies under Simple Temporal Context Variations
von: Khodja, Hichem Ammar, et al.
Veröffentlicht: (2025)
von: Khodja, Hichem Ammar, et al.
Veröffentlicht: (2025)
Constrained latent state modeling: A unifying perspective on representation learning under competing constraints
von: Quellec, Gwenolé
Veröffentlicht: (2026)
von: Quellec, Gwenolé
Veröffentlicht: (2026)
O_FT@EvalLLM2025 : étude comparative de choix de données et de stratégies d'apprentissage pour l'adaptation de modèles de langue à un domaine
von: Rousseau, Ismaël, et al.
Veröffentlicht: (2025)
von: Rousseau, Ismaël, et al.
Veröffentlicht: (2025)
A linguistically-motivated evaluation methodology for unraveling model's abilities in reading comprehension tasks
von: Antoine, Elie, et al.
Veröffentlicht: (2025)
von: Antoine, Elie, et al.
Veröffentlicht: (2025)
Regularising NARX models with multi-task learning
von: Bee, Sarah, et al.
Veröffentlicht: (2025)
von: Bee, Sarah, et al.
Veröffentlicht: (2025)
Emotion Identification for French in Written Texts: Considering their Modes of Expression as a Step Towards Text Complexity Analysis
von: Étienne, Aline, et al.
Veröffentlicht: (2024)
von: Étienne, Aline, et al.
Veröffentlicht: (2024)
Adaptive multi-gradient methods for quasiconvex vector optimization and applications to multi-task learning
von: Minh, Nguyen Anh, et al.
Veröffentlicht: (2024)
von: Minh, Nguyen Anh, et al.
Veröffentlicht: (2024)
DivIL: Unveiling and Addressing Over-Invariance for Out-of- Distribution Generalization
von: Wang, Jiaqi, et al.
Veröffentlicht: (2025)
von: Wang, Jiaqi, et al.
Veröffentlicht: (2025)
Training-free LLM Merging for Multi-task Learning
von: Fu, Zichuan, et al.
Veröffentlicht: (2025)
von: Fu, Zichuan, et al.
Veröffentlicht: (2025)
Low-rank bias, weight decay, and model merging in neural networks
von: Kuzborskij, Ilja, et al.
Veröffentlicht: (2025)
von: Kuzborskij, Ilja, et al.
Veröffentlicht: (2025)
DivControl: Knowledge Diversion for Controllable Image Generation
von: Xie, Yucheng, et al.
Veröffentlicht: (2025)
von: Xie, Yucheng, et al.
Veröffentlicht: (2025)
DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models
von: Alvar, Saeed Ranjbar, et al.
Veröffentlicht: (2025)
von: Alvar, Saeed Ranjbar, et al.
Veröffentlicht: (2025)
NeighborDiv: Training-free Zero-shot Generalist Graph Anomaly Detection via Neighbor Diversity
von: Wei, Kaifeng, et al.
Veröffentlicht: (2026)
von: Wei, Kaifeng, et al.
Veröffentlicht: (2026)
A high-accuracy multi-model mixing retrosynthetic method
von: Xiang, Shang, et al.
Veröffentlicht: (2024)
von: Xiang, Shang, et al.
Veröffentlicht: (2024)
How does the optimizer implicitly bias the model merging loss landscape?
von: Zhang, Chenxiang, et al.
Veröffentlicht: (2025)
von: Zhang, Chenxiang, et al.
Veröffentlicht: (2025)
Beyond Overcorrection: Evaluating Diversity in T2I Models with DivBench
von: Friedrich, Felix, et al.
Veröffentlicht: (2025)
von: Friedrich, Felix, et al.
Veröffentlicht: (2025)
Minimizing robust density power-based divergences for general parametric density models
von: Okuno, Akifumi
Veröffentlicht: (2023)
von: Okuno, Akifumi
Veröffentlicht: (2023)
LaTiM: Longitudinal representation learning in continuous-time models to predict disease progression
von: Zeghlache, Rachid, et al.
Veröffentlicht: (2024)
von: Zeghlache, Rachid, et al.
Veröffentlicht: (2024)
DivQAT: Enhancing Robustness of Quantized Convolutional Neural Networks against Model Extraction Attacks
von: Khaled, Kacem, et al.
Veröffentlicht: (2025)
von: Khaled, Kacem, et al.
Veröffentlicht: (2025)
Joint auto-encoders: a flexible multi-task learning framework
von: Epstein, Baruch, et al.
Veröffentlicht: (2017)
von: Epstein, Baruch, et al.
Veröffentlicht: (2017)
Do different prompting methods yield a common task representation in language models?
von: Davidson, Guy, et al.
Veröffentlicht: (2025)
von: Davidson, Guy, et al.
Veröffentlicht: (2025)
DivLogicEval: A Framework for Benchmarking Logical Reasoning Evaluation in Large Language Models
von: Chung, Tsz Ting, et al.
Veröffentlicht: (2025)
von: Chung, Tsz Ting, et al.
Veröffentlicht: (2025)
Coreset selection for the Sinkhorn divergence and generic smooth divergences
von: Kokot, Alex, et al.
Veröffentlicht: (2025)
von: Kokot, Alex, et al.
Veröffentlicht: (2025)
MergeBench: A Benchmark for Merging Domain-Specialized LLMs
von: He, Yifei, et al.
Veröffentlicht: (2025)
von: He, Yifei, et al.
Veröffentlicht: (2025)
Inductive biases of multi-task learning and finetuning: multiple regimes of feature reuse
von: Lippl, Samuel, et al.
Veröffentlicht: (2023)
von: Lippl, Samuel, et al.
Veröffentlicht: (2023)
Causal vs. Anticausal merging of predictors
von: Mejia, Sergio Hernan Garrido, et al.
Veröffentlicht: (2025)
von: Mejia, Sergio Hernan Garrido, et al.
Veröffentlicht: (2025)
To See a World in a Spark of Neuron: Disentangling Multi-task Interference for Training-free Model Merging
von: Fang, Zitao, et al.
Veröffentlicht: (2025)
von: Fang, Zitao, et al.
Veröffentlicht: (2025)
Cross-talk based multi-task learning for fault classification of machine system influenced by multiple variables
von: Yi, Wonjun, et al.
Veröffentlicht: (2026)
von: Yi, Wonjun, et al.
Veröffentlicht: (2026)
Neuroplasticity-inspired dynamic ANNs for multi-task demand forecasting
von: Żarski, Mateusz, et al.
Veröffentlicht: (2025)
von: Żarski, Mateusz, et al.
Veröffentlicht: (2025)
Robust multi-task boosting using clustering and local ensembling
von: Emami, Seyedsaman, et al.
Veröffentlicht: (2026)
von: Emami, Seyedsaman, et al.
Veröffentlicht: (2026)
Estimating the number of household TV profiles based in customer behaviour using Gaussian mixture model averaging
von: Palma, Gabriel R., et al.
Veröffentlicht: (2025)
von: Palma, Gabriel R., et al.
Veröffentlicht: (2025)
Winner-Take-All bottlenecks enforce disentangled symbolic representations in multi-task learning
von: Gutheil, Julian, et al.
Veröffentlicht: (2026)
von: Gutheil, Julian, et al.
Veröffentlicht: (2026)
FedMerge: Federated Personalization via Model Merging
von: Chen, Shutong, et al.
Veröffentlicht: (2025)
von: Chen, Shutong, et al.
Veröffentlicht: (2025)
MIN-Merging: Merge the Important Neurons for Model Merging
von: Liang, Yunfei
Veröffentlicht: (2025)
von: Liang, Yunfei
Veröffentlicht: (2025)
Rethinking Weight-Averaged Model-merging
von: Wang, Hu, et al.
Veröffentlicht: (2024)
von: Wang, Hu, et al.
Veröffentlicht: (2024)
Learning task-specific predictive models for scientific computing
von: Yin, Jianyuan, et al.
Veröffentlicht: (2025)
von: Yin, Jianyuan, et al.
Veröffentlicht: (2025)
A comprehensive and easy-to-use multi-domain multi-task medical imaging meta-dataset
von: Woerner, Stefano, et al.
Veröffentlicht: (2024)
von: Woerner, Stefano, et al.
Veröffentlicht: (2024)
Rethinking Layer-wise Model Merging through Chain of Merges
von: Buzzega, Pietro, et al.
Veröffentlicht: (2025)
von: Buzzega, Pietro, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Statistical Deficiency for Task Inclusion Estimation
von: Fosse, Loïc, et al.
Veröffentlicht: (2025) -
TelcoLM: collecting data, adapting, and benchmarking language models for the telecommunication domain
von: Barboule, Camille, et al.
Veröffentlicht: (2024) -
Factual Knowledge in Language Models: Robustness and Anomalies under Simple Temporal Context Variations
von: Khodja, Hichem Ammar, et al.
Veröffentlicht: (2025) -
Constrained latent state modeling: A unifying perspective on representation learning under competing constraints
von: Quellec, Gwenolé
Veröffentlicht: (2026) -
O_FT@EvalLLM2025 : étude comparative de choix de données et de stratégies d'apprentissage pour l'adaptation de modèles de langue à un domaine
von: Rousseau, Ismaël, et al.
Veröffentlicht: (2025)