Task Addition and Weight Disentanglement in Closed-Vocabulary Models
Fuente:
arXiv
Saved in:
| Main Authors: | Hazimeh, Adam, Favero, Alessandro, Frossard, Pascal |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Model soups need only one ingredient
by: Abdollahpoorrostam, Alireza, et al.
Published: (2026)
by: Abdollahpoorrostam, Alireza, et al.
Published: (2026)
Backdoor Unlearning by Linear Task Decomposition
by: Abdelraheem, Amel, et al.
Published: (2025)
by: Abdelraheem, Amel, et al.
Published: (2025)
How Compositional Generalization and Creativity Improve as Diffusion Models are Trained
by: Favero, Alessandro, et al.
Published: (2025)
by: Favero, Alessandro, et al.
Published: (2025)
Semantic Document Derendering: SVG Reconstruction via Vision-Language Modeling
by: Hazimeh, Adam, et al.
Published: (2025)
by: Hazimeh, Adam, et al.
Published: (2025)
MEMOIR: Lifelong Model Editing with Minimal Overwrite and Informed Retention for LLMs
by: Wang, Ke, et al.
Published: (2025)
by: Wang, Ke, et al.
Published: (2025)
Pareto Low-Rank Adapters: Efficient Multi-Task Learning with Preferences
by: Dimitriadis, Nikolaos, et al.
Published: (2024)
by: Dimitriadis, Nikolaos, et al.
Published: (2024)
Graph-Dictionary Signal Model for Sparse Representations of Multivariate Data
by: Cappelletti, William, et al.
Published: (2024)
by: Cappelletti, William, et al.
Published: (2024)
LiNeS: Post-training Layer Scaling Prevents Forgetting and Enhances Model Merging
by: Wang, Ke, et al.
Published: (2024)
by: Wang, Ke, et al.
Published: (2024)
PUMA: margin-based data pruning
by: Maroto, Javier, et al.
Published: (2024)
by: Maroto, Javier, et al.
Published: (2024)
Sequential Representation Learning via Static-Dynamic Conditional Disentanglement
by: Simon, Mathieu Cyrille, et al.
Published: (2024)
by: Simon, Mathieu Cyrille, et al.
Published: (2024)
The Physics of Data and Tasks: Theories of Locality and Compositionality in Deep Learning
by: Favero, Alessandro
Published: (2025)
by: Favero, Alessandro
Published: (2025)
Flow based approach for Dynamic Temporal Causal models with non-Gaussian or Heteroscedastic Noises
by: Rahmani, Abdellah, et al.
Published: (2025)
by: Rahmani, Abdellah, et al.
Published: (2025)
Deep End-to-End Survival Analysis with Temporal Consistency
by: Vieyra, Mariana Vargas, et al.
Published: (2024)
by: Vieyra, Mariana Vargas, et al.
Published: (2024)
Causal Temporal Regime Structure Learning
by: Rahmani, Abdellah, et al.
Published: (2023)
by: Rahmani, Abdellah, et al.
Published: (2023)
Sparse Training of Discrete Diffusion Models for Graph Generation
by: Qin, Yiming, et al.
Published: (2023)
by: Qin, Yiming, et al.
Published: (2023)
Localizing Task Information for Improved Model Merging and Compression
by: Wang, Ke, et al.
Published: (2024)
by: Wang, Ke, et al.
Published: (2024)
Generative Modelling of Structurally Constrained Graphs
by: Madeira, Manuel, et al.
Published: (2024)
by: Madeira, Manuel, et al.
Published: (2024)
Bures-Wasserstein Means of Graphs
by: Haasler, Isabel, et al.
Published: (2023)
by: Haasler, Isabel, et al.
Published: (2023)
Scaling Laws for Downstream Task Performance of Large Language Models
by: Isik, Berivan, et al.
Published: (2024)
by: Isik, Berivan, et al.
Published: (2024)
DeFoG: Discrete Flow Matching for Graph Generation
by: Qin, Yiming, et al.
Published: (2024)
by: Qin, Yiming, et al.
Published: (2024)
Fine-Tuning Attention Modules Only: Enhancing Weight Disentanglement in Task Arithmetic
by: Jin, Ruochen, et al.
Published: (2024)
by: Jin, Ruochen, et al.
Published: (2024)
Multi-Task Model Merging via Adaptive Weight Disentanglement
by: Xiong, Feng, et al.
Published: (2024)
by: Xiong, Feng, et al.
Published: (2024)
Bigger Isn't Always Memorizing: Early Stopping Overparameterized Diffusion Models
by: Favero, Alessandro, et al.
Published: (2025)
by: Favero, Alessandro, et al.
Published: (2025)
MOONSHOT : A Framework for Multi-Objective Pruning of Vision and Large Language Models
by: Afriat, Gabriel, et al.
Published: (2026)
by: Afriat, Gabriel, et al.
Published: (2026)
Balancing Symmetry and Efficiency in Graph Flow Matching
by: Honoré, Benjamin, et al.
Published: (2026)
by: Honoré, Benjamin, et al.
Published: (2026)
Behavior Tokens Speak Louder: Disentangled Explainable Recommendation with Behavior Vocabulary
by: Feng, Xinshun, et al.
Published: (2025)
by: Feng, Xinshun, et al.
Published: (2025)
Towards Modeling Learner Performance with Large Language Models
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
rETF-semiSL: Semi-Supervised Learning for Neural Collapse in Temporal Data
by: Xie, Yuhan, et al.
Published: (2025)
by: Xie, Yuhan, et al.
Published: (2025)
Generating Directed Graphs with Dual Attention and Asymmetric Encoding
by: Carballo-Castro, Alba, et al.
Published: (2025)
by: Carballo-Castro, Alba, et al.
Published: (2025)
Inductive Domain Transfer In Misspecified Simulation-Based Inference
by: Senouf, Ortal, et al.
Published: (2025)
by: Senouf, Ortal, et al.
Published: (2025)
Operationalizing Quantized Disentanglement
by: Barin-Pacela, Vitoria, et al.
Published: (2025)
by: Barin-Pacela, Vitoria, et al.
Published: (2025)
Learn from your own latents and not from tokens: A sample-complexity theory
by: Korchinski, Daniel J., et al.
Published: (2026)
by: Korchinski, Daniel J., et al.
Published: (2026)
Shapley-Inspired Feature Weighting in $k$-means with No Additional Hyperparameters
by: Fawley, Richard J., et al.
Published: (2025)
by: Fawley, Richard J., et al.
Published: (2025)
Task Addition in Multi-Task Learning by Geometrical Alignment
by: Yim, Soorin, et al.
Published: (2024)
by: Yim, Soorin, et al.
Published: (2024)
Adaptive Policy Learning to Additional Tasks
by: Hao, Wenjian, et al.
Published: (2023)
by: Hao, Wenjian, et al.
Published: (2023)
Distributional Reduction: Unifying Dimensionality Reduction and Clustering with Gromov-Wasserstein
by: Van Assel, Hugues, et al.
Published: (2024)
by: Van Assel, Hugues, et al.
Published: (2024)
Disentangling and Mitigating the Impact of Task Similarity for Continual Learning
by: Hiratani, Naoki
Published: (2024)
by: Hiratani, Naoki
Published: (2024)
Benchmarking Uncertainty Disentanglement: Specialized Uncertainties for Specialized Tasks
by: Mucsányi, Bálint, et al.
Published: (2024)
by: Mucsányi, Bálint, et al.
Published: (2024)
OSSCAR: One-Shot Structured Pruning in Vision and Language Models with Combinatorial Optimization
by: Meng, Xiang, et al.
Published: (2024)
by: Meng, Xiang, et al.
Published: (2024)
Measuring Orthogonality as the Blind-Spot of Uncertainty Disentanglement
by: de Jong, Ivo Pascal, et al.
Published: (2024)
by: de Jong, Ivo Pascal, et al.
Published: (2024)
Similar Items
-
Model soups need only one ingredient
by: Abdollahpoorrostam, Alireza, et al.
Published: (2026) -
Backdoor Unlearning by Linear Task Decomposition
by: Abdelraheem, Amel, et al.
Published: (2025) -
How Compositional Generalization and Creativity Improve as Diffusion Models are Trained
by: Favero, Alessandro, et al.
Published: (2025) -
Semantic Document Derendering: SVG Reconstruction via Vision-Language Modeling
by: Hazimeh, Adam, et al.
Published: (2025) -
MEMOIR: Lifelong Model Editing with Minimal Overwrite and Informed Retention for LLMs
by: Wang, Ke, et al.
Published: (2025)