Dataless Weight Disentanglement in Task Arithmetic via Kronecker-Factored Approximate Curvature
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Porrello, Angelo, Buzzega, Pietro, Dangel, Felix, Sommariva, Thomas, Salami, Riccardo, Bonicelli, Lorenzo, Calderara, Simone |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rethinking Layer-wise Model Merging through Chain of Merges
von: Buzzega, Pietro, et al.
Veröffentlicht: (2025)
von: Buzzega, Pietro, et al.
Veröffentlicht: (2025)
Distilling Linearized Behavior into Non-Linear Fine-Tuning for Effective Task Arithmetic
von: Sommariva, Thomas, et al.
Veröffentlicht: (2026)
von: Sommariva, Thomas, et al.
Veröffentlicht: (2026)
A Second-Order Perspective on Model Compositionality and Incremental Learning
von: Porrello, Angelo, et al.
Veröffentlicht: (2024)
von: Porrello, Angelo, et al.
Veröffentlicht: (2024)
Modular Embedding Recomposition for Incremental Learning
von: Panariello, Aniello, et al.
Veröffentlicht: (2025)
von: Panariello, Aniello, et al.
Veröffentlicht: (2025)
CLIP with Generative Latent Replay: a Strong Baseline for Incremental Learning
von: Frascaroli, Emanuele, et al.
Veröffentlicht: (2024)
von: Frascaroli, Emanuele, et al.
Veröffentlicht: (2024)
Trajectory Forecasting through Low-Rank Adaptation of Discrete Latent Codes
von: Benaglia, Riccardo, et al.
Veröffentlicht: (2024)
von: Benaglia, Riccardo, et al.
Veröffentlicht: (2024)
How to Train Your Metamorphic Deep Neural Network
von: Sommariva, Thomas, et al.
Veröffentlicht: (2025)
von: Sommariva, Thomas, et al.
Veröffentlicht: (2025)
Intrinsic Training Signals for Federated Learning Aggregation
von: Fiorini, Cosimo, et al.
Veröffentlicht: (2025)
von: Fiorini, Cosimo, et al.
Veröffentlicht: (2025)
Closed-form merging of parameter-efficient modules for Federated Continual Learning
von: Salami, Riccardo, et al.
Veröffentlicht: (2024)
von: Salami, Riccardo, et al.
Veröffentlicht: (2024)
May the Forgetting Be with You: Alternate Replay for Learning with Noisy Labels
von: Millunzi, Monica, et al.
Veröffentlicht: (2024)
von: Millunzi, Monica, et al.
Veröffentlicht: (2024)
Federated Class-Incremental Learning with Hierarchical Generative Prototypes
von: Salami, Riccardo, et al.
Veröffentlicht: (2024)
von: Salami, Riccardo, et al.
Veröffentlicht: (2024)
Kronecker-Factored Approximate Curvature for Physics-Informed Neural Networks
von: Dangel, Felix, et al.
Veröffentlicht: (2024)
von: Dangel, Felix, et al.
Veröffentlicht: (2024)
An Attention-based Representation Distillation Baseline for Multi-Label Continual Learning
von: Menabue, Martin, et al.
Veröffentlicht: (2024)
von: Menabue, Martin, et al.
Veröffentlicht: (2024)
Transporting Task Vectors across Different Architectures without Training
von: Rinaldi, Filippo, et al.
Veröffentlicht: (2026)
von: Rinaldi, Filippo, et al.
Veröffentlicht: (2026)
Kronecker-factored Approximate Curvature (KFAC) From Scratch
von: Dangel, Felix, et al.
Veröffentlicht: (2025)
von: Dangel, Felix, et al.
Veröffentlicht: (2025)
Semantic Residual Prompts for Continual Learning
von: Menabue, Martin, et al.
Veröffentlicht: (2024)
von: Menabue, Martin, et al.
Veröffentlicht: (2024)
Self-Labeling the Job Shop Scheduling Problem
von: Corsini, Andrea, et al.
Veröffentlicht: (2024)
von: Corsini, Andrea, et al.
Veröffentlicht: (2024)
Update Your Transformer to the Latest Release: Re-Basin of Task Vectors
von: Rinaldi, Filippo, et al.
Veröffentlicht: (2025)
von: Rinaldi, Filippo, et al.
Veröffentlicht: (2025)
Towards Robust Knowledge Removal in Federated Learning with High Data Heterogeneity
von: Santi, Riccardo, et al.
Veröffentlicht: (2025)
von: Santi, Riccardo, et al.
Veröffentlicht: (2025)
Gradient-Sign Masking for Task Vector Transport Across Pre-Trained Models
von: Rinaldi, Filippo, et al.
Veröffentlicht: (2025)
von: Rinaldi, Filippo, et al.
Veröffentlicht: (2025)
Understanding and Enforcing Weight Disentanglement in Task Arithmetic
von: Liu, Shangge, et al.
Veröffentlicht: (2026)
von: Liu, Shangge, et al.
Veröffentlicht: (2026)
DOLFIN: Balancing Stability and Plasticity in Federated Continual Learning
von: Moussadek, Omayma, et al.
Veröffentlicht: (2025)
von: Moussadek, Omayma, et al.
Veröffentlicht: (2025)
STAER: Temporal Aligned Rehearsal for Continual Spiking Neural Network
von: Gianferrari, Matteo, et al.
Veröffentlicht: (2026)
von: Gianferrari, Matteo, et al.
Veröffentlicht: (2026)
Accurate and Efficient Low-Rank Model Merging in Core Space
von: Panariello, Aniello, et al.
Veröffentlicht: (2025)
von: Panariello, Aniello, et al.
Veröffentlicht: (2025)
Zero-Shot Synthetic-to-Real Handwritten Text Recognition via Task Analogies
von: Garrido-Munoz, Carlos, et al.
Veröffentlicht: (2026)
von: Garrido-Munoz, Carlos, et al.
Veröffentlicht: (2026)
Mask and Compress: Efficient Skeleton-based Action Recognition in Continual Learning
von: Mosconi, Matteo, et al.
Veröffentlicht: (2024)
von: Mosconi, Matteo, et al.
Veröffentlicht: (2024)
Dataless Knowledge Fusion by Merging Weights of Language Models
von: Jin, Xisen, et al.
Veröffentlicht: (2022)
von: Jin, Xisen, et al.
Veröffentlicht: (2022)
Generalized Fisher-Weighted SVD: Scalable Kronecker-Factored Fisher Approximation for Compressing Large Language Models
von: Chekalina, Viktoriia, et al.
Veröffentlicht: (2025)
von: Chekalina, Viktoriia, et al.
Veröffentlicht: (2025)
Revisiting Scalable Hessian Diagonal Approximations for Applications in Reinforcement Learning
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2024)
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2024)
ABRA: Teleporting Fine-Tuned Knowledge Across Domains for Open-Vocabulary Object Detection
von: Bernardi, Mattia, et al.
Veröffentlicht: (2026)
von: Bernardi, Mattia, et al.
Veröffentlicht: (2026)
Is Multiple Object Tracking a Matter of Specialization?
von: Mancusi, Gianluca, et al.
Veröffentlicht: (2024)
von: Mancusi, Gianluca, et al.
Veröffentlicht: (2024)
DitHub: A Modular Framework for Incremental Open-Vocabulary Object Detection
von: Cappellino, Chiara, et al.
Veröffentlicht: (2025)
von: Cappellino, Chiara, et al.
Veröffentlicht: (2025)
Can LLMs Generate Visualizations with Dataless Prompts?
von: Coelho, Darius, et al.
Veröffentlicht: (2024)
von: Coelho, Darius, et al.
Veröffentlicht: (2024)
Kronecker-Factored Approximate Curvature for Modern Neural Network Architectures
von: Eschenhagen, Runa, et al.
Veröffentlicht: (2023)
von: Eschenhagen, Runa, et al.
Veröffentlicht: (2023)
Monocular Per-Object Distance Estimation with Masked Object Modeling
von: Panariello, Aniello, et al.
Veröffentlicht: (2024)
von: Panariello, Aniello, et al.
Veröffentlicht: (2024)
KFCPO: Kronecker-Factored Approximated Constrained Policy Optimization
von: Lim, Joonyoung, et al.
Veröffentlicht: (2025)
von: Lim, Joonyoung, et al.
Veröffentlicht: (2025)
CAMNet: Leveraging Cooperative Awareness Messages for Vehicle Trajectory Prediction
von: Grasselli, Mattia, et al.
Veröffentlicht: (2025)
von: Grasselli, Mattia, et al.
Veröffentlicht: (2025)
Algebraic Priors for Approximately Equivariant Networks
von: Ali, Riccardo, et al.
Veröffentlicht: (2025)
von: Ali, Riccardo, et al.
Veröffentlicht: (2025)
Selective Attention-based Modulation for Continual Learning
von: Bellitto, Giovanni, et al.
Veröffentlicht: (2024)
von: Bellitto, Giovanni, et al.
Veröffentlicht: (2024)
A New Way: Kronecker-Factored Approximate Curvature Deep Hedging and its Benefits
von: Enkhbayar, Tsogt-Ochir
Veröffentlicht: (2024)
von: Enkhbayar, Tsogt-Ochir
Veröffentlicht: (2024)
Ähnliche Einträge
-
Rethinking Layer-wise Model Merging through Chain of Merges
von: Buzzega, Pietro, et al.
Veröffentlicht: (2025) -
Distilling Linearized Behavior into Non-Linear Fine-Tuning for Effective Task Arithmetic
von: Sommariva, Thomas, et al.
Veröffentlicht: (2026) -
A Second-Order Perspective on Model Compositionality and Incremental Learning
von: Porrello, Angelo, et al.
Veröffentlicht: (2024) -
Modular Embedding Recomposition for Incremental Learning
von: Panariello, Aniello, et al.
Veröffentlicht: (2025) -
CLIP with Generative Latent Replay: a Strong Baseline for Incremental Learning
von: Frascaroli, Emanuele, et al.
Veröffentlicht: (2024)