How does the optimizer implicitly bias the model merging loss landscape?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Chenxiang, Theus, Alexander, Teney, Damien, Orvieto, Antonio, Pang, Jun, Mauw, Sjouke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Spurious Privacy Leakage in Neural Networks
von: Zhang, Chenxiang, et al.
Veröffentlicht: (2025)
von: Zhang, Chenxiang, et al.
Veröffentlicht: (2025)
Bits for Privacy: Evaluating Post-Training Quantization via Membership Inference
von: Zhang, Chenxiang, et al.
Veröffentlicht: (2025)
von: Zhang, Chenxiang, et al.
Veröffentlicht: (2025)
Meta-RL Induces Exploration in Language Agents
von: Jiang, Yulun, et al.
Veröffentlicht: (2025)
von: Jiang, Yulun, et al.
Veröffentlicht: (2025)
Scalable Ensemble Diversification for OOD Generalization and Detection
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2024)
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2024)
Recurrent neural networks: vanishing and exploding gradients are not the end of the story
von: Zucchet, Nicolas, et al.
Veröffentlicht: (2024)
von: Zucchet, Nicolas, et al.
Veröffentlicht: (2024)
Neural Redshift: Random Networks are not Random Functions
von: Teney, Damien, et al.
Veröffentlicht: (2024)
von: Teney, Damien, et al.
Veröffentlicht: (2024)
Mitigating Shortcut Learning with Diffusion Counterfactuals and Diverse Ensembles
von: Scimeca, Luca, et al.
Veröffentlicht: (2023)
von: Scimeca, Luca, et al.
Veröffentlicht: (2023)
Robust training of implicit generative models for multivariate and heavy-tailed distributions with an invariant statistical loss
von: de Frutos, José Manuel, et al.
Veröffentlicht: (2024)
von: de Frutos, José Manuel, et al.
Veröffentlicht: (2024)
Can you Finetune your Binoculars? Embedding Text Watermarks into the Weights of Large Language Models
von: Elhassan, Fay, et al.
Veröffentlicht: (2025)
von: Elhassan, Fay, et al.
Veröffentlicht: (2025)
Design Principles for Sequence Models via Coefficient Dynamics
von: Sieber, Jerome, et al.
Veröffentlicht: (2025)
von: Sieber, Jerome, et al.
Veröffentlicht: (2025)
Recovering implicit physics model under real-world constraints
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
How does Bayesian Sampling help Membership Inference Attacks?
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025)
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025)
Causal vs. Anticausal merging of predictors
von: Mejia, Sergio Hernan Garrido, et al.
Veröffentlicht: (2025)
von: Mejia, Sergio Hernan Garrido, et al.
Veröffentlicht: (2025)
Generalized Linear Mode Connectivity for Transformers
von: Theus, Alexander, et al.
Veröffentlicht: (2025)
von: Theus, Alexander, et al.
Veröffentlicht: (2025)
What model does MuZero learn?
von: He, Jinke, et al.
Veröffentlicht: (2023)
von: He, Jinke, et al.
Veröffentlicht: (2023)
Rolling Ball Optimizer: Learning by ironing out loss landscape wrinkles
von: Belgoumri, Mohammed Djameleddine, et al.
Veröffentlicht: (2025)
von: Belgoumri, Mohammed Djameleddine, et al.
Veröffentlicht: (2025)
OOD-Chameleon: Is Algorithm Selection for OOD Generalization Learnable?
von: Jiang, Liangze, et al.
Veröffentlicht: (2024)
von: Jiang, Liangze, et al.
Veröffentlicht: (2024)
Fréchet regression with implicit denoising and multicollinearity reduction
von: Mansouri, Dou El Kefel, et al.
Veröffentlicht: (2024)
von: Mansouri, Dou El Kefel, et al.
Veröffentlicht: (2024)
Closed-form merging of parameter-efficient modules for Federated Continual Learning
von: Salami, Riccardo, et al.
Veröffentlicht: (2024)
von: Salami, Riccardo, et al.
Veröffentlicht: (2024)
(Almost) Free Modality Stitching of Foundation Models
von: Singh, Jaisidh, et al.
Veröffentlicht: (2025)
von: Singh, Jaisidh, et al.
Veröffentlicht: (2025)
How to systematically develop an effective AI-based bias correction model?
von: Zhou, Xiao, et al.
Veröffentlicht: (2025)
von: Zhou, Xiao, et al.
Veröffentlicht: (2025)
Synergy and Diversity in CLIP: Enhancing Performance Through Adaptive Backbone Ensembling
von: Rodriguez-Opazo, Cristian, et al.
Veröffentlicht: (2024)
von: Rodriguez-Opazo, Cristian, et al.
Veröffentlicht: (2024)
Understanding the differences in Foundation Models: Attention, State Space Models, and Recurrent Neural Networks
von: Sieber, Jerome, et al.
Veröffentlicht: (2024)
von: Sieber, Jerome, et al.
Veröffentlicht: (2024)
How to be fair? A study of label and selection bias
von: Favier, Marco, et al.
Veröffentlicht: (2024)
von: Favier, Marco, et al.
Veröffentlicht: (2024)
Uncertainty modeling for fine-tuned implicit functions
von: Susmelj, Anna, et al.
Veröffentlicht: (2024)
von: Susmelj, Anna, et al.
Veröffentlicht: (2024)
Enhancing robustness of data-driven SHM models: adversarial training with circle loss
von: Yang, Xiangli, et al.
Veröffentlicht: (2024)
von: Yang, Xiangli, et al.
Veröffentlicht: (2024)
Generalized Interpolating Discrete Diffusion
von: von Rütte, Dimitri, et al.
Veröffentlicht: (2025)
von: von Rütte, Dimitri, et al.
Veröffentlicht: (2025)
GDP nowcasting with artificial neural networks: How much does long-term memory matter?
von: Németh, Kristóf, et al.
Veröffentlicht: (2023)
von: Németh, Kristóf, et al.
Veröffentlicht: (2023)
Exploring multimodal implicit behavior learning for vehicle navigation in simulated cities
von: Antonelo, Eric Aislan, et al.
Veröffentlicht: (2025)
von: Antonelo, Eric Aislan, et al.
Veröffentlicht: (2025)
GRASP: Deterministic argument ranking in interaction graphs
von: Misra, Diganta, et al.
Veröffentlicht: (2026)
von: Misra, Diganta, et al.
Veröffentlicht: (2026)
Understanding the dynamics of the frequency bias in neural networks
von: Molina, Juan, et al.
Veröffentlicht: (2024)
von: Molina, Juan, et al.
Veröffentlicht: (2024)
Detecting labeling bias using influence functions
von: Jørgensen, Frida, et al.
Veröffentlicht: (2026)
von: Jørgensen, Frida, et al.
Veröffentlicht: (2026)
Improving the classification of extreme classes by means of loss regularisation and generalised beta distributions
von: Vargas, Víctor Manuel, et al.
Veröffentlicht: (2024)
von: Vargas, Víctor Manuel, et al.
Veröffentlicht: (2024)
On student-teacher deviations in distillation: does it pay to disobey?
von: Nagarajan, Vaishnavh, et al.
Veröffentlicht: (2023)
von: Nagarajan, Vaishnavh, et al.
Veröffentlicht: (2023)
Chemist-aligned retrosynthesis by ensembling diverse inductive bias models
von: Maziarz, Krzysztof, et al.
Veröffentlicht: (2024)
von: Maziarz, Krzysztof, et al.
Veröffentlicht: (2024)
Inducing anxiety in large language models can induce bias
von: Coda-Forno, Julian, et al.
Veröffentlicht: (2023)
von: Coda-Forno, Julian, et al.
Veröffentlicht: (2023)
Text-guided multi-property molecular optimization with a diffusion language model
von: Xiong, Yida, et al.
Veröffentlicht: (2024)
von: Xiong, Yida, et al.
Veröffentlicht: (2024)
Generative method for aerodynamic optimization based on classifier-free guided denoising diffusion probabilistic model
von: Deng, Shisong, et al.
Veröffentlicht: (2025)
von: Deng, Shisong, et al.
Veröffentlicht: (2025)
Bayesian Low-Rank LeArning (Bella): A Practical Approach to Bayesian Neural Networks
von: Doan, Bao Gia, et al.
Veröffentlicht: (2024)
von: Doan, Bao Gia, et al.
Veröffentlicht: (2024)
Towards certifiable AI in aviation: landscape, challenges, and opportunities
von: Bello, Hymalai, et al.
Veröffentlicht: (2024)
von: Bello, Hymalai, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Spurious Privacy Leakage in Neural Networks
von: Zhang, Chenxiang, et al.
Veröffentlicht: (2025) -
Bits for Privacy: Evaluating Post-Training Quantization via Membership Inference
von: Zhang, Chenxiang, et al.
Veröffentlicht: (2025) -
Meta-RL Induces Exploration in Language Agents
von: Jiang, Yulun, et al.
Veröffentlicht: (2025) -
Scalable Ensemble Diversification for OOD Generalization and Detection
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2024) -
Recurrent neural networks: vanishing and exploding gradients are not the end of the story
von: Zucchet, Nicolas, et al.
Veröffentlicht: (2024)