PLeaS -- Merging Models with Permutations and Least Squares
Fuente:
arXiv
Saved in:
| Main Authors: | Nasery, Anshul, Hayase, Jonathan, Koh, Pang Wei, Oh, Sewoong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scalable Fingerprinting of Large Language Models
by: Nasery, Anshul, et al.
Published: (2025)
by: Nasery, Anshul, et al.
Published: (2025)
Insufficient Statistics Perturbation: Stable Estimators for Private Least Squares
by: Brown, Gavin, et al.
Published: (2024)
by: Brown, Gavin, et al.
Published: (2024)
Are Robust LLM Fingerprints Adversarially Robust?
by: Nasery, Anshul, et al.
Published: (2025)
by: Nasery, Anshul, et al.
Published: (2025)
Sampling from Your Language Model One Byte at a Time
by: Hayase, Jonathan, et al.
Published: (2025)
by: Hayase, Jonathan, et al.
Published: (2025)
Data Mixture Inference: What do BPE Tokenizers Reveal about their Training Data?
by: Hayase, Jonathan, et al.
Published: (2024)
by: Hayase, Jonathan, et al.
Published: (2024)
SuperBPE: Space Travel for Language Models
by: Liu, Alisa, et al.
Published: (2025)
by: Liu, Alisa, et al.
Published: (2025)
S4S: Solving for a Diffusion Model Solver
by: Frankel, Eric, et al.
Published: (2025)
by: Frankel, Eric, et al.
Published: (2025)
PEEKABOO: Interactive Video Generation via Masked-Diffusion
by: Jain, Yash, et al.
Published: (2023)
by: Jain, Yash, et al.
Published: (2023)
Merge to Learn: Efficiently Adding Skills to Language Models with Model Merging
by: Morrison, Jacob, et al.
Published: (2024)
by: Morrison, Jacob, et al.
Published: (2024)
Partitioned Least Squares
by: Esposito, Roberto, et al.
Published: (2020)
by: Esposito, Roberto, et al.
Published: (2020)
Anchored Decoding: Provably Reducing Copyright Risk for Any Language Model
by: He, Jacqueline, et al.
Published: (2026)
by: He, Jacqueline, et al.
Published: (2026)
Understanding the Gains from Repeated Self-Distillation
by: Pareek, Divyansh, et al.
Published: (2024)
by: Pareek, Divyansh, et al.
Published: (2024)
Understanding the Gain from Data Filtering in Multimodal Contrastive Learning
by: Pareek, Divyansh, et al.
Published: (2025)
by: Pareek, Divyansh, et al.
Published: (2025)
Multilingual Diversity Improves Vision-Language Representations
by: Nguyen, Thao, et al.
Published: (2024)
by: Nguyen, Thao, et al.
Published: (2024)
Outlier-robust Autocovariance Least Square Estimation via Iteratively Reweighted Least Square
by: Li, Jiahong, et al.
Published: (2026)
by: Li, Jiahong, et al.
Published: (2026)
The Non-Local Model Merging Problem: Permutation Symmetries and Variance Collapse
by: Sharma, Ekansh, et al.
Published: (2024)
by: Sharma, Ekansh, et al.
Published: (2024)
On Least Square Estimation in Softmax Gating Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Communication-Efficient l_0 Penalized Least Square
by: Gong, Chenqi, et al.
Published: (2025)
by: Gong, Chenqi, et al.
Published: (2025)
Ordinary Least Squares as an Attention Mechanism
by: Coulombe, Philippe Goulet
Published: (2025)
by: Coulombe, Philippe Goulet
Published: (2025)
Tensor-Based Foundations of Ordinary Least Squares and Neural Network Regression Models
by: Algarte, Roberto Dias
Published: (2024)
by: Algarte, Roberto Dias
Published: (2024)
Core-elements Subsampling for Alternating Least Squares
by: Xue, Dunyao, et al.
Published: (2025)
by: Xue, Dunyao, et al.
Published: (2025)
Randomization Techniques to Mitigate the Risk of Copyright Infringement
by: Chen, Wei-Ning, et al.
Published: (2024)
by: Chen, Wei-Ning, et al.
Published: (2024)
OML: A Primitive for Reconciling Open Access with Owner Control in AI Model Distribution
by: Cheng, Zerui, et al.
Published: (2024)
by: Cheng, Zerui, et al.
Published: (2024)
Improving Implicit Regularization of SGD with Preconditioning for Least Square Problems
by: Su, Junwei, et al.
Published: (2024)
by: Su, Junwei, et al.
Published: (2024)
$(ε, δ)$-Differentially Private Partial Least Squares Regression
by: Nikzad-Langerodi, Ramin, et al.
Published: (2024)
by: Nikzad-Langerodi, Ramin, et al.
Published: (2024)
A Hybrid Federated Kernel Regularized Least Squares Algorithm
by: Damiani, Celeste, et al.
Published: (2024)
by: Damiani, Celeste, et al.
Published: (2024)
Towards Learning High-Precision Least Squares Algorithms with Sequence Models
by: Liu, Jerry, et al.
Published: (2025)
by: Liu, Jerry, et al.
Published: (2025)
A False Sense of Privacy: Evaluating Textual Data Sanitization Beyond Surface-level Privacy Leakage
by: Xin, Rui, et al.
Published: (2025)
by: Xin, Rui, et al.
Published: (2025)
Least Squares and Marginal Log-Likelihood Model Predictive Control using Normalizing Flows
by: Cramer, Eike
Published: (2024)
by: Cramer, Eike
Published: (2024)
Relaxed Sparsest-Permutation Formulation for Causal Discovery at Scale
by: Oh, Sunmin, et al.
Published: (2026)
by: Oh, Sunmin, et al.
Published: (2026)
Generalization for Least Squares Regression With Simple Spiked Covariances
by: Li, Jiping, et al.
Published: (2024)
by: Li, Jiping, et al.
Published: (2024)
Convex Regression in Multidimensions: Suboptimality of Least Squares Estimators
by: Kur, Gil, et al.
Published: (2020)
by: Kur, Gil, et al.
Published: (2020)
Transformers Don't In-Context Learn Least Squares Regression
by: Hill, Joshua, et al.
Published: (2025)
by: Hill, Joshua, et al.
Published: (2025)
Kernel Recursive Least Squares Dictionary Learning Algorithm
by: Alipoor, Ghasem, et al.
Published: (2025)
by: Alipoor, Ghasem, et al.
Published: (2025)
Randomized Least Squares Value Iteration itself is Joint Differentially Private
by: Lu, Haiyang, et al.
Published: (2026)
by: Lu, Haiyang, et al.
Published: (2026)
High-Dimensional Partial Least Squares: Spectral Analysis and Fundamental Limitations
by: Léger, Victor, et al.
Published: (2025)
by: Léger, Victor, et al.
Published: (2025)
Generalized Least Squares Kernelized Tensor Factorization
by: Lei, Mengying, et al.
Published: (2024)
by: Lei, Mengying, et al.
Published: (2024)
Algebraic and Statistical Properties of the Ordinary Least Squares Interpolator
by: Shen, Dennis, et al.
Published: (2023)
by: Shen, Dennis, et al.
Published: (2023)
Understanding MLP-Mixer as a Wide and Sparse MLP
by: Hayase, Tomohiro, et al.
Published: (2023)
by: Hayase, Tomohiro, et al.
Published: (2023)
Shape Constraints in Symbolic Regression using Penalized Least Squares
by: Martinek, Viktor, et al.
Published: (2024)
by: Martinek, Viktor, et al.
Published: (2024)
Similar Items
-
Scalable Fingerprinting of Large Language Models
by: Nasery, Anshul, et al.
Published: (2025) -
Insufficient Statistics Perturbation: Stable Estimators for Private Least Squares
by: Brown, Gavin, et al.
Published: (2024) -
Are Robust LLM Fingerprints Adversarially Robust?
by: Nasery, Anshul, et al.
Published: (2025) -
Sampling from Your Language Model One Byte at a Time
by: Hayase, Jonathan, et al.
Published: (2025) -
Data Mixture Inference: What do BPE Tokenizers Reveal about their Training Data?
by: Hayase, Jonathan, et al.
Published: (2024)