On the Nystrom Approximation for Preconditioning in Kernel Machines
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Abedsoltan, Amirhesam, Pandit, Parthe, Rademacher, Luis, Belkin, Mikhail |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Fast training of large kernel models with delayed projections
par: Abedsoltan, Amirhesam, et autres
Publié: (2024)
par: Abedsoltan, Amirhesam, et autres
Publié: (2024)
Benign, Tempered, or Catastrophic: A Taxonomy of Overfitting
par: Mallinar, Neil, et autres
Publié: (2022)
par: Mallinar, Neil, et autres
Publié: (2022)
Mirror Descent on Reproducing Kernel Banach Spaces
par: Kumar, Akash, et autres
Publié: (2024)
par: Kumar, Akash, et autres
Publié: (2024)
Context-Scaling versus Task-Scaling in In-Context Learning
par: Abedsoltan, Amirhesam, et autres
Publié: (2024)
par: Abedsoltan, Amirhesam, et autres
Publié: (2024)
Task Generalization With AutoRegressive Compositional Structure: Can Learning From $D$ Tasks Generalize to $D^{T}$ Tasks?
par: Abedsoltan, Amirhesam, et autres
Publié: (2025)
par: Abedsoltan, Amirhesam, et autres
Publié: (2025)
Universality of Kernel Random Matrices and Kernel Regression in the Quadratic Regime
par: Pandit, Parthe, et autres
Publié: (2024)
par: Pandit, Parthe, et autres
Publié: (2024)
Emergence in non-neural models: grokking modular arithmetic via average gradient outer product
par: Mallinar, Neil, et autres
Publié: (2024)
par: Mallinar, Neil, et autres
Publié: (2024)
Feature maps for the Laplacian kernel and its generalizations
par: Ahir, Sudhendu, et autres
Publié: (2025)
par: Ahir, Sudhendu, et autres
Publié: (2025)
Asymptotic convexity of wide and shallow neural networks
par: Borkar, Vivek, et autres
Publié: (2025)
par: Borkar, Vivek, et autres
Publié: (2025)
NysAct: A Scalable Preconditioned Gradient Descent using Nystrom Approximation
par: Seung, Hyunseok, et autres
Publié: (2025)
par: Seung, Hyunseok, et autres
Publié: (2025)
Scalable Kernel Logistic Regression with Nyström Approximation: Theoretical Analysis and Application to Discrete Choice Modelling
par: Martín-Baos, José Ángel, et autres
Publié: (2024)
par: Martín-Baos, José Ángel, et autres
Publié: (2024)
Eigenvectors of the De Bruijn Graph Laplacian: A Natural Basis for the Cut and Cycle Space
par: Philippakis, Anthony, et autres
Publié: (2024)
par: Philippakis, Anthony, et autres
Publié: (2024)
Faster Low-Rank Approximation and Kernel Ridge Regression via the Block-Nyström Method
par: Garg, Sachin, et autres
Publié: (2025)
par: Garg, Sachin, et autres
Publié: (2025)
Breaking Data Symmetry is Needed For Generalization in Feature Learning Kernels
par: Bernal, Marcel Tomàs, et autres
Publié: (2026)
par: Bernal, Marcel Tomàs, et autres
Publié: (2026)
Linear Recursive Feature Machines provably recover low-rank matrices
par: Radhakrishnan, Adityanarayanan, et autres
Publié: (2024)
par: Radhakrishnan, Adityanarayanan, et autres
Publié: (2024)
Interpretable Self-Supervised Learning via Representer Landmarks and Nyström Approximation
par: Zarvandi, Maedeh, et autres
Publié: (2025)
par: Zarvandi, Maedeh, et autres
Publié: (2025)
Nyström Kernel Stein Discrepancy
par: Kalinke, Florian, et autres
Publié: (2024)
par: Kalinke, Florian, et autres
Publié: (2024)
A Scalable Nystrom-Based Kernel Two-Sample Test with Permutations
par: Chatalic, Antoine, et autres
Publié: (2025)
par: Chatalic, Antoine, et autres
Publié: (2025)
Learning in Feature Spaces via Coupled Covariances: Asymmetric Kernel SVD and Nyström method
par: Tao, Qinghua, et autres
Publié: (2024)
par: Tao, Qinghua, et autres
Publié: (2024)
More is Better in Modern Machine Learning: when Infinite Overparameterization is Optimal and Overfitting is Obligatory
par: Simon, James B., et autres
Publié: (2023)
par: Simon, James B., et autres
Publié: (2023)
Nyström Kernel Stein Discrepancy Tests
par: Kalinke, Florian, et autres
Publié: (2026)
par: Kalinke, Florian, et autres
Publié: (2026)
The Nyström method for convex loss functions
par: Della Vecchia, Andrea, et autres
Publié: (2020)
par: Della Vecchia, Andrea, et autres
Publié: (2020)
Laplace Approximation for Bayesian Tensor Network Kernel Machines
par: Saiapin, Albert, et autres
Publié: (2026)
par: Saiapin, Albert, et autres
Publié: (2026)
Catching rationalization in the act: detecting motivated reasoning before and after CoT via activation probing
par: Mirtaheri, Parsa, et autres
Publié: (2026)
par: Mirtaheri, Parsa, et autres
Publié: (2026)
Laplace Approximation For Tensor Train Kernel Machines In System Identification
par: Saiapin, Albert, et autres
Publié: (2025)
par: Saiapin, Albert, et autres
Publié: (2025)
Analysis of Nystrom method with sequential ridge leverage scores
par: Calandriello, Daniele, et autres
Publié: (2026)
par: Calandriello, Daniele, et autres
Publié: (2026)
On-Policy Distillation of Language Models for Autonomous Vehicle Motion Planning
par: Afsharrad, Amirhossein, et autres
Publié: (2026)
par: Afsharrad, Amirhossein, et autres
Publié: (2026)
General and Efficient Steering of Unconditional Diffusion
par: Wang, Qingsong, et autres
Publié: (2026)
par: Wang, Qingsong, et autres
Publié: (2026)
A Gap Between the Gaussian RKHS and Neural Networks: An Infinite-Center Asymptotic Analysis
par: Kumar, Akash, et autres
Publié: (2025)
par: Kumar, Akash, et autres
Publié: (2025)
MAC: An Efficient Gradient Preconditioning using Mean Activation Approximated Curvature
par: Seung, Hyunseok, et autres
Publié: (2025)
par: Seung, Hyunseok, et autres
Publié: (2025)
On the Approximation of Kernel functions
par: Dommel, Paul, et autres
Publié: (2024)
par: Dommel, Paul, et autres
Publié: (2024)
Average gradient outer product as a mechanism for deep neural collapse
par: Beaglehole, Daniel, et autres
Publié: (2024)
par: Beaglehole, Daniel, et autres
Publié: (2024)
xRFM: Accurate, scalable, and interpretable feature learning models for tabular data
par: Beaglehole, Daniel, et autres
Publié: (2025)
par: Beaglehole, Daniel, et autres
Publié: (2025)
Scalable Linearized Laplace Approximation via Surrogate Neural Kernel
par: Ortega, Luis A., et autres
Publié: (2026)
par: Ortega, Luis A., et autres
Publié: (2026)
Catapults in SGD: spikes in the training loss and their impact on generalization through feature learning
par: Zhu, Libin, et autres
Publié: (2023)
par: Zhu, Libin, et autres
Publié: (2023)
Quadratic models for understanding catapult dynamics of neural networks
par: Zhu, Libin, et autres
Publié: (2022)
par: Zhu, Libin, et autres
Publié: (2022)
Fast Spectrum Estimation of Some Kernel Matrices
par: Lepilov, Mikhail
Publié: (2024)
par: Lepilov, Mikhail
Publié: (2024)
A Unifying View of Linear Function Approximation in Off-Policy RL Through Matrix Splitting and Preconditioning
par: Wu, Zechen, et autres
Publié: (2025)
par: Wu, Zechen, et autres
Publié: (2025)
Universal Sequence Preconditioning
par: Marsden, Annie, et autres
Publié: (2025)
par: Marsden, Annie, et autres
Publié: (2025)
Bi-Level Policy Optimization with Nyström Hypergradients
par: Prakash, Arjun, et autres
Publié: (2025)
par: Prakash, Arjun, et autres
Publié: (2025)
Documents similaires
-
Fast training of large kernel models with delayed projections
par: Abedsoltan, Amirhesam, et autres
Publié: (2024) -
Benign, Tempered, or Catastrophic: A Taxonomy of Overfitting
par: Mallinar, Neil, et autres
Publié: (2022) -
Mirror Descent on Reproducing Kernel Banach Spaces
par: Kumar, Akash, et autres
Publié: (2024) -
Context-Scaling versus Task-Scaling in In-Context Learning
par: Abedsoltan, Amirhesam, et autres
Publié: (2024) -
Task Generalization With AutoRegressive Compositional Structure: Can Learning From $D$ Tasks Generalize to $D^{T}$ Tasks?
par: Abedsoltan, Amirhesam, et autres
Publié: (2025)