Salvato in:
| Autori principali: | AlQuabeh, Hilal, de Vazelhes, William, Gu, Bin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2402.01146 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Uncovering the Spectral Bias in Diagonal State Space Models
di: Solozabal, Ruben, et al.
Pubblicazione: (2025)
di: Solozabal, Ruben, et al.
Pubblicazione: (2025)
Constrained Adversarial Perturbation
di: Nishad, Virendra, et al.
Pubblicazione: (2025)
di: Nishad, Virendra, et al.
Pubblicazione: (2025)
Mechanistic Insights into Grokking from the Embedding Layer
di: AlquBoj, H. V., et al.
Pubblicazione: (2025)
di: AlquBoj, H. V., et al.
Pubblicazione: (2025)
New Insight of Variance reduce in Zero-Order Hard-Thresholding: Mitigating Gradient Error and Expansivity Contradictions
di: Yuan, Xinzhe, et al.
Pubblicazione: (2026)
di: Yuan, Xinzhe, et al.
Pubblicazione: (2026)
Zeroth-Order Hard-Thresholding: Gradient Error vs. Expansivity
di: de Vazelhes, William, et al.
Pubblicazione: (2022)
di: de Vazelhes, William, et al.
Pubblicazione: (2022)
Optimization over Sparse Support-Preserving Sets: Two-Step Projection with Global Optimality Guarantees
di: de Vazelhes, William, et al.
Pubblicazione: (2025)
di: de Vazelhes, William, et al.
Pubblicazione: (2025)
Emergence of Primacy and Recency Effect in Mamba: A Mechanistic Point of View
di: Airlangga, Muhammad Cendekia, et al.
Pubblicazione: (2025)
di: Airlangga, Muhammad Cendekia, et al.
Pubblicazione: (2025)
Hard-Thresholding Meets Evolution Strategies in Reinforcement Learning
di: Gao, Chengqian, et al.
Pubblicazione: (2024)
di: Gao, Chengqian, et al.
Pubblicazione: (2024)
Iterative Regularization with k-support Norm: An Important Complement to Sparse Recovery
di: de Vazelhes, William, et al.
Pubblicazione: (2023)
di: de Vazelhes, William, et al.
Pubblicazione: (2023)
The Geometry of Numerical Reasoning: Language Models Compare Numeric Properties in Linear Subspaces
di: El-Shangiti, Ahmed Oumar, et al.
Pubblicazione: (2024)
di: El-Shangiti, Ahmed Oumar, et al.
Pubblicazione: (2024)
Learning Curves of Stochastic Gradient Descent in Kernel Regression
di: Zhang, Haihan, et al.
Pubblicazione: (2025)
di: Zhang, Haihan, et al.
Pubblicazione: (2025)
Learning Associative Memories with Gradient Descent
di: Cabannes, Vivien, et al.
Pubblicazione: (2024)
di: Cabannes, Vivien, et al.
Pubblicazione: (2024)
Partially Lazy Gradient Descent for Smoothed Online Learning
di: Mhaisen, Naram, et al.
Pubblicazione: (2026)
di: Mhaisen, Naram, et al.
Pubblicazione: (2026)
Adaptive Kernel Selection for Stein Variational Gradient Descent
di: Melcher, Moritz, et al.
Pubblicazione: (2025)
di: Melcher, Moritz, et al.
Pubblicazione: (2025)
Weighted Averaged Stochastic Gradient Descent: Asymptotic Normality and Optimality
di: Wei, Ziyang, et al.
Pubblicazione: (2023)
di: Wei, Ziyang, et al.
Pubblicazione: (2023)
Geometrically Inspired Kernel Machines for Collaborative Learning Beyond Gradient Descent
di: Kumar, Mohit, et al.
Pubblicazione: (2024)
di: Kumar, Mohit, et al.
Pubblicazione: (2024)
Natural Gradient Descent for Online Continual Learning
di: Khawand, Joe, et al.
Pubblicazione: (2026)
di: Khawand, Joe, et al.
Pubblicazione: (2026)
Stability-based Generalization Analysis of Randomized Coordinate Descent for Pairwise Learning
di: Wu, Liang, et al.
Pubblicazione: (2025)
di: Wu, Liang, et al.
Pubblicazione: (2025)
Harmonized Gradient Descent for Class Imbalanced Data Stream Online Learning
di: Zhou, Han, et al.
Pubblicazione: (2025)
di: Zhou, Han, et al.
Pubblicazione: (2025)
Learning Operators by Regularized Stochastic Gradient Descent with Operator-valued Kernels
di: Yang, Jia-Qi, et al.
Pubblicazione: (2025)
di: Yang, Jia-Qi, et al.
Pubblicazione: (2025)
Quantum Algorithm for Sparse Online Learning with Truncated Gradient Descent
di: Lim, Debbie, et al.
Pubblicazione: (2024)
di: Lim, Debbie, et al.
Pubblicazione: (2024)
Limit Theorems for Stochastic Gradient Descent with Infinite Variance
di: Blanchet, Jose, et al.
Pubblicazione: (2024)
di: Blanchet, Jose, et al.
Pubblicazione: (2024)
The Power of Random Features and the Limits of Distribution-Free Gradient Descent
di: Karchmer, Ari, et al.
Pubblicazione: (2025)
di: Karchmer, Ari, et al.
Pubblicazione: (2025)
Central Limit Theorems for Stochastic Gradient Descent Quantile Estimators
di: Wei, Ziyang, et al.
Pubblicazione: (2025)
di: Wei, Ziyang, et al.
Pubblicazione: (2025)
Comparing Federated Stochastic Gradient Descent and Federated Averaging for Predicting Hospital Length of Stay
di: Balik, Mehmet Yigit
Pubblicazione: (2024)
di: Balik, Mehmet Yigit
Pubblicazione: (2024)
Feature Averaging: An Implicit Bias of Gradient Descent Leading to Non-Robustness in Neural Networks
di: Li, Binghui, et al.
Pubblicazione: (2024)
di: Li, Binghui, et al.
Pubblicazione: (2024)
Functional Central Limit Theorem for Stochastic Gradient Descent
di: Flamand, Kessang, et al.
Pubblicazione: (2026)
di: Flamand, Kessang, et al.
Pubblicazione: (2026)
Variational Online Mirror Descent for Robust Learning in Schrödinger Bridge
di: Han, Dong-Sig, et al.
Pubblicazione: (2025)
di: Han, Dong-Sig, et al.
Pubblicazione: (2025)
Quantum Natural Stochastic Pairwise Coordinate Descent
di: Sohail, Mohammad Aamir, et al.
Pubblicazione: (2024)
di: Sohail, Mohammad Aamir, et al.
Pubblicazione: (2024)
Metric Learning from Limited Pairwise Preference Comparisons
di: Wang, Zhi, et al.
Pubblicazione: (2024)
di: Wang, Zhi, et al.
Pubblicazione: (2024)
Beyond Cross-Validation: Adaptive Parameter Selection for Kernel-Based Gradient Descents
di: Liu, Xiaotong, et al.
Pubblicazione: (2026)
di: Liu, Xiaotong, et al.
Pubblicazione: (2026)
Unraveling the Gradient Descent Dynamics of Transformers
di: Song, Bingqing, et al.
Pubblicazione: (2024)
di: Song, Bingqing, et al.
Pubblicazione: (2024)
Curl Descent: Non-Gradient Learning Dynamics with Sign-Diverse Plasticity
di: Ninou, Hugo, et al.
Pubblicazione: (2025)
di: Ninou, Hugo, et al.
Pubblicazione: (2025)
Truncated Kernel Stochastic Gradient Descent on Spheres
di: Bai, Jinhui, et al.
Pubblicazione: (2024)
di: Bai, Jinhui, et al.
Pubblicazione: (2024)
Distributed Gradient Descent for Functional Learning
di: Yu, Zhan, et al.
Pubblicazione: (2023)
di: Yu, Zhan, et al.
Pubblicazione: (2023)
Feedback Descent: Open-Ended Text Optimization via Pairwise Comparison
di: Lee, Yoonho, et al.
Pubblicazione: (2025)
di: Lee, Yoonho, et al.
Pubblicazione: (2025)
Trained Mamba Emulates Online Gradient Descent in In-Context Linear Regression
di: Jiang, Jiarui, et al.
Pubblicazione: (2025)
di: Jiang, Jiarui, et al.
Pubblicazione: (2025)
Online Statistical Inference for Contextual Bandits via Stochastic Gradient Descent
di: Chang, Xiangyu, et al.
Pubblicazione: (2022)
di: Chang, Xiangyu, et al.
Pubblicazione: (2022)
The Limit Points of (Optimistic) Gradient Descent in Min-Max Optimization
di: Daskalakis, Constantinos, et al.
Pubblicazione: (2018)
di: Daskalakis, Constantinos, et al.
Pubblicazione: (2018)
Characterizing Dynamical Stability of Stochastic Gradient Descent in Overparameterized Learning
di: Chemnitz, Dennis, et al.
Pubblicazione: (2024)
di: Chemnitz, Dennis, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Uncovering the Spectral Bias in Diagonal State Space Models
di: Solozabal, Ruben, et al.
Pubblicazione: (2025) -
Constrained Adversarial Perturbation
di: Nishad, Virendra, et al.
Pubblicazione: (2025) -
Mechanistic Insights into Grokking from the Embedding Layer
di: AlquBoj, H. V., et al.
Pubblicazione: (2025) -
New Insight of Variance reduce in Zero-Order Hard-Thresholding: Mitigating Gradient Error and Expansivity Contradictions
di: Yuan, Xinzhe, et al.
Pubblicazione: (2026) -
Zeroth-Order Hard-Thresholding: Gradient Error vs. Expansivity
di: de Vazelhes, William, et al.
Pubblicazione: (2022)