CaAdam: Improving Adam optimizer using connection aware methods
Fuente:
arXiv
Saved in:
| Main Authors: | Genet, Remi, Inzirillo, Hugo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Temporal Linear Network for Time Series Forecasting
by: Genet, Remi, et al.
Published: (2024)
by: Genet, Remi, et al.
Published: (2024)
A Temporal Kolmogorov-Arnold Transformer for Time Series Forecasting
by: Genet, Remi, et al.
Published: (2024)
by: Genet, Remi, et al.
Published: (2024)
SigKAN: Signature-Weighted Kolmogorov-Arnold Networks for Time Series
by: Inzirillo, Hugo, et al.
Published: (2024)
by: Inzirillo, Hugo, et al.
Published: (2024)
SigGate: Enhancing Recurrent Neural Networks with Signature-Based Gating Mechanisms
by: Genet, Rémi, et al.
Published: (2025)
by: Genet, Rémi, et al.
Published: (2025)
STAN: Smooth Transition Autoregressive Networks
by: Inzirillo, Hugo, et al.
Published: (2025)
by: Inzirillo, Hugo, et al.
Published: (2025)
TKAN: Temporal Kolmogorov-Arnold Networks
by: Genet, Remi, et al.
Published: (2024)
by: Genet, Remi, et al.
Published: (2024)
LEMs: A Primer On Large Execution Models
by: Genet, Remi, et al.
Published: (2025)
by: Genet, Remi, et al.
Published: (2025)
A Gated Residual Kolmogorov-Arnold Networks for Mixtures of Experts
by: Inzirillo, Hugo, et al.
Published: (2024)
by: Inzirillo, Hugo, et al.
Published: (2024)
Keras Sig: Efficient Path Signature Computation on GPU in Keras 3
by: Genet, Rémi, et al.
Published: (2025)
by: Genet, Rémi, et al.
Published: (2025)
Deep State Space Recurrent Neural Networks for Time Series Forecasting
by: Inzirillo, Hugo
Published: (2024)
by: Inzirillo, Hugo
Published: (2024)
Clustering Digital Assets Using Path Signatures: Application to Portfolio Construction
by: Inzirillo, Hugo
Published: (2024)
by: Inzirillo, Hugo
Published: (2024)
Convergence rates for the Adam optimizer
by: Dereich, Steffen, et al.
Published: (2024)
by: Dereich, Steffen, et al.
Published: (2024)
SOAP: Improving and Stabilizing Shampoo using Adam
by: Vyas, Nikhil, et al.
Published: (2024)
by: Vyas, Nikhil, et al.
Published: (2024)
WarpAdam: A new Adam optimizer based on Meta-Learning approach
by: Pan, Chengxi, et al.
Published: (2024)
by: Pan, Chengxi, et al.
Published: (2024)
Deep Learning for VWAP Execution in Crypto Markets: Beyond the Volume Curve
by: Genet, Remi
Published: (2025)
by: Genet, Remi
Published: (2025)
Recurrent Neural Networks for Dynamic VWAP Execution: Adaptive Trading Strategies with Temporal Kolmogorov-Arnold Networks
by: Genet, Remi
Published: (2025)
by: Genet, Remi
Published: (2025)
VWAP Execution with Signature-Enhanced Transformers: A Multi-Asset Learning Approach
by: Genet, Remi
Published: (2025)
by: Genet, Remi
Published: (2025)
HomeAdam: Adam and AdamW Algorithms Sometimes Go Home to Obtain Better Provable Generalization
by: Huang, Feihu, et al.
Published: (2026)
by: Huang, Feihu, et al.
Published: (2026)
Adam-HNAG: A Convergent Reformulation of Adam with Accelerated Rate
by: Yu, Yaxin, et al.
Published: (2026)
by: Yu, Yaxin, et al.
Published: (2026)
Adam on Local Time: Addressing Nonstationarity in RL with Relative Adam Timesteps
by: Ellis, Benjamin, et al.
Published: (2024)
by: Ellis, Benjamin, et al.
Published: (2024)
On Suppressing Range of Adaptive Stepsizes of Adam to Improve Generalisation Performance
by: Zhang, Guoqiang
Published: (2023)
by: Zhang, Guoqiang
Published: (2023)
Adam with model exponential moving average is effective for nonconvex optimization
by: Ahn, Kwangjun, et al.
Published: (2024)
by: Ahn, Kwangjun, et al.
Published: (2024)
Batch size invariant Adam
by: Wang, Xi, et al.
Published: (2024)
by: Wang, Xi, et al.
Published: (2024)
Tune My Adam, Please!
by: Athanasiadis, Theodoros, et al.
Published: (2025)
by: Athanasiadis, Theodoros, et al.
Published: (2025)
In Search of Adam's Secret Sauce
by: Orvieto, Antonio, et al.
Published: (2025)
by: Orvieto, Antonio, et al.
Published: (2025)
Understanding Adam Optimizer via Online Learning of Updates: Adam is FTRL in Disguise
by: Ahn, Kwangjun, et al.
Published: (2024)
by: Ahn, Kwangjun, et al.
Published: (2024)
Adam-SHANG: A Convergent Adam-Type Method for Stochastic Smooth Convex Optimization
by: Yu, Yaxin, et al.
Published: (2026)
by: Yu, Yaxin, et al.
Published: (2026)
The Implicit Bias of Adam on Separable Data
by: Zhang, Chenyang, et al.
Published: (2024)
by: Zhang, Chenyang, et al.
Published: (2024)
Adam Simplified: Bias Correction Debunked
by: Laing, Sam, et al.
Published: (2025)
by: Laing, Sam, et al.
Published: (2025)
Refresh-Scaling the Memory of Balanced Adam
by: Fernández-Hernández, Alberto, et al.
Published: (2026)
by: Fernández-Hernández, Alberto, et al.
Published: (2026)
FAdam: Adam is a natural gradient optimizer using diagonal empirical Fisher information
by: Hwang, Dongseong
Published: (2024)
by: Hwang, Dongseong
Published: (2024)
Adam symmetry theorem: characterization of the convergence of the stochastic Adam optimizer
by: Dereich, Steffen, et al.
Published: (2025)
by: Dereich, Steffen, et al.
Published: (2025)
From Adam to Adam-Like Lagrangians: Second-Order Nonlocal Dynamics
by: Heredia, Carlos
Published: (2026)
by: Heredia, Carlos
Published: (2026)
On the Implicit Bias of Adam
by: Cattaneo, Matias D., et al.
Published: (2023)
by: Cattaneo, Matias D., et al.
Published: (2023)
Adaptive Preconditioners Trigger Loss Spikes in Adam
by: Bai, Zhiwei, et al.
Published: (2025)
by: Bai, Zhiwei, et al.
Published: (2025)
Promoting Exploration in Memory-Augmented Adam using Critical Momenta
by: Malviya, Pranshu, et al.
Published: (2023)
by: Malviya, Pranshu, et al.
Published: (2023)
Revisiting Adam for Streaming Reinforcement Learning
by: Gogianu, Florin, et al.
Published: (2026)
by: Gogianu, Florin, et al.
Published: (2026)
In-Run Data Shapley for Adam Optimizer
by: Ding, Meng, et al.
Published: (2026)
by: Ding, Meng, et al.
Published: (2026)
Better Embeddings with Coupled Adam
by: Stollenwerk, Felix, et al.
Published: (2025)
by: Stollenwerk, Felix, et al.
Published: (2025)
A Method for Enhancing Generalization of Adam by Multiple Integrations
by: Jin, Long, et al.
Published: (2024)
by: Jin, Long, et al.
Published: (2024)
Similar Items
-
A Temporal Linear Network for Time Series Forecasting
by: Genet, Remi, et al.
Published: (2024) -
A Temporal Kolmogorov-Arnold Transformer for Time Series Forecasting
by: Genet, Remi, et al.
Published: (2024) -
SigKAN: Signature-Weighted Kolmogorov-Arnold Networks for Time Series
by: Inzirillo, Hugo, et al.
Published: (2024) -
SigGate: Enhancing Recurrent Neural Networks with Signature-Based Gating Mechanisms
by: Genet, Rémi, et al.
Published: (2025) -
STAN: Smooth Transition Autoregressive Networks
by: Inzirillo, Hugo, et al.
Published: (2025)