Kourkoutas-Beta: A Sunspike-Driven Adam Optimizer with Desert Flair
Fuente:
arXiv
Saved in:
| Main Author: | Kassinos, Stavros C. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ZetA: A Riemann Zeta-Scaled Extension of Adam for Deep Learning
by: BC, Samiksha
Published: (2025)
by: BC, Samiksha
Published: (2025)
On the Convergence Behavior of Preconditioned Gradient Descent Toward the Rich Learning Regime
by: Jiang, Shuai, et al.
Published: (2026)
by: Jiang, Shuai, et al.
Published: (2026)
Adam Improves Muon: Adaptive Moment Estimation with Orthogonalized Momentum
by: Zhang, Minxin, et al.
Published: (2026)
by: Zhang, Minxin, et al.
Published: (2026)
Local properties of neural networks through the lens of layer-wise Hessians
by: Bolshim, Maxim, et al.
Published: (2025)
by: Bolshim, Maxim, et al.
Published: (2025)
Inter-Layer Hessian Analysis of Neural Networks with DAG Architectures
by: Bolshim, Maxim, et al.
Published: (2026)
by: Bolshim, Maxim, et al.
Published: (2026)
Stochastic Estimation of the Layer-wise Hessian Trace for Monitoring Neural-network Training
by: Bolshim, Maxim, et al.
Published: (2026)
by: Bolshim, Maxim, et al.
Published: (2026)
FlowAdam: Implicit Regularization via Geometry-Aware Soft Momentum Injection
by: Singh, Devender, et al.
Published: (2026)
by: Singh, Devender, et al.
Published: (2026)
Benchmarking Generative AI Against Bayesian Optimization for Constrained Multi-Objective Inverse Design
by: Awan, Muhammad Bilal, et al.
Published: (2025)
by: Awan, Muhammad Bilal, et al.
Published: (2025)
CAO: Curvature-Adaptive Optimization via Periodic Low-Rank Hessian Sketching
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Sparsifying dimensionality reduction of PDE solution data with Bregman learning
by: Heeringa, Tjeerd Jan, et al.
Published: (2024)
by: Heeringa, Tjeerd Jan, et al.
Published: (2024)
Data-induced multiscale losses and efficient multirate gradient descent schemes
by: He, Juncai, et al.
Published: (2024)
by: He, Juncai, et al.
Published: (2024)
Deceptron: Learned Local Inverses for Fast and Stable Physics Inversion
by: Kachhadiya, Aaditya L.
Published: (2025)
by: Kachhadiya, Aaditya L.
Published: (2025)
EB-gMCR: Energy-Based Generative Modeling for Signal Unmixing and Multivariate Curve Resolution
by: Chang, Yu-Tang, et al.
Published: (2025)
by: Chang, Yu-Tang, et al.
Published: (2025)
NeurOptimisation: The Spiking Way to Evolve
by: Cruz-Duarte, Jorge Mario, et al.
Published: (2025)
by: Cruz-Duarte, Jorge Mario, et al.
Published: (2025)
A Dual-Path Generative Framework for Zero-Day Fraud Detection in Banking Systems
by: Ismail, Nasim Abdirahman, et al.
Published: (2026)
by: Ismail, Nasim Abdirahman, et al.
Published: (2026)
Refining Graphical Neural Network Predictions Using Flow Matching for Optimal Power Flow with Constraint-Satisfaction Guarantee
by: Khanal, Kshitiz
Published: (2025)
by: Khanal, Kshitiz
Published: (2025)
i-DEQ: A stable inertial deep equilibrium model for image restoration
by: Clerc, Antonin, et al.
Published: (2026)
by: Clerc, Antonin, et al.
Published: (2026)
Evaluation of Differential Privacy Mechanisms on Federated Learning
by: Varsani, Tejash
Published: (2025)
by: Varsani, Tejash
Published: (2025)
Deep Legendre Transform
by: Minabutdinov, Aleksey, et al.
Published: (2025)
by: Minabutdinov, Aleksey, et al.
Published: (2025)
Ghosts of Softmax: Complex Singularities That Limit Safe Step Sizes in Cross-Entropy
by: Sao, Piyush
Published: (2026)
by: Sao, Piyush
Published: (2026)
JacNet: Learning Functions with Structured Jacobians
by: Lorraine, Jonathan, et al.
Published: (2024)
by: Lorraine, Jonathan, et al.
Published: (2024)
Stabilized Adaptive Loss and Residual-Based Collocation for Physics-Informed Neural Networks
by: Singh, Divyavardhan, et al.
Published: (2026)
by: Singh, Divyavardhan, et al.
Published: (2026)
Sparse Training of Neural Networks based on Multilevel Mirror Descent
by: Lunk, Yannick, et al.
Published: (2026)
by: Lunk, Yannick, et al.
Published: (2026)
Sprecher Networks: A Parameter-Efficient Kolmogorov-Arnold Architecture
by: Hägg, Christian, et al.
Published: (2025)
by: Hägg, Christian, et al.
Published: (2025)
Gradient descent provably escapes saddle points in the training of shallow ReLU networks
by: Cheridito, Patrick, et al.
Published: (2022)
by: Cheridito, Patrick, et al.
Published: (2022)
TED++: Submanifold-Aware Backdoor Detection via Layerwise Tubular-Neighbourhood Screening
by: Le, Nam, et al.
Published: (2025)
by: Le, Nam, et al.
Published: (2025)
Learning Hamiltonian flows from numerical integrators and examples
by: Fang, Rui, et al.
Published: (2025)
by: Fang, Rui, et al.
Published: (2025)
J6: Jacobian-Driven Role Attribution for Multi-Objective Prompt Optimization in LLMs
by: Wu, Yao
Published: (2025)
by: Wu, Yao
Published: (2025)
SigGate-GT: Taming Over-Smoothing in Graph Transformers via Sigmoid-Gated Attention
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
The Non-Linearity Perturbation Threshold: Width Scaling and Landscape Bifurcations in Deep Learning
by: Alexander, Michael
Published: (2026)
by: Alexander, Michael
Published: (2026)
Temporal Anchoring in Deepening Embedding Spaces: Event-Indexed Projections, Drift, Convergence, and an Internal Computational Architecture
by: Alpay, Faruk, et al.
Published: (2025)
by: Alpay, Faruk, et al.
Published: (2025)
Analyzing Closed-loop Training Techniques for Realistic Traffic Agent Models in Autonomous Highway Driving Simulations
by: Bitzer, Matthias, et al.
Published: (2024)
by: Bitzer, Matthias, et al.
Published: (2024)
$δ$-STEAL: LLM Stealing Attack with Local Differential Privacy
by: Dang, Kieu, et al.
Published: (2025)
by: Dang, Kieu, et al.
Published: (2025)
Improving Hyperparameter Optimization with Checkpointed Model Weights
by: Mehta, Nikhil, et al.
Published: (2024)
by: Mehta, Nikhil, et al.
Published: (2024)
Total Generalized Variation regularization closes the gap between neural-eld and classical methods in seismic travel-time tomography
by: Kurosawa, Isao
Published: (2026)
by: Kurosawa, Isao
Published: (2026)
Active Learning for Conditional Generative Compressed Sensing
by: DeLise, Alexander, et al.
Published: (2026)
by: DeLise, Alexander, et al.
Published: (2026)
Physics Informed Differentiable Solvers for Learning Parametric Solution Manifolds in Heterogeneous Physical Systems
by: Panahi, Milad, et al.
Published: (2026)
by: Panahi, Milad, et al.
Published: (2026)
Learning Nonlinear Finite Element Solution Operators using Multilayer Perceptrons and Energy Minimization
by: Larson, Mats G., et al.
Published: (2024)
by: Larson, Mats G., et al.
Published: (2024)
The Age of Sensorial Zero Trust: Why We Can No Longer Trust Our Senses
by: Xavier, Fabio Correa
Published: (2025)
by: Xavier, Fabio Correa
Published: (2025)
A Reinforcement Learning Method for Environments with Stochastic Variables: Post-Decision Proximal Policy Optimization with Dual Critic Networks
by: Felizardo, Leonardo Kanashiro, et al.
Published: (2025)
by: Felizardo, Leonardo Kanashiro, et al.
Published: (2025)
Similar Items
-
ZetA: A Riemann Zeta-Scaled Extension of Adam for Deep Learning
by: BC, Samiksha
Published: (2025) -
On the Convergence Behavior of Preconditioned Gradient Descent Toward the Rich Learning Regime
by: Jiang, Shuai, et al.
Published: (2026) -
Adam Improves Muon: Adaptive Moment Estimation with Orthogonalized Momentum
by: Zhang, Minxin, et al.
Published: (2026) -
Local properties of neural networks through the lens of layer-wise Hessians
by: Bolshim, Maxim, et al.
Published: (2025) -
Inter-Layer Hessian Analysis of Neural Networks with DAG Architectures
by: Bolshim, Maxim, et al.
Published: (2026)