AlphaGrad: Non-Linear Gradient Normalization Optimizer
Fuente:
arXiv
Saved in:
| Main Author: | Sane, Soham |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AM-PPO: (Advantage) Alpha-Modulation with Proximal Policy Optimization
by: Sane, Soham
Published: (2025)
by: Sane, Soham
Published: (2025)
Understanding Transformer Optimization via Gradient Heterogeneity
by: Tomihari, Akiyoshi, et al.
Published: (2025)
by: Tomihari, Akiyoshi, et al.
Published: (2025)
On Non-Linear operators for Geometric Deep Learning
by: Sergeant-Perthuis, Grégoire, et al.
Published: (2022)
by: Sergeant-Perthuis, Grégoire, et al.
Published: (2022)
AlphaEvolve: A coding agent for scientific and algorithmic discovery
by: Novikov, Alexander, et al.
Published: (2025)
by: Novikov, Alexander, et al.
Published: (2025)
Recurrent Neural Networks Learn to Store and Generate Sequences using Non-Linear Representations
by: Csordás, Róbert, et al.
Published: (2024)
by: Csordás, Róbert, et al.
Published: (2024)
Enhancing Cross Entropy with a Linearly Adaptive Loss Function for Optimized Classification Performance
by: Shim, Jae Wan
Published: (2025)
by: Shim, Jae Wan
Published: (2025)
The Alpha-Alternator: Dynamic Adaptation To Varying Noise Levels In Sequences Using The Vendi Score For Improved Robustness and Performance
by: Rezaei, Mohammad Reza, et al.
Published: (2025)
by: Rezaei, Mohammad Reza, et al.
Published: (2025)
Enhancing Neural Network Representations with Prior Knowledge-Based Normalization
by: Faye, Bilal, et al.
Published: (2024)
by: Faye, Bilal, et al.
Published: (2024)
Unveiling the Potential of Spiking Dynamics in Graph Representation Learning through Spatial-Temporal Normalization and Coding Strategies
by: Xu, Mingkun, et al.
Published: (2024)
by: Xu, Mingkun, et al.
Published: (2024)
A Truly Sparse and General Implementation of Gradient-Based Synaptic Plasticity
by: Lohoff, Jamie, et al.
Published: (2025)
by: Lohoff, Jamie, et al.
Published: (2025)
Robust Lagrangian and Adversarial Policy Gradient for Robust Constrained Markov Decision Processes
by: Bossens, David M.
Published: (2023)
by: Bossens, David M.
Published: (2023)
Gradient-Free Training of Spiking Neural Networks via Low-Rank Evolution Strategies
by: Patankar, Dhruv, et al.
Published: (2026)
by: Patankar, Dhruv, et al.
Published: (2026)
Advancing Direct Training for Spiking Neural Networks with Circulate-Firing Neurons and Learnable Gradients
by: Zhou, Feifan, et al.
Published: (2026)
by: Zhou, Feifan, et al.
Published: (2026)
Enabling Robust In-Context Memory and Rapid Task Adaptation in Transformers with Hebbian and Gradient-Based Plasticity
by: Chaudhary, Siddharth
Published: (2025)
by: Chaudhary, Siddharth
Published: (2025)
Gradient-Free Continual Learning in Spiking Neural Networks via Inter-Spike Interval Regularization
by: Roy, Samrendra, et al.
Published: (2026)
by: Roy, Samrendra, et al.
Published: (2026)
Predicting Deterioration in Mild Cognitive Impairment with Survival Transformers, Extreme Gradient Boosting and Cox Proportional Hazard Modelling
by: Musto, Henry, et al.
Published: (2024)
by: Musto, Henry, et al.
Published: (2024)
LASER: Linear Compression in Wireless Distributed Optimization
by: Makkuva, Ashok Vardhan, et al.
Published: (2023)
by: Makkuva, Ashok Vardhan, et al.
Published: (2023)
Multi-Timescale Conductance Spiking Networks: A Sparse, Gradient-Trainable Framework with Rich Firing Dynamics for Enhanced Temporal Processing
by: Fulleda-Garcia, Alex, et al.
Published: (2026)
by: Fulleda-Garcia, Alex, et al.
Published: (2026)
Offline Multi-Objective Optimization
by: Xue, Ke, et al.
Published: (2024)
by: Xue, Ke, et al.
Published: (2024)
Surrogate Benchmarks for Model Merging Optimization
by: Akizuki, Rio, et al.
Published: (2025)
by: Akizuki, Rio, et al.
Published: (2025)
Few-shot Quality-Diversity Optimization
by: Salehi, Achkan, et al.
Published: (2021)
by: Salehi, Achkan, et al.
Published: (2021)
Reinforced In-Context Black-Box Optimization
by: Song, Lei, et al.
Published: (2024)
by: Song, Lei, et al.
Published: (2024)
Automated Algorithm Design for Auto-Tuning Optimizers
by: Willemsen, Floris-Jan, et al.
Published: (2025)
by: Willemsen, Floris-Jan, et al.
Published: (2025)
Why Flow Matching is Particle Swarm Optimization?
by: Ouyang, Kaichen
Published: (2025)
by: Ouyang, Kaichen
Published: (2025)
Multi-Task Optimization over Networks of Tasks
by: Hatzky, Julian, et al.
Published: (2026)
by: Hatzky, Julian, et al.
Published: (2026)
Offline Model-Based Optimization by Learning to Rank
by: Tan, Rong-Xi, et al.
Published: (2024)
by: Tan, Rong-Xi, et al.
Published: (2024)
Universality of Linear Recurrences Followed by Non-linear Projections: Finite-Width Guarantees and Benefits of Complex Eigenvalues
by: Orvieto, Antonio, et al.
Published: (2023)
by: Orvieto, Antonio, et al.
Published: (2023)
Implicit Regularization via Spectral Neural Networks and Non-linear Matrix Sensing
by: Chu, Hong T. M., et al.
Published: (2024)
by: Chu, Hong T. M., et al.
Published: (2024)
EOE: Evolutionary Optimization of Experts for Training Language Models
by: Chen, Yingshi
Published: (2025)
by: Chen, Yingshi
Published: (2025)
On Accelerating Edge AI: Optimizing Resource-Constrained Environments
by: Sander, Jacob, et al.
Published: (2025)
by: Sander, Jacob, et al.
Published: (2025)
Position: Leverage Foundational Models for Black-Box Optimization
by: Song, Xingyou, et al.
Published: (2024)
by: Song, Xingyou, et al.
Published: (2024)
A Simple and Efficient Approach to Batch Bayesian Optimization
by: Zhan, Dawei, et al.
Published: (2024)
by: Zhan, Dawei, et al.
Published: (2024)
Task-free Adaptive Meta Black-box Optimization
by: Wang, Chao, et al.
Published: (2026)
by: Wang, Chao, et al.
Published: (2026)
Multi-scale Topology Optimization using Neural Networks
by: Chen, Hongrui, et al.
Published: (2024)
by: Chen, Hongrui, et al.
Published: (2024)
Generalized Population-Based Training for Hyperparameter Optimization in Reinforcement Learning
by: Bai, Hui, et al.
Published: (2024)
by: Bai, Hui, et al.
Published: (2024)
An Improved Grey Wolf Optimization Algorithm for Heart Disease Prediction
by: Niu, Sihan, et al.
Published: (2024)
by: Niu, Sihan, et al.
Published: (2024)
MO-CAPO: Multi-Objective Cost-Aware Prompt Optimization
by: Büssing, Jan, et al.
Published: (2026)
by: Büssing, Jan, et al.
Published: (2026)
Can We Optimize Deep RL Policy Weights as Trajectory Modeling?
by: Tang, Hongyao
Published: (2025)
by: Tang, Hongyao
Published: (2025)
AutoQD: Automatic Discovery of Diverse Behaviors with Quality-Diversity Optimization
by: Hedayatian, Saeed, et al.
Published: (2025)
by: Hedayatian, Saeed, et al.
Published: (2025)
An Amortized Efficiency Threshold for Comparing Neural and Heuristic Solvers in Combinatorial Optimization
by: Afifi, Sohaib
Published: (2026)
by: Afifi, Sohaib
Published: (2026)
Similar Items
-
AM-PPO: (Advantage) Alpha-Modulation with Proximal Policy Optimization
by: Sane, Soham
Published: (2025) -
Understanding Transformer Optimization via Gradient Heterogeneity
by: Tomihari, Akiyoshi, et al.
Published: (2025) -
On Non-Linear operators for Geometric Deep Learning
by: Sergeant-Perthuis, Grégoire, et al.
Published: (2022) -
AlphaEvolve: A coding agent for scientific and algorithmic discovery
by: Novikov, Alexander, et al.
Published: (2025) -
Recurrent Neural Networks Learn to Store and Generate Sequences using Non-Linear Representations
by: Csordás, Róbert, et al.
Published: (2024)