Deep Legendre Transform
Fuente:
arXiv
Saved in:
| Main Authors: | Minabutdinov, Aleksey, Cheridito, Patrick |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CAO: Curvature-Adaptive Optimization via Periodic Low-Rank Hessian Sketching
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Local properties of neural networks through the lens of layer-wise Hessians
by: Bolshim, Maxim, et al.
Published: (2025)
by: Bolshim, Maxim, et al.
Published: (2025)
Inter-Layer Hessian Analysis of Neural Networks with DAG Architectures
by: Bolshim, Maxim, et al.
Published: (2026)
by: Bolshim, Maxim, et al.
Published: (2026)
Stochastic Estimation of the Layer-wise Hessian Trace for Monitoring Neural-network Training
by: Bolshim, Maxim, et al.
Published: (2026)
by: Bolshim, Maxim, et al.
Published: (2026)
Differentiable Optimization Layers for Guaranteed Fairness in Deep Learning
by: Troxell, David, et al.
Published: (2026)
by: Troxell, David, et al.
Published: (2026)
ZetA: A Riemann Zeta-Scaled Extension of Adam for Deep Learning
by: BC, Samiksha
Published: (2025)
by: BC, Samiksha
Published: (2025)
Escaping Saddle Points via Curvature-Calibrated Perturbations: A Complete Analysis with Explicit Constants and Empirical Validation
by: Alpay, Faruk, et al.
Published: (2025)
by: Alpay, Faruk, et al.
Published: (2025)
"Abuse Risks are Often Inherent to Product Features": Exploring AI Vendors' Bug Bounty and Responsible Disclosure Policies
by: Piao, Yangheran, et al.
Published: (2025)
by: Piao, Yangheran, et al.
Published: (2025)
prunAdag: an adaptive pruning-aware gradient method
by: Porcelli, Margherita, et al.
Published: (2025)
by: Porcelli, Margherita, et al.
Published: (2025)
Benchmarking Generative AI Against Bayesian Optimization for Constrained Multi-Objective Inverse Design
by: Awan, Muhammad Bilal, et al.
Published: (2025)
by: Awan, Muhammad Bilal, et al.
Published: (2025)
TOPSIS-like metaheuristic for LABS problem
by: Urbańczyk, Aleksandra, et al.
Published: (2025)
by: Urbańczyk, Aleksandra, et al.
Published: (2025)
Sparse Training of Neural Networks based on Multilevel Mirror Descent
by: Lunk, Yannick, et al.
Published: (2026)
by: Lunk, Yannick, et al.
Published: (2026)
Sequential, Parallel and Consecutive Hybrid Evolutionary-Swarm Optimization Metaheuristics
by: Urbańczyk, Piotr, et al.
Published: (2025)
by: Urbańczyk, Piotr, et al.
Published: (2025)
Anarchy in the swarm: Testing informed and uninformed diversity-enhancing mechanisms within PSO framework
by: Urbańczyk, Piotr, et al.
Published: (2026)
by: Urbańczyk, Piotr, et al.
Published: (2026)
Refining Graphical Neural Network Predictions Using Flow Matching for Optimal Power Flow with Constraint-Satisfaction Guarantee
by: Khanal, Kshitiz
Published: (2025)
by: Khanal, Kshitiz
Published: (2025)
Functional Similarity Metric for Neural Networks: Overcoming Parametric Ambiguity via Activation Region Analysis
by: Hennadii, Kutomanov
Published: (2026)
by: Hennadii, Kutomanov
Published: (2026)
Socio-cognitive agent-oriented evolutionary algorithm with trust-based optimization
by: Urbańczyk, Aleksandra, et al.
Published: (2025)
by: Urbańczyk, Aleksandra, et al.
Published: (2025)
NeurOptimisation: The Spiking Way to Evolve
by: Cruz-Duarte, Jorge Mario, et al.
Published: (2025)
by: Cruz-Duarte, Jorge Mario, et al.
Published: (2025)
On the Value of Tokeniser Pretraining in Physics Foundation Models
by: Sotoudeh, Hadi, et al.
Published: (2026)
by: Sotoudeh, Hadi, et al.
Published: (2026)
On the Convergence Behavior of Preconditioned Gradient Descent Toward the Rich Learning Regime
by: Jiang, Shuai, et al.
Published: (2026)
by: Jiang, Shuai, et al.
Published: (2026)
i-DEQ: A stable inertial deep equilibrium model for image restoration
by: Clerc, Antonin, et al.
Published: (2026)
by: Clerc, Antonin, et al.
Published: (2026)
Enhancing Model Based Derivative Free Optimization using Direct Search
by: Li, Zijun, et al.
Published: (2026)
by: Li, Zijun, et al.
Published: (2026)
NOVAK: Unified adaptive optimizer for deep neural networks
by: Kavun, Sergii
Published: (2026)
by: Kavun, Sergii
Published: (2026)
Revisiting Non-separable Binary Classification and its Applications in Anomaly Detection
by: Lau, Matthew, et al.
Published: (2023)
by: Lau, Matthew, et al.
Published: (2023)
A Reduction from Delayed to Immediate Feedback for Online Convex Optimization with Improved Guarantees
by: Ryabchenko, Alexander, et al.
Published: (2026)
by: Ryabchenko, Alexander, et al.
Published: (2026)
Decentralized Optimization with Topology-Independent Communication
by: Lin, Ying, et al.
Published: (2025)
by: Lin, Ying, et al.
Published: (2025)
An infeasible interior-point arc-search method with Nesterov's restarting strategy for linear programming problems
by: Iida, Einosuke, et al.
Published: (2023)
by: Iida, Einosuke, et al.
Published: (2023)
PH-VAE: A Polynomial Hierarchical Variational Autoencoder Towards Disentangled Representation Learning
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
Estimating the Event-Related Potential from Few EEG Trials
by: Nørskov, Anders Vestergaard, et al.
Published: (2025)
by: Nørskov, Anders Vestergaard, et al.
Published: (2025)
Tricks and Plug-ins for Gradient Boosting with Transformers
by: Fang, Biyi, et al.
Published: (2025)
by: Fang, Biyi, et al.
Published: (2025)
Adam symmetry theorem: characterization of the convergence of the stochastic Adam optimizer
by: Dereich, Steffen, et al.
Published: (2025)
by: Dereich, Steffen, et al.
Published: (2025)
Uniform a priori bounds and error analysis for the Adam stochastic gradient descent optimization method
by: Dereich, Steffen, et al.
Published: (2026)
by: Dereich, Steffen, et al.
Published: (2026)
Gradient descent provably escapes saddle points in the training of shallow ReLU networks
by: Cheridito, Patrick, et al.
Published: (2022)
by: Cheridito, Patrick, et al.
Published: (2022)
Adam Improves Muon: Adaptive Moment Estimation with Orthogonalized Momentum
by: Zhang, Minxin, et al.
Published: (2026)
by: Zhang, Minxin, et al.
Published: (2026)
Explicit Dropout: Deterministic Regularization for Transformer Architectures
by: Agrawal, Vidhi, et al.
Published: (2026)
by: Agrawal, Vidhi, et al.
Published: (2026)
Optimizing Inference in Transformer-Based Models: A Multi-Method Benchmark
by: Ho, Siu Hang, et al.
Published: (2025)
by: Ho, Siu Hang, et al.
Published: (2025)
Temporal Anchoring in Deepening Embedding Spaces: Event-Indexed Projections, Drift, Convergence, and an Internal Computational Architecture
by: Alpay, Faruk, et al.
Published: (2025)
by: Alpay, Faruk, et al.
Published: (2025)
Kourkoutas-Beta: A Sunspike-Driven Adam Optimizer with Desert Flair
by: Kassinos, Stavros C.
Published: (2025)
by: Kassinos, Stavros C.
Published: (2025)
From Non-Identifiability to Goal-Integrated Decision-Making in Parametric Inverse Optimization
by: Ahmadi, Farzin, et al.
Published: (2026)
by: Ahmadi, Farzin, et al.
Published: (2026)
Gradient Descent Methods for Regularized Optimization
by: Nikolovski, Filip, et al.
Published: (2024)
by: Nikolovski, Filip, et al.
Published: (2024)
Similar Items
-
CAO: Curvature-Adaptive Optimization via Periodic Low-Rank Hessian Sketching
by: Du, Wenzhang
Published: (2025) -
Local properties of neural networks through the lens of layer-wise Hessians
by: Bolshim, Maxim, et al.
Published: (2025) -
Inter-Layer Hessian Analysis of Neural Networks with DAG Architectures
by: Bolshim, Maxim, et al.
Published: (2026) -
Stochastic Estimation of the Layer-wise Hessian Trace for Monitoring Neural-network Training
by: Bolshim, Maxim, et al.
Published: (2026) -
Differentiable Optimization Layers for Guaranteed Fairness in Deep Learning
by: Troxell, David, et al.
Published: (2026)