Exploring the loss landscape of regularized neural networks via convex duality
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Sungyoon, Mishkin, Aaron, Pilanci, Mert |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimal Sets and Solution Paths of ReLU Networks
by: Mishkin, Aaron, et al.
Published: (2023)
by: Mishkin, Aaron, et al.
Published: (2023)
Convex Relaxations of ReLU Neural Networks Approximate Global Optima in Polynomial Time
by: Kim, Sungyoon, et al.
Published: (2024)
by: Kim, Sungyoon, et al.
Published: (2024)
Fast Convex Optimization for Two-Layer ReLU Networks: Equivalent Model Classes and Cone Decompositions
by: Mishkin, Aaron, et al.
Published: (2022)
by: Mishkin, Aaron, et al.
Published: (2022)
Faster Convergence of Stochastic Accelerated Gradient Descent under Interpolation
by: Mishkin, Aaron, et al.
Published: (2024)
by: Mishkin, Aaron, et al.
Published: (2024)
MatRL: Provably Generalizable Iterative Algorithm Discovery via Monte-Carlo Tree Search
by: Kim, Sungyoon, et al.
Published: (2025)
by: Kim, Sungyoon, et al.
Published: (2025)
Randomized Geometric Algebra Methods for Convex Neural Networks
by: Wang, Yifei, et al.
Published: (2024)
by: Wang, Yifei, et al.
Published: (2024)
Optimizer-Induced Mode Connectivity: From AdamW to Muon
by: Zhang, Fangzhao, et al.
Published: (2026)
by: Zhang, Fangzhao, et al.
Published: (2026)
From Complexity to Clarity: Analytical Expressions of Deep Neural Network Weights via Clifford's Geometric Algebra and Convexity
by: Pilanci, Mert
Published: (2023)
by: Pilanci, Mert
Published: (2023)
Convex Distillation: Efficient Compression of Deep Networks via Convex Optimization
by: Varshney, Prateek, et al.
Published: (2024)
by: Varshney, Prateek, et al.
Published: (2024)
A Library of Mirrors: Deep Neural Nets in Low Dimensions are Convex Lasso Models with Reflection Features
by: Zeger, Emi, et al.
Published: (2024)
by: Zeger, Emi, et al.
Published: (2024)
Black Boxes and Looking Glasses: Multilevel Symmetries, Reflection Planes, and Convex Optimization in Deep Networks
by: Zeger, Emi, et al.
Published: (2024)
by: Zeger, Emi, et al.
Published: (2024)
Visualizing the loss landscapes of physics-informed neural networks
by: Rowan, Conor, et al.
Published: (2026)
by: Rowan, Conor, et al.
Published: (2026)
Optimal Scalar Quantization for Matrix Multiplication: Closed-Form Density and Phase Transition
by: Ang, Calvin, et al.
Published: (2026)
by: Ang, Calvin, et al.
Published: (2026)
Analyzing Neural Network-Based Generative Diffusion Models through Convex Optimization
by: Zhang, Fangzhao, et al.
Published: (2024)
by: Zhang, Fangzhao, et al.
Published: (2024)
Spectral Adapter: Fine-Tuning in Spectral Space
by: Zhang, Fangzhao, et al.
Published: (2024)
by: Zhang, Fangzhao, et al.
Published: (2024)
Unveiling Hidden Convexity in Deep Learning: a Sparse Signal Processing Perspective
by: Zeger, Emi, et al.
Published: (2026)
by: Zeger, Emi, et al.
Published: (2026)
Localizing Memorized Regions in Diffusion Models via Coordinate-Wise Curvature Differences
by: Kim, Gwangho, et al.
Published: (2026)
by: Kim, Gwangho, et al.
Published: (2026)
Active Learning of Deep Neural Networks via Gradient-Free Cutting Planes
by: Zhang, Erica, et al.
Published: (2024)
by: Zhang, Erica, et al.
Published: (2024)
Adversarial Training of Two-Layer Polynomial and ReLU Activation Networks via Convex Optimization
by: Kuelbs, Daniel, et al.
Published: (2024)
by: Kuelbs, Daniel, et al.
Published: (2024)
Riemannian Preconditioned LoRA for Fine-Tuning Foundation Models
by: Zhang, Fangzhao, et al.
Published: (2024)
by: Zhang, Fangzhao, et al.
Published: (2024)
A Recovery Guarantee for Sparse Neural Networks
by: Fridovich-Keil, Sara, et al.
Published: (2025)
by: Fridovich-Keil, Sara, et al.
Published: (2025)
Thinking While Listening: Simple Test Time Scaling For Audio Classification
by: Verma, Prateek, et al.
Published: (2025)
by: Verma, Prateek, et al.
Published: (2025)
FlashSketch: Sketch-Kernel Co-Design for Fast Sparse Sketching on GPUs
by: Dwaraknath, Rajat Vadiraj, et al.
Published: (2026)
by: Dwaraknath, Rajat Vadiraj, et al.
Published: (2026)
Adaptive Large Language Models By Layerwise Attention Shortcuts
by: Verma, Prateek, et al.
Published: (2024)
by: Verma, Prateek, et al.
Published: (2024)
Towards Signal Processing In Large Language Models
by: Verma, Prateek, et al.
Published: (2024)
by: Verma, Prateek, et al.
Published: (2024)
Dynamical loss functions shape landscape topography and improve learning in artificial neural networks
by: Pallero, Eduardo Lavin, et al.
Published: (2024)
by: Pallero, Eduardo Lavin, et al.
Published: (2024)
CRONOS: Enhancing Deep Learning with Scalable GPU Accelerated Convex Neural Networks
by: Feng, Miria, et al.
Published: (2024)
by: Feng, Miria, et al.
Published: (2024)
The loss landscape of deep linear neural networks: a second-order analysis
by: Achour, El Mehdi, et al.
Published: (2021)
by: Achour, El Mehdi, et al.
Published: (2021)
AdaPTwin: Low-Cost Adaptive Compression of Product Twins in Transformers
by: Biju, Emil, et al.
Published: (2024)
by: Biju, Emil, et al.
Published: (2024)
Large Language Models Implicitly Learn to See and Hear Just By Reading
by: Verma, Prateek, et al.
Published: (2025)
by: Verma, Prateek, et al.
Published: (2025)
Convex Optimization for Alignment and Preference Learning on a Single GPU
by: Feng, Miria, et al.
Published: (2026)
by: Feng, Miria, et al.
Published: (2026)
Adaptive Inference: Theoretical Limits and Unexplored Opportunities
by: Hor, Soheil, et al.
Published: (2024)
by: Hor, Soheil, et al.
Published: (2024)
Investigating generalization capabilities of neural networks by means of loss landscapes and Hessian analysis
by: Gabdullin, Nikita
Published: (2024)
by: Gabdullin, Nikita
Published: (2024)
Convex Low-resource Accent-Robust Language Detection in Speech Recognition
by: Feng, Miria, et al.
Published: (2026)
by: Feng, Miria, et al.
Published: (2026)
Newton Meets Marchenko-Pastur: Massively Parallel Second-Order Optimization with Hessian Sketching and Debiasing
by: Romanov, Elad, et al.
Published: (2024)
by: Romanov, Elad, et al.
Published: (2024)
The effect of the number of parameters and the number of local feature patches on loss landscapes in distributed quantum neural networks
by: Kawase, Yoshiaki
Published: (2025)
by: Kawase, Yoshiaki
Published: (2025)
Improved Physics-informed neural networks loss function regularization with a variance-based term
by: Hanna, John M., et al.
Published: (2024)
by: Hanna, John M., et al.
Published: (2024)
A lift for input-convex neural network training
by: Siahkoohi, Ali, et al.
Published: (2026)
by: Siahkoohi, Ali, et al.
Published: (2026)
Inconsistency-Aware Minimization: Improving Generalization with Unlabeled Data
by: Kim, Hee-Sung, et al.
Published: (2026)
by: Kim, Hee-Sung, et al.
Published: (2026)
Level Set Teleportation: An Optimization Perspective
by: Mishkin, Aaron, et al.
Published: (2024)
by: Mishkin, Aaron, et al.
Published: (2024)
Similar Items
-
Optimal Sets and Solution Paths of ReLU Networks
by: Mishkin, Aaron, et al.
Published: (2023) -
Convex Relaxations of ReLU Neural Networks Approximate Global Optima in Polynomial Time
by: Kim, Sungyoon, et al.
Published: (2024) -
Fast Convex Optimization for Two-Layer ReLU Networks: Equivalent Model Classes and Cone Decompositions
by: Mishkin, Aaron, et al.
Published: (2022) -
Faster Convergence of Stochastic Accelerated Gradient Descent under Interpolation
by: Mishkin, Aaron, et al.
Published: (2024) -
MatRL: Provably Generalizable Iterative Algorithm Discovery via Monte-Carlo Tree Search
by: Kim, Sungyoon, et al.
Published: (2025)