Adaptive multiple optimal learning factors for neural network training
Fuente:
arXiv
Saved in:
| Main Author: | Challagundla, Jeshwanth |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SI-Agent: An Agentic Framework for Feedback-Driven Generation and Tuning of Human-Readable System Instructions for Large Language Models
by: Challagundla, Jeshwanth
Published: (2025)
by: Challagundla, Jeshwanth
Published: (2025)
Simmering: Sufficient is better than optimal for training neural networks
by: Babayan, Irina, et al.
Published: (2024)
by: Babayan, Irina, et al.
Published: (2024)
Advanced Neural Network Architecture for Enhanced Multi-Lead ECG Arrhythmia Detection through Optimized Feature Extraction
by: Challagundla, Bhavith Chandra
Published: (2024)
by: Challagundla, Bhavith Chandra
Published: (2024)
Target noise: A pre-training based neural network initialization for efficient high resolution learning
by: Wang, Shaowen, et al.
Published: (2026)
by: Wang, Shaowen, et al.
Published: (2026)
Representations learnt by SGD and Adaptive learning rules: Conditions that vary sparsity and selectivity in neural networks
by: Park, Jin Hyun
Published: (2022)
by: Park, Jin Hyun
Published: (2022)
Confidence-gated training for efficient early-exit neural networks
by: Mokssit, Saad, et al.
Published: (2025)
by: Mokssit, Saad, et al.
Published: (2025)
LayerCollapse: Adaptive compression of neural networks
by: Shabgahi, Soheil Zibakhsh, et al.
Published: (2023)
by: Shabgahi, Soheil Zibakhsh, et al.
Published: (2023)
PGR-DRC: Pre-Global Routing DRC Violation Prediction Using Unsupervised Learning
by: Islam, Riadul, et al.
Published: (2025)
by: Islam, Riadul, et al.
Published: (2025)
A comparative analysis of a neural network with calculated weights and a neural network with random generation of weights based on the training dataset size
by: Geidarov, Polad
Published: (2025)
by: Geidarov, Polad
Published: (2025)
Understanding the learned look-ahead behavior of chess neural networks
by: Cruz, Diogo
Published: (2025)
by: Cruz, Diogo
Published: (2025)
Making Sigmoid-MSE Great Again: Output Reset Challenges Softmax Cross-Entropy in Neural Network Classification
by: Tyagi, Kanishka, et al.
Published: (2024)
by: Tyagi, Kanishka, et al.
Published: (2024)
Feature learning as alignment: a structural property of gradient descent in non-linear neural networks
by: Beaglehole, Daniel, et al.
Published: (2024)
by: Beaglehole, Daniel, et al.
Published: (2024)
On permutation-invariant neural networks
by: Kimura, Masanari, et al.
Published: (2024)
by: Kimura, Masanari, et al.
Published: (2024)
Sobolev acceleration for neural networks
by: Oh, Jong Kwon, et al.
Published: (2025)
by: Oh, Jong Kwon, et al.
Published: (2025)
Attention mechanisms in neural networks
by: Hays, Hasi
Published: (2026)
by: Hays, Hasi
Published: (2026)
An algorithmic framework for the optimization of deep neural networks architectures and hyperparameters
by: Keisler, Julie, et al.
Published: (2023)
by: Keisler, Julie, et al.
Published: (2023)
Efficient and provably convergent end-to-end training of deep neural networks with linear constraints
by: Yang, Zonglin, et al.
Published: (2026)
by: Yang, Zonglin, et al.
Published: (2026)
Principles of Lipschitz continuity in neural networks
by: Luo, Róisín
Published: (2026)
by: Luo, Róisín
Published: (2026)
Linearity-based neural network compression
by: Dobler, Silas, et al.
Published: (2025)
by: Dobler, Silas, et al.
Published: (2025)
Astral: training physics-informed neural networks with error majorants
by: Fanaskov, Vladimir, et al.
Published: (2024)
by: Fanaskov, Vladimir, et al.
Published: (2024)
On-site estimation of battery electrochemical parameters via transfer learning based physics-informed neural network approach
by: Yeregui, Josu, et al.
Published: (2025)
by: Yeregui, Josu, et al.
Published: (2025)
A framework for measuring the training efficiency of a neural architecture
by: Cueto-Mendoza, Eduardo, et al.
Published: (2024)
by: Cueto-Mendoza, Eduardo, et al.
Published: (2024)
Towards graph neural networks for provably solving convex optimization problems
by: Qian, Chendi, et al.
Published: (2025)
by: Qian, Chendi, et al.
Published: (2025)
Applying graph neural network to SupplyGraph for supply chain network
by: Han, Kihwan
Published: (2024)
by: Han, Kihwan
Published: (2024)
Can neural networks do arithmetic? A survey on the elementary numerical skills of state-of-the-art deep learning models
by: Testolin, Alberto
Published: (2023)
by: Testolin, Alberto
Published: (2023)
Understanding the dynamics of the frequency bias in neural networks
by: Molina, Juan, et al.
Published: (2024)
by: Molina, Juan, et al.
Published: (2024)
Graph neural networks informed locally by thermodynamics
by: Tierz, Alicia, et al.
Published: (2024)
by: Tierz, Alicia, et al.
Published: (2024)
Graph neural networks and non-commuting operators
by: Velasco, Mauricio, et al.
Published: (2024)
by: Velasco, Mauricio, et al.
Published: (2024)
Outlier-robust neural network training: variation regularization meets trimmed loss to prevent functional breakdown
by: Okuno, Akifumi, et al.
Published: (2023)
by: Okuno, Akifumi, et al.
Published: (2023)
Crystal structure prediction using graph neural combinatorial optimization
by: Gerolymatos, Stavros, et al.
Published: (2026)
by: Gerolymatos, Stavros, et al.
Published: (2026)
Randomness and signal propagation in physics-informed neural networks (PINNs): A neural PDE perspective
by: Tucny, Jean-Michel, et al.
Published: (2025)
by: Tucny, Jean-Michel, et al.
Published: (2025)
Planning in a recurrent neural network that plays Sokoban
by: Taufeeque, Mohammad, et al.
Published: (2024)
by: Taufeeque, Mohammad, et al.
Published: (2024)
Understanding polysemanticity in neural networks through coding theory
by: Marshall, Simon C., et al.
Published: (2024)
by: Marshall, Simon C., et al.
Published: (2024)
Variational autoencoder-based neural network model compression
by: Cheng, Liang, et al.
Published: (2024)
by: Cheng, Liang, et al.
Published: (2024)
Conditional computation in neural networks: principles and research trends
by: Scardapane, Simone, et al.
Published: (2024)
by: Scardapane, Simone, et al.
Published: (2024)
Utilizing Lyapunov Exponents in designing deep neural networks
by: Mittra, Tirthankar
Published: (2024)
by: Mittra, Tirthankar
Published: (2024)
Grey-informed neural network for time-series forecasting
by: Xie, Wanli, et al.
Published: (2024)
by: Xie, Wanli, et al.
Published: (2024)
Elimination-compensation pruning for fully-connected neural networks
by: Ballini, Enrico, et al.
Published: (2026)
by: Ballini, Enrico, et al.
Published: (2026)
Concealed Adversarial attacks on neural networks for sequential data
by: Sokerin, Petr, et al.
Published: (2025)
by: Sokerin, Petr, et al.
Published: (2025)
Dopamine-driven synaptic credit assignment in neural networks
by: Nambusubramaniyan, Saranraj, et al.
Published: (2025)
by: Nambusubramaniyan, Saranraj, et al.
Published: (2025)
Similar Items
-
SI-Agent: An Agentic Framework for Feedback-Driven Generation and Tuning of Human-Readable System Instructions for Large Language Models
by: Challagundla, Jeshwanth
Published: (2025) -
Simmering: Sufficient is better than optimal for training neural networks
by: Babayan, Irina, et al.
Published: (2024) -
Advanced Neural Network Architecture for Enhanced Multi-Lead ECG Arrhythmia Detection through Optimized Feature Extraction
by: Challagundla, Bhavith Chandra
Published: (2024) -
Target noise: A pre-training based neural network initialization for efficient high resolution learning
by: Wang, Shaowen, et al.
Published: (2026) -
Representations learnt by SGD and Adaptive learning rules: Conditions that vary sparsity and selectivity in neural networks
by: Park, Jin Hyun
Published: (2022)