Super Level Sets and Exponential Decay: A Synergistic Approach to Stable Neural Network Training
Fuente:
arXiv
Saved in:
| Main Authors: | Chaudhary, Jatin, Nidhi, Dipak, Heikkonen, Jukka, Merisaari, Haari, Kanth, Rajiv |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pathway-based Progressive Inference (PaPI) for Energy-Efficient Continual Learning
by: Gaurav, Suyash, et al.
Published: (2025)
by: Gaurav, Suyash, et al.
Published: (2025)
Cross-Vendor Reproducibility of Radiomics-based Machine Learning Models for Computer-aided Diagnosis
by: Chaudhary, Jatin, et al.
Published: (2024)
by: Chaudhary, Jatin, et al.
Published: (2024)
Governance-as-a-Service: A Multi-Agent Framework for AI System Compliance and Policy Enforcement
by: Gaurav, Suyash, et al.
Published: (2025)
by: Gaurav, Suyash, et al.
Published: (2025)
Focus Your Attention: Towards Data-Intuitive Lightweight Vision Transformers
by: Gaurav, Suyash, et al.
Published: (2025)
by: Gaurav, Suyash, et al.
Published: (2025)
Training Dynamics of the Cooldown Stage in Warmup-Stable-Decay Learning Rate Scheduler
by: Dremov, Aleksandr, et al.
Published: (2025)
by: Dremov, Aleksandr, et al.
Published: (2025)
Graph Neural Networks at a Fraction
by: Joshi, Rucha Bhalchandra, et al.
Published: (2025)
by: Joshi, Rucha Bhalchandra, et al.
Published: (2025)
Self-Abstraction Learning for Effective and Stable Training of Deep Neural Networks
by: Cho, Wonyong, et al.
Published: (2026)
by: Cho, Wonyong, et al.
Published: (2026)
IDInit: A Universal and Stable Initialization Method for Neural Network Training
by: Pan, Yu, et al.
Published: (2025)
by: Pan, Yu, et al.
Published: (2025)
Super-Exponential Regret for UCT, AlphaGo and Variants
by: Orseau, Laurent, et al.
Published: (2024)
by: Orseau, Laurent, et al.
Published: (2024)
SAMOSA: Sharpness Aware Minimization for Open Set Active learning
by: Kim, Young In, et al.
Published: (2025)
by: Kim, Young In, et al.
Published: (2025)
DASH: Warm-Starting Neural Network Training in Stationary Settings without Loss of Plasticity
by: Shin, Baekrok, et al.
Published: (2024)
by: Shin, Baekrok, et al.
Published: (2024)
Compressing Neural Networks Using Tensor Networks with Exponentially Fewer Variational Parameters
by: Qing, Yong, et al.
Published: (2023)
by: Qing, Yong, et al.
Published: (2023)
MoR: Mixture Of Representations For Mixed-Precision Training
by: Su, Bor-Yiing, et al.
Published: (2025)
by: Su, Bor-Yiing, et al.
Published: (2025)
The Hidden Power of Normalization Layers in Neural Networks: Exponential Capacity Control
by: Than, Khoat
Published: (2025)
by: Than, Khoat
Published: (2025)
Automatically Differentiable Nonlinear Tensor Networks (ADNTNs) for Exponential Compression of Deep Neural Networks
by: Cichocki, Andrzej, et al.
Published: (2026)
by: Cichocki, Andrzej, et al.
Published: (2026)
Studying Cross-cluster Modularity in Neural Networks
by: Golechha, Satvik, et al.
Published: (2025)
by: Golechha, Satvik, et al.
Published: (2025)
Super-Level-Set Regression: Conditional Quantiles via Volume Minimization
by: Braun, Sacha, et al.
Published: (2026)
by: Braun, Sacha, et al.
Published: (2026)
Neural Network-based Vehicular Channel Estimation Performance: Effect of Noise in the Training Set
by: Ngorima, Simbarashe Aldrin, et al.
Published: (2025)
by: Ngorima, Simbarashe Aldrin, et al.
Published: (2025)
What Can You Do When You Have Zero Rewards During RL?
by: Prakash, Jatin, et al.
Published: (2025)
by: Prakash, Jatin, et al.
Published: (2025)
Compelling ReLU Networks to Exhibit Exponentially Many Linear Regions at Initialization and During Training
by: Milkert, Max, et al.
Published: (2023)
by: Milkert, Max, et al.
Published: (2023)
PII-Scope: A Comprehensive Study on Training Data PII Extraction Attacks in LLMs
by: Nakka, Krishna Kanth, et al.
Published: (2024)
by: Nakka, Krishna Kanth, et al.
Published: (2024)
Random-Set Graph Neural Networks
by: Woodley, Tommy, et al.
Published: (2026)
by: Woodley, Tommy, et al.
Published: (2026)
Budgeted Attention Allocation: Cost-Conditioned Compute Control for Efficient Transformers
by: Nidhi, Amrit
Published: (2026)
by: Nidhi, Amrit
Published: (2026)
Set-Valued Sensitivity Analysis of Deep Neural Networks
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
Early-Exit Neural Networks with Nested Prediction Sets
by: Jazbec, Metod, et al.
Published: (2023)
by: Jazbec, Metod, et al.
Published: (2023)
Identifying Backdoored Graphs in Graph Neural Network Training: An Explanation-Based Approach with Novel Metrics
by: Downer, Jane, et al.
Published: (2024)
by: Downer, Jane, et al.
Published: (2024)
LayerPipe2: Multistage Pipelining and Weight Recompute via Improved Exponential Moving Average for Training Neural Networks
by: Unnikrishnan, Nanda K., et al.
Published: (2025)
by: Unnikrishnan, Nanda K., et al.
Published: (2025)
ARDDQN: Attention Recurrent Double Deep Q-Network for UAV Coverage Path Planning and Data Harvesting
by: Kumar, Praveen, et al.
Published: (2024)
by: Kumar, Praveen, et al.
Published: (2024)
Cross-Entropy Optimization for Hyperparameter Optimization in Stochastic Gradient-based Approaches to Train Deep Neural Networks
by: Li, Kevin, et al.
Published: (2024)
by: Li, Kevin, et al.
Published: (2024)
Dimer-Enhanced Optimization: A First-Order Approach to Escaping Saddle Points in Neural Network Training
by: Hu, Yue, et al.
Published: (2025)
by: Hu, Yue, et al.
Published: (2025)
Parallelizing Node-Level Explainability in Graph Neural Networks
by: Llorente, Oscar, et al.
Published: (2026)
by: Llorente, Oscar, et al.
Published: (2026)
Training Neural Networks for Modularity aids Interpretability
by: Golechha, Satvik, et al.
Published: (2024)
by: Golechha, Satvik, et al.
Published: (2024)
Z-Error Loss for Training Neural Networks
by: Godin, Guillaume
Published: (2025)
by: Godin, Guillaume
Published: (2025)
Automatic Stability and Recovery for Neural Network Training
by: Or, Barak
Published: (2026)
by: Or, Barak
Published: (2026)
Gradient-Free Training of Quantized Neural Networks
by: Cohen, Noa, et al.
Published: (2024)
by: Cohen, Noa, et al.
Published: (2024)
Energy Consumption in Parallel Neural Network Training
by: Huber, Philipp, et al.
Published: (2025)
by: Huber, Philipp, et al.
Published: (2025)
VESPO: Variational Sequence-Level Soft Policy Optimization for Stable Off-Policy LLM Training
by: Shen, Guobin, et al.
Published: (2026)
by: Shen, Guobin, et al.
Published: (2026)
WaveGNN: Integrating Graph Neural Networks and Transformers for Decay-Aware Classification of Irregular Clinical Time-Series
by: Hajisafi, Arash, et al.
Published: (2024)
by: Hajisafi, Arash, et al.
Published: (2024)
Q-AGNN: Quantum-Enhanced Attentive Graph Neural Network for Intrusion Detection
by: Chaudhary, Devashish, et al.
Published: (2026)
by: Chaudhary, Devashish, et al.
Published: (2026)
Many Ways to be Right: Rashomon Sets for Concept-Based Neural Networks
by: Feng, Shihan, et al.
Published: (2025)
by: Feng, Shihan, et al.
Published: (2025)
Similar Items
-
Pathway-based Progressive Inference (PaPI) for Energy-Efficient Continual Learning
by: Gaurav, Suyash, et al.
Published: (2025) -
Cross-Vendor Reproducibility of Radiomics-based Machine Learning Models for Computer-aided Diagnosis
by: Chaudhary, Jatin, et al.
Published: (2024) -
Governance-as-a-Service: A Multi-Agent Framework for AI System Compliance and Policy Enforcement
by: Gaurav, Suyash, et al.
Published: (2025) -
Focus Your Attention: Towards Data-Intuitive Lightweight Vision Transformers
by: Gaurav, Suyash, et al.
Published: (2025) -
Training Dynamics of the Cooldown Stage in Warmup-Stable-Decay Learning Rate Scheduler
by: Dremov, Aleksandr, et al.
Published: (2025)