Hyperparameter Transfer Laws for Non-Recurrent Multi-Path Neural Networks
Fuente:
arXiv
Guardado en:
| Autores principales: | Wu, Shenxi, Zhang, Haosong, Ma, Xingjian, Bian, Shirui, Zhang, Yichi, Chen, Xi, Lin, Wei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Arithmetic-Mean $μ$P for Modern Architectures: A Unified Learning-Rate Scale for CNNs and ResNets
por: Zhang, Haosong, et al.
Publicado: (2025)
por: Zhang, Haosong, et al.
Publicado: (2025)
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
por: Zheng, Yaowei, et al.
Publicado: (2026)
por: Zheng, Yaowei, et al.
Publicado: (2026)
UMOEA/D: A Multiobjective Evolutionary Algorithm for Uniform Pareto Objectives based on Decomposition
por: Zhang, Xiaoyuan, et al.
Publicado: (2024)
por: Zhang, Xiaoyuan, et al.
Publicado: (2024)
Time-Warping Recurrent Neural Networks for Transfer Learning
por: Hirschi, Jonathon
Publicado: (2026)
por: Hirschi, Jonathon
Publicado: (2026)
Uncertainty-Aware Graph Neural Networks: A Multi-Hop Evidence Fusion Approach
por: Chen, Qingfeng, et al.
Publicado: (2025)
por: Chen, Qingfeng, et al.
Publicado: (2025)
Understanding the Mechanisms of Fast Hyperparameter Transfer
por: Ghosh, Nikhil, et al.
Publicado: (2025)
por: Ghosh, Nikhil, et al.
Publicado: (2025)
Unraveling Privacy Risks of Individual Fairness in Graph Neural Networks
por: Zhang, He, et al.
Publicado: (2023)
por: Zhang, He, et al.
Publicado: (2023)
Diffusing to the Top: Boost Graph Neural Networks with Minimal Hyperparameter Tuning
por: Lin, Lequan, et al.
Publicado: (2024)
por: Lin, Lequan, et al.
Publicado: (2024)
Deriving Hyperparameter Scaling Laws via Modern Optimization Theory
por: Shulgin, Egor, et al.
Publicado: (2026)
por: Shulgin, Egor, et al.
Publicado: (2026)
On Optimizing Hyperparameters for Quantum Neural Networks
por: Herbst, Sabrina, et al.
Publicado: (2024)
por: Herbst, Sabrina, et al.
Publicado: (2024)
On Non-asymptotic Theory of Recurrent Neural Networks in Temporal Point Processes
por: Chen, Zhiheng, et al.
Publicado: (2024)
por: Chen, Zhiheng, et al.
Publicado: (2024)
Utilizing Causal Network Markers to Identify Tipping Points ahead of Critical Transition
por: Bian, Shirui, et al.
Publicado: (2024)
por: Bian, Shirui, et al.
Publicado: (2024)
The Sensitivity of Variational Bayesian Neural Network Performance to Hyperparameters
por: Koermer, Scott, et al.
Publicado: (2025)
por: Koermer, Scott, et al.
Publicado: (2025)
Maximal Update Parametrization and Zero-Shot Hyperparameter Transfer for Fourier Neural Operators
por: Li, Shanda, et al.
Publicado: (2025)
por: Li, Shanda, et al.
Publicado: (2025)
Hyperparameter Transfer for Dense Associative Memories
por: Holtzman, Roi, et al.
Publicado: (2026)
por: Holtzman, Roi, et al.
Publicado: (2026)
Hyperparameter Transfer with Mixture-of-Expert Layers
por: Jiang, Tianze, et al.
Publicado: (2026)
por: Jiang, Tianze, et al.
Publicado: (2026)
Rethinking the Relationship between Recurrent and Non-Recurrent Neural Networks: A Study in Sparsity
por: Hershey, Quincy, et al.
Publicado: (2024)
por: Hershey, Quincy, et al.
Publicado: (2024)
Generalization and Risk Bounds for Recurrent Neural Networks
por: Cheng, Xuewei, et al.
Publicado: (2024)
por: Cheng, Xuewei, et al.
Publicado: (2024)
Trustworthy Graph Neural Networks: Aspects, Methods and Trends
por: Zhang, He, et al.
Publicado: (2022)
por: Zhang, He, et al.
Publicado: (2022)
AbstainGNN: Teaching Graph Neural Networks to Abstain for Graph Classification
por: Lin, Xixun, et al.
Publicado: (2026)
por: Lin, Xixun, et al.
Publicado: (2026)
Decision-focused Graph Neural Networks for Combinatorial Optimization
por: Liu, Yang, et al.
Publicado: (2024)
por: Liu, Yang, et al.
Publicado: (2024)
Guaranteeing Conservation Laws with Projection in Physics-Informed Neural Networks
por: Baez, Anthony, et al.
Publicado: (2024)
por: Baez, Anthony, et al.
Publicado: (2024)
Bayesian Optimization for Hyperparameters Tuning in Neural Networks
por: Onorato, Gabriele
Publicado: (2024)
por: Onorato, Gabriele
Publicado: (2024)
On Optimal Hyperparameters for Differentially Private Deep Transfer Learning
por: Rehn, Aki, et al.
Publicado: (2025)
por: Rehn, Aki, et al.
Publicado: (2025)
Configuration-to-Performance Scaling Law with Neural Ansatz
por: Zhang, Huaqing, et al.
Publicado: (2026)
por: Zhang, Huaqing, et al.
Publicado: (2026)
Scaling Laws for Black box Adversarial Attacks
por: Liu, Chuan, et al.
Publicado: (2024)
por: Liu, Chuan, et al.
Publicado: (2024)
Enhancing Continuous Domain Adaptation with Multi-Path Transfer Curriculum
por: Liu, Hanbing, et al.
Publicado: (2024)
por: Liu, Hanbing, et al.
Publicado: (2024)
Principled Architecture-aware Scaling of Hyperparameters
por: Chen, Wuyang, et al.
Publicado: (2024)
por: Chen, Wuyang, et al.
Publicado: (2024)
Hyperparameter Transfer Enables Consistent Gains of Matrix-Preconditioned Optimizers Across Scales
por: Qiu, Shikai, et al.
Publicado: (2025)
por: Qiu, Shikai, et al.
Publicado: (2025)
Jump Diffusion-Informed Neural Networks with Transfer Learning for Accurate American Option Pricing under Data Scarcity
por: Sun, Qiguo, et al.
Publicado: (2024)
por: Sun, Qiguo, et al.
Publicado: (2024)
FlexHB: a More Efficient and Flexible Framework for Hyperparameter Optimization
por: Zhang, Yang, et al.
Publicado: (2024)
por: Zhang, Yang, et al.
Publicado: (2024)
Towards Neural Scaling Laws for Time Series Foundation Models
por: Yao, Qingren, et al.
Publicado: (2024)
por: Yao, Qingren, et al.
Publicado: (2024)
Signal-Adaptive Trust Regions for Gradient-Free Optimization of Recurrent Spiking Neural Networks
por: Li, Jinhao, et al.
Publicado: (2026)
por: Li, Jinhao, et al.
Publicado: (2026)
The Unreasonable Effectiveness Of Early Discarding After One Epoch In Neural Network Hyperparameter Optimization
por: Egele, Romain, et al.
Publicado: (2024)
por: Egele, Romain, et al.
Publicado: (2024)
Revisiting Edge Perturbation for Graph Neural Network in Graph Data Augmentation and Attack
por: Liu, Xin, et al.
Publicado: (2024)
por: Liu, Xin, et al.
Publicado: (2024)
Music Emotion Prediction Using Recurrent Neural Networks
por: Chang, Xinyu, et al.
Publicado: (2024)
por: Chang, Xinyu, et al.
Publicado: (2024)
Spectral Pruning for Recurrent Neural Networks
por: Furuya, Takashi, et al.
Publicado: (2021)
por: Furuya, Takashi, et al.
Publicado: (2021)
Investigating Sparsity in Recurrent Neural Networks
por: Darji, Harshil
Publicado: (2024)
por: Darji, Harshil
Publicado: (2024)
Union Subgraph Neural Networks
por: Xu, Jiaxing, et al.
Publicado: (2023)
por: Xu, Jiaxing, et al.
Publicado: (2023)
Graph Neural Networks for Graphs with Heterophily: A Survey
por: Zheng, Xin, et al.
Publicado: (2022)
por: Zheng, Xin, et al.
Publicado: (2022)
Ejemplares similares
-
Arithmetic-Mean $μ$P for Modern Architectures: A Unified Learning-Rate Scale for CNNs and ResNets
por: Zhang, Haosong, et al.
Publicado: (2025) -
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
por: Zheng, Yaowei, et al.
Publicado: (2026) -
UMOEA/D: A Multiobjective Evolutionary Algorithm for Uniform Pareto Objectives based on Decomposition
por: Zhang, Xiaoyuan, et al.
Publicado: (2024) -
Time-Warping Recurrent Neural Networks for Transfer Learning
por: Hirschi, Jonathon
Publicado: (2026) -
Uncertainty-Aware Graph Neural Networks: A Multi-Hop Evidence Fusion Approach
por: Chen, Qingfeng, et al.
Publicado: (2025)