Configuration-to-Performance Scaling Law with Neural Ansatz
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Huaqing, Wen, Kaiyue, Ma, Tengyu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Divide-and-Conquer CoT: RL for Reducing Latency via Parallel Reasoning
by: Mahankali, Arvind, et al.
Published: (2026)
by: Mahankali, Arvind, et al.
Published: (2026)
Scaling Self-Play with Self-Guidance
by: Bailey, Luke, et al.
Published: (2026)
by: Bailey, Luke, et al.
Published: (2026)
Fantastic Pretraining Optimizers and Where to Find Them
by: Wen, Kaiyue, et al.
Published: (2025)
by: Wen, Kaiyue, et al.
Published: (2025)
From Sparse Dependence to Sparse Attention: Unveiling How Chain-of-Thought Enhances Transformer Sample Efficiency
by: Wen, Kaiyue, et al.
Published: (2024)
by: Wen, Kaiyue, et al.
Published: (2024)
Pseudo-Formalization for Automatic Proof Verification
by: Barkallah, Slim, et al.
Published: (2026)
by: Barkallah, Slim, et al.
Published: (2026)
Task Generalization With AutoRegressive Compositional Structure: Can Learning From $D$ Tasks Generalize to $D^{T}$ Tasks?
by: Abedsoltan, Amirhesam, et al.
Published: (2025)
by: Abedsoltan, Amirhesam, et al.
Published: (2025)
Understanding Warmup-Stable-Decay Learning Rates: A River Valley Loss Landscape Perspective
by: Wen, Kaiyue, et al.
Published: (2024)
by: Wen, Kaiyue, et al.
Published: (2024)
On the Neural Feature Ansatz for Deep Neural Networks
by: Tansley, Edward, et al.
Published: (2025)
by: Tansley, Edward, et al.
Published: (2025)
Utility Boundary of Dataset Distillation: Scaling and Configuration-Coverage Laws
by: Luo, Zhengquan, et al.
Published: (2025)
by: Luo, Zhengquan, et al.
Published: (2025)
Neural Neural Scaling Laws
by: Hu, Michael Y., et al.
Published: (2026)
by: Hu, Michael Y., et al.
Published: (2026)
Renormalizable Spectral-Shell Dynamics as the Origin of Neural Scaling Laws
by: Zhang, Yizhou
Published: (2025)
by: Zhang, Yizhou
Published: (2025)
On the Invariance and Generality of Neural Scaling Laws
by: Han, Xing, et al.
Published: (2026)
by: Han, Xing, et al.
Published: (2026)
Do Neural Scaling Laws Exist on Graph Self-Supervised Learning?
by: Ma, Qian, et al.
Published: (2024)
by: Ma, Qian, et al.
Published: (2024)
Non-Asymptotic Length Generalization
by: Chen, Thomas, et al.
Published: (2025)
by: Chen, Thomas, et al.
Published: (2025)
Formal Theorem Proving by Rewarding LLMs to Decompose Proofs Hierarchically
by: Dong, Kefan, et al.
Published: (2024)
by: Dong, Kefan, et al.
Published: (2024)
Unified Neural Network Scaling Laws and Scale-time Equivalence
by: Boopathy, Akhilan, et al.
Published: (2024)
by: Boopathy, Akhilan, et al.
Published: (2024)
Breaking Neural Network Scaling Laws with Modularity
by: Boopathy, Akhilan, et al.
Published: (2024)
by: Boopathy, Akhilan, et al.
Published: (2024)
STP: Self-play LLM Theorem Provers with Iterative Conjecturing and Proving
by: Dong, Kefan, et al.
Published: (2025)
by: Dong, Kefan, et al.
Published: (2025)
On the Optimizer Dependence of Neural Scaling Laws
by: Ramani, Vansh, et al.
Published: (2026)
by: Ramani, Vansh, et al.
Published: (2026)
Towards Neural Scaling Laws on Graphs
by: Liu, Jingzhe, et al.
Published: (2024)
by: Liu, Jingzhe, et al.
Published: (2024)
Analyzing Neural Scaling Laws in Two-Layer Networks with Power-Law Data Spectra
by: Worschech, Roman, et al.
Published: (2024)
by: Worschech, Roman, et al.
Published: (2024)
Explaining Neural Scaling Laws
by: Bahri, Yasaman, et al.
Published: (2021)
by: Bahri, Yasaman, et al.
Published: (2021)
RNNs are not Transformers (Yet): The Key Bottleneck on In-context Retrieval
by: Wen, Kaiyue, et al.
Published: (2024)
by: Wen, Kaiyue, et al.
Published: (2024)
Neural Scaling Laws for Deep Regression
by: Cadez, Tilen, et al.
Published: (2025)
by: Cadez, Tilen, et al.
Published: (2025)
Scaling Laws for Neural Material Models
by: Trikha, Akshay, et al.
Published: (2025)
by: Trikha, Akshay, et al.
Published: (2025)
AlphaZero Neural Scaling and Zipf's Law: a Tale of Board Games and Power Laws
by: Neumann, Oren, et al.
Published: (2024)
by: Neumann, Oren, et al.
Published: (2024)
On Neural Scaling Laws for Weather Emulation through Continual Training
by: Subramanian, Shashank, et al.
Published: (2026)
by: Subramanian, Shashank, et al.
Published: (2026)
Complexity Scaling Laws for Neural Models using Combinatorial Optimization
by: Weissman, Lowell, et al.
Published: (2025)
by: Weissman, Lowell, et al.
Published: (2025)
Scaling Law for Time Series Forecasting
by: Shi, Jingzhe, et al.
Published: (2024)
by: Shi, Jingzhe, et al.
Published: (2024)
Information-Theoretic Foundations for Neural Scaling Laws
by: Jeon, Hong Jun, et al.
Published: (2024)
by: Jeon, Hong Jun, et al.
Published: (2024)
Beyond Scaling Laws: Understanding Transformer Performance with Associative Memory
by: Niu, Xueyan, et al.
Published: (2024)
by: Niu, Xueyan, et al.
Published: (2024)
Comparative Analysis of QNN Architectures for Wind Power Prediction: Feature Maps and Ansatz Configurations
by: Hangun, Batuhan, et al.
Published: (2025)
by: Hangun, Batuhan, et al.
Published: (2025)
Linux Kernel Configurations at Scale: A Dataset for Performance and Evolution Analysis
by: Borges, Heraldo, et al.
Published: (2025)
by: Borges, Heraldo, et al.
Published: (2025)
xLSTM Scaling Laws: Competitive Performance with Linear Time-Complexity
by: Beck, Maximilian, et al.
Published: (2025)
by: Beck, Maximilian, et al.
Published: (2025)
Neural Scaling Laws of Deep ReLU and Deep Operator Network: A Theoretical Study
by: Liu, Hao, et al.
Published: (2024)
by: Liu, Hao, et al.
Published: (2024)
Zero-Shot Performance Prediction for Probabilistic Scaling Laws
by: Schram, Viktoria, et al.
Published: (2025)
by: Schram, Viktoria, et al.
Published: (2025)
Hyperparameter Transfer Laws for Non-Recurrent Multi-Path Neural Networks
by: Wu, Shenxi, et al.
Published: (2026)
by: Wu, Shenxi, et al.
Published: (2026)
Guaranteeing Conservation Laws with Projection in Physics-Informed Neural Networks
by: Baez, Anthony, et al.
Published: (2024)
by: Baez, Anthony, et al.
Published: (2024)
Practical Scaling Laws: Converting Compute into Performance in a Data-Constrained World
by: Bryant, Christopher M., et al.
Published: (2026)
by: Bryant, Christopher M., et al.
Published: (2026)
Beyond Heuristics: Globally Optimal Configuration of Implicit Neural Representations
by: Chen, Sipeng, et al.
Published: (2025)
by: Chen, Sipeng, et al.
Published: (2025)
Similar Items
-
Divide-and-Conquer CoT: RL for Reducing Latency via Parallel Reasoning
by: Mahankali, Arvind, et al.
Published: (2026) -
Scaling Self-Play with Self-Guidance
by: Bailey, Luke, et al.
Published: (2026) -
Fantastic Pretraining Optimizers and Where to Find Them
by: Wen, Kaiyue, et al.
Published: (2025) -
From Sparse Dependence to Sparse Attention: Unveiling How Chain-of-Thought Enhances Transformer Sample Efficiency
by: Wen, Kaiyue, et al.
Published: (2024) -
Pseudo-Formalization for Automatic Proof Verification
by: Barkallah, Slim, et al.
Published: (2026)