On Neural Scaling Laws for Weather Emulation through Continual Training
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Subramanian, Shashank, Kiefer, Alexander, Nigmetov, Arnur, Gholami, Amir, Morozov, Dmitriy, Mahoney, Michael W. |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Persistence-Augmented Neural Networks
par: Wang, Elena Xinyi, et autres
Publié: (2026)
par: Wang, Elena Xinyi, et autres
Publié: (2026)
Distributed Computation of Persistent Cohomology
par: Nigmetov, Arnur, et autres
Publié: (2024)
par: Nigmetov, Arnur, et autres
Publié: (2024)
Towards Foundation Models for Scientific Machine Learning: Characterizing Scaling and Transfer Behavior
par: Subramanian, Shashank, et autres
Publié: (2023)
par: Subramanian, Shashank, et autres
Publié: (2023)
Data-Efficient Operator Learning via Unsupervised Pretraining and In-Context Learning
par: Chen, Wuyang, et autres
Publié: (2024)
par: Chen, Wuyang, et autres
Publié: (2024)
Towards Scaling Law Analysis For Spatiotemporal Weather Data
par: Kiefer, Alexander, et autres
Publié: (2026)
par: Kiefer, Alexander, et autres
Publié: (2026)
Topological potentials guiding protein self-assembly
par: Spirandelli, Ivan, et autres
Publié: (2025)
par: Spirandelli, Ivan, et autres
Publié: (2025)
Wilkins: HPC In Situ Workflows Made Easy
par: Yildiz, Orcun, et autres
Publié: (2024)
par: Yildiz, Orcun, et autres
Publié: (2024)
SciML Agents: Write the Solver, Not the Solution
par: Gaonkar, Saarth, et autres
Publié: (2025)
par: Gaonkar, Saarth, et autres
Publié: (2025)
Determinant Estimation under Memory Constraints and Neural Scaling Laws
par: Ameli, Siavash, et autres
Publié: (2025)
par: Ameli, Siavash, et autres
Publié: (2025)
Examining Fast Radiatively Driven Responses Using Machine-Learning Weather Emulators
par: Mahesh, Ankur, et autres
Publié: (2026)
par: Mahesh, Ankur, et autres
Publié: (2026)
Analyzing and Exploring Training Recipes for Large-Scale Transformer-Based Weather Prediction
par: Willard, Jared D., et autres
Publié: (2024)
par: Willard, Jared D., et autres
Publié: (2024)
Scaling Laws of Global Weather Models
par: Yu, Yuejiang, et autres
Publié: (2026)
par: Yu, Yuejiang, et autres
Publié: (2026)
ETS: Efficient Tree Search for Inference-Time Scaling
par: Hooper, Coleman, et autres
Publié: (2025)
par: Hooper, Coleman, et autres
Publié: (2025)
KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
par: Hooper, Coleman, et autres
Publié: (2024)
par: Hooper, Coleman, et autres
Publié: (2024)
Scaling Laws for Emulation of Stellar Spectra
par: Różański, Tomasz, et autres
Publié: (2025)
par: Różański, Tomasz, et autres
Publié: (2025)
Reward Under Attack: Analyzing the Robustness and Hackability of Process Reward Models
par: Tiwari, Rishabh, et autres
Publié: (2026)
par: Tiwari, Rishabh, et autres
Publié: (2026)
AI and Memory Wall
par: Gholami, Amir, et autres
Publié: (2024)
par: Gholami, Amir, et autres
Publié: (2024)
Evaluating Loss Landscapes from a Topology Perspective
par: Xie, Tiankai, et autres
Publié: (2024)
par: Xie, Tiankai, et autres
Publié: (2024)
Neural Neural Scaling Laws
par: Hu, Michael Y., et autres
Publié: (2026)
par: Hu, Michael Y., et autres
Publié: (2026)
Hierarchical Implicit Neural Emulators
par: Jiang, Ruoxi, et autres
Publié: (2025)
par: Jiang, Ruoxi, et autres
Publié: (2025)
Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations
par: Hägele, Alexander, et autres
Publié: (2024)
par: Hägele, Alexander, et autres
Publié: (2024)
SqueezeLLM: Dense-and-Sparse Quantization
par: Kim, Sehoon, et autres
Publié: (2023)
par: Kim, Sehoon, et autres
Publié: (2023)
Multipole Attention for Efficient Long Context Reasoning
par: Hooper, Coleman, et autres
Publié: (2025)
par: Hooper, Coleman, et autres
Publié: (2025)
XQuant: Breaking the Memory Wall for LLM Inference with KV Cache Rematerialization
par: Tomar, Aditya, et autres
Publié: (2025)
par: Tomar, Aditya, et autres
Publié: (2025)
Beyond Next-Token Prediction: A Performance Characterization of Diffusion versus Autoregressive Language Models
par: Kim, Minseo, et autres
Publié: (2025)
par: Kim, Minseo, et autres
Publié: (2025)
The Journey Matters: Average Parameter Count over Pre-training Unifies Sparse and Dense Scaling Laws
par: Jin, Tian, et autres
Publié: (2025)
par: Jin, Tian, et autres
Publié: (2025)
Complexity Scaling Laws for Neural Models using Combinatorial Optimization
par: Weissman, Lowell, et autres
Publié: (2025)
par: Weissman, Lowell, et autres
Publié: (2025)
Generalizing PDE Emulation with Equation-Aware Neural Operators
par: Zhu, Qian-Ze, et autres
Publié: (2025)
par: Zhu, Qian-Ze, et autres
Publié: (2025)
A Dynamical Model of Neural Scaling Laws
par: Bordelon, Blake, et autres
Publié: (2024)
par: Bordelon, Blake, et autres
Publié: (2024)
On the Invariance and Generality of Neural Scaling Laws
par: Han, Xing, et autres
Publié: (2026)
par: Han, Xing, et autres
Publié: (2026)
Neural Emulator Superiority: When Machine Learning for PDEs Surpasses its Training Data
par: Koehler, Felix, et autres
Publié: (2025)
par: Koehler, Felix, et autres
Publié: (2025)
Visualizing Loss Functions as Topological Landscape Profiles
par: Geniesse, Caleb, et autres
Publié: (2024)
par: Geniesse, Caleb, et autres
Publié: (2024)
Speculative Interaction Agents: Building Real-Time Agents with Asynchronous I/O and Speculative Tool Calling
par: Hooper, Coleman, et autres
Publié: (2026)
par: Hooper, Coleman, et autres
Publié: (2026)
Huge Ensembles Part I: Design of Ensemble Weather Forecasts using Spherical Fourier Neural Operators
par: Mahesh, Ankur, et autres
Publié: (2024)
par: Mahesh, Ankur, et autres
Publié: (2024)
Scaling Laws for Post Training Quantized Large Language Models
par: Xu, Zifei, et autres
Publié: (2024)
par: Xu, Zifei, et autres
Publié: (2024)
Random Matrix Theory for Deep Learning: Beyond Eigenvalues of Linear Models
par: Liao, Zhenyu, et autres
Publié: (2025)
par: Liao, Zhenyu, et autres
Publié: (2025)
Scaling Law for Quantization-Aware Training
par: Chen, Mengzhao, et autres
Publié: (2025)
par: Chen, Mengzhao, et autres
Publié: (2025)
Unified Neural Network Scaling Laws and Scale-time Equivalence
par: Boopathy, Akhilan, et autres
Publié: (2024)
par: Boopathy, Akhilan, et autres
Publié: (2024)
Personalized Counterfactual Framework: Generating Potential Outcomes from Wearable Data
par: Subramanian, Ajan, et autres
Publié: (2025)
par: Subramanian, Ajan, et autres
Publié: (2025)
Configuration-to-Performance Scaling Law with Neural Ansatz
par: Zhang, Huaqing, et autres
Publié: (2026)
par: Zhang, Huaqing, et autres
Publié: (2026)
Documents similaires
-
Persistence-Augmented Neural Networks
par: Wang, Elena Xinyi, et autres
Publié: (2026) -
Distributed Computation of Persistent Cohomology
par: Nigmetov, Arnur, et autres
Publié: (2024) -
Towards Foundation Models for Scientific Machine Learning: Characterizing Scaling and Transfer Behavior
par: Subramanian, Shashank, et autres
Publié: (2023) -
Data-Efficient Operator Learning via Unsupervised Pretraining and In-Context Learning
par: Chen, Wuyang, et autres
Publié: (2024) -
Towards Scaling Law Analysis For Spatiotemporal Weather Data
par: Kiefer, Alexander, et autres
Publié: (2026)