Large Language Model Scaling Laws for Neural Quantum States in Quantum Chemistry

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Knitter, Oliver, Zhao, Dan, Leichenauer, Stefan, Veerapaneni, Shravan
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866908541118316544
author Knitter, Oliver
Zhao, Dan
Leichenauer, Stefan
Veerapaneni, Shravan
author_facet Knitter, Oliver
Zhao, Dan
Leichenauer, Stefan
Veerapaneni, Shravan
contents Scaling laws have been used to describe how large language model (LLM) performance scales with model size, training data size, or amount of computational resources. Motivated by the fact that neural quantum states (NQS) has increasingly adopted LLM-based components, we seek to understand NQS scaling laws, thereby shedding light on the scalability and optimal performance--resource trade-offs of NQS ansatze. In particular, we identify scaling laws that predict the performance, as measured by absolute error and V-score, for transformer-based NQS as a function of problem size in second-quantized quantum chemistry applications. By performing analogous compute-constrained optimization of the obtained parametric curves, we find that the relationship between model size and training time is highly dependent on loss metric and ansatz, and does not follow the approximately linear relationship found for language models.
format Preprint
id arxiv_https___arxiv_org_abs_2509_12679
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Large Language Model Scaling Laws for Neural Quantum States in Quantum Chemistry
Knitter, Oliver
Zhao, Dan
Leichenauer, Stefan
Veerapaneni, Shravan
Machine Learning
Computational Engineering, Finance, and Science
Quantum Physics
Scaling laws have been used to describe how large language model (LLM) performance scales with model size, training data size, or amount of computational resources. Motivated by the fact that neural quantum states (NQS) has increasingly adopted LLM-based components, we seek to understand NQS scaling laws, thereby shedding light on the scalability and optimal performance--resource trade-offs of NQS ansatze. In particular, we identify scaling laws that predict the performance, as measured by absolute error and V-score, for transformer-based NQS as a function of problem size in second-quantized quantum chemistry applications. By performing analogous compute-constrained optimization of the obtained parametric curves, we find that the relationship between model size and training time is highly dependent on loss metric and ansatz, and does not follow the approximately linear relationship found for language models.
title Large Language Model Scaling Laws for Neural Quantum States in Quantum Chemistry
topic Machine Learning
Computational Engineering, Finance, and Science
Quantum Physics
url https://arxiv.org/abs/2509.12679