A unified foundational framework for knowledge injection and evaluation of Large Language Models in Combustion Science

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Yang, Zonglin, Mao, Runze, Wu, Tianhao, Li, Han, Zhou, QingGuo, Chen, Zhi X.
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866918370833596416
author Yang, Zonglin
Mao, Runze
Wu, Tianhao
Li, Han
Zhou, QingGuo
Chen, Zhi X.
author_facet Yang, Zonglin
Mao, Runze
Wu, Tianhao
Li, Han
Zhou, QingGuo
Chen, Zhi X.
contents To advance foundation Large Language Models (LLMs) for combustion science, this study presents the first end-to-end framework for developing domain-specialized models for the combustion community. The framework comprises an AI-ready multimodal knowledge base at the 3.5 billion-token scale, extracted from over 200,000 peer-reviewed articles, 8,000 theses and dissertations, and approximately 400,000 lines of combustion CFD code; a rigorous and largely automated evaluation benchmark (CombustionQA, 436 questions across eight subfields); and a three-stage knowledge-injection pathway that progresses from lightweight retrieval-augmented generation (RAG) to knowledge-graph-enhanced retrieval and continued pretraining. We first quantitatively validate Stage 1 (naive RAG) and find a hard ceiling: standard RAG accuracy peaks at 60%, far surpassing zero-shot performance (23%) yet well below the theoretical upper bound (87%). We further demonstrate that this stage's performance is severely constrained by context contamination. Consequently, building a domain foundation model requires structured knowledge graphs and continued pretraining (Stages 2 and 3).
format Preprint
id arxiv_https___arxiv_org_abs_2603_04452
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle A unified foundational framework for knowledge injection and evaluation of Large Language Models in Combustion Science
Yang, Zonglin
Mao, Runze
Wu, Tianhao
Li, Han
Zhou, QingGuo
Chen, Zhi X.
Computation and Language
Artificial Intelligence
To advance foundation Large Language Models (LLMs) for combustion science, this study presents the first end-to-end framework for developing domain-specialized models for the combustion community. The framework comprises an AI-ready multimodal knowledge base at the 3.5 billion-token scale, extracted from over 200,000 peer-reviewed articles, 8,000 theses and dissertations, and approximately 400,000 lines of combustion CFD code; a rigorous and largely automated evaluation benchmark (CombustionQA, 436 questions across eight subfields); and a three-stage knowledge-injection pathway that progresses from lightweight retrieval-augmented generation (RAG) to knowledge-graph-enhanced retrieval and continued pretraining. We first quantitatively validate Stage 1 (naive RAG) and find a hard ceiling: standard RAG accuracy peaks at 60%, far surpassing zero-shot performance (23%) yet well below the theoretical upper bound (87%). We further demonstrate that this stage's performance is severely constrained by context contamination. Consequently, building a domain foundation model requires structured knowledge graphs and continued pretraining (Stages 2 and 3).
title A unified foundational framework for knowledge injection and evaluation of Large Language Models in Combustion Science
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2603.04452