Large Language Model-Driven Database for Thermoelectric Materials

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Itani, Suman, Zhang, Yibo, Zang, Jiadong
Format: Preprint
Veröffentlicht: 2024
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866915087408693248
author Itani, Suman
Zhang, Yibo
Zang, Jiadong
author_facet Itani, Suman
Zhang, Yibo
Zang, Jiadong
contents Thermoelectric materials provide a sustainable way to convert waste heat into electricity. However, data-driven discovery and optimization of these materials are challenging because of a lack of a reliable database. Here we developed a comprehensive database of 7,123 thermoelectric compounds, containing key information such as chemical composition, structural detail, seebeck coefficient, electrical and thermal conductivity, power factor, and figure of merit (ZT). We used the GPTArticleExtractor workflow, powered by large language models (LLM), to extract and curate data automatically from the scientific literature published in Elsevier journals. This process enabled the creation of a structured database that addresses the challenges of manual data collection. The open access database could stimulate data-driven research and advance thermoelectric material analysis and discovery.
format Preprint
id arxiv_https___arxiv_org_abs_2501_00564
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Large Language Model-Driven Database for Thermoelectric Materials
Itani, Suman
Zhang, Yibo
Zang, Jiadong
Materials Science
Digital Libraries
Thermoelectric materials provide a sustainable way to convert waste heat into electricity. However, data-driven discovery and optimization of these materials are challenging because of a lack of a reliable database. Here we developed a comprehensive database of 7,123 thermoelectric compounds, containing key information such as chemical composition, structural detail, seebeck coefficient, electrical and thermal conductivity, power factor, and figure of merit (ZT). We used the GPTArticleExtractor workflow, powered by large language models (LLM), to extract and curate data automatically from the scientific literature published in Elsevier journals. This process enabled the creation of a structured database that addresses the challenges of manual data collection. The open access database could stimulate data-driven research and advance thermoelectric material analysis and discovery.
title Large Language Model-Driven Database for Thermoelectric Materials
topic Materials Science
Digital Libraries
url https://arxiv.org/abs/2501.00564