On the Limitations of Large Language Models for Conceptual Database Modeling

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Siqueira, Arthur F., Nogueira, Carlos D. S., Farias, Eduarda, Campelo, Claudio E. C., Menezes, Júlia
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866918496909131776
author Siqueira, Arthur F.
Nogueira, Carlos D. S.
Farias, Eduarda
Campelo, Claudio E. C.
Menezes, Júlia
author_facet Siqueira, Arthur F.
Nogueira, Carlos D. S.
Farias, Eduarda
Campelo, Claudio E. C.
Menezes, Júlia
contents This article analyzes the use of Large Language Models (LLMs) as support for the conceptual modeling of relational databases through the automatic generation of Entity-Relationship (ER) diagrams from natural language requirements. The approach combines different language models with prompt engineering techniques to evaluate their ability to identify entities, relationships, and attributes in a conceptually consistent manner. The experimental evaluation involved three LLMs, each subjected to three prompting techniques (Zero-Shot, Chain of Thought, and Chain of Thought + Verifier), applied to the same requirements scenario with progressively increasing complexity. The generated diagrams were qualitatively analyzed through direct comparison with the textual requirements, considering the structural and semantic adherence of the modeled elements. The results indicate that, although LLMs show reasonable performance in less complex scenarios, their reliability decreases as the complexity of the requirements increases, with a rise in inconsistencies, ambiguities, and failures in representing constraints. These findings reinforce that, in their current state, LLMs are not sufficiently mature for reliable use in complex scenarios, and the cost of validation may offset the apparent productivity gains.
format Preprint
id arxiv_https___arxiv_org_abs_2605_11986
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle On the Limitations of Large Language Models for Conceptual Database Modeling
Siqueira, Arthur F.
Nogueira, Carlos D. S.
Farias, Eduarda
Campelo, Claudio E. C.
Menezes, Júlia
Artificial Intelligence
This article analyzes the use of Large Language Models (LLMs) as support for the conceptual modeling of relational databases through the automatic generation of Entity-Relationship (ER) diagrams from natural language requirements. The approach combines different language models with prompt engineering techniques to evaluate their ability to identify entities, relationships, and attributes in a conceptually consistent manner. The experimental evaluation involved three LLMs, each subjected to three prompting techniques (Zero-Shot, Chain of Thought, and Chain of Thought + Verifier), applied to the same requirements scenario with progressively increasing complexity. The generated diagrams were qualitatively analyzed through direct comparison with the textual requirements, considering the structural and semantic adherence of the modeled elements. The results indicate that, although LLMs show reasonable performance in less complex scenarios, their reliability decreases as the complexity of the requirements increases, with a rise in inconsistencies, ambiguities, and failures in representing constraints. These findings reinforce that, in their current state, LLMs are not sufficiently mature for reliable use in complex scenarios, and the cost of validation may offset the apparent productivity gains.
title On the Limitations of Large Language Models for Conceptual Database Modeling
topic Artificial Intelligence
url https://arxiv.org/abs/2605.11986