Large Language Model and Formal Concept Analysis: a comparative study for Topic Modeling

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Boissier, Fabrice, Sen, Monica, Rychkova, Irina
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911416292737024
author Boissier, Fabrice
Sen, Monica
Rychkova, Irina
author_facet Boissier, Fabrice
Sen, Monica
Rychkova, Irina
contents Topic modeling is a research field finding increasing applications: historically from document retrieving, to sentiment analysis and text summarization. Large Language Models (LLM) are currently a major trend in text processing, but few works study their usefulness for this task. Formal Concept Analysis (FCA) has recently been presented as a candidate for topic modeling, but no real applied case study has been conducted. In this work, we compare LLM and FCA to better understand their strengths and weakneses in the topic modeling field. FCA is evaluated through the CREA pipeline used in past experiments on topic modeling and visualization, whereas GPT-5 is used for the LLM. A strategy based on three prompts is applied with GPT-5 in a zero-shot setup: topic generation from document batches, merging of batch results into final topics, and topic labeling. A first experiment reuses the teaching materials previously used to evaluate CREA, while a second experiment analyzes 40 research articles in information systems to compare the extracted topics with the underling subfields.
format Preprint
id arxiv_https___arxiv_org_abs_2602_01933
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Large Language Model and Formal Concept Analysis: a comparative study for Topic Modeling
Boissier, Fabrice
Sen, Monica
Rychkova, Irina
Artificial Intelligence
Topic modeling is a research field finding increasing applications: historically from document retrieving, to sentiment analysis and text summarization. Large Language Models (LLM) are currently a major trend in text processing, but few works study their usefulness for this task. Formal Concept Analysis (FCA) has recently been presented as a candidate for topic modeling, but no real applied case study has been conducted. In this work, we compare LLM and FCA to better understand their strengths and weakneses in the topic modeling field. FCA is evaluated through the CREA pipeline used in past experiments on topic modeling and visualization, whereas GPT-5 is used for the LLM. A strategy based on three prompts is applied with GPT-5 in a zero-shot setup: topic generation from document batches, merging of batch results into final topics, and topic labeling. A first experiment reuses the teaching materials previously used to evaluate CREA, while a second experiment analyzes 40 research articles in information systems to compare the extracted topics with the underling subfields.
title Large Language Model and Formal Concept Analysis: a comparative study for Topic Modeling
topic Artificial Intelligence
url https://arxiv.org/abs/2602.01933