AI Generated Text Detection

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Alikhanov, Adilkhan, Amangeldi, Aidar, Demeubay, Diar, Akhmetzhan, Dilnaz, Moldakhmetov, Nurbek, Polat, Omar, Zharas, Galymzhan
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911357971988480
author Alikhanov, Adilkhan
Amangeldi, Aidar
Demeubay, Diar
Akhmetzhan, Dilnaz
Moldakhmetov, Nurbek
Polat, Omar
Zharas, Galymzhan
author_facet Alikhanov, Adilkhan
Amangeldi, Aidar
Demeubay, Diar
Akhmetzhan, Dilnaz
Moldakhmetov, Nurbek
Polat, Omar
Zharas, Galymzhan
contents The rapid development of large language models has led to an increase in AI-generated text, with students increasingly using LLM-generated content as their own work, which violates academic integrity. This paper presents an evaluation of AI text detection methods, including both traditional machine learning models and transformer-based architectures. We utilize two datasets, HC3 and DAIGT v2, to build a unified benchmark and apply a topic-based data split to prevent information leakage. This approach ensures robust generalization across unseen domains. Our experiments show that TF-IDF logistic regression achieves a reasonable baseline accuracy of 82.87%. However, deep learning models outperform it. The BiLSTM classifier achieves an accuracy of 88.86%, while DistilBERT achieves a similar accuracy of 88.11% with the highest ROC-AUC score of 0.96, demonstrating the strongest overall performance. The results indicate that contextual semantic modeling is significantly superior to lexical features and highlight the importance of mitigating topic memorization through appropriate evaluation protocols. The limitations of this work are primarily related to dataset diversity and computational constraints. In future work, we plan to expand dataset diversity and utilize parameter-efficient fine-tuning methods such as LoRA. We also plan to explore smaller or distilled models and employ more efficient batching strategies and hardware-aware optimization.
format Preprint
id arxiv_https___arxiv_org_abs_2601_03812
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle AI Generated Text Detection
Alikhanov, Adilkhan
Amangeldi, Aidar
Demeubay, Diar
Akhmetzhan, Dilnaz
Moldakhmetov, Nurbek
Polat, Omar
Zharas, Galymzhan
Computation and Language
Artificial Intelligence
The rapid development of large language models has led to an increase in AI-generated text, with students increasingly using LLM-generated content as their own work, which violates academic integrity. This paper presents an evaluation of AI text detection methods, including both traditional machine learning models and transformer-based architectures. We utilize two datasets, HC3 and DAIGT v2, to build a unified benchmark and apply a topic-based data split to prevent information leakage. This approach ensures robust generalization across unseen domains. Our experiments show that TF-IDF logistic regression achieves a reasonable baseline accuracy of 82.87%. However, deep learning models outperform it. The BiLSTM classifier achieves an accuracy of 88.86%, while DistilBERT achieves a similar accuracy of 88.11% with the highest ROC-AUC score of 0.96, demonstrating the strongest overall performance. The results indicate that contextual semantic modeling is significantly superior to lexical features and highlight the importance of mitigating topic memorization through appropriate evaluation protocols. The limitations of this work are primarily related to dataset diversity and computational constraints. In future work, we plan to expand dataset diversity and utilize parameter-efficient fine-tuning methods such as LoRA. We also plan to explore smaller or distilled models and employ more efficient batching strategies and hardware-aware optimization.
title AI Generated Text Detection
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2601.03812