Overview and Prospects of Using Integer Surrogate Keys for Data Warehouse Performance Optimization

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Stumpf, Sviatoslav, Povyshev, Vladislav
Format: Preprint
Veröffentlicht: 2025
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866911274249486336
author Stumpf, Sviatoslav
Povyshev, Vladislav
author_facet Stumpf, Sviatoslav
Povyshev, Vladislav
contents The aim of this paper is to examine and demonstrate how integer-based datetime labels (integer surrogate keys for time) can optimize data-warehouse and time-series performance, proposing practical formats and algorithms and validating their efficiency on real-world workloads. It is shown that replacing standard DATE and TIMESTAMP types with 32- and 64-bit integer formats reduces storage requirements by 30-60 percent and speeds up query execution by 25-40 percent. The paper presents indexing, aggregation, compression, and batching algorithms demonstrating up to an eightfold increase in throughput. Practical examples from finance, telecommunications, IoT, and scientific research confirm the efficiency and versatility of the proposed approach.
format Preprint
id arxiv_https___arxiv_org_abs_2511_14502
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Overview and Prospects of Using Integer Surrogate Keys for Data Warehouse Performance Optimization
Stumpf, Sviatoslav
Povyshev, Vladislav
Databases
Distributed, Parallel, and Cluster Computing
The aim of this paper is to examine and demonstrate how integer-based datetime labels (integer surrogate keys for time) can optimize data-warehouse and time-series performance, proposing practical formats and algorithms and validating their efficiency on real-world workloads. It is shown that replacing standard DATE and TIMESTAMP types with 32- and 64-bit integer formats reduces storage requirements by 30-60 percent and speeds up query execution by 25-40 percent. The paper presents indexing, aggregation, compression, and batching algorithms demonstrating up to an eightfold increase in throughput. Practical examples from finance, telecommunications, IoT, and scientific research confirm the efficiency and versatility of the proposed approach.
title Overview and Prospects of Using Integer Surrogate Keys for Data Warehouse Performance Optimization
topic Databases
Distributed, Parallel, and Cluster Computing
url https://arxiv.org/abs/2511.14502