Clustering country-level all-cause mortality data: a review

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: de Araujo, Pedro Menezes, Gormley, Isobel Claire, Murphy, Thomas Brendan
Formato: Preprint
Publicado: 2025
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866909943575085056
author de Araujo, Pedro Menezes
Gormley, Isobel Claire
Murphy, Thomas Brendan
author_facet de Araujo, Pedro Menezes
Gormley, Isobel Claire
Murphy, Thomas Brendan
contents Mortality data are relevant to demography, public health, and actuarial science. Whilst clustering is increasingly used to explore patterns in such data, no study has reviewed its application to country-level all-cause mortality. This review therefore summarises recent work and addresses key questions: why clustering is used, which mortality data are analysed, which methods are most common, and what main findings emerge. To address these questions, we examine studies applying clustering to country-level all-cause mortality, focusing on mortality indices, data sources, and methodological choices, and we replicate some approaches using Human Mortality Database (HMD) data. Our analysis reveals that clustering is mainly motivated by forecasting and by studying convergence and inequality. Most studies use HMD data from developed countries and rely on k-means, hierarchical, or functional clustering. Main findings include a persistent East-West European division across applications, with clustering generally improving forecast accuracy over single-country models. Overall, this review highlights the methodological range in the literature, summarises clustering results, and identifies gaps, such as the limited evaluation of clustering quality and the underuse of data from countries outside the high-income world.
format Preprint
id arxiv_https___arxiv_org_abs_2512_04831
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Clustering country-level all-cause mortality data: a review
de Araujo, Pedro Menezes
Gormley, Isobel Claire
Murphy, Thomas Brendan
Applications
Mortality data are relevant to demography, public health, and actuarial science. Whilst clustering is increasingly used to explore patterns in such data, no study has reviewed its application to country-level all-cause mortality. This review therefore summarises recent work and addresses key questions: why clustering is used, which mortality data are analysed, which methods are most common, and what main findings emerge. To address these questions, we examine studies applying clustering to country-level all-cause mortality, focusing on mortality indices, data sources, and methodological choices, and we replicate some approaches using Human Mortality Database (HMD) data. Our analysis reveals that clustering is mainly motivated by forecasting and by studying convergence and inequality. Most studies use HMD data from developed countries and rely on k-means, hierarchical, or functional clustering. Main findings include a persistent East-West European division across applications, with clustering generally improving forecast accuracy over single-country models. Overall, this review highlights the methodological range in the literature, summarises clustering results, and identifies gaps, such as the limited evaluation of clustering quality and the underuse of data from countries outside the high-income world.
title Clustering country-level all-cause mortality data: a review
topic Applications
url https://arxiv.org/abs/2512.04831