Clustering Longitudinal Data: A Review of Methods and Software Packages

16Citations
Citations of this article
54Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

Clustering of longitudinal data is becoming increasingly popular in many fields such as social sciences, business, environmental science, medicine and healthcare. However, it is often challenging due to the complex nature of the data, such as dependencies between observations collected over time, missingness, sparsity and non-linearity, making it difficult to identify meaningful patterns and relationships among the data. Despite the increasingly common application of cluster analysis for longitudinal data, many existing methods are still less known to researchers, and limited guidance is provided in choosing between methods and software packages. In this paper, we review several commonly used methods for clustering longitudinal data. These methods are broadly classified into three categories, namely, model-based approaches, algorithm-based approaches and functional clustering approaches. We perform a comparison among these methods and their corresponding R software packages using real-life datasets and simulated datasets under various conditions. Findings from the analyses and recommendations for using these approaches in practice are discussed.

Cite

CITATION STYLE

APA

Lu, Z. (2025). Clustering Longitudinal Data: A Review of Methods and Software Packages. International Statistical Review, 93(3), 425–458. https://doi.org/10.1111/insr.12588

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free