Clustering functional data is a challenging task due to intrinsic infinite-dimensionality and the need for stable, data-adaptive partitioning. In this work, we propose a clustering framework based on Random Projections, which simultaneously performs dimensionality reduction and generates multiple stochastic representations of the original functions. Each projection is clustered independently, and the resulting partitions are then aggregated through an ensemble consensus procedure, enhancing robustness and mitigating the influence of any single projection. To focus on the most informative representations, projections are ranked according to clustering quality criteria, and only a selected subset is retained. In particular, we adopt Gaussian Mixture Models as base clusterers and employ a measure based on the Kullback–Leibler divergence to order the random projections; these choices enable fast computation and eliminate the need to specify the number of clusters a priori. The performance of the proposed methodology is assessed through an extensive simulation study and real-data applications; the obtained results suggest that the proposal represents an effective tool for the clustering of functional data.

Mori, M., Anderlucci, L. (2026). Model-based clustering of functional data via random projection ensembles. ADVANCES IN DATA ANALYSIS AND CLASSIFICATION, 20, 877-901 [10.1007/s11634-026-00684-7].

Model-based clustering of functional data via random projection ensembles

Mori, Matteo
;
Anderlucci, Laura
2026

Abstract

Clustering functional data is a challenging task due to intrinsic infinite-dimensionality and the need for stable, data-adaptive partitioning. In this work, we propose a clustering framework based on Random Projections, which simultaneously performs dimensionality reduction and generates multiple stochastic representations of the original functions. Each projection is clustered independently, and the resulting partitions are then aggregated through an ensemble consensus procedure, enhancing robustness and mitigating the influence of any single projection. To focus on the most informative representations, projections are ranked according to clustering quality criteria, and only a selected subset is retained. In particular, we adopt Gaussian Mixture Models as base clusterers and employ a measure based on the Kullback–Leibler divergence to order the random projections; these choices enable fast computation and eliminate the need to specify the number of clusters a priori. The performance of the proposed methodology is assessed through an extensive simulation study and real-data applications; the obtained results suggest that the proposal represents an effective tool for the clustering of functional data.
2026
Mori, M., Anderlucci, L. (2026). Model-based clustering of functional data via random projection ensembles. ADVANCES IN DATA ANALYSIS AND CLASSIFICATION, 20, 877-901 [10.1007/s11634-026-00684-7].
Mori, Matteo; Anderlucci, Laura
File in questo prodotto:
Eventuali allegati, non sono esposti

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11585/1083962
 Attenzione

Attenzione! I dati visualizzati non sono stati sottoposti a validazione da parte dell'ateneo

Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus ND
  • ???jsp.display-item.citation.isi??? ND
  • OpenAlex ND
social impact