clusterMI
1.6Cluster Analysis with Missing Values by Multiple Imputation
Overview
Allows clustering of incomplete observations by addressing missing values using multiple imputation. For achieving this goal, the methodology consists in three steps, following Audigier and Niang 2022 doi:10.1007/s11634-022-00519-1. I) Missing data imputation using dedicated models. Four multiple imputation methods are proposed, two are based on joint modelling and two are fully sequential methods, as discussed in Audigier et al. (2021) doi:10.48550/arXiv.2106.04424. II) cluster analysis of imputed data sets. Six clustering methods are available (distances-based or model-based), but custom methods can also be easily used. III) Partition pooling. The set of partitions is aggregated using Non-negative Matrix Factorization based method. An associated instability measure is computed by bootstrap (see Fang, Y. and Wang, J., 2012 doi:10.1016/j.csda.2011.09.003). Among applications, this instability measure can be used to choose a number of clusters with missing values. The package also proposes several diagnostic tools to tune the number of imputed data sets, to tune the number of iterations in fully sequential imputation, to check the fit of imputation models, etc.
Install
Health
- OK2026-08-0513 OK · 0 NOTE · 0 WARNING · 0 ERROR · 0 FAILURE
- NOTE2026-08-0112 OK · 1 NOTE · 0 WARNING · 0 ERROR · 0 FAILURE
- OK2026-06-0913 OK · 0 NOTE · 0 WARNING · 0 ERROR · 0 FAILURE
- ERROR2026-06-0812 OK · 0 NOTE · 0 WARNING · 1 ERROR · 0 FAILURE
- OK2026-04-2512 OK · 0 NOTE · 0 WARNING · 0 ERROR · 0 FAILURE
Show 3 earlier snapshots
- NOTE2026-04-2212 OK · 2 NOTE · 0 WARNING · 0 ERROR · 0 FAILURE
- ERROR2026-04-1811 OK · 2 NOTE · 0 WARNING · 1 ERROR · 0 FAILURE
- NOTE2026-03-1012 OK · 2 NOTE · 0 WARNING · 0 ERROR · 0 FAILURE
Documentation
- Examples that run
- 100%
- Documented parameters
- 88%
- Return-value docs
- 100%
- References docs
- 57%
Downloads
Dependencies
Nothing depends on this yet.
Code & Tests
Datasets
People & History
10 releases. Pick two to compare their code metrics. R releases are shown for context.
- RR 4.6.0 released · 2026-04-24
- 1.6Latest
- RR 4.5.0 released · 2025-04-11
- 1.52025-02-24 · diff ↗
- 1.4.02025-02-12 · diff ↗
- 1.32024-12-12 · diff ↗
- 1.2.22024-10-23 · diff ↗
- 1.2.12024-07-07 · diff ↗
- 1.22024-07-04 · diff ↗
- 1.1.12024-05-31 · diff ↗
- 1.1.02024-05-17 · diff ↗
- RR 4.4.0 released · 2024-04-24
- 1.0.02024-03-12
- RR 4.3.0 released · 2023-04-21
Package metadata
- First published
- 2024-03-12
- Total releases
- 10 / 2 yrs
- License
- GPL-2 | GPL-3 OSI
- Minimum R
- ≥ 3.5.0
- Bundled data
- 5.0 KB / 1 file
- Download size
- 1.1 MB
- Installed size
- not tracked yet
- With dependencies
- not tracked yet
Cite
Cite this package
Run in R for the authors' preferred citation:
citation("clusterMI")This is what citation() produces when a package has no citation file of its own. If it prints something else, use that.
Cite the R Observatory
For a number measured here: a download total, a coverage figure, an archival date.
From data release v2026-08-15, which the citation names so these numbers can be found later. More on citing and the projects behind them.