Methods for Clustering Mixed-Type Data

Implements methods for clustering mixed-type data, specifically combinations of continuous and nominal data. Special attention is paid to the often-overlooked problem of equitably balancing the contribution of the continuous and categorical variables. This package implements KAMILA clustering, a novel method for clustering mixed-type data in the spirit of k-means clustering. It does not require dummy coding of variables, and is efficient enough to scale to rather large data sets. Also implemented is Modha-Spangler clustering, which uses a brute-force strategy to maximize the cluster separation simultaneously in the continuous and categorical variables.

======= R package for clustering mixed data. For more information, run


from the R terminal.


Reference manual

It appears you don't have a PDF plugin for this browser. You can click here to download the reference manual.

install.packages("kamila") by Alexander Foss, 4 months ago

Report a bug at

Browse source code at

Authors: Alexander Foss [aut, cre], Marianthi Markatou [aut]

Documentation:   PDF Manual  

GPL-3 | file LICENSE license

Imports stats, abind, KernSmooth, gtools, Rcpp, mclust, plyr

Suggests testthat, clustMD, ggplot2, Hmisc

Linking to Rcpp

See at CRAN