Spark library for generalized K-Means clustering. Supports general Bregman divergences. Suitable for clustering probabilistic data, time series data, high dimensional data, and very large data.
- kullback-leibler-divergence
- clustering
- spark
- entropy
- bregman-divergence
- euclidean-distance
- embeddings
- k-means
- itakura-saito-divergence
- spark-mllib
- cosine-similarity
- similarity-search
Scala versions:
2.10
6
versions found for
massivedatascience-clusterer