I’ve been studying about k-means clustering , and one thing that’s not clear is

Question

0

Asked: May 13, 20262026-05-13T17:17:58+00:00 2026-05-13T17:17:58+00:00

I’ve been studying about k-means clustering , and one thing that’s not clear is

0

I’ve been studying about k-means clustering, and one thing that’s not clear is how you choose the value of k. Is it just a matter of trial and error, or is there more to it?

Report

Leave an answer
Cancel reply

You must login to add an answer.

Need An Account,

1 Answer

Editorial Team · Answer 1 · 2026-05-13T17:17:58+00:00

You can maximize the Bayesian Information Criterion (BIC):

BIC(C | X) = L(X | C) - (p / 2) * log n

where L(X | C) is the log-likelihood of the dataset X according to model C, p is the number of parameters in the model C, and n is the number of points in the dataset.
See “X-means: extending K-means with efficient estimation of the number of clusters” by Dan Pelleg and Andrew Moore in ICML 2000.

Another approach is to start with a large value for k and keep removing centroids (reducing k) until it no longer reduces the description length. See “MDL principle for robust vector quantisation” by Horst Bischof, Ales Leonardis, and Alexander Selb in Pattern Analysis and Applications vol. 2, p. 59-72, 1999.

Finally, you can start with one cluster, then keep splitting clusters until the points assigned to each cluster have a Gaussian distribution. In “Learning the k in k-means” (NIPS 2003), Greg Hamerly and Charles Elkan show some evidence that this works better than BIC, and that BIC does not penalize the model’s complexity strongly enough.

Sign Up

Sign In

Forgot Password

The Archive Base Latest Questions

I’ve been studying about k-means clustering , and one thing that’s not clear is

Leave an answerCancel reply

1 Answer

Leave an answer
Cancel reply