This article is the first in a series of articles looking at the different aspects of k-means clustering, beginning with a discussion on centroid initialization.