Following up on KMeans Clustering Now Running on Elastic MapReduce, Stephen Green has generously documented the steps that was necessary to get an example of k-Means clustering up and running on Amazon’s Elastic MapReduce (EMR) on the Apache Lucene Mahout wiki.
D. Arthur, and S. Vassilvitskii. SODA '07: Proceedings of the eighteenth annual ACM-SIAM symposium on Discrete algorithms, page 1027--1035. Philadelphia, PA, USA, Society for Industrial and Applied Mathematics, (2007)
A. Hotho, A. Maedche, and S. Staab. ICDM '01: Proceedings of the 2001 IEEE International Conference on Data Mining, page 607--608. Washington, DC, USA, IEEE Computer Society, (2001)
I. Yoo, and X. Hu. JCDL '06: Proceedings of the 6th ACM/IEEE-CS joint conference on Digital libraries, page 220--229. New York, NY, USA, ACM Press, (2006)