Following up on KMeans Clustering Now Running on Elastic MapReduce, Stephen Green has generously documented the steps that was necessary to get an example of k-Means clustering up and running on Amazon’s Elastic MapReduce (EMR) on the Apache Lucene Mahout wiki.
A. Hotho, A. Maedche, and S. Staab. ICDM '01: Proceedings of the 2001 IEEE International Conference on Data Mining, page 607--608. Washington, DC, USA, IEEE Computer Society, (2001)
I. Yoo, and X. Hu. JCDL '06: Proceedings of the 6th ACM/IEEE-CS joint conference on Digital libraries, page 220--229. New York, NY, USA, ACM Press, (2006)
Y. Zhao, and G. Karypis. CIKM '02: Proceedings of the eleventh international conference on Information and knowledge management, page 515--524. New York, NY, USA, ACM Press, (2002)