Abstract
Text mining approaches uses feature similarity techniques or distributed keyword searching techniques. But machine learning techniques develop a statistical model to categorize documents by learning from vast amount of medical documents available at pubmed. It is unsupervised techniques. The proposed algorithm enhances the traditional document clustering techniques.and generate accurate and reliable model. We experimented the algorithm with 1000 document data set It showed the significant improvement over other traditional algorithms.