Privacy-preserving smart contracts for fuzzy WordNet-based document representation and clustering using regularised K-means method
Abstract
Key technology for unsupervised intelligent classification of any textual content is the clustering of documents. Prior document knowledge is not required for document clustering, which is an unsupervised method of learning as compared with document classification. For clustering rather than classification, little prior knowledge of the data is needed. The crucial challenges of document clustering are the high dimensionality, measurability, preciseness, extraction of semantic relationships from texts, and meaningful cluster labels. Fuzzy WordNet-based document representation and clustering using the regularised K-means method as an efficient framework is introduced in the present paper with the purpose of improving the quality of document clustering. To estimate the performance of this framework we carried out experiments on different datasets. Experimental results show that this framework improves the quality of document clustering when compared to other existing methods. Furthermore, this system gives generalised and concrete labels for documents and improves the speed of clustering by reducing their size.
Community
0 commentsNo discussion yet
Be the first to share a question or observation.