Research on Chinese Text Clustering Based on Sentence Structure Analysis
YIN Jidong
XIE Chahua
PENG Song
LIU Hong
ZENG Zhaohu
Abstract:Most of the existing K-means clustering algorithms are based on the data carriers,they are difficult to apply to Chi-nese text clustering analysis.In this paper,a new text clustering method based on sentence structure analysis is introduced,which can accurately calculate the senmantic similarity and cluster texts.The method combines the advantages of the improved K-means, the sentence structure analysis is defined for reducing the complexity of the text set to improve the accuracy of the calculation of the semantic similarity between texts. Experimental results show that the method can gain a higher precision(0.96)than some widely used clustering methods.
Keywords:text clusteringK-meanssentence structure analysis
Publication Date:2018-05-02
Online Publishing Date:2025-08-15(First online date of this platform, not the publication date of the document)
Pages:4( 933-935,1067 )
Computer and Digital Engineering

Computer and Digital Engineering

ISTIC
ISSN:1672-9722
Year, Vol.(Issue):2018,46(5)