Chinese Web Page Clustering Algorithm Based on the Suffix Tree
摘要
In this paper, an improved algorithm, named STC-I. is proposed for Chinese Web page clustering based on Chinese language characteristics, which adopts a new unit choice principle and a novel suffix tree construction policy. The experimental results show that the new algorithm keeps advantages of STC, and is better than STC in precision and speed when they are used to cluster Chinese Web page.
引用本文(GB/T 7714)
YANGJian-wu. Chinese Web Page Clustering Algorithm Based on the Suffix Tree[J]. Acta Scientiarum Naturalium Universitatis Sunyatseni, 2004.
引文网络
本站仅收录题录与摘要供学习参考,全文版权归属出版方;如有侵权请联系我们删除。