A RBF network for chinese text classification based on concept feature extraction

  • Authors:
  • Minghu Jiang;Lin Wang;Yinghua Lu;Shasha Liao

  • Affiliations:
  • School of Electronic Eng., Beijing University of Post and Telecommunication, Beijing, China and Lab. of Computational Linguistics, School of Humanities and Social Sciences, Tsinghua University, Be ...;School of Electronic Eng., Beijing University of Post and Telecommunication, Beijing, China and Interdisciplinary Center for Scientific Computing, IWR, University of Heidelberg, Heidelberg, German ...;School of Electronic Eng., Beijing University of Post and Telecommunication, Beijing, China;Lab. of Computational Linguistics, School of Humanities and Social Sciences, Tsinghua University, Beijing, China

  • Venue:
  • ICONIP'06 Proceedings of the 13th international conference on Neural information processing - Volume Part III
  • Year:
  • 2006

Quantified Score

Hi-index 0.00

Visualization

Abstract

The feature selection is an important part in automatic text classification. In this paper, we use a Chinese semantic dictionary -- Hownet to extract the concepts from the word as the feature set, because it can better reflect the meaning of the text. We construct a combined feature set that consists of both sememes and the Chinese words, propose a CHI-MCOR weighing method according to the weighing theories and classification precision. The effectiveness of the competitive network and the Radial Basis Function (RBF) network in text classification are examined. Experimental result shows that if the words are extracted properly, not only the feature dimension is smaller but also the classification precision is higher, the RBF network outperform competitive network for automatic text classification because of the application of supervised learning. Besides its much shorter training time than the BP network's, the RBF network makes precision and recall rates that are almost at the same level as the BP network's.