135
Views
10
CrossRef citations to date
0
Altmetric
Original Articles

Text feature selection algorithm based on Chi-square rank correlation factorization

Pages 153-160 | Received 01 Feb 2015, Published online: 03 Jan 2017
 

Abstract

Feather Selection is an effective method to reduce the dimension of text feature. The existing feature selection methods usually use empirical estimation methods when determining the scale of the feature selection. These methods have achieved good results in some specific corpora. However, it is not easy to promote the further generalization of automatic text categorization due to the insufficient theoretical basis. Therefore, a text feature selection algorithm based on Chi-square rank correlation factorization is proposed based on the comprehensive consideration of the whole and local distribution of text features. Under the condition that the algorithm does not need any prior knowledge, the feature weights are portrayed and the feature selection is completed, fully reflects the characteristics of the probability distribution.

Reprints and Corporate Permissions

Please note: Selecting permissions does not provide access to the full text of the article, please see our help page How do I view content?

To request a reprint or corporate permissions for this article, please click on the relevant link below:

Academic Permissions

Please note: Selecting permissions does not provide access to the full text of the article, please see our help page How do I view content?

Obtain permissions instantly via Rightslink by clicking on the button below:

If you are unable to obtain permissions via Rightslink, please complete and submit this Permissions form. For more information, please visit our Permissions help page.