An Efficient Feature Subset Selection for Improved Stability Using T-Statistic

R. Karthika · 2017

Large amounts of data gets accumulated and stored in the databases in day to day life that are high dimensional in nature. The data mining task is used to excavate the useful information from the high dimensional data. To classify or cluster the high dimensional data, the dimensionality of the data needs to be reduced. Feature selection is used to select the features that are relevant to the analysis and discards the features that are not relevant as well as redundant. There are so many feature subset selection algorithms available. In this paper, we evaluate the stability of the subset of the features selected using a measure called T-Statistic and improve the prediction accuracy of the classifier using Booster.

Read the paper · More papers on PaperTik