Script identification for document images between Chinese and Korean language based on wavelet statistic features at the level of text row
Jin Jingxua · Journal of Yanbian University · 2013
A script identification method between Chinese and Korean language based on wavelet statistic feature is presented.To reduce the dimension and improve calculation efficiency,each 2Drow-document image partitioned from original document image is converted into 1Dprojection signal in both vertical and horizontal direction.1Dwavelet decomposition is implemented on the both projections.Then,wavelet statistic features are calculated for both projections and merged as feature vector of row-document.Effectiveness of wavelet statistic feature vector is evaluated by BP neural network.The experimental results show that the identification accuracy average around 94%.