Grammatical Error Correction: Machine Translation and Classifiers

Alla Rozovskaya, Dan Roth · 2016

We focus on two leading state-of-the-art approaches to grammatical error correction -machine learning classification and machine translation.Based on the comparative study of the two learning frameworks and through error analysis of the output of the state-of-the-art systems, we identify key strengths and weaknesses of each of these approaches and demonstrate their complementarity.In particular, the machine translation method learns from parallel data without requiring further linguistic input and is better at correcting complex mistakes.The classification approach possesses other desirable characteristics, such as the ability to easily generalize beyond what was seen in training, the ability to train without human-annotated data, and the flexibility to adjust knowledge sources for individual error types.Based on this analysis, we develop an algorithmic approach that combines the strengths of both methods.We present several systems based on resources used in previous work with a relative improvement of over 20% (and 7.4 F score points) over the previous state-of-the-art.System Method Performance P R F0.5 CoNLL-2014 top 3 MT 41.62 21.40 35.01 CoNLL-2014 top 2 Classif.41.78 24.88 36.79CoNLL-2014 top 1 MT, rules 39.71 30.10 37.33 Susanto et al. (2014) MT, classif.53.55 19.14 39.39 Miz.& Mats.(2016) MT 45.80 26.60 40.00This work MT, classif.60.17 25.64 47.40

Read the paper · More papers on PaperTik