Post-edited and error annotated machine translation corpus PErr 1.0
Maja Popović, Mihael Arčan · Americanae (AECID Library) · 2016
The PE²rr corpus contains source language texts from different domains along with their automatically generated translations into several morphologically rich languages, their post-edited versions, and error annotations of the performed post-edit operations. The main advantage of the corpus is the fusion of post-editing and error classification tasks, which have usually been seen as two independent tasks, although naturally they are not.