Some Mistakes Of Linguistic Corpususers

Г Ф Лутфуллина · ˜The œEuropean Proceedings of Social & Behavioural Sciences · 2018

The linguistic corpus users must be aware of receiving erroneous data. The first error is related to the word frequency in the diachronic perspective. It is necessary to use special formula in order to correctly calculate the word’s frequency. The second error is related to the grammatical search. It is difficult to set correct search parameters. The third mistake is connected with lexical homonyms. It is necessary to be cautious when you meet all lexical homonyms. The fourth mistake is related to semantic features combination search. If you "play" with semantic features search, you can get ambiguous results. The user of the linguistic corpus should not rely entirely on the search results. There is always some "noise" in search results caused sometimes by contexts shortages. It is necessary to evaluate the received data based on your language competence, to correctly select and set the search parameters. An improperly compiled search can lead to distortion of real results: their exaggeration or understatement.

Read the paper · More papers on PaperTik