Automated semantic tagging using fuzzy grammar fragments

Trevor Martin, Yun Cheng Shen, Ben Azvine · 2008

One of the bottlenecks preventing wider adoption of the semantic Web is the overhead in annotating existing Web content. In cases where we have unstructured text, it is useful to extract fragments of structured data which can then be used as the basis for automatic tagging. A common approach is to use pattern matching (e.g. regular expressions) or more general grammar-based techniques, but these are not robust against small deviations. Fuzzy grammars allow partial matches, and we outline an efficient parsing technique to determine the degree to which a string is parsed by a grammar fragment. A simple application shows the methodpsilas validity.

Read the paper · More papers on PaperTik