The PIA Project: Learning to Semantically Annotate Texts from an Ontology and XML-Instance Data

Nigel Collier, Koichi Takeuchi, Keita Tsuji · 2001

The development of the XML and RDF(S) standards o#er a positive environment for machine learning to enable the automatic XML-annotation of texts that can encourage the extension of Semantic Web applications. After reviewing the current limitations of information extraction technology, specifically its lack of portability to new domains, we introduce the PIA project for automatically XML-annotating domain-based texts using example XML texts and an ontology for supervised training. 1.

Read the paper · More papers on PaperTik