The PIA Project: Learning to Semantically Annotate Texts from an Ontology and XML-Instance Data
Nigel Collier, Koichi Takeuchi, Keita Tsuji · 2001
The development of the XML and RDF(S) standards o#er a positive environment for machine learning to enable the automatic XML-annotation of texts that can encourage the extension of Semantic Web applications. After reviewing the current limitations of information extraction technology, specifically its lack of portability to new domains, we introduce the PIA project for automatically XML-annotating domain-based texts using example XML texts and an ontology for supervised training. 1.