The semantic annotated documents: from HTML to the semantic web
Jason Chi-Shun Hung · 2007
Abstract:-The current circumstance of the Semantic Web is that there is not much of a Semantic Web due to the lack of annotated web pages. There is such a lack because annotating web pages currently does not provide much practical benefit. In this work an automated approach to semantics extraction and annotation on textual data is proposed. Word sense disambiguation technique is used to identify the concepts, and RDF is used to annotate the semantics. A corresponding approach to retrieve data via ontology is also discussed. Finally a framework to integrate and automate these processes is demonstrated. In this fashion all the existing data on the Web can be processed and brought to the Semantic Web.