Research of Metadata Extracting Algorithm for Components from XML-Based in the Semantic Web

Hongbin Wang, Daxin Liu, Wei Sun · 2008

In this paper a novel algorithm to metadata extracting for components from XML-based based on rules in the semantic Web is proposed. This algorithm is based on the challenges facing institutions managing ever-increasing numbers of components over the long term. We consider the current environments and challenges for managing preservation components in the semantic Web. The Dublin core-based metadata is used to manage components in the semantic Web. Design of rule library is discussed after introducing syntax and semantics of the rules. The system named SCDB (software component database) is developed. The system is based on the algorithm aforementioned. The system can perform automatic Dublin core-based metadata extracting algorithm from components which is described by XML documents. The Dublin Core- based metadata extracting algorithm for components based on rules is described in this paper. The future works are discussed in the end.

Read the paper · More papers on PaperTik