Building automatically a business registration ontology

Melania Degeratu, Vasileios Hatzivassiloglou · International Conference on Digital Government Research · 2002

We discuss a domain-independent, corpus based method for dictionary-less automatic extraction of ontological knowledge from domain-specific unannotated documents. We present the architecture, algorithms, and results for ONTOSTRUCT---a new system that uses machine learning and statistical techniques to analyze text sources, discover terms, link equivalent terms into concepts, learn both hierarchical and non-hierarchical conceptual relations, and build an extensive, semantically sound hierarchy of concepts. We report on ONTOSTRUCT's results in constructing a domain-specific ontology for the business registration domain, and evaluate the performance of two of its modules.

Read the paper · More papers on PaperTik