Using The Barton Libraries Dataset As An RDF benchmark

Daniel J. Abadi, Adam Marcus, Samuel R. Madden, Kate Hollenbach · DSpace@MIT (Massachusetts Institute of Technology) · 2007

This report describes the Barton Libraries RDF dataset and Long-well query benchmark that we use for our recent VLDB paper on Scalable Semantic Web Data Management Using Vertical Partition-ing [4]. 1. BARTON DATA The dataset used for this benchmark is taken from the publicly available Barton Libraries dataset [1]. This data is provided by the Simile Project [3], which develops tools for library data manage-ment and interoperability. The data contains records that compose an RDF-formatted dump of the MIT Libraries Barton catalog, con-verted from raw data stored in an old library format standard called MARC (Machine Readable Catalog). Because of the multiple sources the data was derived from and the diverse nature of the data that is cataloged, the structure of the data is quite irregular.

Read the paper · More papers on PaperTik