Energy Data Collection Project Year 1

José Luis Ambite, Yigal Arens, Luis Gravano, Vasilis Hatzivassiloglou, Eduard Hovy, Judith L. Klavans, Andrew Philpot, Usha Ramachandran, Jay Sandhaus, Amit Singla, Brian Whitman, Eduard H. Hovy · 2000

The massive amount of statistical and text data available from Federal Agencies has created a set of daunting challenges to both research and analysis communities. These problems include heterogeneity, size, distribution, and control of terminology. At the Digital Government Research Center we are investigating solutions to three key problems, namely, (1) ontological mappings for terminology standardization; (2) data integration across data bases with high speed query processing; and (3) interfaces for query input and presentation of results. This collaboration between researchers from Columbia University and the Information Sciences Institute of the University of Southern California employs technology developed at both locations, in particular the SENSUS ontology, the SIMS multi-database access planner, the LKB automated dictionary and terminology analysis system, and others. The pilot application targets gasoline data from BLS, EIA, Census, and other agencies. • Introduction: The Digital Government Research Center As access to the web becomes a household commodity, the Government (and in particular Federal Agencies such as the Census Bureau, the Bureau of Labor Statistics, and others) has a mandate to make its

Read the paper · More papers on PaperTik