A Web data extraction approach to harvesting data from online sources
Richi Nayak, Magnus Haugaasen · QUT ePrints (Queensland University of Technology) · 2006
With the Web becoming a main source of data representation, businesses have opportunities to gather data from various independent web sources and condense these data into specialized services.However, there is no unified structure of web pages and therefore extracting data from sources can be a complex task.We present a solution to locate and extract data from a large group of online bookmaker pages to provide a real-time service to deliver price on sporting events.