The MIDAS data-mining project at Stanford
Jeffrey David Ullman · 2003
The article summarizes recent research into data-mining techniques that are in progress at Stanford: 1. The Google search engine: beating Yahoo et al. at their own game. 2. Query flocks: generalizing association rules/market baskets in a query precompiler that uses a relational DBMS effectively. 3. Synthesizing knowledge from the Web: exploiting the Web's redundancy to extract data automatically. 4. Detecting low-frequency events: unlike marketing, where you only care about items that lots of people buy, extracting intelligence from text usually requires looking for a small number of unexpected juxtapositions of terms.