High-performance regular expression scanning on the Cell/B.E. processor

Daniele Paolo Scarpazza, Gregory F. Russell · 2009

Matching regular expressions (regexps) is a very common work-load. For example, tokenization, which consists of recognizing words or keywords in a character stream, appears in every search engine indexer. Tokenization also consumes 30% or more of most XML processors' execution time and represents the first stage of any programming language compiler.

Read the paper · More papers on PaperTik