Finite-horizon quickest search in correlated high-dimensional data
Saeid Balaneshin, Ali Tajer, H. Vincent Poor · 2014
The problem of searching over a large number of data streams for identifying one that holds certain features of interest is considered. The data streams are assumed to be generated by one of two possible statistical distributions with cumulative distribution functions F0and F1and the objective is to identify one sequence generated by F1as quickly as possible, and prior to a pre-specified deadline. Furthermore, it is assumed that the generation of the data streams follows a known dependency kernel such that the likelihood of a sequence being generated by F1depends on the underlying distributions of the other data streams. The optimal sequential sampling strategy is characterized, and numerical evaluations are provided to illustrate the gains of incorporating the information about the dependency structure into the design of the sampling process.