Controlling your TV with gestures

Mingyu Chen, Lily Mummert, Padmanabhan S. Pillai, Alexander G. Hauptmann, Rahul Sukthankar · 2010

Vision-based user interfaces enable natural interaction modalities such as gestures. Such interfaces require computationally intensive video processing at low latency. We demonstrate an application that recognizes gestures to control TV operations. Accurate recognition is achieved by using a new descriptor called MoSIFT, which explicitly encodes optical flow with appearance features. MoSIFT is computationally expensive - a sequential implementation runs 100 times slower than real time. To reduce latency sufficiently for interaction, the application is implemented on a runtime system that exploits the parallelism inherent in video understanding applications.

Read the paper · More papers on PaperTik