Learning grammar of complex activities via deep neural networks

Becky Mashaido · 2021

Motivated by the growing amount of publicly available video data on online streaming servicesand an increased interest in applications that analyze continuous video streams such as autonomous driving , this technical report provides a theoretical insight into deep neural networks for video learning, under label constraints. I build upon previous work in video learning for computer vision, make observations on model performance and propose further mechanisms to help improve our observations.--Author's abstract

Read the paper · More papers on PaperTik