AUDIO CAPTION GENERATION FROM IMAGES USING DEEP LEARNING

Omprakash Jatashankar Yadav, Atharva Jadhav, Abdul Hannan Sunsara, Idris Vohra · International Journal Of Trendy Research In Engineering And Technology · 2021

Visually impaired individuals face various types of difficulties as they cannot visualize the natural environment.To overcome this problem, the proposed system would automatically generate captions for an input images and convert the generated caption to an audio format so that visually impaired individuals can listen to the generated captions.Captioning is performed using Deep Learning algorithm Convolution Neural Network (CNN), Recurrent Neural Network (RNN) and Long Short-Term Memory.

Read the paper · More papers on PaperTik