A CNN and LSTM-based Model for Creating Captions for Photos

Balamuralikrishna Thati, Swathi Voddi, Srikanth Busa, Surendra Surendra, J.N.V.R. Swarup Kumar, Madhusudan Rao · International Journal on Recent and Innovation Trends in Computing and Communication · 2023

Can a machine interpret an image's meaning with the same speed as the human brain when it is seen? This problem was heavily researched by computer vision specialists, who believed it to be unsolvable until recently. It is now possible to develop models that can generate captions for pictures because of advancements in deep learning techniques, accessibility to large datasets, and processing power. This will be accomplished by the Python-based implementation of the article's deep learning convolutional neural network technique and a particular kind of recurrent neural network. Here the proposed model uses CNN and LSTM methods to achieve desired task

Read the paper · More papers on PaperTik