Identifying the concept of Image and Captioning Using Deep Neural Networks
International Journal of Advanced Trends in Computer Science and Engineering · 2020
In recent years, Image captioning has become a challenging artificial intelligence problem.Many researchers have been interested in the field of AI and became an arduous and exciting task.Image captioning automatically generates the textual description consistent with the content observed in a picture, and it is the mixture of two methods, including computer vision and natural language processing.Computer vision is to understand the images' content and natural language processing to understand the image into words in the correct order.Recently, Deep learning methods are achieving better results on caption generation problems.They can define a single end-to-end model to predict a caption when a photograph is given, instead of requiring a pipeline of specifically designed models or sophisticated data preparation.By using deep learning techniques like CNN, RNN accurate descriptions can be predicted.Convolutional Neural Network (CNN) implicitly extract features from the image, and Recurrent Neural Network is used for sentence generation.The developed model was trained to capture the image concept and generate the textual description observed in an image.