A Comparative Analysis of Image Captioning Techniques
Diya Theresa Sunil, Seema Safar, Abhijith Das, M M Amijith, Devika M Joshy · 2023
Image captioning is the task of generating a textual description that accurately represents the content of an image. This task involves combining computer vision techniques, such as object recognition and scene understanding, with natural language processing to produce a human-like description of an image. Over time, various models have been introduced to perform image captioning, all aiming to accurately describe the content of an image. These models have practical applications such as improving the accessibility of multimedia content, assisting individuals with visual impairments, medical image captioning, and enhancing image search and retrieval. This paper explores some of the models and studies their efficiency using different evaluation metrics.