AudioScene: Enhancing Visual Independence Through Scene Recognition

S Vidhyalakshmi, J. M. Gnanasekar · International Research Journal on Advanced Engineering Hub (IRJAEH) · 2024

Visually challenged individuals face numerous challenges in navigating and understanding their surroundings due to their reliance on visual information. This initiative aims to empower them by utilizing cutting-edge technology to enhance their accessibility and independence. By employing deep learning algorithms captions are generated for the image, providing users with information about their surroundings. Leveraging advanced image captioning techniques and datasets like MS COCO, key features are highlighted. The system then delivers this information as audio output to the user, enabling them to navigate with confidence. Ultimately, this innovative solution, AudioScene, offers valuable support to visually impaired individuals, facilitating safer and more informed travel experiences.

Read the paper · More papers on PaperTik