Webcam-Based OCR System with Text-to-Speech Conversion

Rekha R Nair, Tina Babu, S.P. Kishore, J. Poornima, Madduru Sambasivudu, M. Bhargavi · 2024

The project develops an OCR system that converts webcam-captured text to speech, featuring user registration for personalized experiences. Key components include image capture via webcam, text extraction using OCR, database storage, and text-to-speech conversion. The system employs Python, OpenCV for image processing, Tesseract for OCR, and Google Text-to-Speech for audio output. Main objectives are to streamline text digitization and assist visually impaired users. The modular design allows for future enhancements like language translation. Comprehensive testing ensures reliability and effectiveness. The system demonstrates an efficient solution for converting printed text to digital and audio formats, with applications in document digitization, education, and accessibility services. It reduces manual effort in text digitization and improves access to written information for visually impaired individuals, showcasing potential across various fields.

Read the paper · More papers on PaperTik