Merging Speech Recognition Capabilities with Large Language Models for Enhanced Communication and Interaction
Vivek Saurabh, Modi Himabindu, Vijilius Helena Raj, Amit Dutt, Pradeep Kumar Chandra, Hayder Naji Sameer · 2024
In summary Voice recognition and large language models have changed communication. The study presents SECIM, a new paradigm for two-way human-computer communication that feels natural and makes sense in their existing environment. SECIM has a voice synthesis network, dynamic language learning tool, and several attention levels. Comparing the model against six well-known conventional methodologies provided a complete evaluation. Statistics show the model performed better and more efficiently in all areas. In this age of rapid technical growth, speech recognition and big language models have revolutionized human communication. This confluence of cutting-edge technology has allowed many new applications, changing how people and machines connect. The results show that the SECIM might improve NLP, voice recognition, and direct communication in real-world contexts.