AI-synthesized speech : generation and detection

Ehab Abdrabuh · 2022

From speech to images, and videos, advances in machine learning have led to dramatic improvements in the quality and realism of so-called AI-synthesized content. While there are many exciting and interesting applications, this type of content can also be used to create convincing and dangerous fakes. We seek to develop forensic techniques that can distinguish a real human voice from a synthesized voice. We observe that deep neural networks used to synthesize speech introduce specific and unusual artifacts not typically found in human speech. Although not necessarily audible, we develop various detection algorithms to measure these artifacts and be able to differentiate between human and synthesized speech.

Read the paper · More papers on PaperTik