An Overview of the ICASSP Special Session on AI Security and Privacy in Speech and Audio Processing
Zhao Xu Ren, Kun Qian, Tanja Schultz, Björn Wolfgang Schuller · 2023
Perceiving and producing speech and audio signals are the basic ways for humans to communicate with each other and know about the world. Benefiting from the advancement of Big Data, signal processing, and Artificial Intelligence (AI), intelligent machines have been rapidly developed to process speech and audio signals for assisting human life. Deep learning has been demonstrated to achieve excellent performance based on large amounts of data streams. In the meanwhile, the problems of security vulnerability and privacy leakage appear along with the booming technologies. Systems with security and privacy problems can expose users’ personal information to danger and cause users’ distrust. To facilitate technology development in tackling the aforementioned issues, the special session on “AI security and privacy in speech and audio processing" was organised at ICASSP 2023. In this study, we provide a comprehensive overview of the invited high-quality contributions at the special session. We further discuss the current research challenges, and point out potential avenues for future works. This work is expected to summarise the research advancements and inspire more innovative studies in this area.