Robust automated music transcription systems
Andrew Sterian, Gregory H. Wakefield · 1996
We define an automated music transcription system as a sequence of processing modules, beginning with acoustic acquisition and ending with a pitch sequence. This modular approach allows the various system components to be studied and optimized separately. Our own single-voice transcription system is presented as a working example of such modular designs. To assess the performance of such transcription systems under real-world constraints, we propose a single numerical score based on matching known events in the source material to corresponding events in the transcription. We use this score to evaluate our transcription system under several conditions.