Audio Feature Learning with Triplet-Based Embedding Network

Xiaoyu Qi, Deshun Yang, Xiaoou Chen · Proceedings of the AAAI Conference on Artificial Intelligence · 2017

We propose a triplet-based network for audio feature learning for version identification. Existing methods use hand-crafted features for a music as a whole while we learn features by a triplet-based neural network on segment-level, focusing on the most similar parts between music versions. We conduct extensive experiments and demonstrate our merits.

Read the paper · More papers on PaperTik