Histogram of gradients of Time-Frequency Representations for Audio Scene Detection

Alain Rakotomamonjy, Gilles Gasso · IEEE/ACM Transactions on Audio Speech and Language Processing · 2014

Presents our entry to the Detection and Classification of Acoustic Scenes challenge. The approach we propose for classifying acoustic scenes is based on transforming the audio signal into a time-frequency representation and then in extracting relevant features about shapes and evolutions of time-frequency structures. These features are based on histogram of gradients that are subsequently fed to a multi-class linear support vector machines.

Read the paper · More papers on PaperTik