Structured Auto-Encoder with application to Music Genre Recognition

Defferrard, Michaël

Defferrard, Michaël

2015

Formats

Format
BibTeX
MARC
MARCXML
DublinCore
EndNote
NLM
RefWorks
RIS

Files

Abstract

In this work, we present a technique that learns discriminative audio features for Music Information Retrieval (MIR). The novelty of the proposed technique is to design auto-encoders that make use of data structures to learn enhanced sparse data representations. The data structure is borrowed from the Manifold Learning field, that is data are supposed to be sampled from smooth manifolds, which are here represented by graphs of proximities of the input data. As a consequence, the proposed auto-encoders finds sparse data representations that are quite robust w.r.t. perturbations. The model is formulated as a non-convex optimization problem. However, it can be decomposed into iterative sub-optimization problems that are convex and for which well-posed iterative schemes are provided in the context of the Fast Iterative Shrinkage-Thresholding (FISTA) framework. Our numerical experiments show two main results. Firstly, our graph-based auto-encoders improve the classification accuracy by 2% over the auto-encoders without graph structure for the popular GTZAN music dataset. Secondly, our model is significantly more robust as it is 8% more accurate than the standard model in the presence of 10% of perturbations.

Details

Title Structured Auto-Encoder with application to Music Genre Recognition

Author(s) Defferrard, Michaël

Advisor(s)

Vandergheynst, Pierre
Bresson, Xavier
Paratte, Johann

Date 2015

Keywords

auto-encoder; sparse representation; manifold learning; music information retrieval; graph; non-convex optimization

Additional link URL

Laboratories LTS2

Record Appears in Scientific production and competences > STI - School of Engineering > IEM - Institut d'Electricité et de Microtechnique > LTS2 - Signal Processing Laboratory 2
Work produced at EPFL
Student projects

Work type Master's Thesis

Record creation date 2016-04-12

Actions

Preview

Select file: