• DocumentCode
    1066558
  • Title

    Improving Multilabel Analysis of Music Titles: A Large-Scale Validation of the Correction Approach

  • Author

    Pachet, François ; Roy, Pierre

  • Author_Institution
    SONY Comput. Sci. Lab., Paris
  • Volume
    17
  • Issue
    2
  • fYear
    2009
  • Firstpage
    335
  • Lastpage
    343
  • Abstract
    This paper addresses the problem of automatically extracting perceptive information from acoustic signals, in a supervised classification context. Global labels, i.e., atomic information describing a music title in its entirety, such as its genre, mood, main instruments, or type of vocals, are entered by humans. Classifiers are trained to map audio features to these labels. However, the performances of these classifiers on individual labels are rarely satisfactory. In the case we have to predict several labels simultaneously, we introduce a correction scheme to improve these performances. In this scheme-an instance of the classifier fusion paradigm-an extra layer of classifiers is built to exploit redundancies between labels and correct some of the errors coming from the individual acoustic classifiers. We describe a series of experiments aiming at validating this approach on a large-scale database of music and metadata (about 30 000 titles and 600 labels per title). The experiments show that the approach brings statistically significant improvements.
  • Keywords
    acoustic signal processing; feature extraction; learning (artificial intelligence); meta data; music; acoustic signals; feature extraction; multilabel analysis; music; supervised classification; Data mining; Humans; Large-scale systems; Mood; Multiple signal classification; Music; Performance analysis; Robustness; Spatial databases; Speech analysis; Feature extraction; learning systems; music; pattern classification;
  • fLanguage
    English
  • Journal_Title
    Audio, Speech, and Language Processing, IEEE Transactions on
  • Publisher
    ieee
  • ISSN
    1558-7916
  • Type

    jour

  • DOI
    10.1109/TASL.2008.2008734
  • Filename
    4749454