A comprehensive trainable error model for sung music queries

  • Authors:
  • Colin J. Meek;William P. Birmingham

  • Affiliations:
  • Microsoft Corporation, SQL Server, Redmond, WA;Math and Computer Science, Grove City College, Grove City, PA

  • Venue:
  • Journal of Artificial Intelligence Research
  • Year:
  • 2004

Quantified Score

Hi-index 0.00

Visualization

Abstract

We propose a model for errors in sung queries, a variant of the hidden Markov model (HMM). This is a solution to the problem of identifying the degree of similarity between a (typically error-laden) sung query and a potential target in a database of musical works, an important problem in the field of music information retrieval. Similarity metrics are a critical component of "query-by-humming" (QBH) applications which search audio and multimedia databases for strong matches to oral queries. Our model comprehensively expresses the types of error or variation between target and query: cumulative and noncumulative local errors, transposition, tempo and tempo changes, insertions, deletions and modulation. The model is not only expressive, but automatically trainable, or able to learn and generalize from query examples. We present results of simulations, designed to assess the discriminatory potential of the model, and tests with real sung queries, to demonstrate relevance to real-world applications.