Chord Segmentation and Recognition using EM-Trained Hidden Markov Models

dc.contributor.authorAlexander Shehen_US
dc.contributor.authorDaniel P.W. Ellisen_US
dc.contributor.editorHolger H. Hoosen_US
dc.contributor.editorDavid Bainbridgeen_US
dc.date.accessioned2004-10-21T04:26:31Z
dc.date.available2004-10-21T04:26:31Z
dc.date.issued2003-10-26en_US
dc.description.abstractAutomatic extraction of content description from commercial audio recordings has a number of important applications, from indexing and retrieval through to novel musicological analyses based on very large corpora of recorded performances. Chord sequences are a description that captures much of the character of a piece in a compact form and using a modest lexicon. Chords also have the attractive property that a piece of music can (mostly) be segmented into time intervals that consist of a single chord, much as recorded speech can (mostly) be segmented into time intervals that correspond to specific words. In this work, we build a system for automatic chord transcription using speech recognition tools. For features we use ``pitch class profile'' vectors to emphasize the tonal content of the signal, and we show that these features far outperform cepstral coefficients for our task. Sequence recognition is accomplished with hidden Markov models (HMMs) directly analogous to subword models in a speech recognizer, and trained by the same Expectation-Maximization (EM) algorithm. Crucially, this allows us to use as input only the chord sequences for our training examples, without requiring the precise timings of the chord changes --- which are determined automatically during training. Our results on a small set of 20 early Beatles songs show frame-level accuracy of around 75% on a forced-alignment task.en_US
dc.format.extent144037 bytes
dc.format.mimetypeapplication/pdf
dc.identifier.isbn0-9746194-0-Xen_US
dc.identifier.urihttp://jhir.library.jhu.edu/handle/1774.2/26
dc.language.isoen_US
dc.publisherJohns Hopkins Universityen_US
dc.subjectAudioen_US
dc.subjectMusic Analysisen_US
dc.titleChord Segmentation and Recognition using EM-Trained Hidden Markov Modelsen_US
dc.typeArticleen_US
Files
Original bundle
Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
paper.pdf
Size:
140.66 KB
Format:
Adobe Portable Document Format
Collections