We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.SD

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo ScienceWISE logo

Computer Science > Sound

Title: User Specific Adaptation in Automatic Transcription of Vocalised Percussion

Abstract: The goal of this work is to develop an application that enables music producers to use their voice to create drum patterns when composing in Digital Audio Workstations (DAWs). An easy-to-use and user-oriented system capable of automatically transcribing vocalisations of percussion sounds, called LVT - Live Vocalised Transcription, is presented. LVT is developed as a Max for Live device which follows the `segment-and-classify' methodology for drum transcription, and includes three modules: i) an onset detector to segment events in time; ii) a module that extracts relevant features from the audio content; and iii) a machine-learning component that implements the k-Nearest Neighbours (kNN) algorithm for the classification of vocalised drum timbres.
Due to the wide differences in vocalisations from distinct users for the same drum sound, a user-specific approach to vocalised transcription is proposed. In this perspective, a given end-user trains the algorithm with their own vocalisations for each drum sound before inputting their desired pattern into the DAW. The user adaption is achieved via a new Max external which implements Sequential Forward Selection (SFS) for choosing the most relevant features for a given set of input drum sounds.
Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)
Journal reference: Proc. of RecPad-2017, Amadora, Portugal, pp. 19-20, October, 2017
Cite as: arXiv:1811.02406 [cs.SD]
  (or arXiv:1811.02406v1 [cs.SD] for this version)

Submission history

From: Antonio Ramires [view email]
[v1] Tue, 6 Nov 2018 15:23:53 GMT (1360kb,D)

Link back to: arXiv, form interface, contact.