M. Graciarena, M. Delplanche, E. Shriberg, A. Stolcke, and L. Ferrer, “Acoustic front-end optimization for bird species recognition,” in Proc. 2010 IEEE International Conference on Acoustics Speech and Signal Processing (ICASSP 2010), pp. 293–296.
Abstract
The goal of this work was to explore the optimization of the feature extraction module (front-end) parameters to improve bird species recognition. We explored optimizing the spectral and temporal parameters of a Mel cepstrum feature-based front-end, starting from common parameter values used in speech processing experiments. These features were modeled using a Gaussian mixture model (GMM) system. We found an important improvement when increasing the spectral bandwidth and increasing the number of filter banks. We found no improvement when switching the filter bank distribution…
Share this



