Musical Polyphony Estimation
Published in Proceedings of the Audio Engineering Society, Convention 144, 2018
Recommended citation: S. Kareer, and S. Basu, "Musical Polyphony Estimation," Engineering Brief 434, (2018 May.)' http://www.aes.org/e-lib/browse.cfm?elib=19547
Knowing the number of sources present in a mixture is useful for many computer audition problems such as polyphonic music transcription, source separation, and speech enhancement. Most existing algorithms for these applications require the user to provide this number thereby limiting the possibility of complete automatization. In this paper we explore a few probabilistic and machine learning approaches for an autonomous source number estimation. We then propose an implementation of a multi-class classification method using convolutional neural networks for musical polyphony estimation. In addition, we use these results to improve the performance of an instrument classifier based on the same dataset. Our final classification results for both the networks, prove that this method is a promising starting point for further advancements in unsupervised source counting and separation algorithms for music and speech.