Weakly Supervised Multi-Pitch Estimation Using Cross-Version Alignment

Krause M, Strahl S, Müller M (2023)


Publication Language: English

Publication Type: Conference contribution, Conference Contribution

Publication year: 2023

Publisher: ISMIR

Pages Range: 289-296

Conference Proceedings Title: Proceedings of the International Society for Music Information Retrieval Conference (ISMIR)

Event location: Mailand IT

DOI: 10.5281/ZENODO.10265279

Abstract

Multi-pitch estimation (MPE), the task of detecting active pitches within a polyphonic music recording, has garnered significant research interest in recent years. Most state-of-the-art approaches for MPE are based on deep networks trained using pitch annotations as targets. The success of current methods is therefore limited by the difficulty of obtaining large amounts of accurate annotations. In this paper, we propose a novel technique for learning MPE without any pitch annotations at all. Our approach exploits multiple recorded versions of a musical piece as surrogate targets. Given one version of a piece as input, we train a network to minimize the distance between its output and time-frequency representations of other versions of that piece. Since all versions are based on the same musical score, we hypothesize that the learned output corresponds to pitch estimates. To further ensure that this hypothesis holds, we incorporate domain knowledge about overtones and noise levels into the network. Overall, our method replaces strong pitch annotations with weaker and easier-to-obtain cross-version targets. In our experiments, we show that our proposed approach yields viable multi-pitch estimates and outperforms two baselines.

Authors with CRIS profile

How to cite

APA:

Krause, M., Strahl, S., & Müller, M. (2023). Weakly Supervised Multi-Pitch Estimation Using Cross-Version Alignment. In Proceedings of the International Society for Music Information Retrieval Conference (ISMIR) (pp. 289-296). Mailand, IT: ISMIR.

MLA:

Krause, Michael, Sebastian Strahl, and Meinard Müller. "Weakly Supervised Multi-Pitch Estimation Using Cross-Version Alignment." Proceedings of the International Society for Music Information Retrieval Conference (ISMIR), Mailand ISMIR, 2023. 289-296.

BibTeX: Download