Separated examples by vocal ensemble separation methods using our jaCappella corpus

Tomohiko Nakamura, Shinnosuke Takamichi, Naoko Tanji (The University of Tokyo), Satoru Fukayama (AIST), Hiroshi Saruwatari (The University of Tokyo)

This is the demo page of our ICASSP 2023 paper [1]. We show separated examples by vocal ensemble separation methods using our jaCappella corpus. The project page of the corpus is here.


Mixture Voice part Ground Truth X-UMX [2] DPTNet [3] MRDLA [4]
Dongurikorokoro
(どんぐりころころ in Japanese)
Genre: bossa nova
Lead vocal
Soprano
Alto
Tenor
Bass
Vocal percussion
Hiraitahiraita
(ひらいたひらいた in Japanese)
Genre: enka
Lead vocal
Soprano
Alto
Tenor
Bass
Vocal percussion
Otamajakushi
(お玉じゃくし in Japanese)
Genre: jazz
Lead vocal
Soprano
Alto
Tenor
Bass
Vocal percussion

References

[1] {{ get_bibentry_html("TNakamura202306ICASSP", lang="en") }}
[2] https://github.com/asteroid-team/asteroid/tree/master/egs/musdb18/X-UMX
[3] S. Sarkar, E. Benetos, and M. Sandler, "Vocal harmony separation using time-domain neural networks," in Proc. INTERSPEECH, 2021, pp. 3515–3519.
[4] {{ get_bibentry_html("TNakamura202104IEEEACMTASLP", lang="en") }}