6 ms·
they made an extended abstract for ismir: http://archives.ismir.net/ismir2019/latebreaking/000036.pdf http://archives.ismir.net/ismir2019/latebreaking/000036.pd
by czr 7y ago
they made an extended abstract for ismir: http://archives.ismir.net/ismir2019/latebreaking/000036.pdf http://archives.ismir.net/ismir2019/latebreaking/000036.pdf
methodology is a separate u-net per instrument type to predict a soft mask in spectrogram space (time x frequency), then they apply that mask to the input audio. fairly standard.