asteroid.models.sudormrf module¶

class asteroid.models.sudormrf.SuDORMRFImprovedNet(n_src, bn_chan=128, num_blocks=16, upsampling_depth=4, mask_act='relu', in_chan=None, fb_name='free', kernel_size=21, n_filters=512, stride=None, **fb_kwargs)[source]¶

Bases: asteroid.models.base_models.BaseTasNet

Improved SuDORMRF separation model, as described in [1].

Parameters:

n_src (int) – Number of sources in the input mixtures.
bn_chan (int, optional) – Number of bins in the bottleneck layer and the UNet blocks.
num_blocks (int) – Number of of UBlocks.
upsampling_depth (int) – Depth of upsampling.
mask_act (str) – Name of output activation.
in_chan (int, optional) – Number of input channels, should be equal to n_filters.
fb_name (str, className) – Filterbank family from which to make encoder and decoder. To choose among ['free', 'analytic_free', 'param_sinc', 'stft'].
n_filters (int) – Number of filters / Input dimension of the masker net.
kernel_size (int) – Length of the filters.
stride (int, optional) – Stride of the convolution. If None (default), set to kernel_size // 2.
**fb_kwargs (dict) – Additional kwards to pass to the filterbank creation.

References

[1] : “Sudo rm -rf: Efficient Networks for Universal Audio Source Separation”,: Tzinis et al. MLSP 2020.

class asteroid.models.sudormrf.SuDORMRFNet(n_src, bn_chan=128, num_blocks=16, upsampling_depth=4, mask_act='softmax', in_chan=None, fb_name='free', kernel_size=21, n_filters=512, stride=None, **fb_kwargs)[source]¶

Bases: asteroid.models.base_models.BaseTasNet

SuDORMRF separation model, as described in [1].

Parameters:

n_src (int) – Number of sources in the input mixtures.
bn_chan (int, optional) – Number of bins in the bottleneck layer and the UNet blocks.
num_blocks (int) – Number of of UBlocks.
upsampling_depth (int) – Depth of upsampling.
mask_act (str) – Name of output activation.
in_chan (int, optional) – Number of input channels, should be equal to n_filters.
fb_name (str, className) – Filterbank family from which to make encoder and decoder. To choose among ['free', 'analytic_free', 'param_sinc', 'stft'].
n_filters (int) – Number of filters / Input dimension of the masker net.
kernel_size (int) – Length of the filters.
stride (int, optional) – Stride of the convolution. If None (default), set to kernel_size // 2.
**fb_kwargs (dict) – Additional kwards to pass to the filterbank creation.

References

[1] : “Sudo rm -rf: Efficient Networks for Universal Audio Source Separation”,: Tzinis et al. MLSP 2020.