Sparse Gaussian process Audio Source Separation Using Spectrum Priors in the Time-Domain

10/30/2018
by   Pablo A. Alvarado, et al.
2

Gaussian process (GP) audio source separation is a time-domain approach that circumvents the inherent phase approximation issue of spectrogram based methods. Furthermore, through its kernel, GPs elegantly incorporate prior knowledge about the sources into the separation model. Despite these compelling advantages, the computational complexity of GP inference scales cubically with the number of audio samples. As a result, source separation GP models have been restricted to the analysis of short audio frames. We introduce an efficient application of GPs to time-domain audio source separation, without compromising performance. For this purpose, we used GP regression, together with spectral mixture kernels, and variational sparse GPs. We compared our method with LD-PSDTF (positive semi-definite tensor factorization), KL-NMF (Kullback-Leibler non-negative matrix factorization), and IS-NMF (Itakura-Saito NMF). Results show the proposed method outperforms these techniques.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
10/31/2019

End-to-end Non-Negative Autoencoders for Sound Source Separation

Discriminative models for source separation have recently been shown to ...
research
02/09/2018

Complex ISNMF: a Phase-Aware Model for Monaural Audio Source Separation

This paper introduces a phase-aware probabilistic model for audio source...
research
09/20/2017

Neural Network Alternatives to Convolutive Audio Models for Source Separation

Convolutive Non-Negative Matrix Factorization model factorizes a given a...
research
11/18/2017

Separake: Source Separation with a Little Help From Echoes

It is commonly believed that multipath hurts various audio processing al...
research
04/24/2023

Adversarial Generative NMF for Single Channel Source Separation

The idea of adversarial learning of regularization functionals has recen...
research
11/15/2021

Overview and Introduction to Development of Non-Ergodic Earthquake Ground-Motion Models

This paper provides an overview and introduction to the development of n...
research
05/19/2017

Efficient Learning of Harmonic Priors for Pitch Detection in Polyphonic Music

Automatic music transcription (AMT) aims to infer a latent symbolic repr...

Please sign up or login with your details

Forgot password? Click here to reset