Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Janusz Klejsa

Audio Decoding by Inverse Problem Solving

Sep 12, 2024

Pedro J. Villasana T., Lars Villemoes, Janusz Klejsa, Per Hedelin

Figure 1 for Audio Decoding by Inverse Problem Solving

Figure 2 for Audio Decoding by Inverse Problem Solving

Figure 3 for Audio Decoding by Inverse Problem Solving

Figure 4 for Audio Decoding by Inverse Problem Solving

Abstract:We consider audio decoding as an inverse problem and solve it through diffusion posterior sampling. Explicit conditioning functions are developed for input signal measurements provided by an example of a transform domain perceptual audio codec. Viability is demonstrated by evaluating arbitrary pairings of a set of bitrates and task-agnostic prior models. For instance, we observe significant improvements on piano while maintaining speech performance when a speech model is replaced by a joint model trained on both speech and piano. With a more general music model, improved decoding compared to legacy methods is obtained for a broad range of content types and bitrates. The noisy mean model, underlying the proposed derivation of conditioning, enables a significant reduction of gradient evaluations for diffusion posterior sampling, compared to methods based on Tweedie's mean. Combining Tweedie's mean with our conditioning functions improves the objective performance. An audio demo is available at https://dpscodec-demo.github.io/.

* 5 pages, 4 figures, audio demo available at https://dpscodec-demo.github.io/, pre-review version submitted to ICASSP 2025

Via

Access Paper or Ask Questions

Distribution Preserving Source Separation With Time Frequency Predictive Models

Mar 10, 2023

Pedro J. Villasana T., Janusz Klejsa, Lars Villemoes, Per Hedelin

Figure 1 for Distribution Preserving Source Separation With Time Frequency Predictive Models

Figure 2 for Distribution Preserving Source Separation With Time Frequency Predictive Models

Figure 3 for Distribution Preserving Source Separation With Time Frequency Predictive Models

Figure 4 for Distribution Preserving Source Separation With Time Frequency Predictive Models

Abstract:We provide an example of a distribution preserving source separation method, which aims at addressing perceptual shortcomings of state-of-the-art methods. Our approach uses unconditioned generative models of signal sources. Reconstruction is achieved by means of mix-consistent sampling from a distribution conditioned on a realization of a mix. The separated signals follow their respective source distributions, which provides an advantage when separation results are evaluated in a listening test.

* 5 pages, 4 figures, pre-review version submitted to EUSIPCO 2023

Via

Access Paper or Ask Questions