Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Alvin W. Y. Su

Melody Infilling with User-Provided Structural Context

Oct 06, 2022

Chih-Pin Tan, Alvin W. Y. Su, Yi-Hsuan Yang

Figure 1 for Melody Infilling with User-Provided Structural Context

Figure 2 for Melody Infilling with User-Provided Structural Context

Figure 3 for Melody Infilling with User-Provided Structural Context

Figure 4 for Melody Infilling with User-Provided Structural Context

Abstract:This paper proposes a novel Transformer-based model for music score infilling, to generate a music passage that fills in the gap between given past and future contexts. While existing infilling approaches can generate a passage that connects smoothly locally with the given contexts, they do not take into account the musical form or structure of the music and may therefore generate overly smooth results. To address this issue, we propose a structure-aware conditioning approach that employs a novel attention-selecting module to supply user-provided structure-related information to the Transformer for infilling. With both objective and subjective evaluations, we show that the proposed model can harness the structural information effectively and generate melodies in the style of pop of higher quality than the two existing structure-agnostic infilling models.

Via

Access Paper or Ask Questions

Music Score Expansion with Variable-Length Infilling

Nov 11, 2021

Chih-Pin Tan, Chin-Jui Chang, Alvin W. Y. Su, Yi-Hsuan Yang

Figure 1 for Music Score Expansion with Variable-Length Infilling

Figure 2 for Music Score Expansion with Variable-Length Infilling

Figure 3 for Music Score Expansion with Variable-Length Infilling

Abstract:In this paper, we investigate using the variable-length infilling (VLI) model, which is originally proposed to infill missing segments, to "prolong" existing musical segments at musical boundaries. Specifically, as a case study, we expand 20 musical segments from 12 bars to 16 bars, and examine the degree to which the VLI model preserves musical boundaries in the expanded results using a few objective metrics, including the Register Histogram Similarity we newly propose. The results show that the VLI model has the potential to address the expansion task.

* Going to published as a late-breaking demo paper at ISMIR 2021

Via

Access Paper or Ask Questions