Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:SMITIN: Self-Monitored Inference-Time INtervention for Generative Music Transformers

Apr 02, 2024

Junghyun Koo, Gordon Wichern, Francois G. Germain, Sameer Khurana, Jonathan Le Roux

Figure 1 for SMITIN: Self-Monitored Inference-Time INtervention for Generative Music Transformers

Figure 2 for SMITIN: Self-Monitored Inference-Time INtervention for Generative Music Transformers

Figure 3 for SMITIN: Self-Monitored Inference-Time INtervention for Generative Music Transformers

Figure 4 for SMITIN: Self-Monitored Inference-Time INtervention for Generative Music Transformers

Share this with someone who'll enjoy it:

Abstract:We introduce Self-Monitored Inference-Time INtervention (SMITIN), an approach for controlling an autoregressive generative music transformer using classifier probes. These simple logistic regression probes are trained on the output of each attention head in the transformer using a small dataset of audio examples both exhibiting and missing a specific musical trait (e.g., the presence/absence of drums, or real/synthetic music). We then steer the attention heads in the probe direction, ensuring the generative model output captures the desired musical trait. Additionally, we monitor the probe output to avoid adding an excessive amount of intervention into the autoregressive generation, which could lead to temporally incoherent music. We validate our results objectively and subjectively for both audio continuation and text-to-music applications, demonstrating the ability to add controls to large generative models for which retraining or even fine-tuning is impractical for most musicians. Audio samples of the proposed intervention approach are available on our demo page http://tinyurl.com/smitin .

View paper on

Share this with someone who'll enjoy it:

Title:SMITIN: Self-Monitored Inference-Time INtervention for Generative Music Transformers

Paper and Code