VampNet

Northwestern University and Descript Inc ¡ 2023

A masked acoustic token modeling approach to music generation.

Application types: Audio-to-audio
Architecture:Non-autoregressive transformer

Essential categories

CategoryStatusNotes
Source code✔︎ OpenSystem source code is available and accessible.
Training data✘ ClosedDataset is not available nor is it completely described.
Model weights✔︎ OpenModel checkpoints are available. The weights of the models are licensed https://creativecommons.org/licenses/by-nc-sa/4.0/deed.ml
Code documentation✔︎ OpenCodebase is well documented.
Training procedure✔︎ OpenDescribed in the article, with multiple details on architecture, hyperparameters and GPU usage.
Evaluation procedure~ Partial
Research paper✔︎ OpenAccepted at ISMIR 2023.
Licensing✔︎ OpenMIT License.

Desirable categories

CategoryStatusNotes
Model card∅ Not included
Datasheet∅ Not included
Package∅ Not included
User-oriented application⭐ IncludedGradio UI is provided in the codebase.
Supplementary material page⭐ IncludedComplementary demo page with sound examples and demonstration of the model's capabilities is available.

Raw YAML file with complete evaluation.