GANsynth

Google Magenta · 2019

Algorithm for synthesizing audio with generative adversarial networks.

Application types: Audio synthesisMIDI-to-audio
Architecture:GAN

Essential categories

CategoryStatusNotes
Source code✔︎ OpenSource code is available.
Training data~ PartialDataset is available, but the reduced version that they use and the newly created test/train splits are not provided.
Model weights✔︎ OpenTwo pretrained checkpoints are provided.
Code documentation✔︎ OpenCodebase is documented, including high-level instructions and configuration files.
Training procedure✔︎ OpenTraining procedure is fully documented, including hardware requirements and model configuration.
Evaluation procedure~ PartialEvaluation metrics are described but exact implementations are not referenced.
Research paper✔︎ OpenAccepted at ICLR 2019.
Licensing✔︎ OpenSystem is covered by Apache License.

Desirable categories

CategoryStatusNotes
Model card∅ Not includedVery few details are given of the pretrained checkpoints.
Datasheet∅ Not includedNot available.
Package⭐ IncludedAvailable within the magenta pip package.
User-oriented application∅ Not includedNot available.
Supplementary material page⭐ IncludedDemo page is available, including sound examples of the model's capabilities.

Raw YAML file with complete evaluation.