Music ControlNet

Carnegie Mellon University and Adobe Research · 2023

Generative model that offers multiple precise, time-varying controls over generated audio.

Application types: Melody-to-musicText-to-music
Architecture:Diffusion

Essential categories

CategoryStatusNotes
Source code✘ ClosedNo official code repository available.
Training data✘ ClosedVery limited description of the data used for training, and no sources provided.
Model weights✘ ClosedNo model weights provided.
Code documentation✘ ClosedNo source code available so no documentation.
Training procedure✔︎ OpenTraining procedure is described in detail in the paper.
Evaluation procedure~ PartialEvaluation procedure is described, but details on the specific implementation used for some of the evaluation metrics are not given, limiting the reproducibility of results.
Research paper✔︎ OpenAccepted at IEEE/ACM Transactions on Audio, Speech, and Languange Processing (open access).
Licensing✘ ClosedNot available.

Desirable categories

CategoryStatusNotes
Model card∅ Not includedNo model card available.
Datasheet∅ Not includedNo datasheet available.
Package∅ Not includedNo package available.
User-oriented application∅ Not includedNo user interface nor real-time implementation available.
Supplementary material page⭐ IncludedA complementary website with sound examples and demo of the model is available.

Raw YAML file with complete evaluation.