This model is smaller in filesize (
2.13GB
VS
14.6GB
) due to being lower precision and having ema weights and some things stuff stripped.
Generated spectrograms will be different from the ones in the 14.6GB model
There is no noticable quality difference between the original and the small model
The small model is easier to load on low cpu RAM, for example: If you have only 16GB of RAM, loading the large model could have some issues, like in my case, my pc froze for a few seconds.
The small model loads faster than the large model
The large model is probably better for training, but i have had great success with training LoRA on the small model.
Riffusion (Original readme)
Riffusion is an app for real-time music generation with stable diffusion.
This repository contains the model files, including:
a diffusers formated library
a compiled checkpoint file
a traced unet for improved inference speed
a seed image library for use with riffusion-app
Riffusion v1 Model
Riffusion is a latent text-to-image diffusion model capable of generating spectrogram images given any text input. These spectrograms can be converted into audio clips.
Model Description:
This is a model that can be used to generate and modify images based on text prompts. It is a
Latent Diffusion Model
that uses a fixed, pretrained text encoder (
CLIP ViT-L/14
) as suggested in the
Imagen paper
.
Direct Use
The model is intended for research purposes only. Possible research areas and
tasks include
Generation of artworks, audio, and use in creative processes.
Applications in educational or creative tools.
Research on generative models.
Datasets
The original Stable Diffusion v1.5 was trained on the
LAION-5B
dataset using the
CLIP text encoder
, which provided an amazing starting point with an in-depth understanding of language, including musical concepts. The team at LAION also compiled a fantastic audio dataset from many general, speech, and music sources that we recommend at
LAION-AI/audio-dataset
.
Fine Tuning
Check out the
diffusers training examples
from Hugging Face. Fine tuning requires a dataset of spectrogram images of short audio clips, with associated text describing them. Note that the CLIP encoder is able to understand and connect many words even if they never appear in the dataset. It is also possible to use a
dreambooth
method to get custom styles.
Citation
If you build on this work, please cite it as follows:
@article{Forsgren_Martiros_2022,
author = {Forsgren, Seth* and Martiros, Hayk*},
title = {{Riffusion - Stable diffusion for real-time music generation}},
url = {https://riffusion.com/about},
year = {2022}
}
Runs of GitMylo riffusion-model-v1-small on huggingface.co
6
Total runs
0
24-hour runs
1
3-day runs
4
7-day runs
0
30-day runs
More Information About riffusion-model-v1-small huggingface.co Model
riffusion-model-v1-small huggingface.co is an AI model on huggingface.co that provides riffusion-model-v1-small's model effect (), which can be used instantly with this GitMylo riffusion-model-v1-small model. huggingface.co supports a free trial of the riffusion-model-v1-small model, and also provides paid use of the riffusion-model-v1-small. Support call riffusion-model-v1-small model through api, including Node.js, Python, http.
riffusion-model-v1-small huggingface.co is an online trial and call api platform, which integrates riffusion-model-v1-small's modeling effects, including api services, and provides a free online trial of riffusion-model-v1-small, you can try riffusion-model-v1-small online for free by clicking the link below.
GitMylo riffusion-model-v1-small online free url in huggingface.co:
riffusion-model-v1-small is an open source model from GitHub that offers a free installation service, and any user can find riffusion-model-v1-small on GitHub to install. At the same time, huggingface.co provides the effect of riffusion-model-v1-small install, users can directly use riffusion-model-v1-small installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
riffusion-model-v1-small install url in huggingface.co: