PyTorch
Diffusers
audio
music
minimax-music-3
rvq
reference-audio
comfyui

How to help?

#1
by johndpope - opened

Are you looking for smarter losses?
More training data? Maybe h3 minimax can regurgitate audio for training?
Could we overfit to a given song?

SimpleTuner org

yeah we kinda hit a ceiling with synthetic data and real audio requires finetuning the LM, if you experiment with getting vocals into the model, we have solved style for the most part

i push some code for v5 - https://huggingface.co/johndpope/open-rvq-encoder-minimax-music3/tree/main

it has wandb - i highly recommend -

Project: https://wandb.ai/snoozie/open-rvq-minimax-music3-v5
Cover/listen run: https://wandb.ai/snoozie/open-rvq-minimax-music3-v5/runs/f5phji1d

UPDATE -
wandb no longer recommended - they updated to a 60/mth paid model.

Sign up or log in to comment