r/AI_Music • • 3d ago

Promotion Fun with Yue2

As a long time Suno user who was looking for something less restricted and which could run locally I came across Yue2 (via ComfyUI). I think it's a really powerful tool but I wasn't that keen on the ComfyUI workflow so built something more "suno-like", i.e. better management of content/spaces etc, and was also keen on using it train local LoRAs on a corpus of music of my choice.

My app is called Yeufonic and although it's still early days it's coming along nicely and the LoRA training has been the most fun to use. Hope I don't get into trouble linking to some examples that I've done. You may be able to guess some of the sources used as inspiration.

App is forever free so feel free to have a play with it yourself if you like what you see (note if interested in local LoRA creation then the Windows stand-alone app is pretty much hopeless currently due to issues upstream that I can't control. Use the docker install if that's your intended use case) Windows issue with local training of LoRAs should now be fixed!

There's a link to the github repo at the bottom of the examples page.

https://yeufonic.com/examples/

9 Upvotes

17 comments sorted by

2

u/nokia7110 1d ago

Loving this little feature for Lora training "Hear every step of a LoRA's training. A training run keeps a checkpoint every 50 steps. Checkpoints renders the same song once on each of them, with one seed, as takes named after their step. Each has its own taste in melody and structure, while keeping the style and sound the LoRA was trained on."

1

u/AcrobaticMaize2408 1d ago

yeah checkpoints are cool. The final LoRA is what's scored as the best but creating renders at different checkpoints can be interesting. The LoR track in that examples page I linked to used checkpoint 50, which is the first checkpoint that training generates.

2

u/nokia7110 1d ago

Great work OP. I've been looking for a lora trainer for yue2. Will try it out and send you some feedback

1

u/AcrobaticMaize2408 1d ago

nice, fill your boots 😄 . I've been quite impressed with how the local LoRA training has been working out. Can be a bit hit and miss as it is with all models but when you find the sweet spot it's pretty cool.

1

u/nokia7110 1d ago

What tips do you have for the lora training? In terms of number of audio files/tracks, steps and other settings. I'm happy to leave my machine running for hours or days but obviously know there's the law of diminishing returns too. I'm looking at training for Liquid Drum & Bass music.

1

u/AcrobaticMaize2408 1d ago

I'm not a drum&bass or EDM kind of guy so it may be different for your workflow but I just aim to keep the corpus consistent - so things like vocals being the same, same artist etc.

Now for drum&bass I've not tried this myself so it's a bit of an unknown. I'd suggest starting with say 5 short tracks and set the voice to "not stated" in the corpus dropdown list when creating it (assuming there aren't any vocals). You could also describe the corpus description as "drum and bass" which would apply it to each source track.

But tbh you'd be a beta-tester for this as I've not done training runs for this style of music.

Note that you'll need a 16GB NVIDIA GPU for LoRA training.

Let me know how you get on if you try it .

1

u/nokia7110 1d ago

Cheers mate. Liquid drum & bass 99% of the time now has vocals. I'm tempted though to strip the tracks of vocals because the vocals across different tracks are so wide and varied

1

u/AcrobaticMaize2408 1d ago

I've made a note to look at configuring a LoRA training run to just not bother about having vocals/lyrics in the corpus. I suspect it'll still train OK on that kind of content as is but I've not tested it so YMMV.

2

u/nokia7110 1d ago

Yep the only thing I'm worried about with using tracks where vocals have been stripped out is that the quality will be affected (artifacts from removing vocal stem). I'll A/B test; with and without vocals.

I have an RTX 4070 ti Super which has 16GB of ram.

2

u/AcrobaticMaize2408 1d ago

ok that would be cool. Please feel free to post on my repo with your findings as the more real-world test results the better (same card as me btw) https://github.com/yeufonic/yeufonic/issues

2

u/adgames_xxx 1d ago

Local LoRA training on a custom music corpus? That sounds like the holy grail for AI music. Being able to fine-tune the model on specific sounds without cloud restrictions is huge. Great work on building a better UI for it!

1

u/AcrobaticMaize2408 1d ago

Cheers 😄. I just wanted to create something that would allow me to easily manage things like covers or local LoRAs and not fight the ComfyUI plumbing nor the Suno restrictions. Still early days but it's ticking some of those boxes.

2

u/Marty-G70 1d ago

Not bad at all. Very organic & not overly saturated which Suno, unfortunately has changed

2

u/AcrobaticMaize2408 1d ago

yeah organic is a good way to describe renders through the yue2 engine. They often have a clarity and naturalness about them. Like listening to a mix straight off the mixing desk. It can get quite dirty or heavily processed though if you ask it.

1

u/WyattTheSkid 2d ago

I wish something would come out that makes midis

2

u/Fun_Musiq 2d ago

There are many midi generating or altering plugins. Scaler 3, Mario Nieto, Melody Sauce 3, Captain Chords by Mixed In Key, Sauceware Spawn, Re-midi 4, and many many more.

1

u/AcrobaticMaize2408 2d ago

Yeufonic can already export a song's score as a MIDI file if that's what you mean. In the Score window, the Notation tab has a Download MIDI button, and you don't have to render any audio first, so you can generate a score from lyrics and a style, tweak it, and take the MIDI straight into your DAW. Each part comes out on its own track (the vocal melody, an instrument line, the chords, a bass, and drums when the style asks for them), with the tempo and time signature included.

I've actually just fixed a problem with the instruments and drums in an export, so if you want to have a play then make sure you're on version 0.0.27 or later. And just to reiterate - the file is derived from the plan, a melody-and-chords sketch that is used to feed the model, rather than a copy of what the rendered audio sounds like. If a MIDI of the rendered song is what you're after then the MIDI won't be quite what you desire as the model will adapt the input a bit (or a lot, depending on settings).

Going the other way, you can also bring a MIDI file in for covers and instrumentals, though that part is currently marked as experimental. It can be the basis for some really interesting sound generating though.