@haffner Thanks.

#1
by vectionlabs - opened

@haffner You are the first person to create a heretic version of our models. Thank you! What did you find in our models that pushed you to make heretic variants?

Owner

Answered in the other model's thread already. But I'm looking forward to trying this one - and maybe vl-1-coder to see if it can help improve my frontend work πŸ˜…
But currently I'm more into gamedev / system level C development and am thinking to train/fine-tune a model on a very specific dataset that I would have to prep first. Sadly I don't have any experience with this process yet.

We think that for the frontend you can get more satisfaction from Salience, thanks to the vision, and further training on VL-1-Coder data and a better dataset

Oh, I'm from Vectionlabs' team

@haffner Will you also be making the Heretic models from the Salience-1.5 family? Check our post if you want!

Owner

@haffner Will you also be making the Heretic models from the Salience-1.5 family? Check our post if you want!

I could give it a shot, but I think they will no longer fit in my vram at which point I can't because it would take forever. It needs to fit the entire safetensors fully into vram (maybe with bnb_4bit). But if you're interested in having these available, have a look here:
https://github.com/p-e-w/heretic

It's a surprisingly simple process. Previously I've run 5000 iterations, but usually you'll have a very decent result within the first 500 trials already. Also depends a bit on the underlying model architecture, of course.

Sign up or log in to comment