QAT version

#1
by engrtipusultan - opened

Is there any chance to provide QAT quantization aware training model like gemma4 or LFM or gpt oss so that 4 bit quantization can have better quality.

And 14B or 16B for those 24GB RAM machines were 9GB is to little and 35B is too big :)

Sign up or log in to comment