Would you update a new version with MTP? (It's finally supported in llama.cpp)

#1
by battataka - opened

Hi,

I like your custom quantization to provide better quality (I use IQ3_S on dual RTX 3090), would you update a new GGUF with MTP layers included?

Sign up or log in to comment