Thank you so much, asking about model performance

#1
by MahouOfficial - opened

this the only GGUF unsonsred repo for now, great work.
can you give some spesifications about model performance "if possible"

MahouOfficial changed discussion title from Thank you so much, asking about model performance and the quantazation level. to Thank you so much, asking about model performance

I also wonder if DeepSeek-V4-Flash-Q4-mxfp4-0731.gguf is the full quality one/lossless for maximum performance? Or will a quant with less KL divergence be uploaded?

Edit:
I checked the previous version of DS4 abliterated and it was "Huihui-DeepSeek-V4-Flash-BF16-abliterated-ds4-Q4_K.gguf" around 165GB and the current one is "DeepSeek-V4-Flash-Q4-mxfp4-0731.gguf" which is 156 GB, that's 9GB less so, are we missing some precision?

In terms of storage usage, Q4 > MXFP4

God bless you huihui.

Sign up or log in to comment