Quant type

#1
by Theory-of-mind - opened

Hello. Could you please tell me if these are regular static quants (without iMatrix or other optimizations)?

Hello. These quantisations were created from the same BF16 conversion to GGUF. We used llama-cpp’s ‘convert_hf’ Python script for the original GGUF conversion, then their llama-quantize executable to create the quantisations. No additional changes were made to the base model, and no special conversion/quantisation arguments were used.

Thank you - that's exactly what I needed.

Sign up or log in to comment