steampunque commited on
Commit
34d861b
·
verified ·
1 Parent(s): 33d2e39

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +3 -3
README.md CHANGED
@@ -10,7 +10,7 @@ tags:
10
  - 4-bit
11
  ---
12
 
13
- ## Llama.cpp hybrid layer quantization of Mistral-Small-3.2-24B-Instruct-2506 by mistralai
14
 
15
  Original model: https://huggingface.co/mistralai/Mistral-Small-3.2-24B-Instruct-2506
16
 
@@ -55,8 +55,8 @@ A full set of benchmarks for the model will eventually be given here: https://hu
55
  ## Download the file from below:
56
  | Link | Type | Size/e9 B | Notes |
57
  |------|------|-----------|-------|
58
- | [Mistral-Small-3.2-24B-Instruct-2506.Q4_K_H.gguf](https://huggingface.co/steampunque/Mistral-Small-3.2-24B-Instruct-2506-Hybrid-GGUF/resolve/main/Mistral-Small-3.2-24B-Instruct-2506.Q4_K_H.gguf) | Q4_K_H | 12.7e9 B | ~IQ4_XS quality/size |
59
- | [Mistral-Small-3.2-24B-Instruct-2506.mmproj.gguf](https://huggingface.co/steampunque/Mistral-Small-3.2-24B-Instruct-2506-Hybrid-GGUF/resolve/main/Mistral-Small-3.2-24B-Instruct-2506.mmproj.gguf) | mmproj | 0.88e9 B | multimedia projector |
60
 
61
  A discussion thread about the hybrid layer quant approach can be found here on the llama.cpp git repository:
62
 
 
10
  - 4-bit
11
  ---
12
 
13
+ ## Mixed Precision GGUF layer quantization of Mistral-Small-3.2-24B-Instruct-2506 by mistralai
14
 
15
  Original model: https://huggingface.co/mistralai/Mistral-Small-3.2-24B-Instruct-2506
16
 
 
55
  ## Download the file from below:
56
  | Link | Type | Size/e9 B | Notes |
57
  |------|------|-----------|-------|
58
+ | [Mistral-Small-3.2-24B-Instruct-2506.Q4_K_H.gguf](https://huggingface.co/steampunque/Mistral-Small-3.2-24B-Instruct-2506-MP-GGUF/resolve/main/Mistral-Small-3.2-24B-Instruct-2506.Q4_K_H.gguf) | Q4_K_H | 12.7e9 B | ~IQ4_XS quality/size |
59
+ | [Mistral-Small-3.2-24B-Instruct-2506.mmproj.gguf](https://huggingface.co/steampunque/Mistral-Small-3.2-24B-Instruct-2506-MP-GGUF/resolve/main/Mistral-Small-3.2-24B-Instruct-2506.mmproj.gguf) | mmproj | 0.88e9 B | multimedia projector |
60
 
61
  A discussion thread about the hybrid layer quant approach can be found here on the llama.cpp git repository:
62