steampunque commited on
Commit
6808d13
·
verified ·
1 Parent(s): 9021f11

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +3 -3
README.md CHANGED
@@ -10,7 +10,7 @@ tags:
10
  - 6-bit
11
  ---
12
 
13
- ## GGUF hybrid layer quantization of Voxtral-Mini-3B-2512 by mistralai
14
 
15
  Original model: https://huggingface.co/mistralai/Voxtral-Mini-3B-2507
16
 
@@ -98,8 +98,8 @@ A full set of audio benchmarks for the model is given here: https://huggingface.
98
  ## Download the file from below:
99
  | Link | Type | Size/e9 B | Notes |
100
  |------|------|-----------|-------|
101
- | [Voxtral-Mini-3B-2507.Q6_K_H.gguf](https://huggingface.co/steampunque/Voxtral-Mini-3B-2507-Hybrid-GGUF/resolve/main/Voxtral-Mini-3B-2507.Q6_K_H.gguf) | Q6_K_H | 3.2e9 B | ~Q6_K size |
102
- | [Voxtral-Mini-3B-2507.mmproj.gguf](https://huggingface.co/steampunque/Voxtral-Mini-3B-2507-Hybrid-GGUF/resolve/main/Voxtral-Mini-3B-2517.mmproj.gguf) | F16 | 1.3e9 B | multimedia projector |
103
 
104
  A discussion thread about the hybrid layer quant approach can be found here on the llama.cpp git repository:
105
 
 
10
  - 6-bit
11
  ---
12
 
13
+ ## Mixed Precision GGUF layer quantization of Voxtral-Mini-3B-2512 by mistralai
14
 
15
  Original model: https://huggingface.co/mistralai/Voxtral-Mini-3B-2507
16
 
 
98
  ## Download the file from below:
99
  | Link | Type | Size/e9 B | Notes |
100
  |------|------|-----------|-------|
101
+ | [Voxtral-Mini-3B-2507.Q6_K_H.gguf](https://huggingface.co/steampunque/Voxtral-Mini-3B-2507-MP-GGUF/resolve/main/Voxtral-Mini-3B-2507.Q6_K_H.gguf) | Q6_K_H | 3.2e9 B | ~Q6_K size |
102
+ | [Voxtral-Mini-3B-2507.mmproj.gguf](https://huggingface.co/steampunque/Voxtral-Mini-3B-2507-MP-GGUF/resolve/main/Voxtral-Mini-3B-2517.mmproj.gguf) | F16 | 1.3e9 B | multimedia projector |
103
 
104
  A discussion thread about the hybrid layer quant approach can be found here on the llama.cpp git repository:
105