finalize pinned NVFP4A16 checkpoint metadata
Browse files
README.md
CHANGED
|
@@ -33,7 +33,7 @@ Verified on shard 49, expert 0, real K3 w1/w2/w3 weights: all packed bytes are i
|
|
| 33 |
- Lossless over-span policy: `fail; never silently requantize`
|
| 34 |
- Tensor data size: 1.497 TiB
|
| 35 |
- Size change versus source tensor data: +5.45%
|
| 36 |
-
-
|
| 37 |
|
| 38 |
This is **NVFP4A16**, not calibrated W4A4 NVFP4. Kimi K3 was trained
|
| 39 |
with MXFP8 activations; FP4 activation quantization would require a separate
|
|
|
|
| 33 |
- Lossless over-span policy: `fail; never silently requantize`
|
| 34 |
- Tensor data size: 1.497 TiB
|
| 35 |
- Size change versus source tensor data: +5.45%
|
| 36 |
+
- Companion artifact: [`GrEarl/Kimi-K3-NVFP4A16-Requantized`](https://huggingface.co/GrEarl/Kimi-K3-NVFP4A16-Requantized)
|
| 37 |
|
| 38 |
This is **NVFP4A16**, not calibrated W4A4 NVFP4. Kimi K3 was trained
|
| 39 |
with MXFP8 activations; FP4 activation quantization would require a separate
|