GrEarl commited on
Commit
8c22f32
·
verified ·
1 Parent(s): 3957e75

finalize pinned NVFP4A16 checkpoint metadata

Browse files
Files changed (1) hide show
  1. README.md +1 -1
README.md CHANGED
@@ -33,7 +33,7 @@ Verified on shard 49, expert 0, real K3 w1/w2/w3 weights: all packed bytes are i
33
  - Lossless over-span policy: `fail; never silently requantize`
34
  - Tensor data size: 1.497 TiB
35
  - Size change versus source tensor data: +5.45%
36
- - Planned companion artifact (not yet published): `GrEarl/Kimi-K3-NVFP4A16-Requantized`
37
 
38
  This is **NVFP4A16**, not calibrated W4A4 NVFP4. Kimi K3 was trained
39
  with MXFP8 activations; FP4 activation quantization would require a separate
 
33
  - Lossless over-span policy: `fail; never silently requantize`
34
  - Tensor data size: 1.497 TiB
35
  - Size change versus source tensor data: +5.45%
36
+ - Companion artifact: [`GrEarl/Kimi-K3-NVFP4A16-Requantized`](https://huggingface.co/GrEarl/Kimi-K3-NVFP4A16-Requantized)
37
 
38
  This is **NVFP4A16**, not calibrated W4A4 NVFP4. Kimi K3 was trained
39
  with MXFP8 activations; FP4 activation quantization would require a separate