Fredred89 commited on
Commit
592d782
·
verified ·
1 Parent(s): 83a32f9

Upload PREDATOR_4A_4BIT_CHAMPION_README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. PREDATOR_4A_4BIT_CHAMPION_README.md +18 -0
PREDATOR_4A_4BIT_CHAMPION_README.md ADDED
@@ -0,0 +1,18 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ # Qwen3.5-27B-GLM5.1-Distill-v1-4A-4BIT-CHAMPION
2
+
3
+ **New Predator champion (2026-06-05).** Beats 4A BEST by 0.0056 PPL (chunks=30, Z=22.0, p<1e-19, 30/30 chunks improved).
4
+
5
+ ## Result
6
+ - PPL chunks=30: 6.2342 (was 6.2398 for 4A BEST on original F16)
7
+ - File size: 16.98 GB (SAME as 4A BEST)
8
+ - Same tensor-type-file as 4A BEST
9
+ - Different weights (rotated by Chrysalis calibrated for 4-bit)
10
+
11
+ ## How
12
+ - Pre-rotated F16 with Chrysalis butterflies calibrated for 4-bit quantization loss (not 2-bit)
13
+ - Quantized with PREDATOR_4A_BEST.tensor_types.quant.txt
14
+ - 544 matrices × 50 SGD steps × 2.5 min/layer = 158.8 min on VSI 2x L40S
15
+
16
+ ## See also
17
+ - 4A BEST (the previous champion, beats Q4_K_M by 0.0369 PPL): Fredred89/Qwen3.5-27B-GLM5.1-Distill-v1-GGUF
18
+ - Methodology: 4-bit Chrysalis calibration finds gentler rotations than 2-bit; at 4-bit they HELP (small but real)