DavidAU commited on
Commit
68eb8d2
·
verified ·
1 Parent(s): c703303

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +12 -6
README.md CHANGED
@@ -41,16 +41,16 @@ base_model:
41
  ---
42
 
43
  <small><b><font color="red">Important:</font></b> Tuned, and tweaked to match the legendary Qwen 3.6 27B FF711 (2300+ likes, 4 million+ downloads)
44
- this fine tune matches the stability and power at "arc-c" 709: (118 pts higher than Qwen 3.8 27B) (The OpenAI, Claude and Gemini "zone of intelligence")
45
- in 8 bit and 701 arc-c in 4 bit. This version is called TWIN-TURBO because it drastically reduces thinking tokens (by 1/2 to as LOW as 1/20),
46
  yet maintains output detail and quality. In other words while "reg" Qwen3.8 27B is thinking about "formatting" for a few 1000 tokens, this model is already done and waiting for more.
47
- This repo contains both "regular" and "MTP" Neo and NEO MAX GGUF quants.
48
 
49
  <B>STRONGER/CTRL:</B> Now with 5 reasoning modes (2 new - Spoon / Einstein), and 5 instruct modes (2 new - Spoon / Einstein, all use ZERO REASONING TOKENS)
50
  all switchable on the fly via API, direct and "in chat" (yes - model control at the chat/message level).
51
  </small>
52
 
53
- <h2>Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-NM-DAU-NEO-MTP-GGUF</h2>
54
 
55
  <img src="launch-arm.gif" style="float:right; padding:10px;">
56
 
@@ -65,16 +65,22 @@ created using the COLD FUSION AND FABLE FUSION 711 methods of training.
65
  This is a high detail focused model, with tuning specific to address over reasoning/over thinking and excessive token consumption
66
  THEN to take the model to the next level.
67
 
68
- This model is also LIGHTLY uncensored (see section below), a much more uncensored version will be released shortly.
69
-
70
  This model (both 4 bit and 8 bit) exceeds the base Qwen 3.8 27B in ALL critical 7 benchmarks AND exceeds all 7 benchmarks for Qwen3.6-35B-A3B, Qwen 3.6 27B, and Qwen 3.5 27B.
71
 
72
  The 700s plus "intelligence club" is reserved for OpenAI, Claude and Gemini closed source models.
73
 
74
  Considering that "just" 4 bit (1/4 full precision) is already at Arc-C of 701... a few people are going to have nightmares for a while.
75
 
 
 
 
 
 
 
76
  Quick sample; snippet ("Why choose me to help your creative writing?"), Q4KS , "spoon" reasoning mode, non imatrix, (4 bit; 1/4 full precision):
77
 
 
 
78
  (will be adding full, long examples shortly at the bottom of the page // there are also user examples in the "community" section too.)
79
 
80
  <small>
 
41
  ---
42
 
43
  <small><b><font color="red">Important:</font></b> Tuned, and tweaked to match the legendary Qwen 3.6 27B FF711 (2300+ likes, 4 million+ downloads)
44
+ this fine tune matches the stability and power at "arc-c" 699: (108 pts higher than Qwen 3.8 27B) (The OpenAI, Claude and Gemini "zone of intelligence")
45
+ in 8 bit and 692 arc-c in 4 bit. This version is called TWIN-TURBO because it drastically reduces thinking tokens (by 1/2 to as LOW as 1/20),
46
  yet maintains output detail and quality. In other words while "reg" Qwen3.8 27B is thinking about "formatting" for a few 1000 tokens, this model is already done and waiting for more.
47
+ This repo contains both "regular" and "MTP" Neo and NEO MAX GGUF quants; and this is the highly uncensored version.
48
 
49
  <B>STRONGER/CTRL:</B> Now with 5 reasoning modes (2 new - Spoon / Einstein), and 5 instruct modes (2 new - Spoon / Einstein, all use ZERO REASONING TOKENS)
50
  all switchable on the fly via API, direct and "in chat" (yes - model control at the chat/message level).
51
  </small>
52
 
53
+ <h2>Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-ULTRA-HERETIC-Uncensored-NM-DAU-NEO-MTP-GGUF</h2>
54
 
55
  <img src="launch-arm.gif" style="float:right; padding:10px;">
56
 
 
65
  This is a high detail focused model, with tuning specific to address over reasoning/over thinking and excessive token consumption
66
  THEN to take the model to the next level.
67
 
 
 
68
  This model (both 4 bit and 8 bit) exceeds the base Qwen 3.8 27B in ALL critical 7 benchmarks AND exceeds all 7 benchmarks for Qwen3.6-35B-A3B, Qwen 3.6 27B, and Qwen 3.5 27B.
69
 
70
  The 700s plus "intelligence club" is reserved for OpenAI, Claude and Gemini closed source models.
71
 
72
  Considering that "just" 4 bit (1/4 full precision) is already at Arc-C of 701... a few people are going to have nightmares for a while.
73
 
74
+ This model is STRONGLY uncensored (see section below). If you need slightly higher performance, but less "uncensored" see this version:
75
+
76
+ https://huggingface.co/DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-NM-DAU-NEO-MTP-GGUF
77
+
78
+ ---
79
+
80
  Quick sample; snippet ("Why choose me to help your creative writing?"), Q4KS , "spoon" reasoning mode, non imatrix, (4 bit; 1/4 full precision):
81
 
82
+ ---
83
+
84
  (will be adding full, long examples shortly at the bottom of the page // there are also user examples in the "community" section too.)
85
 
86
  <small>