Image-Text-to-Text
GGUF
English
Chinese
unsloth
fine tune
heretic
uncensored
abliterated
ara
MTP GGUF Quants
Regular GGUF Quants
qwen3_8
qwen3_6
qwen3_5
multi-stage tuned
thinking
reasoning
all use cases
coder
creative
creative writing
all genres
story
writing
fiction
roleplaying
bfloat16
multi-stage-tune
multi-state-merge
Instructions to use DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-ULTRA-HERETIC-Uncensored-NM-DAU-NEO-MTP-GGUF with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Local Apps Settings
- Unsloth Desktop
Update README.md
Browse files
README.md
CHANGED
|
@@ -41,16 +41,16 @@ base_model:
|
|
| 41 |
---
|
| 42 |
|
| 43 |
<small><b><font color="red">Important:</font></b> Tuned, and tweaked to match the legendary Qwen 3.6 27B FF711 (2300+ likes, 4 million+ downloads)
|
| 44 |
-
this fine tune matches the stability and power at "arc-c"
|
| 45 |
-
in 8 bit and
|
| 46 |
yet maintains output detail and quality. In other words while "reg" Qwen3.8 27B is thinking about "formatting" for a few 1000 tokens, this model is already done and waiting for more.
|
| 47 |
-
This repo contains both "regular" and "MTP" Neo and NEO MAX GGUF quants.
|
| 48 |
|
| 49 |
<B>STRONGER/CTRL:</B> Now with 5 reasoning modes (2 new - Spoon / Einstein), and 5 instruct modes (2 new - Spoon / Einstein, all use ZERO REASONING TOKENS)
|
| 50 |
all switchable on the fly via API, direct and "in chat" (yes - model control at the chat/message level).
|
| 51 |
</small>
|
| 52 |
|
| 53 |
-
<h2>Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-
|
| 54 |
|
| 55 |
<img src="launch-arm.gif" style="float:right; padding:10px;">
|
| 56 |
|
|
@@ -65,16 +65,22 @@ created using the COLD FUSION AND FABLE FUSION 711 methods of training.
|
|
| 65 |
This is a high detail focused model, with tuning specific to address over reasoning/over thinking and excessive token consumption
|
| 66 |
THEN to take the model to the next level.
|
| 67 |
|
| 68 |
-
This model is also LIGHTLY uncensored (see section below), a much more uncensored version will be released shortly.
|
| 69 |
-
|
| 70 |
This model (both 4 bit and 8 bit) exceeds the base Qwen 3.8 27B in ALL critical 7 benchmarks AND exceeds all 7 benchmarks for Qwen3.6-35B-A3B, Qwen 3.6 27B, and Qwen 3.5 27B.
|
| 71 |
|
| 72 |
The 700s plus "intelligence club" is reserved for OpenAI, Claude and Gemini closed source models.
|
| 73 |
|
| 74 |
Considering that "just" 4 bit (1/4 full precision) is already at Arc-C of 701... a few people are going to have nightmares for a while.
|
| 75 |
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 76 |
Quick sample; snippet ("Why choose me to help your creative writing?"), Q4KS , "spoon" reasoning mode, non imatrix, (4 bit; 1/4 full precision):
|
| 77 |
|
|
|
|
|
|
|
| 78 |
(will be adding full, long examples shortly at the bottom of the page // there are also user examples in the "community" section too.)
|
| 79 |
|
| 80 |
<small>
|
|
|
|
| 41 |
---
|
| 42 |
|
| 43 |
<small><b><font color="red">Important:</font></b> Tuned, and tweaked to match the legendary Qwen 3.6 27B FF711 (2300+ likes, 4 million+ downloads)
|
| 44 |
+
this fine tune matches the stability and power at "arc-c" 699: (108 pts higher than Qwen 3.8 27B) (The OpenAI, Claude and Gemini "zone of intelligence")
|
| 45 |
+
in 8 bit and 692 arc-c in 4 bit. This version is called TWIN-TURBO because it drastically reduces thinking tokens (by 1/2 to as LOW as 1/20),
|
| 46 |
yet maintains output detail and quality. In other words while "reg" Qwen3.8 27B is thinking about "formatting" for a few 1000 tokens, this model is already done and waiting for more.
|
| 47 |
+
This repo contains both "regular" and "MTP" Neo and NEO MAX GGUF quants; and this is the highly uncensored version.
|
| 48 |
|
| 49 |
<B>STRONGER/CTRL:</B> Now with 5 reasoning modes (2 new - Spoon / Einstein), and 5 instruct modes (2 new - Spoon / Einstein, all use ZERO REASONING TOKENS)
|
| 50 |
all switchable on the fly via API, direct and "in chat" (yes - model control at the chat/message level).
|
| 51 |
</small>
|
| 52 |
|
| 53 |
+
<h2>Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-ULTRA-HERETIC-Uncensored-NM-DAU-NEO-MTP-GGUF</h2>
|
| 54 |
|
| 55 |
<img src="launch-arm.gif" style="float:right; padding:10px;">
|
| 56 |
|
|
|
|
| 65 |
This is a high detail focused model, with tuning specific to address over reasoning/over thinking and excessive token consumption
|
| 66 |
THEN to take the model to the next level.
|
| 67 |
|
|
|
|
|
|
|
| 68 |
This model (both 4 bit and 8 bit) exceeds the base Qwen 3.8 27B in ALL critical 7 benchmarks AND exceeds all 7 benchmarks for Qwen3.6-35B-A3B, Qwen 3.6 27B, and Qwen 3.5 27B.
|
| 69 |
|
| 70 |
The 700s plus "intelligence club" is reserved for OpenAI, Claude and Gemini closed source models.
|
| 71 |
|
| 72 |
Considering that "just" 4 bit (1/4 full precision) is already at Arc-C of 701... a few people are going to have nightmares for a while.
|
| 73 |
|
| 74 |
+
This model is STRONGLY uncensored (see section below). If you need slightly higher performance, but less "uncensored" see this version:
|
| 75 |
+
|
| 76 |
+
https://huggingface.co/DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-NM-DAU-NEO-MTP-GGUF
|
| 77 |
+
|
| 78 |
+
---
|
| 79 |
+
|
| 80 |
Quick sample; snippet ("Why choose me to help your creative writing?"), Q4KS , "spoon" reasoning mode, non imatrix, (4 bit; 1/4 full precision):
|
| 81 |
|
| 82 |
+
---
|
| 83 |
+
|
| 84 |
(will be adding full, long examples shortly at the bottom of the page // there are also user examples in the "community" section too.)
|
| 85 |
|
| 86 |
<small>
|