jackasda211233 commited on
Commit
e8cc7cd
·
verified ·
1 Parent(s): a00669c

Add prominent vision banner with clickable mmproj download links

Browse files
Files changed (1) hide show
  1. README.md +17 -6
README.md CHANGED
@@ -23,6 +23,10 @@ base_model:
23
 
24
  # Qwen3.6-27B AEON RYS MaxThinkCoder IQ4_NL GGUF for ik-llama
25
 
 
 
 
 
26
  > Hard runtime requirement: use the custom AEON ik-llama fork:
27
  >
28
  > https://github.com/noonr48/qwen36-aeon-ik-llama
@@ -32,24 +36,31 @@ base_model:
32
  > **Newer fine-tunes built on this base — [SignalLatch](https://huggingface.co/jackasda211233/Qwen3.6-27B-AEON-RYS-SignalLatch-GGUF) (a behaviour fine-tune) and [PatchCode](https://huggingface.co/jackasda211233/Qwen3.6-27B-AEON-RYS-Agentic-Coder-PatchCode-GGUF) (an agentic-coder distil on top of SignalLatch).** This repo is the non-finetuned base.
33
 
34
 
 
35
  ## Vision Support (mmproj)
36
 
37
- These models support image input. Qwen3.6-27B is natively a vision-language model the included **mmproj** (multimodal projector) files enable image understanding when used with the `--mmproj` flag in llama.cpp / ik-llama.
38
 
39
  The projector is extracted from the official [Qwen/Qwen3.6-27B](https://huggingface.co/Qwen/Qwen3.6-27B) base model. Since text fine-tuning does not modify the vision encoder, one projector works across all three RYS variants (base, SignalLatch, PatchCode).
40
 
41
- | File | Precision | Size | Use |
 
 
42
  |---|---|---:|---|
43
- | `mmproj-Qwen3.6-27B-base-f32.gguf` | F32 (full precision) | 1.8 GB | Maximum accuracy, most VRAM |
44
- | `mmproj-Qwen3.6-27B-base-f16.gguf` | F16 (half precision) | 885 MB | **Recommended** best balance |
45
- | `mmproj-Qwen3.6-27B-base-q8_0.gguf` | Q8_0 (8-bit quantized) | 601 MB | Smallest, minimal quality loss |
 
 
46
 
47
  ### Usage
48
 
49
  Add `--mmproj` to your llama-server command:
50
 
51
  ```bash
52
- ./build/bin/llama-server -m Qwen3.6-27B-AEON-RYS-Agentic-Coder-PatchCode.IQ4_NL.gguf --mmproj mmproj-Qwen3.6-27B-base-f16.gguf --jinja -ngl 999 -c 200000
 
 
53
  ```
54
 
55
  Then send images via the standard OpenAI-compatible API:
 
23
 
24
  # Qwen3.6-27B AEON RYS MaxThinkCoder IQ4_NL GGUF for ik-llama
25
 
26
+
27
+ > **👁️ Vision Support Added** — This model now supports image input! Download a [mmproj projector file](#vision-support-mmproj) from the file list and add `--mmproj` to enable vision. See the [Vision Support section](#vision-support-mmproj) below for details.
28
+
29
+
30
  > Hard runtime requirement: use the custom AEON ik-llama fork:
31
  >
32
  > https://github.com/noonr48/qwen36-aeon-ik-llama
 
36
  > **Newer fine-tunes built on this base — [SignalLatch](https://huggingface.co/jackasda211233/Qwen3.6-27B-AEON-RYS-SignalLatch-GGUF) (a behaviour fine-tune) and [PatchCode](https://huggingface.co/jackasda211233/Qwen3.6-27B-AEON-RYS-Agentic-Coder-PatchCode-GGUF) (an agentic-coder distil on top of SignalLatch).** This repo is the non-finetuned base.
37
 
38
 
39
+
40
  ## Vision Support (mmproj)
41
 
42
+ > **This model supports vision/image input.** Qwen3.6-27B is natively a vision-language model. Download one of the mmproj (multimodal projector) files below and pass it with `--mmproj` to enable image understanding.
43
 
44
  The projector is extracted from the official [Qwen/Qwen3.6-27B](https://huggingface.co/Qwen/Qwen3.6-27B) base model. Since text fine-tuning does not modify the vision encoder, one projector works across all three RYS variants (base, SignalLatch, PatchCode).
45
 
46
+ ### Download a projector
47
+
48
+ | File | Precision | Size | Link |
49
  |---|---|---:|---|
50
+ | `mmproj-Qwen3.6-27B-base-f32.gguf` | F32 (full precision) | 1.8 GB | [⬇ Download](https://huggingface.co/jackasda211233/Qwen3.6-27B-AEON-RYS-15-20-GGUF/resolve/main/mmproj-Qwen3.6-27B-base-f32.gguf) |
51
+ | `mmproj-Qwen3.6-27B-base-f16.gguf` | F16 (half precision) | 885 MB | [⬇ Download](https://huggingface.co/jackasda211233/Qwen3.6-27B-AEON-RYS-15-20-GGUF/resolve/main/mmproj-Qwen3.6-27B-base-f16.gguf) |
52
+ | `mmproj-Qwen3.6-27B-base-q8_0.gguf` | Q8_0 (8-bit quantized) | 601 MB | [⬇ Download](https://huggingface.co/jackasda211233/Qwen3.6-27B-AEON-RYS-15-20-GGUF/resolve/main/mmproj-Qwen3.6-27B-base-q8_0.gguf) |
53
+
54
+ **Recommended:** `mmproj-Qwen3.6-27B-base-f16.gguf` — best balance of quality and size.
55
 
56
  ### Usage
57
 
58
  Add `--mmproj` to your llama-server command:
59
 
60
  ```bash
61
+ ./build/bin/llama-server -m Qwen3.6-27B-AEON-RYS-Agentic-Coder-PatchCode.IQ4_NL.gguf \
62
+ --mmproj mmproj-Qwen3.6-27B-base-f16.gguf \
63
+ --jinja -ngl 999 -c 200000
64
  ```
65
 
66
  Then send images via the standard OpenAI-compatible API: