jackasda211233 commited on
Commit
bac519f
·
verified ·
1 Parent(s): 50e547c

Clarify three-3090 benchmark as experimental

Browse files
Files changed (1) hide show
  1. README.md +1 -0
README.md CHANGED
@@ -25,6 +25,7 @@ quantized_by: jackasda211233
25
 
26
  - Specialized inference/runtime fork for this deployment: [noonr48/rys-splice-ik-llama](https://github.com/noonr48/rys-splice-ik-llama)
27
  - The recent fastpath benchmarking documented there was run on **three RTX 3090s**.
 
28
 
29
  An uncensored, coding-focused Qwen3.5-27B with RYS (Repeat Your Self) layer duplication, built via a novel **splice method** and quantized with a **custom reasoning-focused importance matrix**.
30
 
 
25
 
26
  - Specialized inference/runtime fork for this deployment: [noonr48/rys-splice-ik-llama](https://github.com/noonr48/rys-splice-ik-llama)
27
  - The recent fastpath benchmarking documented there was run on **three RTX 3090s**.
28
+ - **Experimental benchmark note:** that inference comparison is an experimental, session-specific result from one three-RTX-3090 deployment and should be treated as directional rather than a broad performance guarantee.
29
 
30
  An uncensored, coding-focused Qwen3.5-27B with RYS (Repeat Your Self) layer duplication, built via a novel **splice method** and quantized with a **custom reasoning-focused importance matrix**.
31