pure-reasoning-7b-230726-GGUF

GGUF build of breitburg/pure-reasoning-7b-230726, a thinking model that emits a <think>...</think> block then an answer. Converted 23 July 2026 with Unsloth. The ChatML chat template is embedded in the GGUF, so pass --jinja.

Example usage:

  • llama-cli -hf breitburg/pure-reasoning-7b-230726-GGUF:Q8_0 --jinja

Available files

  • pure-reasoning-7b-230726.Q8_0.gguf โ€” ~7.2 GB, higher quality
  • pure-reasoning-7b-230726.Q4_K_M.gguf โ€” ~4 GB, smaller/faster

Trained 2x faster with Unsloth.

Downloads last month
197
GGUF
Model size
7B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

4-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Collection including breitburg/pure-reasoning-7b-230726-GGUF