Qwopus3.6-27B-v2-heretic-GGUF

GGUF conversion of Qwopus3.6-27B-v2-heretic.

Files

  • Qwopus3.6-27B-v2-heretic-F16.gguf
  • Qwopus3.6-27B-v2-heretic-Q4_K_M.gguf

Recommended runtime

For llama.cpp or llama-server, use Q4_K_M for normal inference.

For vLLM, GGUF support is experimental. Prefer the HF safetensors model when possible. If using GGUF with vLLM, use the Q4_K_M file with the base tokenizer and base HF config.

Downloads last month
59
GGUF
Model size
27B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for snowman0919/Qwopus3.6-27B-v2-heretic-GGUF

Quantized
(59)
this model