bottlecapai/ThinkingCap-Qwen3.6-27B-GGUF Image-Text-to-Text โข 27B โข Updated 11 days ago โข 41.1k โข 261
AlexCaswen/Nemotron-3-Super-120B-A12B-MTP-GGUF Text Generation โข 4B โข Updated 15 days ago โข 263 โข 1
deepseek-ai/DeepSeek-V4-Flash-0731 Text Generation โข 304B โข Updated 24 days ago โข 3.27M โข โข 3.69k
Running Agents 90 Accurate GGUF Memory Calculator ๐ 90 Calculate memory for GGUF models using GPU layers + context
view post Post 2991 Good news, llama.cpp seems to be close to supporting MTP on qwen models. Bad news, every single gguf will have to be redone when it is. See translation 1 reply ยท ๐ 15 15 + Reply