Qwen3.8-27B
Collection
8 items • Updated
How to use airagrp/Qwen3.8-27B-MTP-bf16 with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir Qwen3.8-27B-MTP-bf16 airagrp/Qwen3.8-27B-MTP-bf16
This is MTP head only full model at airagrp/Qwen3.8-27B-bf16
Get around 20tps on M5 max
Quantized
Base model
Qwen/Qwen3.8-27B