Qwen3.6_9-12B_A2B_MoE

#47
by Konti201203 - opened

Given the success of the 35B-A3B MoE architecture, are there plans for a smaller MoE version in the 9B-12B range for the Qwen 3.6 series? A model with ~2B active parameters would be a game-changer for local edge deployment and real-time applications.

35B 对 GPU 显存压力较大,如果有 20B A3B 也是不错的

Sign up or log in to comment