Inkling 975B · Alis MLX Dynamic
Receipt-sealed layer-local ALIS-DWQ MLX quants of Inkling 975B multimodal MoE, built on 2x512GB M3 Ultra. Quality 3.7 / capacity 2.7 bpw / Q6 soon.
Image-Text-to-Text • 952B • Updated • 250k • • 1.76kNote Base model — 975B-class multimodal MoE (Apache-2.0), 66 hybrid layers, 256 routed + 2 shared experts, vision + audio towers.
avlp12/Inkling-975B-Alis-MLX-Dynamic-3.7bpw
Image-Text-to-Text • 947B • Updated • 122 • 2Note Quality / golden-spot tier: 3.71 bpw, 409 GiB — experts 3-bit, entire non-expert skeleton exact BF16. Certified DWQ pass, 1 accepted layer improvement. Single 512 GB Mac.
avlp12/Inkling-975B-Alis-MLX-Dynamic-6.6bpw
Image-Text-to-Text • Updated • 23Note Performance / maximum-fidelity tier: 6.60 bpw, 728 GiB two-box pipeline build — experts 6-bit alis-clip, everything else byte-exact BF16. Ships as stage0/stage1 artifacts for 2x512GB serving.
avlp12/Inkling-975B-Alis-MLX-Dynamic-2.7bpw
Image-Text-to-Text • 947B • Updated • 33Note Capacity / size-optimal tier: 2.72 bpw, 299 GiB — 2-bit expert floor with targeted 3/4/6/8-bit promotions. Certified DWQ pass, 7 accepted layer improvements. Fits 320 GB+ Macs.