DeepSeek V4 @ IQ3XXS on M1 Ultra 128GB- 6 tok/s -> 16 tok/s in LM Studio after patch

#24
by jeevacation2 - opened

https://github.com/noreff/lmstudio-dsv4-patch
M1 Ultra 128GB, Unsloth UD-IQ3_XXS, wired limit at 120GB. I was at 5-6 tok/s before the patch. Getting 15-16 tok/s now with the patched engine, and the output seems to have improved. Big thanks to this guy.

jeevacation2 changed discussion status to closed
jeevacation2 changed discussion status to open

Sign up or log in to comment