Adrian Ionescu
adr-insc
AI & ML interests
None yet
Organizations
None yet
Official llama.cpp slower than the https://github.com/am17an/llama.cpp.git
1
#30 opened 2 months ago
by
adr-insc
Qwen 3.6 is much slower than Qwen 3.5 for the same quatization: UD-Q4_K_XL
10
#8 opened 3 months ago
by
adr-insc
Qwen 3.6 is much slower than Qwen 3.5 for the same quatization: UD-Q4_K_XL
10
#8 opened 3 months ago
by
adr-insc
Qwen 3.6 is much slower than Qwen 3.5 for the same quatization: UD-Q4_K_XL
10
#8 opened 3 months ago
by
adr-insc