GLM-5.3-Flash
GGUF quantization(s) of zai-org/GLM-5.3-Flash-BF16. Made with timkhronos' llama.cpp#27773 - requires that PR until support is merged.
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support
Model tree for ddh0/GLM-5.3-Flash-GGUF
Base model
zai-org/GLM-5.3-Flash