--- base_model: unsloth/gemma-4-12B-it-GGUF tags: - gguf - wllama --- # gemma-4-12b-it-Q4_K_M — split GGUF for wllama `gemma-4-12b-it-Q4_K_M.gguf` from [unsloth/gemma-4-12B-it-GGUF](https://huggingface.co/unsloth/gemma-4-12B-it-GGUF) (credit to the original model authors and quantizer), split into <2 GB shards with `llama-gguf-split` so it can be loaded by [wllama](https://github.com/ngxson/wllama) in the browser. Load the first shard; wllama resolves the rest automatically.