--- language: - en library_name: mlx license: gemma pipeline_tag: image-text-to-text base_model: google/gemma-4-12B-it-qat-q4_0-unquantized tags: - mlx - mxfp4 - qat - quantized - apple-silicon - gemma4 - vision - audio ---
# OsaurusAI/gemma-4-12B-it-qat-MXFP4 MXFP4 MLX bundle converted from [google/gemma-4-12B-it-qat-q4_0-unquantized](https://huggingface.co/google/gemma-4-12B-it-qat-q4_0-unquantized). Decoder linears are quantized with MLX `mxfp4` at group size 32; embeddings, norms, and Gemma 4 early-fusion media embedders are preserved as fp16 passthrough. ## Bundle | Field | Value | |---|---| | Source | `google/gemma-4-12B-it-qat-q4_0-unquantized` | | Architecture | `gemma4_unified` / `Gemma4UnifiedForConditionalGeneration` | | Text layers | 48 total (8 full attention, 40 sliding attention) | | Hidden size | 3840 | | Quantization | `mxfp4`, bits=4, group_size=32 | | Quantized weights | 328 tensors with matching `.scales` sidecars | | Shards | 7 safetensors shards | | Indexed weight bytes | 7.37 GiB | | Processor | `processor_class=Gemma4UnifiedProcessor, image_seq_length=280, audio_seq_length=750, audio_ms_per_token=40, video_processor=present` | ## Modalities | Path | Status | |---|---| | Text | Preserved | | Vision | model_type=gemma4_unified_vision, patch_size=16 | | Audio | model_type=gemma4_unified_audio | | Video | no video_config | Audio encoder/config is present and preserved. No `video_config` is present in the source config; the processor file includes a video processor block, but this card does not claim a verified video runtime path. ## Tokenizer And Template | Field | Value | |---|---| | BOS token/id | `