G4-Moonlight-Dusk-26B-A4B-APEX-GGUF

APEX (Adaptive Precision for EXpert Models) quantizations of Vortex5/G4-Moonlight-Dusk-26B-A4B.

I used the script from LocalAI team | APEX Project | Technical Report

I also used the latest chat_template.jinja(2026-07-22 , if google ever update the template again , override the template with the latest) from Google , which fixes a lot of issue , for RP specifically thinking will not longer leak out .

imatrix data is from mradermacher

Downloads last month
774
GGUF
Model size
26B params
Architecture
gemma4
Hardware compatibility
Log In to add your hardware

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for debais/G4-Moonlight-Dusk-26B-A4B-APEX-GGUF

Quantized
(3)
this model