exllamav3 quantizations of TheDrummer's Artemis-31B-v1.2.

Quantized using 12414d0 of the dev branch. Requires exllamav3-v1.5.1 or above.

2.25bpw_h6 11.761 GiB
4.00bpw_h6 17.730 GiB (this model)
6.00bpw_h8 24.877 GiB

Original model card follows below.


Model Card WIP

Usage

config-v1q

Downloads last month
56
Safetensors
Model size
10B params
Tensor type
BF16
·
F16
·
I16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for MikeRoz/Artemis-v1.2-4.00bpw-h6-exl3

Quantized
(20)
this model