exllamav3 quantizations of TheDrummer's Artemis-31B-v1.2.

Quantized using 12414d0 of the dev branch. Requires exllamav3-v1.5.1 or above.

2.25bpw_h6 11.761 GiB
4.00bpw_h6 17.730 GiB
6.00bpw_h8 24.877 GiB (this model)

Original model card follows below.


Model Card WIP

Usage

config-v1q

Downloads last month
21
Safetensors
Model size
13B params
Tensor type
BF16
·
F16
·
I16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for MikeRoz/Artemis-v1.2-6.00bpw-h8-exl3

Quantized
(20)
this model