Qwen3.8-27B-NQ-SVD-IQ2

โš ๏ธ EXPERIMENTAL REPOSITORY โ€” the model is NOT ready yet and is under active development.

Quantization of Qwen3.8-27B using SVD low-rank factorization with quantized factors, produced with NeuralQuant (NQ).

  • NQ โ€” NeuralQuant (our quantizer engine)
  • SVD โ€” low-rank factorization of Linear weights: W โ‰ˆ AยทB + E (all factors are quantized: A โ†’ IQ2/IQ3, B โ†’ IQ2, residual E โ†’ IQ1/ternary)
  • IQ2 โ€” 2-bit IQ codebooks for the low-rank factors

Current status: the SVD support in the NeuralQuant library is being implemented. Model weights will be published once the model is validated.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for agiws/Qwen3.8-27B-NQ-SVD-IQ2

Base model

Qwen/Qwen3.8-27B
Finetuned
(263)
this model