Text Generation
Transformers
GGUF
PyTorch
nvidia
nemotron-3.5
imatrix
conversational

[Quant Request] Fastino-Nemotron-3.5-Lightning-Finance (Fastino + NVIDIA)

#8
by hemono - opened

Hi Unsloth team!

NVIDIA and Fastino Labs just officially released fastino/Fastino-Nemotron-3.5-Lightning-Finance, a specialized financial MoE fine-tune built on top of Nemotron 3.5 Lightning (128K context, LoRA r32, achieving +50pp on FinQA and 74.8% on TAT-QA).

Since you already have the full GGUF conversion pipeline and custom llama.cpp kernels ready for Nemotron 3.5 Lightning 30B-A3B, could you please release official Unsloth Dynamic GGUF quants (UD-Q4_K_XL / Q4_K_M) for this finance model?

It would be super helpful for the entire open-source finance & quant community!

Model Link: https://huggingface.co/fastino/Fastino-Nemotron-3.5-Lightning-Finance

Thank you as always for your unmatched tooling and support!

Sign up or log in to comment