Text Generation
Transformers
GGUF
PyTorch
nvidia
nemotron-3.5
imatrix
conversational

Nemotron: wrong number of tensors - expected 417, got 408

#5
by AIPeso - opened

Hi,
I am trying to load Nemotron-3.5 in LM Studio on Windows 11

NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF IQ4_NL (downloaded 2026.08.20) and it is not loaded - just an error:

Failed to load model.

error loading model: done_getting_tensors: wrong number of tensors; expected 417, got 408

Is it an issue with the model?

Thank you!

I am getting the same error. Is the issue resolved?

pbhogan1 - I believe you need to update LM Studio. I just had to build a new version of llama.cpp to run this model. My older build (9400s) gave the same error- it doesn't know what to do with the mtp layers.

I solved this problem by re-downloading the GGUF file again.

Sign up or log in to comment