voixful-nemotron-speech-streaming-en-0.6b

Apple Core AI (Palette4-quantized) conversion of nvidia/nemotron-speech-streaming-en-0.6b — a cache-aware streaming RNN-T for lowest-latency on-device English dictation on Apple Silicon. Used by the Voixful engine and PrivoVoice.

License & attribution

Weights derive from nvidia/nemotron-speech-streaming-en-0.6b, © NVIDIA, under the NVIDIA Open Model License, and are redistributed under those terms with attribution. The Voixful conversion pipeline is Apache-2.0.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for AustinJiangH/voixful-nemotron-speech-streaming-en-0.6b

Finetuned
(12)
this model