ai-pc-benchmarks-transcription-intel-openvino

OpenVINO IR bundle for nvidia/parakeet-tdt-0.6b-v3, used by the AI PC Benchmarks Intel transcription workload.

Contents

  • model/config.json β€” model type, feature size, subsampling.
  • model/manifest.json β€” build metadata, reference performance numbers.
  • model/vocab.txt β€” SentencePiece vocab, 8193 tokens, <blk> at id 8192.
  • model/ir/preprocessor_fused_n240000.{xml,bin} β€” fixed-chunk preprocessor (15 s @ 16 kHz).
  • model/ir/encoder_t1501.{xml,bin} β€” W8A8-quantized Conformer encoder.
  • model/ir/decoder_joint_e1_t1.{xml,bin} β€” W8A8-quantized TDT decoder + joint.

Runtime

OpenVINO >= 2026.1. The IRs compile to Intel NPU, GPU, or CPU on first use and are cached via CACHE_DIR. No prebuilt NPU blobs ship here β€” the benchmark compiles them on-device the first time a given device is used.

License

CC-BY-4.0, inherited from the base model.

Downloads last month
4
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for mweinbach1/ai-pc-benchmarks-transcription-intel-openvino

Finetuned
(71)
this model