LynnReal Flash

Code: LynnReal-AI/LynnReal-Omni

A 42-block joint video/audio DiT with three denoiser evaluations. Supports text-to-video and native first/last keyframes. Middle blocks use spatial stride-two token selection with full-resolution residual restoration.

The complete transformer uses trained W8A8 projections: INT8 weights, per-token INT8 activations, INT32 accumulation, and BF16 outputs. Sensitive projections remain BF16. Packed weights and per-channel scales are stored without requantization in indexed Hugging Face safetensors shards.

Load using the release model.flash_int8.load_transformer or the launchers in script/sample/flash/int8. This custom packed format requires that loader; a generic Diffusers from_pretrained does not install the INT8 operators. Shared conditioning and codec components are included in this model bundle. See LICENSE and NOTICE for upstream terms.

Downloads last month
53
Safetensors
Model size
28B params
Tensor type
F32
BF16
I8
Inference Providers NEW
This model isn't deployed by any Inference Provider. 馃檵 Ask for provider support