LongCat-Video-Avatar 1.5 — MLX
Apple MLX port of Meituan's audio-driven video diffusion. Source + recipe: github.com/xocialize/longcat-avatar-mlx
Text-to-Video • Updated • 6Note Recommended. DMD step-distillation LoRA pre-merged into the DiT. 8-step inference, smallest load step, simplest pipeline code. Use this unless you specifically need to toggle between 50-step base and 8-step distilled paths.
mlx-community/LongCat-Video-Avatar-1.5-bf16
Text-to-Video • UpdatedNote Alternate. Base bf16 weights + separate DMD LoRA file. Lets you switch between the 50-step base inference path and the 8-step DMD distilled path at runtime by toggling the LoRA.
meituan-longcat/LongCat-Video-Avatar-1.5
Updated • 2.06k • 818Note Upstream PyTorch reference checkpoint. The MLX weights in this collection are converted from here by recipes/convert_longcat_avatar.py in the GitHub repo.
meituan-longcat/LongCat-Video
Text-to-Video • Updated • 1.8k • • 568Note Base LongCat-Video model. Source for the umT5-XXL text encoder, Wan 2.1 VAE, and SentencePiece tokenizer that the Avatar 1.5 overlay reuses.
-
mlx-community/LongCat-Video-Avatar-1.5-q4-dmd-merged
Text-to-Video • Updated • 2 -
mlx-community/LongCat-Video-Avatar-1.5-q8-dmd-merged
Text-to-Video • Updated • 1