--- license: apache-2.0 base_model: Qwen/Qwen3-VL-30B-A3B-Instruct pipeline_tag: image-text-to-text language: - en tags: - image-text-to-text - vision - multimodal - zen - zenlm - hanzo --- > **Status:** Superseded by [zenlm/zen3-vl](https://huggingface.co/zenlm/zen3-vl). Weights remain available for reproducibility. # zen-vl-30b-agent > Superseded by [zenlm/zen3-vl](https://huggingface.co/zenlm/zen3-vl) — canonical name. Vision-language agent model for image understanding, OCR, and visual reasoning (30B total / 3B active (MoE)). Repackaged from [Qwen/Qwen3-VL-30B-A3B-Instruct](https://huggingface.co/Qwen/Qwen3-VL-30B-A3B-Instruct) (apache-2.0, Alibaba Qwen). **Not trained from scratch** — a permissively-licensed redistribution for the OSS-clean Zen model line. ## Specs | Property | Value | |----------|-------| | Parameters | 30B total / 3B active (MoE) | | Architecture | Qwen3-VL | | Modality | text + image + video | ## License `apache-2.0`. Upstream: **Qwen/Qwen3-VL-30B-A3B-Instruct** by Alibaba Qwen (apache-2.0).