Qwen3-VL 30B-A3B

Apache 2.0

Alibaba Β· 31B (3B active) Β· Mixture of Experts

Efficient vision MoE β€” 3B active, strong temporal & document understanding Check if your GPU or Mac can run Qwen3-VL 30B-A3B locally β€” 17.3 GB min, 28.9 GB recommended.

2025-09256K context

Mixture of Experts

Total experts: 128
Active experts: 8
Active params: 3.0B

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K210.4 GBlowβ€”
Q3_K_M314.4 GBmoderateβ€”
Q4_K_M416.4 GBgoodβ€”
Q5_K_M520.3 GBgoodβ€”
Q6_K624.3 GBexcellentβ€”
Q8_0832.3 GBexcellentβ€”
F161664 GBlosslessβ€”

Can I run Qwen3-VL 30B-A3B locally?

Can I run Qwen3-VL 30B-A3B locally?
Qwen3-VL 30B-A3B needs about 17.3 GB of memory at a minimum and 28.9 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Qwen3-VL 30B-A3B need?
At Q4_K_M, Qwen3-VL 30B-A3B uses about 16.4 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.