Qwen3-VL 8B

Apache 2.0

Alibaba Β· 8.8B Β· Dense

The community-favourite local VLM β€” superb OCR, receipts & captioning Check if your GPU or Mac can run Qwen3-VL 8B locally β€” 4.9 GB min, 8.2 GB recommended.

2025-10256K context

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K23.3 GBlowβ€”
Q3_K_M34.4 GBmoderateβ€”
Q4_K_M45 GBgoodβ€”
Q5_K_M56.1 GBgoodβ€”
Q6_K67.3 GBexcellentβ€”
Q8_089.5 GBexcellentβ€”
F161618.5 GBlosslessβ€”

Can I run Qwen3-VL 8B locally?

Can I run Qwen3-VL 8B locally?
Qwen3-VL 8B needs about 4.9 GB of memory at a minimum and 8.2 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Qwen3-VL 8B need?
At Q4_K_M, Qwen3-VL 8B uses about 5 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.