Back to models
Qwen

Alibaba

Qwen models you can run locally

Alibaba's Qwen family spans compact on-device models, multimodal assistants, reasoning systems and large mixture-of-experts models, plus open image generation with Qwen Image and Z-Image and Wan for local video.

34 models·Qwen, Z-Image, WanOfficial website ↗

Available models

Ordered from the lightest model to the largest.

Qwen 3 0.6B

0.6B · Dense

0.8GB

Ultra-light Qwen 3 model for constrained devices

32K ctx

Qwen 3.5 0.8B

0.8B · Dense

0.9GB

Ultra-tiny model for embedded and edge

32K ctx

Wan 2.1 T2V 1.3B

1.3B · Dense

1.2GB

Tiny open text-to-video — 480p clips on 8GB consumer GPUs

4K ctx

Qwen 2.5 Coder 1.5B

1.5B · Dense

1.3GB

Ultra-lightweight coding model

32K ctx

Qwen 3 1.7B

1.7B · Dense

1.4GB

Compact multilingual Qwen 3

32K ctx

Qwen 3.5 2B

2B · Dense

1.5GB

Small multimodal Qwen 3.5

32K ctx

Qwen 3.5 4B

4B · Dense

2.5GB

Small multimodal Qwen 3.5

32K ctx

Qwen3-VL 4B

4.4B · Dense

2.8GB

Compact dedicated vision-language model — OCR & image chat on edge

256K ctx

Wan 2.2 TI2V 5B

5B · Dense

3.1GB

Unified text/image-to-video — the local sweet spot under Apache 2.0

4K ctx

Z-Image Turbo

6B · Dense

3.6GB

8-step distilled image model — photorealism and bilingual text on 16GB cards

4K ctx

Qwen 2.5 Coder 7B

7B · Dense

4.1GB

Dedicated coding model

128K ctx

Qwen 3 8B

8B · Dense

4.6GB

Qwen 3 with thinking mode support

128K ctx

Qwen3-VL 8B

8.8B · Dense

5GB

The community-favourite local VLM — superb OCR, receipts & captioning

256K ctx

Qwen 3.5 9B

9B · Dense

5.1GB

Multimodal Qwen 3.5 mid-size

32K ctx

Qwen 3 14B

14B · Dense

7.7GB

Strong all-rounder with thinking mode

128K ctx

Qwen Image 2512

20B · Dense

10.7GB

Open text-to-image with strong English and Chinese typography

4K ctx

Qwen 3.8 27B

27B · Dense

14.3GB

Flagship dense Qwen 3.8 — native multimodal all-rounder with video understanding

256K ctx

Wan 2.2 T2V A14B

27B · 14B active · MoE

14.3GB

Flagship open Wan 2.2 — 14B-active MoE for photoreal text-to-video

4K ctx

Qwen 3.5 27B

27.8B · Dense

14.7GB

Flagship native multimodal Qwen 3.5

256K ctx

Qwen 3.6 27B

27.8B · Dense

14.7GB

Flagship dense Qwen 3.6 — native multimodal all-rounder

256K ctx

Qwen 3 30B-A3B

30B · 3.3B active · MoE

15.9GB

MoE with only 3.3B active — extremely efficient

128K ctx

Qwen 3 Coder 30B-A3B

30B · 3B active · MoE

15.9GB

Efficient agentic coding MoE — 3B active, 256K context

256K ctx

Qwen3-VL 30B-A3B

31B · 3B active · MoE

16.4GB

Efficient vision MoE — 3B active, strong temporal & document understanding

256K ctx

Qwen 3 32B

32B · Dense

16.9GB

Qwen 3 flagship dense model

128K ctx

Qwen 3.5 35B-A3B

35B · 3B active · MoE

18.4GB

Efficient multimodal MoE with 3B active

256K ctx

Qwen 3.6 35B-A3B

36B · 3B active · MoE

18.9GB

Big-model quality at 3B-active speed — the mid-hardware sweet spot

256K ctx

Qwen 3 Next 80B-A3B

80B · 3B active · MoE

41.5GB

High-sparsity MoE — extreme low activation ratio for fast inference at 80B scale

256K ctx

Qwen 3 Coder Next 80B-A3B

80B · 3B active · MoE

41.5GB

Ultra-efficient agentic coding MoE optimized for tool-calling coding agents

256K ctx

Qwen 3.5 122B-A10B

122B · 10B active · MoE

63GB

Large multimodal MoE

256K ctx

Qwen 3 235B-A22B

235B · 22B active · MoE

120.9GB

Massive MoE with 22B active — frontier quality

128K ctx

Qwen 3 VL 235B-A22B

235B · 22B active · MoE

120.9GB

Flagship vision-language MoE — frontier multimodal reasoning and agentic GUI control

256K ctx

Qwen 3.5 397B-A17B

397B · 17B active · MoE

203.9GB

Largest multimodal Qwen 3.5 MoE

256K ctx

Qwen 3 Coder 480B

480B · 35B active · MoE

246.4GB

Largest open coding MoE — 35B active

256K ctx

Qwen 3.8 2.4T-A95B

2.4T · 95B active · MoE

1229.8GB

Frontier Qwen 3.8 MoE — 95B active, 1M context

1024K ctx