Qwen 3 Coder 480B

Apache 2.0

Alibaba Β· 480B (35B active) Β· Mixture of Experts

Largest open coding MoE β€” 35B active Check if your GPU or Mac can run Qwen 3 Coder 480B locally β€” 268.2 GB min, 447 GB recommended.

2025-07256K context

Mixture of Experts

Total experts: 128
Active experts: 8
Active params: 35.0B

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K2154.2 GBlowβ€”
Q3_K_M3215.6 GBmoderateβ€”
Q4_K_M4246.4 GBgoodβ€”
Q5_K_M5307.8 GBgoodβ€”
Q6_K6369.3 GBexcellentβ€”
Q8_08492.2 GBexcellentβ€”
F1616984 GBlosslessβ€”

About this model

Qwen 3 logo Qwen3-Coder is the most agentic code model to date in the Qwen series.

Get started

480B

Cloud

ollama run qwen3-coder:480b-cloud

Local

ollama run qwen3-coder:480b

Running locally requires a minimum of 250GB of memory or unified memory.

30B

ollama run qwen3-coder:30b

Overview

qwen3-coder:30b offers 30B total parameters with only 3.3B activated, delivering strong performance while maintaining efficiency.

  • Exceptional agentic capabilities for real-world software engineering tasks through advanced long-horizon reinforcement learning on SWE-Bench and similar benchmarks.
  • Long context support with 256K tokens natively and up to 1M tokens using extrapolation methods, optimized for repository-scale understanding.
  • Scaled pretraining on 7.5T tokens (70% code ratio) while preserving strong general and mathematical abilities.
  • Execution-driven reinforcement learning that significantly boosts code execution success rates across diverse real-world coding tasks.

image.png

Reference

Can I run Qwen 3 Coder 480B locally?

Can I run Qwen 3 Coder 480B locally?
Qwen 3 Coder 480B needs about 268.2 GB of memory at a minimum and 447 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Qwen 3 Coder 480B need?
At Q4_K_M, Qwen 3 Coder 480B uses about 246.4 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.