A newer version is available: GLM-5.3View β†’

GLM-5.2

MIT

Z.ai Β· 753B (40B active) Β· Mixture of Experts

Frontier open-weight coder β€” top SWE-bench, 1M context Check if your GPU or Mac can run GLM-5.2 locally β€” 420.8 GB min, 701.3 GB recommended.

2026-061024K context

Mixture of Experts

Total experts: 256
Active experts: 8
Active params: 40.0B

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K2241.6 GBlowβ€”
Q3_K_M3338 GBmoderateβ€”
Q4_K_M4386.2 GBgoodβ€”
Q5_K_M5482.6 GBgoodβ€”
Q6_K6579.1 GBexcellentβ€”
Q8_08771.9 GBexcellentβ€”
F16161543.3 GBlosslessβ€”

Can I run GLM-5.2 locally?

Can I run GLM-5.2 locally?
GLM-5.2 needs about 420.8 GB of memory at a minimum and 701.3 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does GLM-5.2 need?
At Q4_K_M, GLM-5.2 uses about 386.2 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.