OLMo 2 32B

Apache 2.0

Allen AI Β· 32B Β· Dense

Fully open research model by Allen AI Check if your GPU or Mac can run OLMo 2 32B locally β€” 17.9 GB min, 29.8 GB recommended.

2025-034K context

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K210.7 GBlowβ€”
Q3_K_M314.8 GBmoderateβ€”
Q4_K_M416.9 GBgoodβ€”
Q5_K_M521 GBgoodβ€”
Q6_K625.1 GBexcellentβ€”
Q8_0833.3 GBexcellentβ€”
F161666.1 GBlosslessβ€”

Can I run OLMo 2 32B locally?

Can I run OLMo 2 32B locally?
OLMo 2 32B needs about 17.9 GB of memory at a minimum and 29.8 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does OLMo 2 32B need?
At Q4_K_M, OLMo 2 32B uses about 16.9 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.