DeepSeek
Modelos de DeepSeek que puedes ejecutar localmente
DeepSeek builds open-weight reasoning and coding models, including efficient distilled variants and large mixture-of-experts systems.
Modelos disponibles
Ordenados del modelo más ligero al más grande.
DeepSeek R1 1.5B
1.5B · Densa
Tiny reasoning model distilled from R1
64K ctx
DeepSeek R1 Distill 7B
7B · Densa
R1 reasoning distilled into Qwen 7B
64K ctx
DeepSeek R1 Distill 14B
14B · Densa
R1 reasoning distilled into Qwen 14B
64K ctx
DeepSeek R1 Distill 32B
32B · Densa
R1 reasoning distilled into Qwen 32B — sweet spot
64K ctx
DeepSeek V4 Flash
158B · 13B active · MoE
Efficient long-context V4 — 13B active, 1M context
1024K ctx
DeepSeek R1
671B · 37B active · MoE
Massive MoE reasoning model — 37B active
64K ctx
DeepSeek V3.2
685B · 37B active · MoE
State-of-the-art MoE — 37B active params
128K ctx
DeepSeek V4 Pro
1.6T · 49B active · MoE
Flagship V4 MoE — 49B active, 1M context
1024K ctx