AI Potluck
Model components / Base / pretrained models

DeepSeek-V4-Pro

DeepSeek

DeepSeek-V4-Pro, 1.6T total / 49B active MoE, 1M-token context, hybrid CSA+HCA attention. Released April 24 2026 under MIT, alongside V4-Flash (284B). V4-Pro-Max = max-reasoning mode. Verified live June 2026.

DeepSeek-V4-Pro, 1.6T total / 49B active MoE, 1M-token context, hybrid CSA+HCA attention. Released April 24 2026 under MIT, alongside V4-Flash (284B). V4-Pro-Max = max-reasoning mode. Verified live June 2026. Consolidated on 2026-07-29 from deepseek-v3-2, deepseek-v3-base, deepseek-v4-pro-base; openness follows deepseek-v4-pro, the release that currently governs. See sources/slug_aliases.yaml.

Openness

3 high confidence
3.0
weights
open(MIT, on HF, both Pro and Flash)
data
closed
code
partial(inference/serving
license
MIT(permissive, no use restrictions)

MIT weights across the V4 line, training corpus not released, inference and serving code only, so 3. V4-Pro governs and both it and V4-Pro-Base are plain MIT, which keeps the tier clear of an unsettled question: the retired deepseek-v3-base record carried a compound `code MIT + model DeepSeek-Model-License`, and whether that license's acceptable-use restrictions count as use-restricting is open in issue #117. The tier does not depend on that ruling. Adoption 5 is carried from V3.2, the most downloaded release.

Adoption

5 high confidence
5.0

ATOM Report (Apr 2026): China open models reached ~1.15B cumulative HF downloads with DeepSeek a primary driver; DeepSeek-V3.2 widely served (OpenRouter, multiple inference providers) with sustained post-launch production usage and frontier-at-10x-lower-cost positioning. Real usage well into the >10M-equivalent band across web app + API + derivatives.

Capability

5 high confidence
5.0

SWE-bench Verified 80.6% (trails Claude Opus 4.6 by ~0.2), MMLU-Pro 87.5, GPQA Diamond 90.1, LiveCodeBench Pass@1 93.5 (reported best of any model). V4-Pro-Max described as the strongest open source model available, at/near closed frontier on coding and reasoning.

Unchanged since 2026-07-30 (last edited, not re-checked)