LLM Releases
← Catalog

DeepSeek-V4-Pro-0813

Available
DeepSeekProprietary

The dated GA build behind DeepSeek's 'deepseek-v4-pro' API id, pinned on the official pricing table with an OpenRouter listing dated Aug 12 2026 — the Pro-tier counterpart to the way V4-Flash graduated as DeepSeek-V4-Flash-0731. It is a quiet version pin (no separate launch post or benchmark card), keeping the V4-Pro architecture and the 1M-token context / ~384K max-output window. List pricing is cache-heavy: $0.435 input cache-miss / $0.003625 cache-hit / $0.87 output per Mtok, with concurrency 500 (vs Flash's 2500); DeepSeek warns a significant, undated API price increase is coming. Thinking is on by default at effort 'high' (requested medium/xhigh both collapse to high). Architecture figures (1.6T total / 49B active MoE, hybrid long-context attention) are carried over from the April 2026 V4-Pro preview and are not independently reconfirmed for the 0813 build; Hugging Face still hosts only the April preview weights (MIT), with no confirmed separate 0813 open-weight repo, so this row is recorded as proprietary/API-only.

Specifications

License
Proprietary
Weights
Not released
Architecture
Mixture-of-Experts
Parameters
1.6T · 49B active
Context window
1M tokens
Max output
393K tokens
Knowledge cutoff
Price (in / out, $/M)
$0.435 / $0.87
Modalities
TextCode

Benchmarks

No benchmark scores recorded yet. Spotted some? Submit a correction.

Vendor-reported figures are claims until independently verified. See methodology.