RETIRED MODEL

DeepSeek V4 Pro: 1.6T MoE flagship with 1M context

1.6T total parameters, 49B active per token. 1,000,000 token context. Strong reasoning and coding. $1.74/$3.48 per million tokens — fraction of US frontier pricing.

Try DeepSeek V4 Pro on Council See it in a council

DeepSeek V4 Pro is no longer part of the Council AI lineup. It was superseded by DeepSeek V4 Pro (0813), which is available on every paid plan. The specs below are kept for reference.

This model is no longer in the Council AI lineup

DeepSeek V4 Pro has been retired from Council AI. Council AI now runs DeepSeek V4 Pro (0813) in its place. Existing conversations that used it still render correctly; it can no longer be added to a new council.

Specs

ProviderDeepSeek
ArchitectureMixture-of-Experts (MoE)
Parameters1.6T total / 49B active
Context window1,000,000 tokens
MultimodalText only
Web searchNone
Price (in / out per 1M)$1.74 / $3.48 (standard, post May 31 2026)

Best at

Where it loses

About the pricing

DeepSeek launched V4 Pro with a temporary discount that expired May 31, 2026. The $1.74 / $3.48 per million tokens shown here is the standard post-launch rate. Earlier published numbers below the launch discount no longer apply. Council AI bundles V4 Pro inside monthly plan budgets at the standard rate.

Frequently asked questions

What does '49B active' mean in MoE?

DeepSeek V4 Pro routes each token through only a subset of its 1.6T total parameters — about 49B activate per forward pass. This is how MoE architectures get frontier-scale quality at much lower inference cost than a dense model of equivalent quality would need.

How much does DeepSeek V4 Pro cost?

$1.74 per million input tokens and $3.48 per million output. This is the standard rate post the May 31, 2026 launch-discount expiry.

Is DeepSeek V4 Pro open-weights?

DeepSeek has historically released open-weight variants of its models. Check DeepSeek's release notes for the current open-weight posture of V4 Pro specifically; Council AI uses the hosted API.

How does V4 Pro compare to V4 Flash?

V4 Pro is the flagship — slower, smarter, and the one to pick when you want maximum capability. V4 Flash is the cheap, fast sibling for high-volume work where you don't need the heavy reasoning.

Why include DeepSeek in a council?

Two reasons: (1) a non-US-aligned voice often catches things GPT/Claude/Gemini all miss because of similar training-data biases, and (2) the price makes it cheap to add as a third opinion.

Can DeepSeek V4 Pro really handle 1M tokens?

Yes — natively. Coherence across the full window is competitive with Gemini 3 Pro on long-context benchmarks like RULER.

Are there data-residency concerns?

DeepSeek is headquartered in China. If your compliance posture excludes Chinese providers, swap in Mistral Large 3 (EU) or stick to US-headquartered providers.