DeepSeek V4 Pro

DeepSeekopen-sourceOpen Source

Left preview and reached GA on 12–13 Aug 2026 as the '0813' checkpoint — same 1.6T MoE / 49B-active architecture as the April launch, but re-post-trained and, notably, repriced upward. The promotional $0.435/$0.87 rate is gone: GA pricing is $0.66 input / $1.98 output per 1M tokens, a roughly 52% input and 128% output increase, which DeepSeek framed as stepping back from a pure race-to-zero strategy. Still MIT-licensed weights on Hugging Face, 1M context. Vendor-reported scores are strong — LiveCodeBench 93.5%, SWE-bench Verified 80.6% (matching Gemini 3.1 Pro), GPQA Diamond 90.1%, Codeforces 3206 — but treat them cautiously: contamination-resistant independent benchmarks (DeepSWE) and Code Arena both show a wider gap to closed frontier models than DeepSeek's own numbers suggest, and third-party Arena Elo for the 0813 build is still preliminary (~1458–1465) as of 17 Aug 2026.

Context Window

1M

tokens

Max Output

33K

tokens

Input Price

$0.66

per 1M tokens

Output Price

$1.98

per 1M tokens

Speed

62

tokens/sec

Released

Aug 2026

2026-08-13

Blended Cost

$1.32

per 1M tokens

Value Score

67.4

quality per $

Capabilities

ChatFunction CallingCode GenerationReasoning

Benchmarks

Quality Index
89
MMLU Pro
88.2
HumanEval (Coding)
91.6
MATH
87.9
Arena ELO
1461

Price vs Quality

6080100$1.00$10.00$100.00Blended cost per 1M tokens (log scale) →GPT-4o — $6.25/1M, quality 85GPT-4o Mini — $0.375/1M, quality 72o3 Mini — $2.75/1M, quality 88o3 — $5.00/1M, quality 94GPT-4.1 — $5.00/1M, quality 89Claude Opus 4 — $45.00/1M, quality 91Claude Sonnet 4 — $9.00/1M, quality 88Claude 3.5 Haiku — $2.40/1M, quality 75Gemini 2.5 Pro — $5.63/1M, quality 92Gemini 2.0 Flash — $0.250/1M, quality 74Llama 4 Maverick — $0.400/1M, quality 80Llama 4 Scout — $0.275/1M, quality 71Mistral Large 2 — $4.00/1M, quality 79Codestral — $0.600/1M, quality 76DeepSeek V3 — $0.685/1M, quality 86DeepSeek R1 — $1.37/1M, quality 91Grok 3 — $9.00/1M, quality 87Grok 3 Mini — $0.400/1M, quality 78Command R+ — $6.25/1M, quality 68Amazon Nova Pro — $2.00/1M, quality 70Qwen 2.5 72B — $0.600/1M, quality 80Qwen 2.5 Coder 32B — $0.300/1M, quality 74Sonar Pro — $9.00/1M, quality 78GPT-5.5 — $17.50/1M, quality 97GPT-5.5 Pro — $105.00/1M, quality 96Claude Opus 4.7 — $15.00/1M, quality 96Gemini 3.1 Pro — $7.00/1M, quality 96DeepSeek V4 Flash — $0.210/1M, quality 84Qwen 3.6 Plus — $3.50/1M, quality 86Grok 4.3 — $1.88/1M, quality 93Kimi K2.6 — $2.11/1M, quality 92Kimi K3 — $9.00/1M, quality 97Claude Sonnet 4.6 — $9.00/1M, quality 90GLM-5.1 — $2.03/1M, quality 88Claude Opus 4.8 — $15.00/1M, quality 98Qwen 3.7 Max — $5.00/1M, quality 94Claude Fable 5 — $30.00/1M, quality 100MiniMax M3 — $1.50/1M, quality 89GPT-5.6 Sol — $17.50/1M, quality 98GPT-5.6 Terra — $7.00/1M, quality 93GPT-5.6 Luna — $0.700/1M, quality 83Gemini 3.6 Flash — $2.25/1M, quality 90Gemini 3.5 Flash-Lite — $1.40/1M, quality 79Claude Sonnet 5 — $9.00/1M, quality 93GLM-5.2 — $2.90/1M, quality 91Claude Opus 5 — $15.00/1M, quality 99Claude Mythos 5 — $30.00/1M, quality 100Grok 4.5 — $4.00/1M, quality 94Qwen3.8 Max — $4.00/1M, quality 96DeepSeek V4 Pro — $1.32/1M, quality 89DeepSeek V4 Pro
Cheaper is left, more capable is up. Nothing in the directory is both cheaper and at least as capable. Self-host models are left off the price axis — their cost depends on the hardware you run them on, not on tokens.

Context Window in Context

  • DeepSeek V4 Pro1M
  • GPT-4.11M
  • Gemini 2.5 Pro1M
  • Gemini 2.0 Flash1M
  • Llama 4 Maverick1M
  • GPT-5.51M
  • GPT-5.5 Pro1M
  • Claude Opus 4.71M
Bar length is logarithmic — each full step is 10× the tokens. Shown against the models with the nearest context windows, not the directory extremes.

Pricing History

Apr 24, 2026$1.74 / $3.48(launch)
May 31, 2026$0.435 / $0.87
Aug 13, 2026$0.66 / $1.98current

Open Source: Licensed under MIT

Noch nicht auf unserem Gateway

DeepSeek V4 Pro ist open-weight, du kannst es also selbst betreiben. Swfte Connect macht aus deinem Deployment einen verwalteten Endpoint, der dieselbe API spricht wie jedes andere Modell in diesem Verzeichnis.

Die Anfrage
curl https://api.swfte.com/agents/v2/gateway/chat/completions \
  -H "Authorization: Bearer $SWFTE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek:deepseek-v4-pro-0813",
    "messages": [{"role": "user", "content": "Hello"}]
  }'
Dieses Modell mit Swfte Connect bereitstellen

About DeepSeek V4 Pro

DeepSeek V4 Pro is a open-source AI model by DeepSeek, released on August 13, 2026. It supports a context window of 1000K tokens and can generate up to 33K output tokens.

At $0.66 per million input tokens and $1.98 per million output tokens, its blended cost of $1.32/1M tokens puts it in the mid-range pricing tier. Its value score of 67.4 reflects the balance of quality and cost.

DeepSeek V4 Pro is available as an open-source model under the MIT license, meaning you can self-host it for predictable costs or use it through API providers like Swfte Connect.

Using DeepSeek V4 Pro with Swfte

Access DeepSeek V4 Pro through Swfte Connect, our unified LLM gateway. Connect gives you a single API for 50+ models, with automatic routing, cost optimization, and fallback handling. You can also try DeepSeek V4 Pro in our AI Playground before integrating.

Deploy a model with Swfte Connect

One gateway, every provider, per-token cost visibility. Swap models without touching your code.