DeepSeek logo

DeepSeek: DeepSeek V4 Pro

Open MoE flagship with million-token context for coding and long agent runs

Source-linked deepseek-thinking Open weights Released 2026-04-24
API record Report
InputT
OutputT
Input price$0.435/M
Output price$0.87/M
Context1M
Max output384K
Providers63
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering DeepSeek V4 Pro
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
AnyAPI deepseek/deepseek-v4-pro 1M 384K ReasoningToolsJSON Docs ↗
Model Oracle AI deepseek-v4-pro 1M 384K ReasoningToolsJSON Docs ↗
Alibaba Token Plan (China) deepseek-v4-pro 1M 384K ReasoningToolsJSON Docs ↗
Volcengine Ark Coding Plan deepseek-v4-pro 1M 384K ReasoningToolsJSON Docs ↗
SenseNova (China) deepseek-v4-pro 1.04858M 65.536K ReasoningToolsJSON Docs ↗
UnoRouter deepseek-v4-pro:free 1M 384K ReasoningToolsJSON Docs ↗
Kenari deepseek-v4-pro 1M 384K ReasoningToolsJSON Docs ↗
SCNet Token Plan DeepSeek-V4-Pro 1M 384K ReasoningToolsJSON Docs ↗
Alibaba Token Plan deepseek-v4-pro 1M 384K ReasoningToolsJSON Docs ↗
Umans AI Coding Plan umans-deepseek-v4-pro-0813 1.04858M 393.215K ReasoningToolsJSON Docs ↗
routing.run deepseek-v4-pro 1M 64K $0.348 $0.696 ReasoningToolsJSON Docs ↗
CrofAI deepseek-v4-pro 1M 131.072K $0.35 $0.8 $0.003 ReasoningToolsJSON Docs ↗
EBCloud DeepSeek-V4-Pro 1M 384K $0.429 $0.857 ReasoningToolsJSON Docs ↗
Auriko deepseek-v4-pro 1M 384K $0.435 $0.87 $0.004 ReasoningToolsJSON Docs ↗
TokenGo deepseek/deepseek-v4-pro 1M 384K $0.435 $0.87 ReasoningToolsJSON Docs ↗
Pioneer deepseek-ai/DeepSeek-V4-Pro 1M 131.072K $0.435 $0.87 $0.004 ReasoningToolsJSON Docs ↗
Hugging Face deepseek-ai/DeepSeek-V4-Pro 1.04858M 393.216K $0.435 $0.87 $0.004 ReasoningToolsJSON Docs ↗
ZenMux deepseek/deepseek-v4-pro 1M 384K $0.435 $0.87 $0.004 ReasoningToolsJSON Docs ↗
DeepSeek deepseek-v4-pro 1M 384K $0.435 $0.87 $0.004 ReasoningToolsJSON Docs ↗
Nvidia deepseek-ai/deepseek-v4-pro 1.04858M 393.216K $0.435 $0.87 $0.004 ReasoningToolsJSON Docs ↗
DevPass (LLM Gateway) deepseek-v4-pro 1.05M 384K $0.435 $0.87 $0.004 ReasoningToolsJSON Docs ↗
OpenCode Go deepseek-v4-pro 1M 384K $0.435 $0.87 $0.004 ReasoningToolsJSON Docs ↗
Alibaba (China) deepseek-v4-pro 1M 384K $0.435 $0.87 $0.004 ReasoningToolsJSON Docs ↗
LLM Gateway deepseek/deepseek-v4-pro 1.05M 393.216K $0.435 $0.87 $0.004 ReasoningTools Docs ↗
Vivgrid deepseek-v4-pro 1M 384K $0.435 $0.87 $0.004 ReasoningToolsJSON Docs ↗
Modelis deepseek-v4-pro 1M 384K $0.435 $0.87 ReasoningToolsJSON Docs ↗
OrcaRouter deepseek/deepseek-v4-pro 1M 384K $0.442 $0.884 $0.06 ReasoningToolsJSON Docs ↗
Vercel AI Gateway deepseek/deepseek-v4-pro 1M 384K $0.66 $1.98 $0.022 ReasoningToolsJSON Docs ↗
Merge Gateway deepseek/deepseek-v4-pro 1M 384K $0.66 $1.98 $0.022 ReasoningToolsJSON Docs ↗
Vancine deepseek-v4-pro 1M 384K $0.66 $1.98 $0.022 ReasoningToolsJSON Docs ↗
above.dev deepseek-v4-pro 1M 384K $0.726 $2.178 $0.024 ReasoningToolsJSON Docs ↗
UnoRouter deepseek-v4-pro 1M 384K $0.9 $1.8 ReasoningToolsJSON Docs ↗
OpenRouter deepseek/deepseek-v4-pro 1.04858M 384K $0.955 $1.911 $0.08 ReasoningToolsJSON Docs ↗
ai& deepseek-ai/deepseek-v4-pro 1.04858M 384K $1 $2.5 ReasoningToolsJSON Docs ↗
Neuralwatt deepseek-v4-pro 1.04856M 393.216K $1 $3 $0.1 ReasoningToolsJSON Docs ↗
NanoGPT deepseek/deepseek-v4-pro 1.04858M 384K $1.1 $2.2 $0.11 ReasoningToolsJSON Docs ↗
CrossModel deepseek/deepseek-v4-pro 1M 384K $1.215 $3.645 $0.041 ReasoningToolsJSON Docs ↗
Deep Infra deepseek-ai/DeepSeek-V4-Pro 1.04858M 16.384K $1.3 $2.6 $0.1 ReasoningToolsJSON Docs ↗
Eden AI deepseek/deepseek-v4-pro 1M 384K $1.32 $3.96 $0.044 ReasoningToolsJSON Docs ↗
Requesty deepseek-v4-pro 1M 131.072K $1.32 $3.96 $0.044 ReasoningToolsJSON Docs ↗
Umans AI umans-deepseek-v4-pro-0813 1.04858M 393.215K $1.32 $3.96 $0.044 ReasoningToolsJSON Docs ↗
Ofox deepseek/deepseek-v4-pro 1M 384K $1.32 $3.96 $0.044 ReasoningToolsJSON Docs ↗
GMI Cloud deepseek-ai/DeepSeek-V4-Pro 1.04858M 384K $1.392 $2.784 $0.116 ReasoningToolsJSON Docs ↗
Kilo Gateway deepseek/deepseek-v4-pro 1.024M 384K $1.6 $3.2 $0.135 ReasoningToolsJSON Docs ↗
NovitaAI deepseek/deepseek-v4-pro 1.04858M 393.216K $1.6 $3.2 $0.135 ReasoningToolsJSON Docs ↗
Jalapeno Cloud DeepSeek-V4-Pro 1.04858M 384K $1.6 $3.38 ReasoningToolsJSON Docs ↗
EmpirioLabs AI deepseek-v4-pro 1M 393.216K $1.65 $3.3 $1.65 ReasoningToolsJSON Docs ↗
Cortecs deepseek-v4-pro 1.04858M 1.04858M $1.73 $3.46 $0.432 ReasoningTools Docs ↗
DigitalOcean deepseek-v4-pro 1.04858M 393.216K $1.74 $3.48 ReasoningToolsJSON Docs ↗
Fireworks AI accounts/fireworks/models/deepseek-v4-pro 1M 384K $1.74 $3.48 $0.145 ReasoningToolsJSON Docs ↗
ClinePass cline-pass/deepseek-v4-pro 1M 384K $1.74 $3.48 $0.015 ReasoningToolsJSON Docs ↗
OpenCode Zen deepseek-v4-pro 1M 384K $1.74 $3.84 $0.145 ReasoningToolsJSON Docs ↗
Baseten deepseek-ai/DeepSeek-V4-Pro 1.04858M 262.144K $1.74 $3.48 $0.145 ReasoningToolsJSON Docs ↗
HPC-AI deepseek/deepseek-v4-pro 1.002M 128K $1.74 $3.48 $0.145 ReasoningToolsJSON Docs ↗
Impossibl deepseek/deepseek-v4-pro 1M 384K $1.74 $3.48 $0.145 ReasoningToolsJSON Docs ↗
Abacus deepseek-ai/DeepSeek-V4-Pro 1M 32.768K $1.74 $3.48 $0.15 ReasoningToolsJSON Docs ↗
SiliconFlow deepseek-ai/DeepSeek-V4-Pro 1M 384K $1.74 $3.48 $0.145 ReasoningToolsJSON Docs ↗
Cloudflare AI Gateway deepseek/deepseek-v4-pro 131.072K 384K $1.74 $3.48 $0.145 ReasoningToolsJSON Docs ↗
Arcee deepseek/deepseek-v4-pro 512K 384K $1.74 $3.48 $0.2 ReasoningToolsJSON Docs ↗
FastRouter deepseek/deepseek-v4-pro 1M 384K $1.74 $3.48 ReasoningToolsJSON Docs ↗
Nebius Token Factory deepseek-ai/DeepSeek-V4-Pro 1.04858M 1.04858M $1.75 $3.5 $0.15 ReasoningToolsJSON Docs ↗
TensorX deepseek/deepseek-v4-pro 1.04858M 384K $1.75 $3.5 $0.438 ReasoningToolsJSON Docs ↗
Charm Hyper deepseek-v4-pro 1M 384K $2.4 $4.8 $0.2 ReasoningToolsJSON Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $0.348/M

Across 53 priced providers

Median input $1.215/M

Midpoint of listed rates

Highest input $2.4/M

6.9× the lowest listed rate

Output range $0.696 – $4.8

Per million output tokens

routing.run $0.348/MLowest
CrofAI $0.35/M
EBCloud $0.429/M
Auriko $0.435/M
TokenGo $0.435/M
Pioneer $0.435/M
Hugging Face $0.435/M
ZenMux $0.435/M
DeepSeek $0.435/M
Nvidia $0.435/M
DevPass (LLM Gateway) $0.435/M
OpenCode Go $0.435/M
Alibaba (China) $0.435/M
LLM Gateway $0.435/M
Vivgrid $0.435/M
Modelis $0.435/M
OrcaRouter $0.442/M
Vercel AI Gateway $0.66/M
Merge Gateway $0.66/M
Vancine $0.66/M
above.dev $0.726/M
UnoRouter $0.9/M
OpenRouter $0.955/M
ai& $1/M
Neuralwatt $1/M
NanoGPT $1.1/M
CrossModel $1.215/M
Deep Infra $1.3/M
Eden AI $1.32/M
Requesty $1.32/M
Umans AI $1.32/M
Ofox $1.32/M
GMI Cloud $1.392/M
Kilo Gateway $1.6/M
NovitaAI $1.6/M
Jalapeno Cloud $1.6/M
EmpirioLabs AI $1.65/M
Cortecs $1.73/M
DigitalOcean $1.74/M
Fireworks AI $1.74/M
ClinePass $1.74/M
OpenCode Zen $1.74/M
Baseten $1.74/M
HPC-AI $1.74/M
Impossibl $1.74/M
Abacus $1.74/M
SiliconFlow $1.74/M
Cloudflare AI Gateway $1.74/M
Arcee $1.74/M
FastRouter $1.74/M
Nebius Token Factory $1.75/M
TensorX $1.75/M
Charm Hyper $2.4/M
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.

Context limits also differ by provider, from 131.072K to 1.05M tokens. Compare the provider table above before choosing on price alone.

Specification

Capabilities

Recorded from the source catalog and provider listings.

Reasoning Yes
Tool calling Yes
Structured output Yes
× Attachments No
× Vision input No
Open weights Yes
Creator
DeepSeek
Model family
deepseek-thinking
Knowledge cutoff
2025-05
License
Not documented
Release date
2026-04-24
Model ID
deepseek/deepseek-v4-pro

Weights: Hugging Face ↗

Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
Benchmark profile
Coding IndexCoding Index Agentic IndexAgentic Index IntelligenceIntelligence Index Design Arena: w…Design Arena: website Design Arena: c…Design Arena: codecategories Design Arena: g…Design Arena: gamedev Design Arena: d…Design Arena: dataviz Design Arena: u…Design Arena: uicomponent This model — Coding Index: 59.4 index (68.5th percentile) This model — Agentic Index: 27.7 index (66.1th percentile) This model — Intelligence Index: 30.9 index (60.9th percentile) This model — Design Arena: website: 1247 elo (63.1th percentile) This model — Design Arena: codecategories: 1256 elo (65.0th percentile) This model — Design Arena: gamedev: 1252 elo (67.5th percentile) This model — Design Arena: dataviz: 1218 elo (54.8th percentile) This model — Design Arena: uicomponent: 1237 elo (58.4th percentile)

Each axis is this model's percentile among the 8 benchmarks it has published results for, measured against every other model with a score on that same benchmark. Percentiles are used because benchmarks do not share a scale — a 60 on one is not a 60 on another. Hover any point for the raw score.

Coding Index index Source ↗
Coding Index #65 of 205
Agentic Index index Source ↗
Agentic Index #66 of 193
Intelligence Index index Source ↗
Intelligence Index #75 of 192
Design Arena: website elo Source ↗
Design Arena: website #65 of 175
Design Arena: codecategories elo Source ↗
Design Arena: codecategories #59 of 167
Design Arena: gamedev elo Source ↗
Design Arena: gamedev #53 of 166
Design Arena: dataviz elo Source ↗
Design Arena: dataviz #75 of 165
Design Arena: uicomponent elo Source ↗
Design Arena: uicomponent #67 of 161
Design Arena: 3d elo Source ↗
Design Arena: 3d #38 of 157
Design Arena: svg elo Source ↗
Design Arena: svg #69 of 116
Design Arena: asciiart elo Source ↗
Design Arena: asciiart #59 of 99
Design Arena: fullstack elo Source ↗
Design Arena: fullstack #67 of 67
Design Arena: webapps elo Source ↗
Design Arena: webapps #67 of 67
Design Arena: godotgamedev elo Source ↗
Design Arena: godotgamedev #41 of 45
SWE-Bench Verified resolved Source ↗
SWE-Bench Verified #5 of 32
Artificial Analysis Coding Agent Index average pass@1 · Claude Code Source ↗
Artificial Analysis Coding Agent Index #5 of 6
SWE-Atlas Codebase QnA pass@1 · Claude Code Source ↗
SWE-Atlas Codebase QnA #5 of 6
SWE-Bench Pro pass@1 · Claude Code Source ↗
SWE-Bench Pro #4 of 6
Terminal-Bench — 2.1 pass@1 · Claude Code Source ↗
Terminal-Bench #4 of 6

Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.

Cost against capability

Price and performance

Listed input price plotted against Coding Index, the benchmark with the widest published coverage that this model appears in.

0 43 86 $0.01 $0.1 $1 $10 Claude Opus 5 — $5/M, score 78.0 OpenAI: GPT-5.1 (batch) — $0.625/M, score 49.4 Anthropic: Claude Haiku 4.5 (batch) — $0.5/M, score 43.9 OpenAI: GPT-5 (batch) — $0.625/M, score 37.8 Gemini 3.1 Flash Lite Preview — $0.25/M, score 34.7 Gemma 3 27B IT — $0.08/M, score 10.1 OpenAI: GPT-4.1 Mini (batch) — $0.2/M, score 20.2 Gemma 4 31B IT — $0.09/M, score 43.4 Mistral: Ministral 3 8B 2512 (batch) — $0.075/M, score 9.7 GPT-5.4 — $2.5/M, score 71.1 GPT-5.6 Luna — $0.2/M, score 71.4 Gemini 3.5 Flash Lite — $0.3/M, score 49.3 Mistral: Mistral Large 3 2512 (batch) — $0.25/M, score 20.1 OpenAI: o3 Mini High (batch) — $0.55/M, score 16.3 Qwen: Qwen3.8 Max (0902) — $2/M, score 71.8 Gemini 3.5 Flash — $1.5/M, score 70.1 Z.ai: GLM 5.3 Flash (batch) — $0.075/M, score 71.5 Gemma 3 12B IT — $0.05/M, score 5.8 GPT-4.1 mini — $0.4/M, score 20.2 DeepSeek V4 Pro — $0.435/M, score 59.4 Claude Sonnet 5 — $2/M, score 71.5 Mistral: Ministral 3 8B 2512 — $0.15/M, score 9.7 OpenAI: o1 (batch) — $7.5/M, score 39.7 Z.ai: GLM 5.1 — $0.966/M, score 55.8 Gemini 3.7 Flash — $0.75/M, score 76.1 LongCat-2.0 — $0.3/M, score 45.3 Anthropic: Claude Opus 4.8 (batch) — $2.5/M, score 74.3 Mistral: Mistral Large 3 2512 — $0.5/M, score 20.1 DeepSeek V3.2 — $0.18/M, score 44.2 OpenAI: GPT-5.5 (batch) — $2.5/M, score 74.9 Anthropic: Claude Opus 4.7 — $5/M, score 73.6 Google: Gemini 3.5 Flash (batch) — $0.75/M, score 70.1 OpenAI: GPT-5.4 Nano (batch) — $0.1/M, score 56.1 Google: Gemini 2.5 Pro (batch) — $0.625/M, score 33.3 Inception: Mercury 2 — $0.25/M, score 31.1 OpenAI: GPT-4.1 Nano (batch) — $0.05/M, score 11.1 Gemini 3.6 Flash — $0.75/M, score 69.2 Anthropic: Claude Sonnet 4.5 — $3/M, score 52.1 OpenAI: GPT-4o-mini (batch) — $0.075/M, score 11.4 Qwen: Qwen3.5-9B (batch) — $0.17/M, score 28.7 GPT-5.4 mini — $0.75/M, score 56.1 Gemma 4 26B A4B IT — $0.042/M, score 39.3 Claude Fable 5 — $10/M, score 76.5 Mistral: Mistral Small 4 (batch) — $0.075/M, score 26.6 Gemini 2.5 Pro — $1.25/M, score 33.3 Z.ai: GLM 5.3 Flash — $0.15/M, score 71.5 OpenAI: GPT-3.5 Turbo (batch) — $0.25/M, score 10.7 MiniMax: MiniMax M3 — $0.3/M, score 58.6 DeepSeek: DeepSeek V4 Pro 0813 (batch) — $0.66/M, score 68.8 Thinking Machines: Inkling Small (batch) — $0.5/M, score 52.9 Anthropic: Claude Fable 5.1 — $10/M, score 81.6 Google: Gemma 4 31B (batch) — $0.39/M, score 43.4 Qwen: Qwen3.8 2.4T A95B (batch) — $2/M, score 71.9 Mistral: Devstral 2 2512 — $0.4/M, score 31.3 MoonshotAI: Kimi K2.7 Code (batch) — $0.95/M, score 60.8 DeepSeek-R1 — $0.7/M, score 24.6 Inkling — $1.87/M, score 52.1 OpenAI: gpt-oss-20b (batch) — $0.05/M, score 20.7 Muse Spark 1.1 — $1.25/M, score 71.3 MiMo-V2.5 — $0.14/M, score 56.8 Google: Gemini 3.5 Flash Lite (batch) — $0.15/M, score 49.3 Kimi K2 Thinking — $0.4/M, score 21.0 Mistral: Mistral Medium 3.5 (batch) — $0.75/M, score 46.9 Qwen: Qwen3.8 Max (0803) — $2/M, score 68.9 Mistral: Mistral Medium 3.1 (batch) — $0.2/M, score 20.5 GPT-5.1 — $1.25/M, score 49.4 Qwen: Qwen3.8 27B — $0.42/M, score 68.1 Anthropic: Claude Fable 5.1 (batch) — $5/M, score 81.6 Google: Gemini 3.8 Flash (batch) — $0.375/M, score 76.3 Google: Gemini 3.7 Flash (batch) — $0.375/M, score 76.1 inclusionAI: Ling 3.0 Flash VL — $0.06/M, score 57.0 Z.ai: GLM 5.3 (batch) — $0.7/M, score 74.8 DeepSeek: DeepSeek V4 Flash 0731 (batch) — $0.11/M, score 69.1 MoonshotAI: Kimi K3 (batch) — $3/M, score 76.2 Inkling Small — $0.45/M, score 52.9 GPT-5.5 — $5/M, score 74.9 Anthropic: Claude Sonnet 5 (batch) — $1/M, score 71.5 Anthropic: Claude Fable 5 (batch) — $5/M, score 76.5 Nemotron 3 Super 120B A12B — $0.2/M, score 37.7 IBM: Granite 4.2 8B — $0.06/M, score 22.4 Thinking Machines: Inkling (batch) — $1/M, score 52.1 GPT OSS 120B — $0.03/M, score 30.4 GPT-4 — $30/M, score 13.1 SpaceXAI: Grok 4.6 — $2/M, score 76.8 DeepSeek: DeepSeek V3.1 Terminus — $0.27/M, score 43.5 GPT-5 Mini — $0.25/M, score 15.6 Gemma 3 4B IT — $0.04/M, score 2.7 Claude Opus 5 (batch) — $2.5/M, score 78.0 OpenAI: GPT-6 Astra (batch) — $5/M, score 76.9 OpenAI: GPT-5.6 Terra (batch) — $1/M, score 76.7 OpenAI: GPT-5.4 (batch) — $1.25/M, score 71.1 DeepSeek V4 Flash — $0.15/M, score 52.0 Kimi K2.7 Code — $0.95/M, score 60.8 DeepSeek V4 Pro 0813 — $0.442/M, score 68.8 Z.ai: GLM 5.2 (batch) — $0.7/M, score 68.8 Google: Gemini 3.6 Flash (batch) — $0.375/M, score 69.2 Nemotron 3 Nano 30B A3B — $0.05/M, score 14.4 Z.ai: GLM 5.3 — $1.4/M, score 74.8 Anthropic: Claude Opus 4.8 — $5/M, score 74.3 SpaceXAI: Grok 4.3 — $1.25/M, score 42.2 GPT-5.6 Terra — $2/M, score 76.7 MiniMax: MiniMax M3 (batch) — $0.3/M, score 58.6 Gemini 3.8 Flash — $0.75/M, score 76.3 NVIDIA: Nemotron 3 Ultra (batch) — $0.6/M, score 49.3 MiniMax: MiniMax M2.7 — $0.3/M, score 52.6 SpaceXAI: Grok 4.5 — $2/M, score 72.4 Mistral: Mistral Medium 3.1 — $0.4/M, score 20.5 Kimi K2.6 — $0.95/M, score 61.8 Kimi K3 — $3/M, score 76.2 Anthropic: Claude Sonnet 4.5 (batch) — $1.5/M, score 52.1 OpenAI: gpt-oss-120b (batch) — $0.15/M, score 30.4 GPT-3.5-turbo — $0.5/M, score 10.7 Solar Pro 4 — $0.3/M, score 52.7 Gemini 3.1 Pro Preview — $2/M, score 68.8 Anthropic: Claude Opus 4.7 (batch) — $2.5/M, score 73.6 MiMo-V2.5-Pro — $0.435/M, score 60.2 Google: Gemini 3.1 Pro Preview (batch) — $1/M, score 68.8 Anthropic: Claude Sonnet 4.6 — $3/M, score 63.0 Anthropic: Claude Sonnet 4.6 (batch) — $1.5/M, score 63.0 DeepSeek V4 Flash 0731 — $0.05/M, score 69.1 GPT-5 — $1.25/M, score 37.8 OpenAI: GPT-5 Mini (batch) — $0.125/M, score 15.6 GPT-6 Astra — $10/M, score 76.9 inclusionAI: Ling 3.0 Flash — $0.021/M, score 50.6 Qwen: Qwen3.7 Plus — $0.32/M, score 55.9 SpaceXAI: Grok 4.3 (batch) — $1/M, score 42.2 GPT-4 Turbo — $10/M, score 21.5 Nemotron 3 Ultra 550B A55B — $0.5/M, score 49.3 Anthropic: Claude Sonnet 4 — $3/M, score 37.6 OpenAI: GPT-5.6 Luna (batch) — $0.1/M, score 71.4 OpenAI: GPT-5.4 Mini (batch) — $0.375/M, score 56.1 Qwen: Qwen3.8 2.4T A95B — $2/M, score 71.9 OpenAI: GPT-5.6 Sol (batch) — $1/M, score 77.4 Qwen: Qwen3.6 Plus — $0.325/M, score 54.5 Z.ai: GLM 4.6 — $0.43/M, score 45.8 GPT-4o (2024-05-13) — $5/M, score 24.2 Muse Spark 1.2 — $1.25/M, score 72.2 Qwen: Qwen3 Next 80B A3B Thinking — $0.15/M, score 17.4 GPT-5.6 Sol — $4/M, score 77.4 OpenAI: GPT-4 Turbo (batch) — $5/M, score 21.5 GPT-5.4 nano — $0.2/M, score 56.1 Nemotron 3.5 Lightning 30B A3B — $0.05/M, score 26.8 Hy3 preview — $0.066/M, score 58.8 GPT-4.1 nano — $0.1/M, score 11.1 GPT OSS 20B — $0.02/M, score 20.7 GPT-4o mini — $0.15/M, score 11.4 o1 — $15/M, score 39.7 Kimi K2.5 — $0.3/M, score 46.8 Trinity Large Thinking — $0.25/M, score 25.8 Qwen: Qwen3 30B A3B Thinking 2507 — $0.2/M, score 12.1 Z.ai: GLM 5.2 — $0.966/M, score 68.8 Qwen: Qwen3.7 Max — $1.475/M, score 66.0 SpaceXAI: Grok Build 0.1 — $1/M, score 51.5 Qwen: Qwen3.6 35B A3B — $0.1/M, score 41.9 Qwen: Qwen3.6 27B — $0.3/M, score 53.7 inclusionAI: Ling-2.6-flash — $0.01/M, score 25.3 Mistral: Mistral Small 4 — $0.15/M, score 26.6 Kwaipilot: KAT-Coder-Pro V2 — $0.3/M, score 59.5 Qwen: Qwen3.5-9B — $0.1/M, score 28.7 Qwen: Qwen3.5-35B-A3B — $0.312/M, score 37.0 Qwen: Qwen3.5-122B-A10B — $0.26/M, score 45.7 Upstage: Solar Pro 3 — $0.15/M, score 16.2 Z.ai: GLM 4.7 — $0.4/M, score 45.3 Amazon: Nova 2 Lite — $0.3/M, score 23.0 Mistral: Ministral 3 3B 2512 — $0.1/M, score 4.8 Google: Gemma 3n 4B — $0.06/M, score 3.2 Meta: Llama 4 Maverick — $0.2/M, score 16.3 OpenAI: o3 Mini High — $1.1/M, score 16.3 Meta: Llama 3.3 70B Instruct — $0.1/M, score 11.9 Meta: Llama 3.1 8B Instruct — $0.05/M, score 5.4 Nex AGI: Nex-N2-Pro — $0.25/M, score 59.1 Mistral: Mistral Medium 3.5 — $1.5/M, score 46.9 inclusionAI: Ring-2.6-1T — $0.075/M, score 42.8 IBM: Granite 4.1 8B — $0.05/M, score 9.5 Qwen: Qwen3.5 397B A17B — $0.55/M, score 48.2 Qwen: Qwen3 Coder Next — $0.12/M, score 36.2 Mistral: Ministral 3 14B 2512 — $0.2/M, score 14.4 Anthropic: Claude Haiku 4.5 — $1/M, score 43.9 Qwen: Qwen3 235B A22B Thinking 2507 — $0.23/M, score 22.1 Qwen: Qwen3 8B — $0.117/M, score 9.0 Qwen: Qwen3 14B — $0.227/M, score 13.8 Qwen: Qwen3 32B — $0.08/M, score 15.3 Meta: Llama 4 Scout — $0.1/M, score 8.2 DeepSeek: DeepSeek V3 0324 — $0.25/M, score 21.2 Cohere: Command A — $2.5/M, score 27.8 Step 3.7 Flash — $0.185/M, score 39.6 DeepSeek V4 Pro Input price per million tokens (log scale) Index

The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.

Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
Context Length1024000 → 1000000
Price Completion1.909476 → 0.87
Price Prompt0.954738 → 0.435
Context Length1000000 → 1024000
Price Completion0.87 → 1.909476
Price Prompt0.435 → 0.954738
Context Length1024000 → 1000000
Price Completion1.909476 → 0.87
Price Prompt0.954738 → 0.435
Context Length1000000 → 1024000
Provenance

Sources & verification

Every figure on this page traces back to one of these records.

Methodology
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/deepseek/deepseek-v4-pro
curl
curl "https://model.kyssta.lol/api/v1/models/deepseek/deepseek-v4-pro"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is DeepSeek V4 Pro?

Open MoE flagship with million-token context for coding and long agent runs. It is published by DeepSeek and catalogued here from Models.dev.

How much does DeepSeek V4 Pro cost?

Listed input pricing starts at $0.348 per million tokens from routing.run, rising to $2.4 across 53 listed providers.

What is the context length of DeepSeek V4 Pro?

DeepSeek V4 Pro accepts up to 1M tokens of context and returns up to 384K output tokens.

Does DeepSeek V4 Pro support tool calling and structured output?

Provider catalogs list support for tool calling, structured output, and reasoning.

Which providers serve DeepSeek V4 Pro?

62 providers list this model: AnyAPI, Model Oracle AI, Alibaba Token Plan (China), Volcengine Ark Coding Plan, SenseNova (China), UnoRouter and 56 more.

Are the weights for DeepSeek V4 Pro open?

Yes. The weights are published and downloadable from Hugging Face.

When was DeepSeek V4 Pro released?

The catalog records a release date of 2026-04-24, last verified Sep 11, 2026.

More models from DeepSeek

View all →