ColinBuilds.com

ColinBuilds.com · Current official DeepSeek API pricing

DeepSeek-V4-Pro

Input / 1M
$0.435
Output / 1M
$0.87
Current official context window
1,000,000 tokens

Current official DeepSeek API pricing for deepseek-v4-pro; calculator uses cache-miss input ($0.435/1M). Cache-hit input is listed separately at $0.003625/1M. · as of July 7, 2026

Compare models

Same per-1M input/output rates as above. Enter your monthly token volumes to compare bills.

DeepSeek-V4-Pro

Input: 1M × $0.435 = $0.435
Output: 1M × $0.87 = $0.87
Total
$1.31

GPT-4o

Input: 1M × $2.50 = $2.50
Output: 1M × $10.00 = $10.00
Total
$12.50

Difference

$11.20

GPT-4o costs $11.20 more than DeepSeek-V4-Pro.

Benchmark Matrix

Reported model metrics from published sources.

Active Parameters

49B

Reported active parameter count at inference time.

Popular comparisons

30 total featuring DeepSeek-V4-Pro.

Open compare hub →

DeepSeek-V4-Pro is a current official DeepSeek API route for the V4 family, with a 1M context window, 384K maximum output, and official pricing that separates cache-hit and cache-miss input tokens.

What this model is

DeepSeek-V4-Pro is the current official DeepSeek API model ID from the DeepSeek-V4 Preview release. The exact API model ID is deepseek-v4-pro, served from https://api.deepseek.com for OpenAI-format calls and https://api.deepseek.com/anthropic for Anthropic-format calls.

DeepSeek’s V4 Preview announcement and model card source the open-weight architecture and parameter claims. DeepSeek’s API docs source the hosted API ID, pricing, context, and max-output claims. The current pricing table does not split prices by thinking mode, so this page does not invent a separate thinking price.

Pricing notes

DeepSeek lists prices per 1M tokens. For DeepSeek-V4-Pro, the official pricing page lists $0.003625 per 1M input tokens on a cache hit, $0.435 per 1M input tokens on a cache miss, and $0.87 per 1M output tokens.

The calculator on this page uses the cache-miss input price as the default input rate, because that is the safer comparison when you do not yet know whether prompts will hit cache. Cache-hit pricing is still important for repeated-context workloads, where the effective input cost can be much lower.

DeepSeek also warns that product prices may vary and recommends checking the pricing page regularly. Treat this as a current checked price, not a permanent guarantee.

Benchmarks and specs

DeepSeek’s official pricing table lists DeepSeek-V4-Pro with a 1M context length and a maximum output of 384K tokens. The same table lists JSON Output, Tool Calls, Chat Prefix Completion, and FIM Completion in non-thinking mode.

The V4 preview release states that DeepSeek-V4-Pro has 1.6T total parameters and 49B activated parameters, and that the V4 series uses open weights under the MIT license. This page does not list benchmark scores until a named evaluation source and test setup are verified.

Best fit

DeepSeek-V4-Pro is best for builders who want the stronger current DeepSeek API route, especially where long context, tool use, and reasoning mode matter more than the lowest possible token price.

It is also a useful comparison point against V4-Flash: both share the same official context headline, but V4-Pro is the larger route and has higher official token prices.

Sources