We use essential cookies to make our site work. With your consent, we may also use non-essential cookies to improve user experience and analyze website traffic…

DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

deepseek-ai logo

deepseek-ai/

DeepSeek-V4-Pro-0813

$1.30

in

$2.60

out

$0.10

cached

/ 1M tokens

TierInputOutputCached input
Priority (1.5×)Learn More
$1.95$3.90$0.15
Flex (0.8×)Learn More
$1.04$2.08$0.08

per 1M tokens

Prompt cache retentionLearn More
Cache write
Retain for 5m (1.25×)
$1.625
Retain for 1h (2×)
$2.60

per 1M tokens

Retained in whole blocks of 1,024 tokens; the remainder is billed as standard input. Reuse within the window bills at the cached input rate.

**Shown at the standard tier. Priority and Flex scale these rates the same way they scale input and output.

DeepSeek-V4-Pro-0813 is the official release of DeepSeek-V4-Pro, superseding the preview version, with greatly enhanced agentic capabilities and performance improvements that are especially pronounced in production environments. It is built on the DeepSeek-V4-Pro (Preview) model structure, with a DSpark speculative decoding module attached.

Deploy Private Endpoint
Supports Priority Tier
Supports Flex Tier
Public
fp4
1,048,576
JSON
Function
deepseek-ai/DeepSeek-V4-Pro-0813 cover image
demoapi

i3gfj82o

2026-08-14T21:31:04+00:00