DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

deepseek-ai logo

deepseek-ai/

DeepSeek-V4.1-Flash

$0.14

in

$0.42

out

$0.004

cached

/ 1M tokens

$0.20

in

$0.60

out

$0.006

cached

/ 1M tokens

30% off
TierInputOutputCached input
Priority (1.5×)Learn More
$0.21$0.63$0.0063
Flex (0.8×)Learn More
$0.112$0.336$0.0034

per 1M tokens — promotional pricing, 30% off already applied

Introduce DeepSeek-V4.1-Flash, a multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens.

Deploy Private Endpoint
Supports Priority Tier
Supports Flex Tier
Public
Zero retention
fp8
1,048,576
JSON
Function
Multimodal
deepseek-ai/DeepSeek-V4.1-Flash cover image