DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

deepseek-ai logo

deepseek-ai/

DeepSeek-V4.1-Flash

$0.30

in

$1.20

out

$0.06

cached

/ 1M tokens

TierInputOutputCached input
Priority (1.5×)Learn More
$0.45$1.80$0.09
Flex (0.8×)Learn More
$0.24$0.96$0.048

per 1M tokens

Introduce DeepSeek-V4.1-Flash, a multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens.

Deploy Private Endpoint
Supports Priority Tier
Supports Flex Tier
Public
fp8
1,048,576
JSON
Function
Multimodal
deepseek-ai/DeepSeek-V4.1-Flash cover image