DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

deepseek-ai logo

deepseek-ai/

DeepSeek-V4-Flash-Vision-Exp

$0.216

in

$0.647

out

$0.007

cached

/ 1M tokens

$0.44

in

$1.32

out

$0.014

cached

/ 1M tokens

51% off
TierInputOutputCached input
Priority (1.5×)Learn More
$0.3234$0.9702$0.0103
Flex (0.8×)Learn More
$0.1725$0.5174$0.0055

per 1M tokens — promotional pricing, 51% off already applied

DeepSeek-V4-Flash-Vision-Exp is DeepSeek's experimental multimodal model in the V4-Flash family, adding visual understanding to the V4-Flash architecture. It serves a 1M-token (1,048,576) context window and supports image input with visual grounding, tool calling, structured/JSON output, and configurable reasoning effort (low/high/max, or disabled).

Deploy Private Endpoint
Supports Priority Tier
Supports Flex Tier
Public
Zero retention
fp8
1,048,576
JSON
Function
Multimodal
deepseek-ai/DeepSeek-V4-Flash-Vision-Exp cover image