DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

deepseek-ai logo

deepseek-ai/

DeepSeek-V4-Flash-Vision-Exp

$0.216

in

$0.647

out

$0.069

cached

/ 1M tokens

$0.44

in

$1.32

out

$0.14

cached

/ 1M tokens

51% off
TierInputOutputCached input
Priority (1.5×)Learn More
$0.3234$0.9702$0.1029
Flex (0.8×)Learn More
$0.1725$0.5174$0.0549

per 1M tokens — promotional pricing, 51% off already applied

DeepSeek-V4-Flash-Vision-Exp is DeepSeek's experimental multimodal model in the V4-Flash family, adding visual understanding to the V4-Flash architecture. It serves a 1M-token (1,048,576) context window and supports image input with visual grounding, tool calling, structured/JSON output, and configurable reasoning effort (low/high/max, or disabled).

Deploy Private Endpoint
Supports Priority Tier
Supports Flex Tier
Public
fp8
1,048,576
JSON
Function
Multimodal
deepseek-ai/DeepSeek-V4-Flash-Vision-Exp cover image
api
deepseek-ai/DeepSeek-V4-Flash-Vision-Exp cover image
DeepSeek V4 Flash Vision Exp

Ask me anything

0.00s

You need to log in to use this model

Log In

Settings