DeepInfra raises $107M Series B to scale the inference cloud — read the announcement
deepseek-ai/
$0.14
in
$0.42
out
$0.004
cached
/ 1M tokens
$0.20
in
$0.60
out
$0.006
cached
/ 1M tokens
| Tier | Input | Output | Cached input |
|---|---|---|---|
Priority (1.5×)Learn More | $0.21 | $0.63 | $0.0063 |
Flex (0.8×)Learn More | $0.112 | $0.336 | $0.0034 |
per 1M tokens — promotional pricing, 30% off already applied
Introduce DeepSeek-V4.1-Flash, a multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens.

© 2026 DeepInfra. All rights reserved.