DeepInfra raises $107M Series B to scale the inference cloud — read the announcement
deepseek-ai/
$0.30
in
$1.20
out
$0.06
cached
/ 1M tokens
| Tier | Input | Output | Cached input |
|---|---|---|---|
Priority (1.5×)Learn More | $0.45 | $1.80 | $0.09 |
Flex (0.8×)Learn More | $0.24 | $0.96 | $0.048 |
per 1M tokens
Introduce DeepSeek-V4.1-Flash, a multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens.

© 2026 DeepInfra. All rights reserved.