DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

XiaomiMiMo/
$0.133
in
$0.266
out
$0.003
cached
/ 1M tokens
$0.14
in
$0.28
out
$0.003
cached
/ 1M tokens
| Tier | Input | Output | Cached input |
|---|---|---|---|
Priority (1.5×)Learn More | $0.1995 | $0.399 | $0.004 |
Flex (0.8×)Learn More | $0.1064 | $0.2128 | $0.0021 |
per 1M tokens — promotional pricing, 5% off already applied
MiMo-V2.5 is a native omnimodal model with strong agentic capabilities, supporting text, image, video, and audio understanding within a unified architecture. Built upon the MiMo-V2-Flash backbone and extended with dedicated vision and audio encoders, it delivers robust performance across multimodal perception, long-context reasoning, and agentic workflows.

© 2026 DeepInfra. All rights reserved.