DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

XiaomiMiMo logo

XiaomiMiMo/

MiMo-V2.5

$0.133

in

$0.266

out

$0.003

cached

/ 1M tokens

$0.14

in

$0.28

out

$0.003

cached

/ 1M tokens

5% off
TierInputOutputCached input
Priority (1.5×)Learn More
$0.1995$0.399$0.004
Flex (0.8×)Learn More
$0.1064$0.2128$0.0021

per 1M tokens — promotional pricing, 5% off already applied

MiMo-V2.5 is a native omnimodal model with strong agentic capabilities, supporting text, image, video, and audio understanding within a unified architecture. Built upon the MiMo-V2-Flash backbone and extended with dedicated vision and audio encoders, it delivers robust performance across multimodal perception, long-context reasoning, and agentic workflows.

Deploy Private Endpoint
Supports Priority Tier
Supports Flex Tier
Public
Zero retention
fp8
262,144
JSON
Function
Multimodal
XiaomiMiMo/MiMo-V2.5 cover image