We use essential cookies to make our site work. With your consent, we may also use non-essential cookies to improve user experience and analyze website traffic…

DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

inclusionAI/

Ling-3.0-flash

$0.075

in

$0.22

out

$0.015

cached

/ 1M tokens

The model prioritizes token efficiency and agentic inference at production scale, stretching what developers can achieve within limited token, latency, and serving-cost budgets.

inclusionAI/Ling-3.0-flash cover image