We use essential cookies to make our site work. With your consent, we may also use non-essential cookies to improve user experience and analyze website traffic…

DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

Qwen logo

Qwen/

Qwen3.8-2.4T-A95B

$1.80

in

$5.40

out

$0.18

cached

/ 1M tokens

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total. It is suited for coding, research, complex reasoning, and agentic workflows.

Deploy Private Endpoint
Public
fp4
262,144
JSON
Function
Qwen/Qwen3.8-2.4T-A95B cover image
demoapi

LxrYQARc

2026-08-12T18:24:34+00:00