DeepInfra raises $107M Series B to scale the inference cloud — read the announcement
Qwen/
$2.00
in
$6.00
out
$0.20
cached
/ 1M tokens
| Tier | Input | Output | Cached input |
|---|---|---|---|
Flex (0.8×)Learn More | $1.60 | $4.80 | $0.16 |
per 1M tokens
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total. It is suited for coding, research, complex reasoning, and agentic workflows.

LxrYQARc
2026-08-12T18:24:34+00:00
© 2026 DeepInfra. All rights reserved.