Qwen3.8 2.4T A95B

AlibabaOpen weightsReleased Aug 12, 2026

Qwen3.8 2.4T A95B is not ranked on LMArena yet; output price $6.00 per 1M tokens.

Key metrics

Arena score–Not yet rated
Intelligence39.9Variant: default
Coding71.9Artificial Analysis
List price in / out$2.00 / $6.00USD per 1M tokens
Lowest route$2.00 / $6.00Cheapest OpenRouter route
Context1.05Mtokens
Output speed36tok/sAA measured median
TTFT / first answer1.8 / 58sReasoning model: includes thinking

More from Alibaba

#ModelArenaIntel.Output $tok/s
12Qwen3.8 Max1483.345.4$6.0036
19Qwen3.7 Max1475.729.5$7.50–
38Qwen3.7 Plus1454.825.2$1.6055
43Qwen3.6 Max1446.928.4$7.80–
51Qwen3.8 27B1441.133.7$3.0047
65Qwen3 Max1412.815.6$6.00–

Pricing

  • Official API list price: $2.00 input and $6.00 output per 1M tokens (USD).
  • At a typical 3:1 input-to-output ratio, the blended cost is about $3.00 per 1M tokens.
  • Among 78 models with an output price, it is the #40 cheapest.
  • The cheapest OpenRouter route charges $2.00 input / $6.00 output.

Sources & methodology

Qwen3.8 2.4T A95B (Alibaba): Intelligence Index 39.9, $2.00 in / $6.00 out per 1M tokens, 1.05M context. Data updated Oct 9, 2026.

  • LMArena — Arena score, 95% confidence interval, votes and rank.
  • Artificial Analysis — Intelligence / Coding Index, output speed and latency (medians).
  • OpenRouter — Route prices, context length and listing date.

Arena scores use the Bradley-Terry model (not Elo); ± is the half-width of the 95% confidence interval. Latency figures are Artificial Analysis medians; for reasoning models, time to first token and to first answer include the thinking phase. Data syncs daily at 14:20 (UTC+8).

FAQ

What rank is Qwen3.8 2.4T A95B right now?

Qwen3.8 2.4T A95B has not been scored by LMArena yet, so it is not ranked here.

What does Qwen3.8 2.4T A95B cost via API?

Input $2.00 / output $6.00 per 1M tokens (official API list price; the OpenRouter quote when no official price is available). At a typical 3:1 input-to-output ratio, the blended cost is about $3.00 per 1M tokens.

How fast is Qwen3.8 2.4T A95B?

Artificial Analysis measures about 36 output tokens per second, 1.8 s to the first token and 58 s to the first answer token; it is a reasoning model, so latency includes thinking time.

Is Qwen3.8 2.4T A95B open source? Can it be self-hosted?

Qwen3.8 2.4T A95B is an open-weights model (license: —); you can download the weights and self-host.

What is the context length of Qwen3.8 2.4T A95B?

1.05M (1,048,576 tokens, as listed on OpenRouter).

Related leaderboards