Qwen3.8-Flash-Next

AlibabaReleased Aug 26, 2026

Qwen3.8-Flash-Next is not ranked on LMArena yet; output price $0.47 per 1M tokens.

Key metrics

Arena score–Not yet rated
Intelligence39.8Variant: default
Coding73.1Artificial Analysis
List price in / out$0.15 / $0.47USD per 1M tokens
Lowest route–Cheapest OpenRouter route
Context–tokens
Output speed53tok/sAA measured median
TTFT / first answer1.4 / 39sReasoning model: includes thinking

More from Alibaba

#ModelArenaIntel.Output $tok/s
12Qwen3.8 Max1483.345.4$6.0036
19Qwen3.7 Max1475.729.5$7.50–
38Qwen3.7 Plus1454.825.2$1.6055
43Qwen3.6 Max1446.928.4$7.80–
51Qwen3.8 27B1441.133.7$3.0047
65Qwen3 Max1412.815.6$6.00–

Pricing

  • Official API list price: $0.15 input and $0.47 output per 1M tokens (USD).
  • At a typical 3:1 input-to-output ratio, the blended cost is about $0.23 per 1M tokens.
  • Among 78 models with an output price, it is the #4 cheapest.

Sources & methodology

Qwen3.8-Flash-Next (Alibaba): Intelligence Index 39.8, $0.15 in / $0.47 out per 1M tokens. Data updated Oct 9, 2026.

  • LMArena — Arena score, 95% confidence interval, votes and rank.
  • Artificial Analysis — Intelligence / Coding Index, output speed and latency (medians).
  • OpenRouter — Route prices, context length and listing date.

Arena scores use the Bradley-Terry model (not Elo); ± is the half-width of the 95% confidence interval. Latency figures are Artificial Analysis medians; for reasoning models, time to first token and to first answer include the thinking phase. Data syncs daily at 14:20 (UTC+8).

FAQ

What rank is Qwen3.8-Flash-Next right now?

Qwen3.8-Flash-Next has not been scored by LMArena yet, so it is not ranked here.

What does Qwen3.8-Flash-Next cost via API?

Input $0.15 / output $0.47 per 1M tokens (official API list price; the OpenRouter quote when no official price is available). At a typical 3:1 input-to-output ratio, the blended cost is about $0.23 per 1M tokens.

How fast is Qwen3.8-Flash-Next?

Artificial Analysis measures about 53 output tokens per second, 1.4 s to the first token and 39 s to the first answer token; it is a reasoning model, so latency includes thinking time.

Is Qwen3.8-Flash-Next open source? Can it be self-hosted?

Qwen3.8-Flash-Next is a closed model, available only via official APIs or hosted services.

Related leaderboards