I

Ling 3.1 Flash

InclusionAIReleased Oct 1, 2026

Ling 3.1 Flash is not ranked on LMArena yet; output price $0.90 per 1M tokens.

Key metrics

Arena score–Not yet rated
Intelligence41.1Variant: default
Coding–Artificial Analysis
List price in / out$0.30 / $0.90USD per 1M tokens
Lowest route$0 / $0100% below list price
Context262Ktokens
Output speed215tok/sAA measured median
TTFT / first answer1.1 / 10sReasoning model: includes thinking

Pricing

  • Official API list price: $0.30 input and $0.90 output per 1M tokens (USD).
  • At a typical 3:1 input-to-output ratio, the blended cost is about $0.45 per 1M tokens.
  • Among 78 models with an output price, it is the #12 cheapest.
  • The cheapest OpenRouter route charges $0 input / $0 output. That route is over 60% below list price — usually a promotion or a single provider, so check its reliability.

Sources & methodology

Ling 3.1 Flash (InclusionAI): Intelligence Index 41.1, $0.30 in / $0.90 out per 1M tokens, 262K context. Data updated Oct 9, 2026.

  • LMArena — Arena score, 95% confidence interval, votes and rank.
  • Artificial Analysis — Intelligence / Coding Index, output speed and latency (medians).
  • OpenRouter — Route prices, context length and listing date.

Arena scores use the Bradley-Terry model (not Elo); ± is the half-width of the 95% confidence interval. Latency figures are Artificial Analysis medians; for reasoning models, time to first token and to first answer include the thinking phase. Data syncs daily at 14:20 (UTC+8).

FAQ

What rank is Ling 3.1 Flash right now?

Ling 3.1 Flash has not been scored by LMArena yet, so it is not ranked here.

What does Ling 3.1 Flash cost via API?

Input $0.30 / output $0.90 per 1M tokens (official API list price; the OpenRouter quote when no official price is available). At a typical 3:1 input-to-output ratio, the blended cost is about $0.45 per 1M tokens.

How fast is Ling 3.1 Flash?

Artificial Analysis measures about 215 output tokens per second, 1.1 s to the first token and 10 s to the first answer token; it is a reasoning model, so latency includes thinking time.

Is Ling 3.1 Flash open source? Can it be self-hosted?

Ling 3.1 Flash is a closed model, available only via official APIs or hosted services.

What is the context length of Ling 3.1 Flash?

262K (262,144 tokens, as listed on OpenRouter).

Related leaderboards