Skip to content

Qwen 3.8 Flash

Alibaba · released 2026 · proprietary API

Alibaba's Flash tier for high-volume work, on the same multimodal API as the Max rows.

Context window

1M

1,000,000 tokens

Input price

$0.15

per 1M tokens

Output price

$0.47

per 1M tokens

Blended

$0.23

3:1 in:out mix

Weights

Proprietary

API only

Released

2026

Alibaba

Long-context / tier pricing — International pricing; the Global region bills $0.113/$0.382 for the same model. Batch inference is 50% off, and a context-caching discount applies separately. Listed prices are the base rate for prompts up to ~200K tokens.

Where it ranks

Of 42 models tracked.

Cost to run

#3(tied)

of 42

Input price

#3(tied)

of 42

Context window

#10(tied)

of 42

See the full Frontier Index →

Compare with

The nearest models by blended cost to run.

Put it to work

Specs from BitByteCore's hand-verified Frontier Ledger (v2026-10-01), this row verified 22 Sep 2026 against Alibaba's official pricing ↗. Model pricing moves fast, so confirm with Alibaba before you rely on a figure.