Get the app

Qwen3.8 Flash: price, context window and capabilities

Qwen3.8 Flash is a model from Qwen (Alibaba) (qwen/qwen3.8-flash). It supports reasoning. It supports tool use. It costs $0.150 per million input tokens and $0.470 per million output tokens, with a 1M-token context window.

$0.150input / 1M tokens
$0.470output / 1M tokens
1Mcontext window
text+image+videoaccepts

What it costs in practice

Using the listed prices: a chat turn of 2,000 input and 500 output tokens costs about $0.0005, so 10,000 such turns cost about $5.35. Summarising a 100,000-token document into 2,000 tokens costs about $0.016. A long agent run that reads 1 million tokens and writes 50,000 costs about $0.173.

How it compares

Among the 40 models we track, Qwen3.8 Flash ranks #7 by input price and #5 by output price (1 is cheapest), and #19 by context window (1 is largest). Models with the same tool and reasoning support that cost less per output token: DeepSeek V4 Pro 0423 ($0.418). Models with a larger context window at the same or lower output price: Llama 4 Scout (1.3M), DeepSeek V4 Pro 0423 (1.0M).

More from Qwen (Alibaba)

Prices and limits come from OpenRouter's public model catalogue, last synced 4 Oct 2026. Provider pricing can differ from OpenRouter's and changes often; confirm on the provider's own page before you commit. See the guide to choosing a model or the full comparison.

The daily AI brief, on your phone.

Feed, daily deep-dive and bytes — readable offline, with push alerts for the topics you follow.

Get it on Google Play