Qwen3.6 Flash
$1.00
estimated usage cost
1,000 requests × (1,000 uncached input × $0.25 + 500 output × $1.50) / 1,000,000 = $1.00
base rates: up to 262,144 input tokens per request, inclusive.
Text generation / Qwen
Qwen3.6 Flash is a text generation model from Qwen, available through the Impossibl API. Its published context window is 1,048,576 tokens. The advertised input rate is $0.25 per million input tokens. The advertised output rate is $1.50 per million output tokens. Additional pricing brackets apply when the input exceeds a published context threshold.
api id: qwen/qwen3.6-flash
published base api rates · usd per 1m tokens
| total input | input | output | cache read |
|---|---|---|---|
| over 262,144 tokens | $1.00 | $4.00 | $0.20 |
crossing a threshold applies that bracket to the whole request. cache creation can have a separate charge.
rates are published by the impossibl api. balance-purchase fees, taxes, tool charges, and workspace-specific pricing are separate.
unknown means the public catalog does not establish that fact. a model's documented capabilities can differ from what a particular api supports.
estimate text token usage for your workload. all amounts are in usd.
$1.00
estimated usage cost
1,000 requests × (1,000 uncached input × $0.25 + 500 output × $1.50) / 1,000,000 = $1.00
base rates: up to 262,144 input tokens per request, inclusive.
estimates use published impossibl api usage rates. configured billing adjustments, funding fees, taxes, and custom workspace rates are excluded and may change the final amount.
image and audio usage, cache creation, and additional tool charges are excluded. reasoning tokens count as billable output, even when they are not visible in the response.
Benchmark results are not available for this model yet.
Qwen3.6 Flash is a text-generation model from Qwen, available through the Impossibl API.
Qwen3.6 Flash costs $0.25 per million input tokens and $1.50 per million output tokens through the Impossibl API. Cache reads cost $0.05 per million tokens. These base rates apply through 262,144 total input tokens per request. Above 262,144 input tokens, the long-context rates in the pricing table apply to the whole request. All rates are in USD.
Qwen3.6 Flash has a 1,048,576-token context window through the Impossibl API.
Qwen3.6 Flash accepts text as input. Qwen3.6 Flash returns text.
Qwen3.6 Flash is available through the Impossibl API. Use qwen/qwen3.6-flash as the model identifier with /v1/chat/completions. The quickstart explains how to create an API key and send your first request.
Qwen released Qwen3.6 Flash on April 16, 2026.
this page describes Qwen3.6 Flash as listed by the impossibl public api. prices and availability come from that catalog, refreshed here every five minutes. this is an impossibl offering, not a comparison of independent hosting-provider quotes.
model descriptions and release dates, where available, use the creator sources linked in the faq. api prices and limits use the impossibl catalog. benchmark results, where available, are credited to Artificial Analysis in the benchmarks section. report a correction →