DeepSeek V3.2 Speciale vs Qwen3.7 Flash

rates, context and capabilities side by side on the impossibl api, and what the same workload costs on each.

the models at a glance: base rates, context window and modalities
specificationDeepSeek V3.2 Speciale

deepseek/deepseek-v3-2-speciale

Qwen3.7 Flash

qwen/qwen3.7-flash

input · usd / 1m tok$0.58$0.03
output · usd / 1m tok$1.68$0.13
context window163.84ktokens1mtokens
modalities · in → out

takes text in, gives text out

takes text and image in, gives text out

pricing

base api rates by model
base api rates · usd per 1m tokensDeepSeek V3.2 Speciale

deepseek/deepseek-v3-2-speciale

Qwen3.7 Flash

qwen/qwen3.7-flash

input$0.58$0.03over 32k · $0.10over 256k · $0.20
output$1.68$0.13over 32k · $0.40over 256k · $0.80
cache readnot published$0.006over 32k · $0.02over 256k · $0.04
cache write · default ttlnot published$0.038over 32k · $0.125over 256k · $0.25

brackets are by input tokens per request: a request over a threshold is billed at that bracket's rates for all of its tokens. cache creation can have a separate charge. the estimate below prices one workload at these rates on every model.

specifications

published specifications by model
as publishedDeepSeek V3.2 Speciale

deepseek/deepseek-v3-2-speciale

Qwen3.7 Flash

qwen/qwen3.7-flash

creatorDeepSeekQwen
model typeText generationText generation
context window163,840tokens1,000,000tokens
input

takes text in, gives text out

takes text and image in, gives text out

outputtexttext
tool callingnot verifiednot verified
releaseddec 1, 2025not published
endpoint/v1/chat/completions/v1/chat/completions

unknown means the public catalog does not establish that fact. a model's documented capabilities can differ from what a particular api supports. see the model documentation for integration details.

benchmarks

benchmark scores by model, primary tested configuration
primary configurationDeepSeek V3.2 Speciale

deepseek/deepseek-v3-2-speciale

Qwen3.7 Flash

qwen/qwen3.7-flash

no benchmarks yet

intelligence index · gateway rank14.5#71 of 90not measured
omniscience-17.7#63 of 86not measured
gpqa87.1%#44 of 88
not measured
humanity’s last exam28.7%#45 of 89
not measured
aime 202596.7%
not measured
livecodebench89.6%
not measured
aa-lcr70.0%
not measured
ifbench63.9%
not measured
terminal-bench hard34.8%
not measured
critpt7.4%
not measured
tau20.0%
not measured

DeepSeek V3.2 Speciale source ↗

each column describes the tested configuration named under the model; performance and costs can differ on impossibl. a score one model was not measured on is left blank rather than scored zero. indices are ranked against every model on the gateway with one. last checked 2026-09-11.

estimate your cost

one token workload, priced on every model at the published rates, in usd. change the numbers to match yours.

token workload

total input, including cached tokens

include billable reasoning tokens

the same token usage in each request

a subset of total input, never additional tokens

DeepSeek V3.2 Speciale

deepseek/deepseek-v3-2-speciale

estimated usage cost

$1.42

1,000 requests × (1,000 uncached input × $0.58 + 500 output × $1.68) / 1,000,000 = $1.42

base rates; no context pricing tiers are published.

Qwen3.7 Flash

qwen/qwen3.7-flash

estimated usage cost

$0.095

1,000 requests × (1,000 uncached input × $0.03 + 500 output × $0.13) / 1,000,000 = $0.095

base rates: up to 32,000 input tokens per request, inclusive.

estimates use published impossibl api usage rates. configured billing adjustments, funding fees, taxes, and custom workspace rates are excluded and may change the final amount.

image and audio usage, cache creation, and additional tool charges are excluded. reasoning tokens count as billable output, even when they are not visible in the response.

summary

DeepSeek V3.2 Speciale and Qwen3.7 Flash are available through the Impossibl API. they share the /v1/chat/completions endpoint; the model identifier selects the model for a request.

DeepSeek V3.2 Speciale is created by DeepSeek. It has a 163,840-token context window. Base API pricing (USD): $0.58 per million input tokens; $1.68 per million output tokens.

Qwen3.7 Flash is created by Qwen. It has a 1,000,000-token context window. Base API pricing (USD): $0.03 per million input tokens; $0.13 per million output tokens. Above 32,000 input tokens per request, context-tier rates apply to the whole request.

DeepSeek V3.2 SpecialeQwen3.7 Flash

this page compares DeepSeek V3.2 Speciale vs Qwen3.7 Flash as listed by the impossibl public api. prices and availability come from that catalog, refreshed here every five minutes. this is an impossibl offering, not a comparison of independent hosting-provider quotes.

view the source catalog ↗comparison as markdown ↗report a correction →