DeepSeek V4.1 Flash
$0.90
estimated usage cost
1,000 requests × (1,000 uncached input × $0.30 + 500 output × $1.20) / 1,000,000 = $0.90
base rates; no context pricing tiers are published.
the model directory
compare api pricing, context windows, and capabilities side by side.
use the same workload to understand the cost before choosing a model.
prices are published impossibl api usage rates in usd. compare the billing units and context tiers alongside each price. unpublished capabilities are marked explicitly.
| specification | DeepSeek V4.1 Flash deepseek/deepseek-v4.1-flash | OpenAI o3-mini openai/o3-mini |
|---|---|---|
| creator | DeepSeek | OpenAI |
| model type | Text generation | Text generation |
| input modalities | text, image | text |
| output modalities | text | text |
| base input price | $0.30 / 1m tokens | $1.10 / 1m tokens |
| base output price | $1.20 / 1m tokens | $4.40 / 1m tokens |
| cached input price | $0.006 / 1m tokens | $0.55 / 1m tokens |
| context window | 1.05m tokens | 200k tokens |
| tool calling | not published | not published |
| context pricing tiers | none published | none published |
| endpoints |
|
|
| retirement | not announced | not announced |
estimate text token usage across the same workload. all amounts are in usd.
$0.90
estimated usage cost
1,000 requests × (1,000 uncached input × $0.30 + 500 output × $1.20) / 1,000,000 = $0.90
base rates; no context pricing tiers are published.
$3.30
estimated usage cost
1,000 requests × (1,000 uncached input × $1.10 + 500 output × $4.40) / 1,000,000 = $3.30
base rates; no context pricing tiers are published.
estimates use published impossibl api usage rates. configured billing adjustments, funding fees, taxes, and custom workspace rates are excluded and may change the final amount.
image and audio usage, cache creation, and additional tool charges are excluded. reasoning tokens count as billable output, even when they are not visible in the response.
DeepSeek V4.1 Flash and OpenAI o3-mini are available through the Impossibl API. They share the /v1/chat/completions endpoint. Set the model identifier to select the model for a supported request.
DeepSeek V4.1 Flash is created by DeepSeek. It has a 1,048,576-token context window. Base API pricing (USD): $0.30 per million input tokens; $1.20 per million output tokens.
OpenAI o3-mini is created by OpenAI. It has a 200,000-token context window. Base API pricing (USD): $1.10 per million input tokens; $4.40 per million output tokens.