google/gemini-3.1-pro-preview
estimated usage cost
$8.00
1,000 requests × (1,000 uncached input × $2.00 + 500 output × $12.00) / 1,000,000 = $8.00
base rates: up to 200,000 input tokens per request, inclusive.
rates, context and capabilities side by side on the impossibl api, and what the same workload costs on each.
| specification | Gemini 3.1 Pro (preview) google/gemini-3.1-pro-preview | MiMo V2.5 Pro xiaomi/mimo-v2.5-pro |
|---|---|---|
| input · usd / 1m tok | $2.00 | $0.435 |
| output · usd / 1m tok | $12.00 | $0.87 |
| context window | 1mtokens | 1.02mtokens |
| modalities · in → out | takes text, image, audio and video in, gives text out | takes text in, gives text out |
| base api rates · usd per 1m tokens | Gemini 3.1 Pro (preview) google/gemini-3.1-pro-preview | MiMo V2.5 Pro xiaomi/mimo-v2.5-pro |
|---|---|---|
| input | $2.00over 200k · $4.00 | $0.435 |
| output | $12.00over 200k · $18.00 | $0.87 |
| cache read | $0.20over 200k · $0.40 | $0.004 |
brackets are by input tokens per request: a request over a threshold is billed at that bracket's rates for all of its tokens. cache creation can have a separate charge. the estimate below prices one workload at these rates on every model.
| as published | Gemini 3.1 Pro (preview) google/gemini-3.1-pro-preview | MiMo V2.5 Pro xiaomi/mimo-v2.5-pro |
|---|---|---|
| creator | Xiaomi | |
| model type | Text generation | Text generation |
| context window | 1,000,000tokens | 1,024,000tokens |
| input | takes text, image, audio and video in, gives text out | takes text in, gives text out |
| output | text | text |
| tool calling | not verified | not verified |
| released | feb 19, 2026 | not published |
| endpoint | /v1/chat/completions | /v1/chat/completions |
unknown means the public catalog does not establish that fact. a model's documented capabilities can differ from what a particular api supports. see the model documentation for integration details.
| primary configuration | Gemini 3.1 Pro (preview) google/gemini-3.1-pro-preview | MiMo V2.5 Pro xiaomi/mimo-v2.5-pro no benchmarks yet |
|---|---|---|
| intelligence index · gateway rank | 30.4#34 of 90 | not measured |
| coding index | 68.8#27 of 60 | not measured |
| agentic index | 10.3#41 of 49 | not measured |
| omniscience | 31.9#5 of 86 | not measured |
| gpqa | 94.1%#5 of 88 | not measured |
| humanity’s last exam | 47.0%#10 of 89 | not measured |
| terminal-bench v2.1 | 73.8%#32 of 57 | not measured |
| tau2 | 95.6% | not measured |
| mmmu-pro | 82.4% | not measured |
| aa-lcr | 82.0% | not measured |
| ifbench | 77.1% | not measured |
| harvey lab | 58.9% | not measured |
| scicode | 58.7% | not measured |
| terminal-bench hard | 53.8% | not measured |
| analystagent | 41.3% | not measured |
| automationbench partial | 35.4% | not measured |
| apex-agents | 32.0% | not measured |
| itbench | 30.3% | not measured |
| tau banking | 21.4% | not measured |
| gdp.pdf all-pass | 17.8% | not measured |
| critpt | 17.7% | not measured |
| multilingual lcr | 15.6% | not measured |
| terminal-bench v4.0 | 4.0% | not measured |
Gemini 3.1 Pro (preview) source ↗
each column describes the tested configuration named under the model; performance and costs can differ on impossibl. a score one model was not measured on is left blank rather than scored zero. indices are ranked against every model on the gateway with one. last checked 2026-09-11.
one token workload, priced on every model at the published rates, in usd. change the numbers to match yours.
google/gemini-3.1-pro-preview
estimated usage cost
$8.00
1,000 requests × (1,000 uncached input × $2.00 + 500 output × $12.00) / 1,000,000 = $8.00
base rates: up to 200,000 input tokens per request, inclusive.
xiaomi/mimo-v2.5-pro
estimated usage cost
$0.87
1,000 requests × (1,000 uncached input × $0.435 + 500 output × $0.87) / 1,000,000 = $0.87
base rates; no context pricing tiers are published.
estimates use published impossibl api usage rates. configured billing adjustments, funding fees, taxes, and custom workspace rates are excluded and may change the final amount.
image and audio usage, cache creation, and additional tool charges are excluded. reasoning tokens count as billable output, even when they are not visible in the response.
Gemini 3.1 Pro (preview) and MiMo V2.5 Pro are available through the Impossibl API. they share the /v1/chat/completions endpoint; the model identifier selects the model for a request.
Gemini 3.1 Pro (preview) is created by Google. It has a 1,000,000-token context window. Base API pricing (USD): $2.00 per million input tokens; $12.00 per million output tokens. Above 200,000 input tokens per request, context-tier rates apply to the whole request.
MiMo V2.5 Pro is created by Xiaomi. It has a 1,024,000-token context window. Base API pricing (USD): $0.435 per million input tokens; $0.87 per million output tokens.
Gemini 3.1 Pro (preview) →MiMo V2.5 Pro →
this page compares Gemini 3.1 Pro (preview) vs MiMo V2.5 Pro as listed by the impossibl public api. prices and availability come from that catalog, refreshed here every five minutes. this is an impossibl offering, not a comparison of independent hosting-provider quotes.
view the source catalog ↗comparison as markdown ↗report a correction →