Text generation / Google

Gemini 2.5 Flash

Gemini 2.5 Flash is a text generation model from Google, available through the Impossibl API. Its published context window is 1,000,000 tokens. The advertised input rate is $0.30 per million input tokens. The advertised output rate is $2.50 per million output tokens.

api id: google/gemini-2.5-flash

model as markdown ↗

how much does gemini 2.5 flash cost?

published base api rates · usd per 1m tokens

input
$0.30
output
$2.50
cache read
$0.03
audio input
$1.00

rates are published by the impossibl api. balance-purchase fees, taxes, tool charges, and workspace-specific pricing are separate.

specifications & capabilities

creator
Google
model type
Text generation
context window
1m tokens
maximum output
not published in the catalog
input types
text, image, audio, video
output types
text
tool calling
not verified in the catalog

unknown means the public catalog does not establish that fact. a model's documented capabilities can differ from what a particular api supports.

published endpoint

/v1/chat/completions

see the model documentation for integration details.

estimate your usage

estimate your api cost

estimate text token usage for your workload. all amounts are in usd.

token workload

total input, including cached tokens

include billable reasoning tokens

the same token usage in each request

a subset of total input, never additional tokens

Gemini 2.5 Flash

$1.55

estimated usage cost

1,000 requests × (1,000 uncached input × $0.30 + 500 output × $2.50) / 1,000,000 = $1.55

base rates; no context pricing tiers are published.

estimates use published impossibl api usage rates. configured billing adjustments, funding fees, taxes, and custom workspace rates are excluded and may change the final amount.

image and audio usage, cache creation, and additional tool charges are excluded. reasoning tokens count as billable output, even when they are not visible in the response.

compare more Text generation models

independent benchmarks

Results from Artificial Analysis. Each configuration is listed separately. These evaluations describe the tested model configuration; performance and costs can differ on Impossibl. Last checked 2026-09-11.

Gemini 2.5 Flash (Non-reasoning)
BenchmarkScore
Intelligence Index9.85
Omniscience Index-42.57
GPQA68.28%
Humanity’s Last Exam4.73%
CritPt1.43%
Terminal-Bench Hard12.12%
AA-LCR49.90%
Multilingual LCR8.33%
MMMU-Pro65.49%
IFBench38.98%
Tau214.91%
LiveCodeBench49.52%
AIME 202560.33%
intelligence index (API)9.9

Unmeasured scores are omitted. API values retain their source units.

view source ↗

all source data and breakdowns (JSON) ↗

Gemini 2.5 Flash (Reasoning)
BenchmarkScore
Intelligence Index13.11
Omniscience Index-29.80
GPQA78.99%
Humanity’s Last Exam12.14%
CritPt1.14%
Terminal-Bench Hard13.64%
AA-LCR65.33%
Multilingual LCR8.89%
MMMU-Pro69.08%
IFBench50.27%
Tau231.58%
LiveCodeBench69.52%
AIME 202573.33%
intelligence index (API)13.1

Unmeasured scores are omitted. API values retain their source units.

view source ↗

all source data and breakdowns (JSON) ↗

frequently asked questions

What is Gemini 2.5 Flash?

Gemini 2.5 Flash is a text-generation model from Google, available through the Impossibl API.

How much does Gemini 2.5 Flash cost?

Gemini 2.5 Flash costs $0.30 per million input tokens and $2.50 per million output tokens through the Impossibl API. Cache reads cost $0.03 per million tokens. Audio input costs $1.00 per million audio input tokens. All rates are in USD.

view the pricing table ↑

What is the context length of Gemini 2.5 Flash?

Gemini 2.5 Flash has a 1,000,000-token context window through the Impossibl API.

What inputs and outputs does Gemini 2.5 Flash support?

Gemini 2.5 Flash accepts text, image, audio, and video as input. Gemini 2.5 Flash returns text.

How do I use Gemini 2.5 Flash through an API?

Gemini 2.5 Flash is available through the Impossibl API. Use google/gemini-2.5-flash as the model identifier with /v1/chat/completions. The quickstart explains how to create an API key and send your first request.

follow the api quickstart →

When was Gemini 2.5 Flash released?

Google released Gemini 2.5 Flash on June 17, 2025.

sources & coverage

this page describes Gemini 2.5 Flash as listed by the impossibl public api. prices and availability come from that catalog, refreshed here every five minutes. this is an impossibl offering, not a comparison of independent hosting-provider quotes.

view the source catalog ↗

model descriptions and release dates, where available, use the creator sources linked in the faq. api prices and limits use the impossibl catalog. benchmark results, where available, are credited to Artificial Analysis in the benchmarks section. report a correction →