# Gemini 3.5 Flash

Gemini 3.5 Flash is a text generation model from Google, available through the Impossibl API. Its published context window is 1,000,000 tokens. The advertised input rate is $1.50 per million input tokens. The advertised output rate is $9.00 per million output tokens.

- Model identifier: ` google/gemini-3.5-flash `
- Creator: Google
- Model type: Text generation
- Canonical model page: [Gemini 3.5 Flash](<https://impossibl.com/google/gemini-3.5-flash>)
- API availability: Listed as serving through Impossibl.
- Creator release date: 2026-05-19

## Capabilities and limits

- Context window: 1,000,000 tokens
- Maximum output tokens: Not published in the public catalog; this is separate from the context window.
- Inputs: text, image, audio, video
- Outputs: text
- Tool calling: Not published

## API access

Use ` google/gemini-3.5-flash ` as the model identifier through the Impossibl API.

Published API endpoints:

- ` /v1/chat/completions `

[API quickstart](<https://impossibl.com/docs/agent-quickstart>) · [Models and billing documentation](<https://impossibl.com/docs/models>)

## Published API prices

All rates are in USD. An unpublished rate is unknown, not zero. Workspace-specific pricing, taxes, balance-purchase fees, billing adjustments, and tool charges can affect the amount paid.

| Usage | Published base rate | Billing unit |
| --- | --- | --- |
| Input | $1.50 | per million input tokens |
| Output | $9.00 | per million output tokens |
| Cache read | $0.15 | per million cached input tokens |
| Cache write (default retention) | Not published | per million tokens |
| Cache write (one-hour retention) | Not published | per million tokens |
| Audio input | Not published | per million audio input tokens |

## Independent benchmarks

Source: [Artificial Analysis](<https://artificialanalysis.ai/>). Last checked: 2026-09-11T11:21:18.106Z.
Configurations are evaluated separately. API scores retain their original units. Unmeasured scores are omitted.

### Gemini 3.5 Flash (high)

Source: [Evaluation source](<https://artificialanalysis.ai/models/gemini-3-5-flash>)

| Benchmark | Score |
| --- | --- |
| Intelligence Index | 32.98 |
| Omniscience Index | 21.18 |
| GPQA | 92.22% |
| Humanity’s Last Exam | 42.68% |
| SciCode | 53.94% |
| CritPt | 13.14% |
| Terminal-Bench v2.1 | 78.65% |
| Terminal-Bench v4.0 | 6.57% |
| Terminal-Bench Hard | 40.91% |
| AA-LCR | 73.33% |
| Multilingual LCR | 18.33% |
| Harvey LAB | 82.08% |
| GDPval Elo | 1259.06 |
| GDP.pdf All-pass | 19.80% |
| AutomationBench Partial | 42.08% |
| AnalystAgent | 45.00% |
| MMMU-Pro | 84.28% |
| IFBench | 76.33% |
| Tau2 | 95.32% |
| Tau Banking | 32.16% |
| APEX-Agents | 47.05% |
| ITBench | 40.35% |
| coding index (API) | 70.1 |
| agentic index (API) | 27.3 |
| intelligence index (API) | 33 |

### Gemini 3.5 Flash (medium)

Source: [Evaluation source](<https://artificialanalysis.ai/models/gemini-3-5-flash-medium>)

| Benchmark | Score |
| --- | --- |
| Intelligence Index | 33.63 |
| Omniscience Index | 20.82 |
| GPQA | 92.12% |
| Humanity’s Last Exam | 41.33% |
| CritPt | 10.86% |
| Terminal-Bench Hard | 39.39% |
| AA-LCR | 74.33% |
| MMMU-Pro | 83.87% |
| IFBench | 74.56% |
| Tau2 | 95.61% |
| intelligence index (API) | 33.6 |

### Gemini 3.5 Flash (minimal)

Source: [Evaluation source](<https://artificialanalysis.ai/models/gemini-3-5-flash-minimal>)

| Benchmark | Score |
| --- | --- |
| Intelligence Index | 23.85 |
| Omniscience Index | 0.67 |
| GPQA | 82.83% |
| Humanity’s Last Exam | 24.10% |
| CritPt | 1.43% |
| Terminal-Bench Hard | 46.21% |
| AA-LCR | 61.33% |
| MMMU-Pro | 80.12% |
| IFBench | 47.28% |
| Tau2 | 58.77% |
| intelligence index (API) | 23.8 |

Full benchmark data and breakdowns: [Complete benchmark snapshot](<https://api.impossibl.com/v1/models/google/gemini-3.5-flash/benchmarks>).


## Availability and retirement

The model is listed as serving. No retirement is announced in the current public catalog.

## Frequently asked questions

### What is Gemini 3.5 Flash?

Gemini 3.5 Flash is a text-generation model from Google, available through the Impossibl API.

### How much does Gemini 3.5 Flash cost?

Gemini 3.5 Flash costs $1.50 per million input tokens and $9.00 per million output tokens through the Impossibl API. Cache reads cost $0.15 per million tokens. All rates are in USD.

### What is the context length of Gemini 3.5 Flash?

Gemini 3.5 Flash has a 1,000,000-token context window through the Impossibl API.

### What inputs and outputs does Gemini 3.5 Flash support?

Gemini 3.5 Flash accepts text, image, audio, and video as input. Gemini 3.5 Flash returns text.

### How do I use Gemini 3.5 Flash through an API?

Gemini 3.5 Flash is available through the Impossibl API. Use google/gemini-3.5-flash as the model identifier with /v1/chat/completions. The quickstart explains how to create an API key and send your first request.

### When was Gemini 3.5 Flash released?

Google released Gemini 3.5 Flash on May 19, 2026.

Sources: [Creator release announcement](<https://ai.google.dev/gemini-api/docs/changelog>)

## Sources and related links

Prices, limits, capabilities, endpoints, and availability come from the public Impossibl catalog. Creator documentation supplies descriptions and release dates where available; it does not override the published API offering.

- [Public Impossibl model catalog](<https://api.impossibl.com/v1/models>)
- [Full model directory as JSON](<https://impossibl.com/models.json>)
- [Gemini 3.5 Flash: model details and pricing](<https://impossibl.com/google/gemini-3.5-flash>)
- [Model directory](<https://impossibl.com/models>)
