deepseek/deepseek-v4.1-flashnew model
deepseek v4.1 flash joins the gateway.
$ impossibl changelog
deepseek v4.1 flash joins the gateway.
openai/gpt-image-2.5-flare is live on /v1/images/generations at $5.00 input and $30.00 output per 1M tokens.
openai/gpt-image-2.5-sunburst is live on /v1/images/generations at $5.00 input and $30.00 output per 1M tokens.
gpt-6 astra joins the gateway with a 1,050,000 token context window and text plus image input.
muse spark 1.3 and muse spark 1.3 contributor join the ai api.
workspaces join the console: invite your team, give each member a role — owner, admin, developer, or viewer — and switch between the workspaces you belong to from the top of the console. keys, balance, and usage belong to the workspace.
gemini 3.8 flash joins the ai api with a 1m-token context window. it is promo-priced at $0.75 / $3.75 per million tokens through 2026-12-31.
claude fable 5.1 joins the ai api with a 1m-token context window and adaptive thinking. cache reads bill at $0.25 per million tokens, a quarter of fable 5's rate.
glm 5.3 flash joins the ai api.
qwen3.8 27b joins the ai api.
glm 5.3 joins the ai api.
deepseek v4 pro 0813 joins the ai api.
gemini 3.7 flash joins the gateway.
grok 4.6 joins the gateway.
qwen3.8 2.4t a95b joins the ai api.
muse glimmer 30b joins the ai api.
nemotron 3.5 lightning joins the ai api.
take your request log with you. filter the logs room, press export csv, and the whole filtered window walks itself into a spreadsheet — not just the rows on screen.
a cheap request no longer reads as free. costs under a cent keep their digits down to the credit, so a log of small calls stops printing as a column of $0.00.
qwen3.8 max joins the ai api.
deepseek v4 flash 0731 joins the ai api.
gemini 3.1 flash tts brings 30 steerable voices to /v1/audio/speech, billed per token like every other gemini model.
gemini robotics-er 2 preview is now available.
grok voice tts brings five expressive voices and mp3, wav or pcm output to the new /v1/audio/speech endpoint.
send audio and video to gemini models on chat completions, the responses api, and the gemini-native endpoint, with audio billed at the provider's audio rate.
gpt-transcribe brings high-accuracy speech-to-text for completed files and streamed transcripts to /v1/audio/transcriptions.
the ai api now turns text into vectors on /v1/embeddings, starting with openai/text-embedding-3-small and openai/text-embedding-3-large.
the ai api now generates images, starting with openai/gpt-image-2 on /v1/images/generations.
claude opus 5 joins the ai api with a 1m-token context window and native anthropic fallback.
hunyuan 3 joins the ai api through atlascloud.
qwen 3.7 plus joins the ai api through fireworks.
atlascloud joins the gateway, serving mimo 2.5 and minimax m3.
gemini 3.5 flash-lite is now available.
gemini 3.6 flash is now available.
every request is now logged. browse them in the dashboard, or have agents pull their own from /v1/requests.
mimo 2.5 joins the ai api through atlascloud.
minimax m3 joins the ai api through atlascloud.
usage and balance update live in the dashboard as requests come in.
grok 4.5 joins the gateway.
alibaba's qwen models join the gateway.
qwen's 3.8 max preview joins the gateway.
failed routes no longer slow down every request. in our benchmarks, failover times dropped from 6 seconds to under 1 millisecond.
recharge when your balance runs low.
moonshot joins the gateway.
a better start.
two context tiers: 64k and 256k.
three fast-inference providers join the gateway.
one-liner migration now supported
luna joins the gateway.
sol joins the gateway.
terra joins the gateway.
xai's native web_search and x_search tools.
v4 and v3.2 join the gateway.
bring your own provider keys now supported
fable 5 redeployed.
humans can claim agent-created accounts and merge balances. usage stays at provider price.
sonnet 5 launches.
glm models join the gateway.
api keys, credit top-ups and the model catalog.
anthropic messages support, with image input and tool calling on every surface.
impossibl init