meituan/longcat-2.0
estimated usage cost
$0.90
1,000 requests × (1,000 uncached input × $0.30 + 500 output × $1.20) / 1,000,000 = $0.90
base rates; no context pricing tiers are published.
rates, context and capabilities side by side on the impossibl api, and what the same workload costs on each.
| specification | LongCat 2.0 meituan/longcat-2.0 | Inkling Small thinkingmachines/inkling-small |
|---|---|---|
| input · usd / 1m tok | $0.30 | $0.58 |
| output · usd / 1m tok | $1.20 | $1.44 |
| context window | 1.05mtokens | 65.54ktokens |
| modalities · in → out | takes text in, gives text out | takes text and image in, gives text out |
| base api rates · usd per 1m tokens | LongCat 2.0 meituan/longcat-2.0 | Inkling Small thinkingmachines/inkling-small |
|---|---|---|
| input | $0.30 | $0.58 |
| output | $1.20 | $1.44 |
| cache read | $0.006 | $0.116 |
the estimate below prices one workload at these rates on every model.
| as published | LongCat 2.0 meituan/longcat-2.0 | Inkling Small thinkingmachines/inkling-small |
|---|---|---|
| creator | meituan | Thinking Machines |
| model type | Text generation | Text generation |
| context window | 1,048,756tokens | 65,536tokens |
| input | takes text in, gives text out | takes text and image in, gives text out |
| output | text | text |
| tool calling | not verified | not verified |
| endpoint | /v1/chat/completions | /v1/chat/completions |
unknown means the public catalog does not establish that fact. a model's documented capabilities can differ from what a particular api supports. see the model documentation for integration details.
one token workload, priced on every model at the published rates, in usd. change the numbers to match yours.
meituan/longcat-2.0
estimated usage cost
$0.90
1,000 requests × (1,000 uncached input × $0.30 + 500 output × $1.20) / 1,000,000 = $0.90
base rates; no context pricing tiers are published.
thinkingmachines/inkling-small
estimated usage cost
$1.30
1,000 requests × (1,000 uncached input × $0.58 + 500 output × $1.44) / 1,000,000 = $1.30
base rates; no context pricing tiers are published.
estimates use published impossibl api usage rates. configured billing adjustments, funding fees, taxes, and custom workspace rates are excluded and may change the final amount.
image and audio usage, cache creation, and additional tool charges are excluded. reasoning tokens count as billable output, even when they are not visible in the response.
LongCat 2.0 and Inkling Small are available through the Impossibl API. they share the /v1/chat/completions endpoint; the model identifier selects the model for a request.
LongCat 2.0 is created by meituan. It has a 1,048,756-token context window. Base API pricing (USD): $0.30 per million input tokens; $1.20 per million output tokens.
Inkling Small is created by Thinking Machines. It has a 65,536-token context window. Base API pricing (USD): $0.58 per million input tokens; $1.44 per million output tokens.
this page compares LongCat 2.0 vs Inkling Small as listed by the impossibl public api. prices and availability come from that catalog, refreshed here every five minutes. this is an impossibl offering, not a comparison of independent hosting-provider quotes.
view the source catalog ↗comparison as markdown ↗report a correction →