# Llama 3.3 70B Instruct vs Grok 4.5

Canonical HTML page: [Llama 3.3 70B Instruct vs Grok 4.5](<https://impossibl.com/compare/meta/llama-3.3-70b/xai/grok-4.5>)
Markdown alternative: [Llama 3.3 70B Instruct vs Grok 4.5 as Markdown](<https://impossibl.com/compare/meta/llama-3.3-70b/xai/grok-4.5.md>)

Compare API pricing, context windows, and capabilities side by side. Prices are published API usage rates in USD; check billing units and context tiers alongside each price.

## Specifications and pricing

| Specification | Llama 3.3 70B Instruct | Grok 4.5 |
| --- | --- | --- |
| API identifier | meta/llama-3.3-70b | xai/grok-4.5 |
| Creator | Meta | xAI |
| Model type | Text generation | Text generation |
| Inputs | text | text, image |
| Outputs | text | text |
| Base input price (USD) | $0.71 per million input tokens | $2.00 per million input tokens |
| Base output price (USD) | $0.71 per million output tokens | $6.00 per million output tokens |
| Cache read price (USD) | Not published | $0.30 per million cached input tokens |
| Cache write price (USD) | Not published | Not published |
| One-hour cache write price (USD) | Not published | Not published |
| Audio input token price (USD) | Not published | Not published |
| Context window | 128,000 tokens | 500,000 tokens |
| Tool calling | Not published | Not published |
| Endpoints | /v1/chat/completions | /v1/chat/completions |
| Context pricing tiers (USD per million tokens) | None published | Over 200,000 input tokens: $4.00 input / $12.00 output |
| Retirement | Not announced in the public catalog | Not announced in the public catalog |

When total input, including cached input, strictly exceeds a context threshold, the highest qualifying tier applies to the whole request. Consult each model's complete Markdown for tier-specific cache rates; an unpublished tier cache price does not establish the base cache rate.

## Llama 3.3 70B Instruct vs Grok 4.5: side-by-side summary

Llama 3.3 70B Instruct and Grok 4.5 are available through the Impossibl API.

They share the `/v1/chat/completions` endpoint. Set the model identifier to select the model for a supported request.

Llama 3.3 70B Instruct is created by Meta. It has a 128,000-token context window. Base API pricing (USD): $0.71 per million input tokens; $0.71 per million output tokens.

Grok 4.5 is created by xAI. It has a 500,000-token context window. Base API pricing (USD): $2.00 per million input tokens; $6.00 per million output tokens. Above 200,000 input tokens per request, context-tier rates apply to the whole request.

## Model details

- [Llama 3.3 70B Instruct: full model facts as Markdown](<https://impossibl.com/meta/llama-3.3-70b.md>); [HTML model page](<https://impossibl.com/meta/llama-3.3-70b>)
- [Grok 4.5: full model facts as Markdown](<https://impossibl.com/xai/grok-4.5.md>); [HTML model page](<https://impossibl.com/xai/grok-4.5>)

For compatible text or embedding models with token pricing, the HTML comparison page provides a shared workload calculator. It excludes cache creation, tool calls, taxes, balance-purchase fees, and workspace-specific adjustments.

## Sources and related links

Prices, capabilities, limits, and availability describe the public Impossibl API offering. An unpublished fact is unknown; a published zero price is zero. Compare matching billing units and check context tiers before estimating costs.

- [Public Impossibl model catalog](<https://api.impossibl.com/v1/models>)
- [Full model directory as JSON](<https://impossibl.com/models.json>)
- [Model directory as Markdown](<https://impossibl.com/models.md>)
- [API quickstart](<https://impossibl.com/docs/agent-quickstart>)
- [Model documentation](<https://impossibl.com/docs/models>)
