# Nemotron 3.5 Lightning

Nemotron 3.5 Lightning is a text generation model from NVIDIA, available through the Impossibl API. Its published context window is 262,144 tokens. The advertised input rate is $0.05 per million input tokens. The advertised output rate is $0.20 per million output tokens.

- Model identifier: ` nvidia/nemotron-3.5-lightning `
- Creator: NVIDIA
- Model type: Text generation
- Canonical model page: [Nemotron 3.5 Lightning](<https://impossibl.com/nvidia/nemotron-3.5-lightning>)
- API availability: Listed as serving through Impossibl.
- Creator release date: 2026-08-11

## Capabilities and limits

- Context window: 262,144 tokens
- Maximum output tokens: Not published in the public catalog; this is separate from the context window.
- Inputs: text
- Outputs: text
- Tool calling: Not published

## API access

Use ` nvidia/nemotron-3.5-lightning ` as the model identifier through the Impossibl API.

Published API endpoints:

- ` /v1/chat/completions `

[API quickstart](<https://impossibl.com/docs/agent-quickstart>) · [Models and billing documentation](<https://impossibl.com/docs/models>)

## Published API prices

All rates are in USD. An unpublished rate is unknown, not zero. Workspace-specific pricing, taxes, balance-purchase fees, billing adjustments, and tool charges can affect the amount paid.

| Usage | Published base rate | Billing unit |
| --- | --- | --- |
| Input | $0.05 | per million input tokens |
| Output | $0.20 | per million output tokens |
| Cache read | $0.01 | per million cached input tokens |
| Cache write (default retention) | Not published | per million tokens |
| Cache write (one-hour retention) | Not published | per million tokens |
| Audio input | Not published | per million audio input tokens |

## Independent benchmarks

Source: [Artificial Analysis](<https://artificialanalysis.ai/>). Last checked: 2026-09-11T11:52:20.926Z.
Configurations are evaluated separately. API scores retain their original units. Unmeasured scores are omitted.

### Nemotron 3.5 Lightning

Source: [Evaluation source](<https://artificialanalysis.ai/models/nemotron-3-5-lightning>)

| Benchmark | Score |
| --- | --- |
| Intelligence Index | 13.64 |
| Omniscience Index | -17.72 |
| GPQA | 74.34% |
| Humanity’s Last Exam | 10.57% |
| SciCode | 32.06% |
| CritPt | 0.00% |
| Terminal-Bench v2.1 | 24.34% |
| Terminal-Bench v4.0 | 0.51% |
| AA-LCR | 60.33% |
| Multilingual LCR | 1.11% |
| GDPval Elo | 766.89 |
| GDP.pdf All-pass | 3.00% |
| AutomationBench Partial | 0.84% |
| Tau Banking | 8.87% |
| coding index (API) | 26.8 |
| agentic index (API) | 6.1 |
| intelligence index (API) | 13.6 |

Full benchmark data and breakdowns: [Complete benchmark snapshot](<https://api.impossibl.com/v1/models/nvidia/nemotron-3.5-lightning/benchmarks>).


## Availability and retirement

The model is listed as serving. No retirement is announced in the current public catalog.

## Frequently asked questions

### What is Nemotron 3.5 Lightning?

Nemotron 3.5 Lightning is a text-generation model from NVIDIA, available through the Impossibl API.

### How much does Nemotron 3.5 Lightning cost?

Nemotron 3.5 Lightning costs $0.05 per million input tokens and $0.20 per million output tokens through the Impossibl API. Cache reads cost $0.01 per million tokens. All rates are in USD.

### What is the context length of Nemotron 3.5 Lightning?

Nemotron 3.5 Lightning has a 262,144-token context window through the Impossibl API.

### What inputs and outputs does Nemotron 3.5 Lightning support?

Nemotron 3.5 Lightning accepts text as input. Nemotron 3.5 Lightning returns text.

### How do I use Nemotron 3.5 Lightning through an API?

Nemotron 3.5 Lightning is available through the Impossibl API. Use nvidia/nemotron-3.5-lightning as the model identifier with /v1/chat/completions. The quickstart explains how to create an API key and send your first request.

### When was Nemotron 3.5 Lightning released?

NVIDIA released Nemotron 3.5 Lightning on August 11, 2026.

Sources: [Creator release announcement](<https://blogs.nvidia.com/blog/nemotron-lightning-switchyard-rtx-dgx/>)

## Sources and related links

Prices, limits, capabilities, endpoints, and availability come from the public Impossibl catalog. Creator documentation supplies descriptions and release dates where available; it does not override the published API offering.

- [Public Impossibl model catalog](<https://api.impossibl.com/v1/models>)
- [Full model directory as JSON](<https://impossibl.com/models.json>)
- [Nemotron 3.5 Lightning: model details and pricing](<https://impossibl.com/nvidia/nemotron-3.5-lightning>)
- [Model directory](<https://impossibl.com/models>)
