← All models

Nemotron Lightning 3.5 30B A3B

nemotron-lightning-3p5-30b-a3b

Small, fast open-weight MoE model for high-volume agentic execution

Intelligence
23.6
Context
262K
Training cutoff
--
Added
2026-08-27
Inputs
Text only

Providers

One route serves Nemotron Lightning 3.5 30B A3B. Every figure below is measured on that route over the selected window; a cell with too little traffic behind it says so rather than showing a number.

ProviderContextMax outputInput/1MOutput/1Mp50 TTFTp90 TTFTp50 latencyp90 latencyp50 output rateEligible callsErrorsUptimeEvidence
Fireworks AI only route
accounts/fireworks/models/nemotron-lightning-3p5-30b-a3b
Text only
262K262,144$0.05$0.20------------0%warming · 30m

TTFT is the median time to the first visible answer token over real ModelRelay requests, with a fixed prompt standing in for models that have no traffic yet. Speed ranks by it. Intelligence is the Artificial Analysis Intelligence Index.

Pin a request to one provider

Routing follows the priority order above and fails over on error. To require a single provider, add a provider object to the request body on the OpenAI- and Anthropic-compatible endpoints. Setting allow_fallbacks to false makes the constraint strict: the request fails rather than routing elsewhere.

Fireworks AI
{"provider": {"only": ["fireworks"]}}