fireworks:nvidia/nemotron-3.5-lightning-30b-a3b

Common Name: NVIDIA Nemotron 3.5 Lightning 30B A3B

Fireworks
-50%On SaleSupportedTool InvocationSupportedReasoning
CompareTry in Chat

NVIDIA Nemotron 3.5 Lightning 30B A3B for efficient reasoning and agentic workloads.

Specifications

Context
262.1K
Maximum Output
262.1K
Inputtext
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Availability Trend (24h)

Pricing

Input$0.025/MTokens
Cached Input$0.005/MTokens
Output$0.10/MTokens

Performance Metrics (24h)

Similar Models

$0.15/$0.60/M-50%
ctx512Kmax512Kavail—tps—
InOutCap

MiniMax's reasoning model for long-context coding and agentic tasks.

$0.075/$0.30/M-50%
ctx131Kmax33Kavail—tps—
InOutCap

OpenAI's open-weight 120B model for production, agentic tasks, and high-reasoning use cases.

$0.30/$1.20/M-50%
ctx262Kmax—avail—tps—
InOutCap

NVIDIA's frontier-scale reasoning model for complex agents, long-context analysis, code, math, and science.

$1.50/$7.50/M-50%
ctx1.0Mmax131Kavail—tps—
InOutCap

Fireworks' specialized model built on Kimi K3 with shorter reasoning traces for efficient agentic workloads.