NVIDIA logo

Nemotron 3.5 Lightning 30B A3B

Nemotronv3.5 Lightning 30B A3BCurrent
byNVIDIANVIDIA(enterprise)
Released August 11, 2026
11
OpenTools Scorenormalized /100
20.8
Valuescore / $M
OpenTools Score
Context1.0M tokens
Price (In / Out)$0.10/M / $0.95/M
Training CutoffMay 2026 post-training; September 2025 pre-training
CategoryLarge Language Model
Max Output32.8K tokens

About Nemotron 3.5 Lightning 30B A3B

Nemotron 3.5 Lightning is NVIDIA open-weight 30B-A3B reasoning model for fast, long-running agents. Its hybrid Mamba-2, mixture-of-experts, and attention architecture supports function calling, coding, tool use, long context, and efficient local or serverless deployment.

Capabilities

textreasoningcodingtool usefunction callinglong contextagenticlocal deploymentopen weights

Input Modalities

text

Output Modalities

text

Technical Details

API Identifier
nvidia/nemotron-3.5-lightning-30b-a3b
Category
Large Language Model
Context Window
1,048,576 tokens
Max Output Tokens
32,768 tokens

Tags

reasoningagenticlong-contextopen-weightfunction-callingmoenvfp4

Benchmarks

Performance scores for Nemotron 3.5 Lightning 30B A3B across standard benchmarks.

Artificial Analysis Intelligence Index v4.1.1Artificial Analysis · Aug 2026
24%

Pricing

Token pricing for Nemotron 3.5 Lightning 30B A3B API usage.

Input Tokens

$0.10/M

per million tokens

Output Tokens

$0.95/M

per million tokens

Pricing Calculator

Input cost$0.10
Output cost$0.48
Estimated monthly cost$0.58

NVIDIA build lists serverless NIM pricing at $0.10 input and $0.95 output per million tokens. Self-hosted weights are also available; infrastructure costs vary.

Competing Models

Same pricing tier — direct alternatives to Nemotron 3.5 Lightning 30B A3B

Mid-Range
82
40.1
$0.95/M / $3.15/MCompare
53
70.1
$0.30/M / $1.20/MCompare
65
32.4
$1.00/M / $3.00/MCompare
$1.20/M / $4.00/MCompare
$0.72/M / $2.30/MCompare
79
27.3
$1.40/M / $4.40/MCompare
66
23.9
$1.25/M / $4.25/MCompare
34
52.7
$0.43/M / $0.87/MCompare