Nemotron 3.5 Lightning

NVIDIA🇺🇸 United States
active
Context window1000K tokens

Version History

3.5 Lightningmajor

Nemotron 3.5 Lightning succeeds Nemotron 3 Nano 30B A3B, gaining nine points on the Intelligence Index (15 to 24) and large jumps on agentic benchmarks (GDPval-AA v2 Elo up 334 points, Terminal-Bench v2.1 up from 7% to 24.3%) while becoming the fastest model in its comparison class at nearly 670 tokens per second.

Coverage

model releaseNVIDIA

Nvidia Releases Nemotron 3.5 Lightning: A 31.6B-Parameter Open Model Built for Speed, Not Peak Intelligence

Nvidia's Nemotron 3.5 Lightning, a 31.6B-parameter open-weight model with only 3.6B active parameters, matches OpenAI's gpt-oss-120b on the Artificial Analysis Intelligence Index while delivering the fastest throughput in its class at nearly 670 tokens per second. The model posts especially large gains on agentic benchmarks, beating both gpt-oss-120b and the larger Nemotron 3 Super.

3 min read