NVIDIA Nemotron 3.5 Lightning 30B-A3B (BF16)

NVIDIA🇺🇸 United States
active
Context window1000K tokens

Version History

3.5-Lightning-30B-A3B-BF16minor

NVIDIA released the full-precision BF16 reference weights for Nemotron 3.5 Lightning, a 30B-parameter hybrid Mamba-MoE-Attention model with 3B active parameters and up to 1M token context. This release is intended for fine-tuning and quantization rather than direct deployment, with a companion NVFP4 checkpoint available for optimized inference.

Benchmark Scores

Full leaderboard →
51.6%
SWE-bench Verified

Coverage

model releaseNVIDIA

NVIDIA Releases Nemotron 3.5 Lightning: 30B MoE Model with 1M Token Context and 3B Active Parameters

NVIDIA released the full-precision BF16 reference weights for Nemotron 3.5 Lightning, a 30B-parameter Mixture-of-Experts model with only 3B active parameters and support for up to 1 million tokens of context. The model uses a hybrid Mamba-2, MoE, and Attention architecture and is licensed under OpenMDW-1.1 for commercial use.

2 min read