FireTofu
Nvidia's open-weight Nemotron 3.5 Lightning prioritizes speed over maximum intelligence

Technology · en

Nvidia's open-weight Nemotron 3.5 Lightning prioritizes speed over maximum intelligence

The Decoder · Aug 11, 2026, 3:07 PM UTC

Nvidia's Nemotron 3.5 Lightning is an open-weights model with just 3.6 billion active parameters that matches OpenAI's gpt-oss-120b on the Intelligence Index despite being four times smaller. At nearly 670 tokens per second, it's also the fastest model in the comparison, showing Nvidia is betting on efficiency over raw size.…