Nvidia's open-weight Nemotron 3.5 Lightning prioritizes speed over maximum intelligence Nvidia's open-weight Nemotron 3.5 Lightning, with 3.6 billion active parameters, matches OpenAI's gpt-oss-120b on the Intelligence Index despite being four times smaller, and at nearly 670 tokens per second it is the fastest model in the comparison, signaling Nvidia's focus on efficiency over raw size. Nvidia's Nemotron 3.5 Lightning is an open-weights model with just 3.6 billion active parameters that matches OpenAI's gpt-oss-120b on the Intelligence Index despite being four times smaller. At nearly 670 tokens per second, it's also the fastest model in the comparison, showing Nvidia is betting on efficiency over raw size. The article Nvidia's open-weight Nemotron 3.5 Lightning prioritizes speed over maximum intelligence https://the-decoder.com/nvidias-open-weight-nemotron-3-5-lightning-prioritizes-speed-over-maximum-intelligence/ appeared first on The Decoder https://the-decoder.com .