
A 30B-class open model that lands near gpt-oss-120b on agentic at roughly a quarter the parameters changes the cost calculus for high-volume execution steps. The license is permissive enough to actually deploy.
“It retains the same hybrid Mamba-Transformer architecture and small size from Nemotron 3 Nano, but makes substantial gains in intelligence and agentic performance.”
ArtificialAnlys
“Major intelligence jump: Nemotron 3.5 Lightning scores 24 on the Artificial Analysis Intelligence Index, a +9 point improvement over Nemotron 3 Nano (15). This puts it in line with OpenAI's gpt-oss-120b (24) and only just behind Nemotron 3 Super (26), a model ~4x its size”
ArtificialAnlys
“In pre-release testing of a DeepInfra endpoint serving the final NVFP4 weights, we measured median output speeds of nearly 670 tokens per second, much faster than those models are served in the market today”
ArtificialAnlys
“Meaningful agentic gains: the largest improvements over Nemotron 3 Nano come on agentic evaluations in GDPval-AA v2 (+334 ELO, moving past gpt-oss-120b and Nemotron 3 Super) and Terminal-Bench v2.1 (24% vs 7%).”
ArtificialAnlys
“Near-lossless NVFP4 quantization: as with prior Nemotron releases, the model ships in NVFP4 alongside BF16 weights. We measured the NVFP4 variant at 24 on the Intelligence Index and saw minimal degradation compared to the higher-precision weights”
ArtificialAnlys
postDeepSeek V4 Pro 0813 scores 53 on the Artificial Analysis Intelligence Index, 8
postGoogle has released Gemini 3.7 Flash, improving 4 points over Gemini 3.6 Flash a
blogNVIDIA Nemotron 3.5 Lightning and NeMo Switchyard Deliver Faster, Smarter, More Efficient Agentic AIKari BriskiSign in to comment.
Loading comments…