A 2.4T-parameter open release is not locally runnable as shipped: 4.89TB of weights with no 4-bit quant from the lab, and even 4-bit needs roughly four 512GB machines.
articleServe Qwen3.8-2.4T-A95B, a 2.4T-Parameter Model, with Configurable Reasoning on NVIDIA GB300 NVL72Michelle Horton
postNVIDIA has just released the first Nemotron 3.5 model: Nemotron 3.5 Lightning,…Artificial AnalysisChecking sign-in…
Loading comments…