← All IntelClip / OtherNeumotron's design philosophy: 'faster models are smarter models'
From Local Models: Trust, Control, Optimization — Carter Abdallah, NVIDIA · ≈4:49
“we kind of have this mantra that like faster models are smarter models”
“in order for AI to continue to grow and be useful to everybody, it has to be done in the open”
What’s in it
- Explains philosophy behind Nvidia's open Nemotron model family
- Argues inference speed is a core design goal, not an afterthought
- Flags need for models that run on local hardware, not just data centers
Clip transcript
strongly as well. But the Neumotron family of models is focused on being as open as humanly possible. So we we have this understanding or belief that in order for AI to continue to grow and be useful to everybody, it has to be done in the open so that we can build off of each other, we can compound on each other. And part of what we do because team green, this is always true, is we we think that the the rate that you can squeeze tokens out of models is very important. So we kind of have this mantra that like faster models are smarter models and so a lot of the decisions we make when designing a model like Neumotron is built around how fast can we make it go. As especially you are going to see in the next however many months local AI take off, we we need to make sure that models are well supported on uh hardware that doesn't just exist in massive buildings, you know, thousands of kilometers away from you. Uh and so that's uh you know, for AI to be very useful, it should be quick uh and and open. So, that's kind of the the vibe of
Comments
Sign in to comment.
Loading comments…