
Sutton, the father of reinforcement learning, argues that no amount of scaling solves on-the-job continual learning, and that a new experience-based architecture will supersede today's train-then-deploy paradigm — essential perspective for anyone betting their stack on LLM scaling.
“Reinforcement learning is about understanding your world, whereas large language models are about mimicking people, doing what people say you should do.”
Richard Sutton
“You can’t have prior knowledge if you don’t have ground truth, because the prior knowledge is supposed to be a hint or an initial belief about what the truth is.”
Richard Sutton
“Squirrels don’t go to school. Squirrels can learn all about the world.”
Richard Sutton
“Gradient descent will not make you generalize well. It will make you solve the problem.”
Richard Sutton
“I do think succession to digital intelligence or augmented humans is inevitable.”
Richard Sutton
Checking sign-in…
Loading comments…