
The METR time-horizon chart explained — what the most-cited AI capability curve actually measures and how to read it.
“tasks that take humans twelve hours to do, we predict Speaker 7: that it will succeed at those tasks around fifty percent Speaker 7: of the time.”
Joel Becker
“the simple answer is, literally, we get humans Speaker 7: to sit down and complete the tasks that we give Speaker 7: to AIS and as close to identical conditions as possible.”
Joel Becker
“METER is sort of specialized Speaker 6: in specifically assessing how autonomous are AI systems, what is Speaker 6: the scale and like length and difficulty of tasks that Speaker 6: they're able to do by themselves, partially because we think Speaker 6: it sets the stakes for conversations about AI misalignment.”
Chris Painter
“we have approximately three, Speaker 7: although it varies quite a lot across tasks. Human baselines Speaker 7: per tasks, so you know, typically we're ever going over Speaker 7: something like three.”
Joel Becker
“the first time Speaker 2: I saw this chart or version of this chart, what Speaker 2: I assume, and I suspect others assume, is that it Speaker 2: was able to go off and work on a task Speaker 2: for eleven hours and fifty nine minutes then come back Speaker 2: with an answers. But apparently it's not that.”
Odd Lots
Checking sign-in…
Loading comments…