First independent look at how Meta's coding actually performs on varied build tasks, which is what decides whether it is worth a slot in someone's toolchain.
“These agents remain active throughout each session rather than being spawned for individual tasks, which is something that I have not seen as often.”
Bijan Bowen
“part of testing these AIs is basically just giving them imperfect real life things exactly like this”
Bijan Bowen
“I'm not huge on like testing things that are partially discounted cuz that's not the actual genuine price they're going to be.”
Bijan Bowen
“I think pricing-wise, this may be a little too weak to be competitive, at least in terms of the performance noticed in these specific tests.”
Bijan Bowen
“I kind of like seeing that because the models are independent and have their own styles, I guess could be said.”
Bijan Bowen
Checking sign-in…
Loading comments…