← All IntelClip / AI Tools
Result: solo run shipped a non-functional game
From Fable 5 & Qwen 27B – Traycer Multi-Agent Hands-On Test! · ≈25:57
“it's definitely a very interesting AB test of how this model does when given that big prompt in its entirety versus when given a smarter model to orchestrate it”
Bijan Bowen
“I've tried some of the controls here for throttle and things like that, and we don't really get anything functional”
Bijan Bowen
“this one seemed like it had actually implemented some more realism in terms of the physics than I would have expected”
Bijan Bowen
“I think that Tracer Desktop with this multi-agent coordination capability is really fantastic”
Bijan Bowen
What’s in it
- The head-to-head payoff — after 32 minutes the solo small model produced unusable controls and no sound, while the orchestrated run of the same model produced working flight physics.
Clip transcript
So, after about 32 minutes, it did produce a result for us. Now, I haven't looked at this yet. I don't know if this is going to be any good, but nonetheless, it does seem that it is running right now on port 8080. So, let's see what Quen 3.6 entirely by itself did. Oh, wow. Okay. Um, it's definitely a very interesting AB test of how this model does when given that big prompt in its entirety versus when given a smarter model to orchestrate it. Now, some things I'm going to notice are that actually like this did a better job with the terrain. At least it has more terrain. Actually, I don't know that I would say that. Um, I don't want this terrain for like a flying game. There's no room to take off here. Uh, there's also no sound. All right. Yeah. And unfortunately, I've tried some of the controls here for throttle and things like that, and we don't really get anything functional. So, this was definitely interesting to see cuz unfortunately that just didn't work at all. And then the one where we had Fable orchestrating Quen and we told it don't change any of the code. You just guide it to do things correctly. I also want to note that I know this may sound a little weird, but this one seemed like it had actually implemented some more realism in terms of the physics than I would have expected. And we did kind of see that when it was going through the planning stages and it was giving some like thrust values and other like force and physics calculations. So, just a really interesting AB test to see how this performs. And really, that's probably going to bring us into the conclusion here. I think that Tracer Desktop with this multi-agent coordination capability is really fantastic. The way you can even just move things around seamlessly, the way it has support for so many different types of agents, it makes things really simple and it handles like if you tell an agent, okay, go and spawn a sub agent in this, it doesn't. And it just makes things easier. And this was really quite enjoyable to use, especially when wanting to compare how models perform with other models. So if we did an entire video just testing how a Quen model performs with a bunch of different Asians controlling it. So like Fable 5 controlling Quen 27B or Opus controlling 27B. This would make that a really really seamless video to do and is definitely something to think about in the future. So that is probably going
Recommended reads
Comments
Checking sign-in…
Loading comments…