Clip transcript
check this again just prior to finishing this recording, but we started with about $8.50 of free credit and we're now at $14 spent. So, that means we spent probably around $23 assuming that is the accurate final tally for today's test. So, it's pricey and I wanted to run it without that very heavily discounted mode that allows them to use your prompts for training because it just gives a more realistic feel for what this is actually going to cost to use. I'm not huge on like testing things that are partially discounted cuz that's not the actual genuine price they're going to be. So, I prefer to test things that way, I guess should be said. Okay, and we can see this is still creeping up. So, we'll probably wait to go back to final cost. I think overall I'm not blown away, but we have to remember that Muse Spark 1.0 came out, then 1.1, and now we're at 1.2. This is basically like an entire new generation of product of model from Meta. They're still early in terms of competing with the likes of Anthropic, xAI, Google, OpenAI, and whoever else may come along. So, the leaps are definitely noticeable. I think one of the main things that I noticed that was actually a significant increase over the previous version was the Steve the PC Repairman game. Now, it is possible that because I tried this with a previous version that it was I doubt it was made better specifically because of this test. So, this could probably just reflect some genuine increase in capability on the part of the model. It put in some fun Easter eggs and things of the sort. Basically, the flight combat simulator result was just flat-out disappointing, I would say. properly show the plane models. Everything was inverted. The watch website was acceptable and kind of this site almost like accurately reflects my experience with this model where like it's acceptable. But again, I don't know for the price if acceptable is going to cut it given that the competition really is heating up. There's going to be a new Grok that comes out soon, and I would assess it's probably going to beat this. I would pretty heavily assume that. Also, we have a bunch of other labs kind of throwing their hats in the rings recently as well. Even places like Inkling from Thinking Machines and things. Now, I think this is probably a better coder than that model, but still, there is a bunch of stuff coming out, and I think pricing-wise, this may be a little too weak to be competitive, at least in terms of the performance noticed in these specific tests. And that's my genuine honest opinion having paid for this, albeit not a very expensive amount, at least in the scheme of things. But, it did okay. I just wasn't blown away with any one specific