
A gain driven by taking 4x more turns per task is a different purchase than a gain in per-turn capability: it lands as latency and spend in production. This is the kind of read a headline alone would hide.
postQwen3.8 Max costs $1.14 per Intelligence Index task, more than double Qwen3.7 Ma
postAlibaba's Qwen3.8 Max scores 56 on the Artificial Analysis Intelligence Index at
postAA-Omniscience regresses 10 points from Qwen3.7 Max (+14 to +4), reversing its p
postOnly two labs occupy the Time per Task Pareto frontier: all frontier models undeSign in to comment.
Loading comments…