Cognition launches new SWE-2 model, Rivaling Fable 5.1 and GPT-Astra
Source
seelos
Author
seelos
Date
Terms in this piece · Glossary
benchmark — A standard public test set for comparing AI models — the shared scoreboards behind every "model X beats model Y" claim.
Why it matters
SWE-2 shows a viable recipe for shrinking the cost of near-frontier coding agents: same-league benchmarkA standard public test set for comparing AI models — the shared scoreboards behind every "model X beats model Y" claim.Full definition → scores as larger models at a fraction of the price and far fewer turns per task.