Claude Sonnet 5.5 is now available in Devin Desktop and Devin CLI.
On FrontierCode 1.1 Main, it scores 64.4%, a significant improvement from Sonnet 5 (56.2%), surpassing Fable 5.1 at extra high reasoning effort.
token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
benchmark — A standard public test set for comparing AI models — the shared scoreboards behind every "model X beats model Y" claim.
Why it matters
Sonnet 5.5 is available in Devin Desktop and CLI. Cognition reports 64.4% on its FrontierCode benchmarkA standard public test set for comparing AI models — the shared scoreboards behind every "model X beats model Y" claim.Full definition → against 56.2% for Sonnet 5, which helps when choosing a model for coding agents.