open weights — A model whose trained parameters are published for anyone to download and run — unlike API-only models you can access but never possess.
agent harness — The scaffolding around a model that turns it into a working agent — the loop, the tools it can call, and the rules for when to stop.
benchmark — A standard public test set for comparing AI models — the shared scoreboards behind every "model X beats model Y" claim.
Why it matters
Gives a concrete test for open-model parity — whether it actually works inside an agentic agent harnessThe scaffolding around a model that turns it into a working agent — the loop, the tools it can call, and the rules for when to stop.Full definition → at low price points — instead of benchmarkA standard public test set for comparing AI models — the shared scoreboards behind every "model X beats model Y" claim.Full definition → deltas, and puts the current gap at five to six months.