eval — A repeatable test for AI quality — a set of tasks plus scoring — used the way software teams use test suites, because model output is too variable to judge by eyeballing.
Why it matters
Apache 2.0 was written for source code and leaves weights, configs, evalA repeatable test for AI quality — a set of tasks plus scoring — used the way software teams use test suites, because model output is too variable to judge by eyeballing.Full definition → and data ambiguous. OpenMDW-1.1 covers the whole model distribution under one license, and it now applies retroactively to every Trinity release.