An openly licensed model sized to run on one GPU, benchmarked on MMLU-Pro and IFEval against peers with comparable VRAM needs, which lets teams keep in their own infrastructure instead of routing data to an external API.
Checking sign-in…
Loading comments…