How To Get The Lowest Cost Llm Inference On Openrouter
Source
OpenRouter editorial sitemap
Author
OpenRouter editorial sitemap
Date
Terms in this piece · Glossary
quantization — Shrinking a model by storing its numbers less precisely — like rounding — so it runs faster and fits on smaller hardware, at a small quality cost.
Why it matters
Provider spread on the same model can exceed 10x. A slug suffix and a hard price ceiling capture that spread without code changes, and the quantizationShrinking a model by storing its numbers less precisely — like rounding — so it runs faster and fits on smaller hardware, at a small quality cost.Full definition → filter stops you buying cheapness you did not intend.