
If you're choosing a budget tier, GPT-5.6 Luna at $0.20/$1.20 per million tokens is now cheaper than Gemini 3.1 Flash-Lite and Claude Haiku 4.5, and the underlying cost cuts come partly from AI-driven kernel optimization rather than just margin compression — a signal of where cost curves are headed.
“We also used GPT‑5.6 Sol to optimize the model’s forward pass: the computation that transforms inputs into next-token predictions.”
“With Codex, GPT‑5.6 Sol autonomously rewrote and optimized our production kernels, the core code that executes the mathematical operations that make up the model.”
“That Luna price drop completely changes the landscape with respect to lower priced models.”
“I've switched it over to Luna.”
Checking sign-in…
Loading comments…