New: visual benchmarks for understanding the capabilities of image models 🌅 See how every model performs across a variety of challenging prompts: https://t.co/P98jIE8dsn
Can the model arrange multiple input images into a scene? Hold up the correct number of fingers? Remove the 2nd cup from a row? Sort a grid of images for each model on OpenRouter by price and generation time to understand the tradeoffs with their capabilities.

Start using the image generation api: https://t.co/o2VCkW7azQ See the full announcement: https://t.co/zDhrCSUxQX
We'll expand this with more scenarios and update it every time a new model launches. What other prompts do you want us to run?
Image quality is hard to read off a spec sheet. Seeing every model's output on the same hard prompts, next to price and latency, turns model selection into inspection rather than guesswork.
Checking sign-in…
Loading comments…