DeepSeek-V4-Flash offers strong intelligence-per-dollar on OpenRouter, reportedly outperforming the larger MiniMax M3 model at a fraction of the cost — but be aware output quality swings noticeably depending on the reasoning-effort setting you choose.
Terms in this piece · Glossary
token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
Key quotes
“It's 304 billion parameters - 167GB on Hugging Face - but it appears to punch well above its weight.”
“Artificial Analysis rank it ahead of MiniMax M3 - a 428B model.”
“It's $0.14/million input and $0.27/million output pricing means this may currently be the best value-per-intelligence model out there.”