
Today, we're introducing not one but TWO new models, striking the balance between efficiency and quality to enable you to build production AI agents. — Gemini 3.6 Flash: Addresses efficiency feedback we received from Gemini 3.5 Flash with upgrades in coding, knowledge work, and multimodal tasks faster, more accurately, and with substantially fewer tokens per task — Gemini 3.5 Flash-Lite: Our fastest, most cost-effective 3.5-class model yet built for agentic workflows, hitting ~350 output tokens/sec with improved coding and overall quality Start building with these today via the Gemini API in @GoogleAIStudio or try them out in the @GeminiApp
per completed task matters more than price per token in loops, and Flash-Lite's roughly 350 output tokens per second sets a new floor for latency-sensitive steps.
Checking sign-in…
Loading comments…